IP Library Granted Patent US 12,475,893
Granted Patent B2
US 12,475,893 · App. 18/010,541 · Granted Nov 18, 2025

Eyeglass augmented reality speech to text device and method

Inventor: Alexander G. Westner (Somerville, MA)
Assignee: XanderGlasses, Inc.
G10L15/26G02B27/0101G02B27/017G06F40/58G10L15/22H04R1/08H04R3/005H04R29/008H04W4/80G02B2027/0138G02B2027/0178
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,475,893
App. No.
18/010,541
Filed
Dec 15, 2022
Granted
Nov 18, 2025
Kind
B2
Art Unit
2692
USPC
704/235
Abstract

A method and apparatus to assist people with hearing loss. An augmented reality device with microphones and a display captures speech of a person talking to the wearer of the device and displays real-time captions in the wearer's field of view, while optionally not captioning the wearer's own speech. The microphone system in this apparatus inverts the use of microphones in augmented reality devices by analyzing and processing environmental sounds while ignoring the wearer's own voice.

Claims (45)

1 . A device, comprising:

a body;

at least two microphones systems disposed in the body comprised of a first system comprising at least one microphone positioned outwardly to target a non-wearer and a second microphone system, comprising at least one microphone positioned inwardly to target a wearer of the device;

a processor configured to process signals from the at least two microphone systems; and

a display positioned in a field of view of the wearer;

wherein the at least two systems emit signals having comparatively different signal power profiles enabling distinguishing of audible voice of the wearer from other sounds;

wherein processing signals comprises performing speech to text conversion on the received speech audio of a non-wearer comprising:

sending received speech audio from the device to a connected device;

performing speech to text conversion on the connected device; and

sending the text data to the device from the connected device; and

wherein the display renders text based on audible voice of the non-wearer that is captured on the first microphone system.

2 . The device of claim 1 , where the second microphone system captures voice commands for the device.

3 . The device of claim 1 , where the second microphone system is used as a voice input for another device connected wirelessly.

4 . The device of claim 1 , wherein the device uses signal power comparisons to distinguish between the audible voice of the wearer and the other sounds.

5 . The device of claim 4 , where two such devices are located on each side of eyeglasses and the microphones from each device together form a microphone array to capture sounds.

6 . The device of claim 1 , wherein the rendered text includes a translation of speech from one language into text of a different language.

7 . The device of claim 1 , wherein the rendered text is extended to capture and represent additional characteristics and information from a received audible voice, comprising inflections, emphasis, emotional valence, and recognized voices.

8 . The device of claim 1 , wherein the rendered text also captures and displays speech from the second microphone system.

9 . The device of claim 1 , wherein a real-time audio volume level is rendered on the display as a level meter, indicating a volume of the audible voice of the wearer as captured by the second microphone system.

10 . The device of claim 9 , wherein the level meter indicates when the wearer is speaking too quietly or too loudly, where the first microphone system receives and measures an ambient sound level as an input into the level meter.

11 . The device of claim 1 , further comprising a wireless transceiver.

12 . The device of claim 11 , wherein the wireless transceiver comprises a short-range wireless transceiver.

13 . The device of claim 1 , further comprising a camera.

14 . The device of claim 1 , further comprising one or more mounting mechanisms configured to mount the body to eyeglasses.

15 . The device of claim 1 , wherein the device is attached to eyeglasses.

16 . A method of providing speech to text conversion, the method comprising:

providing a device comprising:

a body;

at least two microphones systems disposed in the body comprised of a first system comprising at least one microphone positioned outwardly to target a non-wearer and a second microphone system, comprising at least one microphone positioned inwardly to target a wearer of the device;

a processor configured to process signals from the at least two microphone systems; and

a display positioned in a field of view of the wearer;

wherein the at least two systems emit signals having comparatively different signal power profiles enabling distinguishing of audible voice of the wearer from other sounds; and

wherein the display renders text based on audible voice of the non-wearer that is captured on the first microphone system;

receiving speech audio on the first microphone system;

receiving audio on the second microphone system; and

comparing the signal power profiles of the one or more microphones of the second microphone system and the one or more microphones of the first system to determine when the wearer or a non-wearer is speaking;

wherein, when the one or more wearer directed microphones of the second system are louder than the one or more microphones of the first system, then the device determines the wearer is speaking distinguishing their speech from the speech of a non-wearer and when the one or more non-wearer directed microphones of the first system are louder, then the device determines that the non-wearer is speaking;

performing speech to text conversion on the received speech audio of a non-wearer comprising:

sending received speech audio from the device to a connected device;

performing speech to text conversion on the connected device; and

sending the text data to the device from the connected device; and

displaying text for the speech audio of the non-wearer via the display of the device.

17 . The method of claim 16 , wherein the speech audio of the wearer is also converted to text and displayed.

18 . The method of claim 16 , wherein the device further comprises one or more mounting mechanisms configured to mount the body to eyeglasses.

19 . The method of claim 16 , wherein the device is attached to eyeglasses.

Assignments (2)
CHANGE OF NAME Recorded Sep 5, 2024
From: SPARK23 CORP.
To: XANDERGLASSES, INC.
Reel/Frame 068850/0128 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 16, 2023
From: WESTNER, ALEXANDER G.
To: SPARK23 CORP
Reel/Frame 063002/0355 →
Continuity (2)
Provisional Application 63074210 · Sep 3, 2020
Related Publication 20230238001A1 · Jul 27, 2023
References Cited (56)
US 6629076B1 · Haken · 2003 [cited by examiner]
US 7775675B2 · Hamm · 2010 [cited by applicant]
US 7785197B2 · Smith · 2010 [cited by examiner]
US 8369549B2 · Bartkowiak · 2013 [cited by examiner]
US 9280972B2 · McCulloch · 2016 [cited by examiner]
US 10074381B1 · Cowburn · 2018 [cited by examiner]
US 10289205B1 · Sumter · 2019 [cited by examiner]
US 10418034B1 · Mondragon · 2019 [cited by examiner]
US 10645492B2 · McElveen et al. · 2020 [cited by applicant]
US 10796106B2 · Kim · 2020 [cited by examiner]
US 10997943B2 · Tao et al. · 2021 [cited by applicant]
US 11172101B1 · Boozer et al. · 2021 [cited by applicant]
US 11184699B2 · Han · 2021 [cited by examiner]
US 11322959B1 · Ardisana, II et al. · 2022 [cited by applicant]
US 11468896B2 · Mondragon · 2022 [cited by examiner]
US 11500226B1 · Muske · 2022 [cited by examiner]
US 11538189B1 · Sztuk et al. · 2022 [cited by applicant]
US 11664022B2 · Kim · 2023 [cited by examiner]
US 12137331B2 · Shmukler · 2024 [cited by examiner]
US 20060025214A1 · Smith · 2006 [cited by applicant]
US 20110237295A1 · Bartkowiak et al. · 2011 [cited by applicant]
US 20120206334A1 · Osterhout · 2012 [cited by examiner]
US 20130223653A1 · Chang · 2013 [cited by examiner]
US 20130346168A1 · Zhou · 2013 [cited by examiner]
US 20140123008A1 · Goldstein · 2014 [cited by examiner]
US 20140337023A1 · McCulloch et al. · 2014 [cited by applicant]
US 20150036856A1 · Pruthi · 2015 [cited by examiner]
US 20150118961A1 · Petit · 2015 [cited by examiner]
US 20160078020A1 · Sumita et al. · 2016 [cited by applicant]
US 20180322861A1 · Ibrahim · 2018 [cited by examiner]
US 20190068529A1 · Mullins · 2019 [cited by examiner]
US 20190138603A1 · Daley · 2019 [cited by examiner]
US 20190196191A1 · Oh · 2019 [cited by examiner]
US 20190251176A1 · Cheng · 2019 [cited by examiner]
US 20190272800A1 · Tao et al. · 2019 [cited by applicant]
US 20190349662A1 · Lindahl · 2019 [cited by examiner]
US 20200034113A1 · Holst, III · 2020 [cited by examiner]
US 20200143807A1 · Ko · 2020 [cited by applicant]
US 20200194028A1 · Lipman · 2020 [cited by examiner]
US 20200265839A1 · McAnallan · 2020 [cited by examiner]
US 20210018752A1 · Sheng · 2021 [cited by examiner]
US 20210165975A1 · Osterhout · 2021 [cited by examiner]
US 20210312940A1 · Lipman · 2021 [cited by examiner]
US 20220254019A1 · Connor · 2022 [cited by examiner]
US 20220270595A1 · Kroehl et al. · 2022 [cited by applicant]
US 20220284254A1 · Sprague · 2022 [cited by examiner]
US 20230188894A1 · Shmukler et al. · 2023 [cited by applicant]
US 20230225637A1 · Lee et al. · 2023 [cited by applicant]
US 20230238001A1 · Westner · 2023 [cited by examiner]
US 20250149042A1 · Westner · 2025 [cited by applicant]
WO WO2009017797A2 · 2009 [cited by applicant]
WO WO2013050749A1 · 2013 [cited by applicant]
WO 2022051097A1 · 2022 [cited by applicant]
WO 2025096057A1 · 2025 [cited by applicant]
International Search Report from International Application No. PCT/US2021/046669; dated: Nov. 23, 2021. [cited by applicant]
International Search Report from International Application No. PCT/US2024/045374, dated: Nov. 22, 2024. [cited by applicant]