IP Library Granted Patent US 12,249,342
Granted Patent B2
US 12,249,342 · App. 18/491,157 · Granted Mar 11, 2025

Visualizing auditory content for accessibility

Inventor: Ron Zass (Kiryat Tivon, IL)
G10L21/0364G06F40/10G10L17/00H04R1/406H04R3/005H04R25/407
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,249,342
App. No.
18/491,157
Granted
Mar 11, 2025
Kind
B2
Abstract

Systems, methods and non-transitory computer readable media for processing audio and visually presenting information are provided. Audio data captured from an environment of a wearer of a wearable apparatus may be obtained. The audio data may be analyzed to obtain textual information. The audio data may be analyzed to associate different portions of the textual information with different speakers. A head mounted display system may be used to present each portion of the textual information in a presentation region associated with the speaker associated with the portion of the textual information. In one example, it may be determined that a part of the textual information is associated with speech that do not involve the user, and presenting the part may be avoided. In one example, a part of the textual information may be associated with speech produced by a user, and presenting the part may be avoided.

Claims (57)

1. A non-transitory computer readable medium storing data and computer implementable instructions that when executed by at least one processor cause the at least one processor to perform a method for processing audio and visually presenting information, the method comprising:

obtaining audio data captured by one or more audio sensors from an environment of a wearer of a wearable apparatus;

analyzing the audio data to obtain textual information;

analyzing the audio data to associate different portions of the textual information with different speakers;

determining that a part of the textual information is associated with speech that do not involve the user;

avoiding presenting the part of the textual information associated with the speech that do not involve the user; and

using a head mounted display system to present each portion of the textual information in a presentation region associated with the speaker associated with the portion of the textual information.

2. The non-transitory computer readable medium of claim 1 , wherein the method further comprises:

analyzing the audio data to identify a nonverbal sound; and

generating description of the nonverbal sound.

3. The non-transitory computer readable medium of claim 1 , wherein the method further comprises:

analyzing the audio data to identify a melody; and

generating description of the melody.

4. The non-transitory computer readable medium of claim 1 , wherein the head mounted display system is an augmented reality display system.

5. The non-transitory computer readable medium of claim 4 , wherein the association of the presentation regions with the speakers is configured to overlay the portions of the textual information associated with a speaker over the speaker in the augmented reality display system.

6. The non-transitory computer readable medium of claim 1 , wherein the method further comprises:

analyzing the audio data to identify an item in the audio data;

representing the item using a graphical symbol; and

displaying the graphical symbol in conjunction with the textual information.

7. The non-transitory computer readable medium of claim 1 , wherein the method further comprises displaying the different portions of the textual information using different sets of visual display parameters.

8. The non-transitory computer readable medium of claim 1 , wherein the method further comprises:

analyzing the audio data to identify different voice properties in different parts of the audio data;

associating different parts of the textual information with different voice properties; and

associating the different parts of the textual information with different textual formats based on the association of the different parts of the textual information with the different voice properties.

9. The non-transitory computer readable medium of claim 1 , wherein the method further comprises associating the different portions of the textual information with different textual formats based on the association of the different portions of the textual information with the different speakers.

10. The non-transitory computer readable medium of claim 1 , wherein the method further comprises:

determining that a portion of a speech is said in a specific linguistic tone;

selecting a visual display parameter for a part of the textual information associated with the portion of the speech based on the specific linguistic tone; and

displaying the part of the textual information using the selected visual display parameter.

11. The non-transitory computer readable medium of claim 1 , wherein the method further comprises:

analyzing the audio data to identify different voice properties in different parts of the audio data;

associating different parts of the textual information with different voice properties;

determining information based on particular voice properties; and

presenting the information determined based on the particular voice properties along the part of the textual information associated with the particular voice properties.

12. The non-transitory computer readable medium of claim 1 , wherein the method further comprises visually presenting information associated with a speaker in conjunction with a portion of the textual information associated with the speaker.

13. The non-transitory computer readable medium of claim 1 , wherein the association of the presentation regions with the speakers is based on spatial orientation of the speakers.

14. The non-transitory computer readable medium of claim 1 , wherein the association of the presentation regions with the speakers is based on positions of the speakers.

15. The non-transitory computer readable medium of claim 1 , wherein the method further comprises:

associating a part of the textual information with speech produced by a user; and

avoiding presenting the part of the textual information associated with the speech produced by the user.

16. The non-transitory computer readable medium of claim 1 , wherein the speech that do not involve the user is a conversation that do not involve the user.

17. The non-transitory computer readable medium of claim 1 , wherein the speech that do not involve the user is a speech not directed at the user.

18. A system for processing audio and visually presenting information, the system comprising:

at least one processing unit configured to:

obtain audio data captured by one or more audio sensors from an environment of a wearer of a wearable apparatus;

analyze the audio data to obtain textual information;

analyze the audio data to associate different portions of the textual information with different speakers;

determine that a part of the textual information is associated with speech that do not involve the user;

avoid presenting the part of the textual information associated with the speech that do not involve the user; and

use a head mounted display system to present each portion of the textual information in a presentation region associated with the speaker associated with the portion of the textual information.

19. A non-transitory computer readable medium storing data and computer implementable instructions that when executed by at least one processor cause the at least one processor to perform a method for processing audio and visually presenting information, the method comprising:

obtaining audio data captured by one or more audio sensors from an environment of a wearer of a wearable apparatus;

analyzing the audio data to obtain textual information;

analyzing the audio data to associate different portions of the textual information with different speakers;

associating a part of the textual information with speech produced by a user;

avoiding presenting the part of the textual information associated with the speech produced by the user; and

using a head mounted display system to present each portion of the textual information in a presentation region associated with the speaker associated with the portion of the textual information.

Assignments (3)
SECURITY INTEREST Recorded Feb 19, 2026
From: RPX CORPORATION
To: BARINGS FINANCE LLC, AS COLLATERAL AGENT
Reel/Frame 073831/0310 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 12, 2026
From: ARGSQUARE, LTD
To: RPX CORPORATION
Reel/Frame 073432/0806 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 2, 2024
From: ZASS, RON, DR.
To: ARGSQUARE LTD
Reel/Frame 069442/0513 →
Continuity (7)
Continuation 17396657 · Aug 7, 2021
Continuation 16670841 · Oct 31, 2019
Continuation 15650916 · Jul 16, 2017
Provisional Application 62460783 · Feb 18, 2017
Provisional Application 62444709 · Jan 10, 2017
Provisional Application 62363261 · Jul 16, 2016
Related Publication 20240055014A1 · Feb 15, 2024
References Cited (116)
US 5940798A · Houde · 1999 [cited by applicant]
US 6707921B2 · Moore · 2004 [cited by applicant]
US 6882971B2 · Craner · 2005 [cited by applicant]
US 7844454B2 · Coles · 2010 [cited by examiner]
US 8183997B1 · Wong · 2012 [cited by applicant]
US 8259954B2 · Shaffer · 2012 [cited by applicant]
US 8326338B1 · Vasilevsky · 2012 [cited by applicant]
US 8326625B2 · Adibi · 2012 [cited by examiner]
US 8441356B1 · Tedesco et al. · 2013 [cited by applicant]
US 8639508B2 · Zhao · 2014 [cited by examiner]
US 8660849B2 · Gruber · 2014 [cited by examiner]
US 8719016B1 · Ziv et al. · 2014 [cited by applicant]
US 8768711B2 · Ativanichayaphong · 2014 [cited by examiner]
US 8825468B2 · Jacobsen · 2014 [cited by applicant]
US 9213705B1 · Story · 2015 [cited by applicant]
US 9280973B1 · Soyannwo · 2016 [cited by examiner]
US 9282284B2 · Kajarekar et al. · 2016 [cited by applicant]
US 9355648B2 · Tsujikawa · 2016 [cited by applicant]
US 9412375B2 · Xiang · 2016 [cited by applicant]
US 9536536B2 · Hetherington et al. · 2017 [cited by applicant]
US 9641942B2 · Strelcyk et al. · 2017 [cited by applicant]
US 9697819B2 · Lv · 2017 [cited by examiner]
US 9746916B2 · Kim et al. · 2017 [cited by applicant]
US 9749583B1 · Fineberg et al. · 2017 [cited by applicant]
US 9799336B2 · Dzik · 2017 [cited by applicant]
US 9870357B2 · Cohen · 2018 [cited by applicant]
US 9894320B2 · Uchiyama et al. · 2018 [cited by applicant]
US 10084920B1 · Gainsboro et al. · 2018 [cited by applicant]
US 10134401B2 · Ziv · 2018 [cited by examiner]
US 10241741B2 · Laaksonen et al. · 2019 [cited by applicant]
US 10268447B1 · Dodge et al. · 2019 [cited by applicant]
US 10269372B1 · Fiedler · 2019 [cited by applicant]
US 10297257B2 · Tsujikawa · 2019 [cited by examiner]
US 10433052B2 · Zass · 2019 [cited by examiner]
US 10446166B2 · Dickins · 2019 [cited by applicant]
US 10516938B2 · Zass · 2019 [cited by examiner]
US 10593332B2 · Ziv · 2020 [cited by examiner]
US 10692485B1 · Grizzel · 2020 [cited by applicant]
US 10803852B2 · Yamamoto · 2020 [cited by applicant]
US 10878802B2 · Yamamoto · 2020 [cited by applicant]
US 10878819B1 · Chavez · 2020 [cited by examiner]
US 11195542B2 · Zass · 2021 [cited by examiner]
US 11837249B2 · Zass · 2023 [cited by examiner]
US 20020103649A1 · Basson et al. · 2002 [cited by applicant]
US 20020173957A1 · Kawane · 2002 [cited by examiner]
US 20030018475A1 · Basu et al. · 2003 [cited by applicant]
US 20040186712A1 · Coles · 2004 [cited by examiner]
US 20060167691A1 · Tuli · 2006 [cited by applicant]
US 20060238877A1 · Ashkenazi et al. · 2006 [cited by applicant]
US 20070172805A1 · Paul · 2007 [cited by applicant]
US 20080201141A1 · Abramov et al. · 2008 [cited by applicant]
US 20090306981A1 · Cromack et al. · 2009 [cited by applicant]
US 20090319265A1 · Wittenstein · 2009 [cited by examiner]
US 20100174533A1 · Pakhomov · 2010 [cited by applicant]
US 20100250257A1 · Hirose et al. · 2010 [cited by applicant]
US 20100262419A1 · De Bruijn et al. · 2010 [cited by applicant]
US 20100280336A1 · Giftakis et al. · 2010 [cited by applicant]
US 20110087491A1 · Wittenstein et al. · 2011 [cited by applicant]
US 20110228914A1 · Zourzouvillys et al. · 2011 [cited by applicant]
US 20120020503A1 · Endo et al. · 2012 [cited by applicant]
US 20120053929A1 · Hsia et al. · 2012 [cited by applicant]
US 20120116772A1 · Jones et al. · 2012 [cited by applicant]
US 20120128186A1 · Endo et al. · 2012 [cited by applicant]
US 20120128683A1 · Shantha · 2012 [cited by applicant]
US 20120215532A1 · Foo et al. · 2012 [cited by applicant]
US 20120265537A1 · Deshmukh et al. · 2012 [cited by applicant]
US 20130080168A1 · Iida et al. · 2013 [cited by applicant]
US 20140081634A1 · Forutanpour · 2014 [cited by applicant]
US 20140122077A1 · Nishikawa et al. · 2014 [cited by applicant]
US 20140163960A1 · Dimitriadis et al. · 2014 [cited by applicant]
US 20140236596A1 · Martinez · 2014 [cited by applicant]
US 20140247926A1 · Gainsboro et al. · 2014 [cited by applicant]
US 20140379352A1 · Gondi · 2014 [cited by applicant]
US 20150011842A1 · Steinberg-Shapira · 2015 [cited by applicant]
US 20150036856A1 · Pruthi et al. · 2015 [cited by applicant]
US 20150058013A1 · Pakhomov et al. · 2015 [cited by applicant]
US 20150073309A1 · Pracar et al. · 2015 [cited by applicant]
US 20150092007A1 · Koborita · 2015 [cited by examiner]
US 20150099946A1 · Sahin · 2015 [cited by applicant]
US 20150269672A1 · Bhuyan · 2015 [cited by examiner]
US 20150302867A1 · Tomlin et al. · 2015 [cited by applicant]
US 20150334346A1 · Cheatham et al. · 2015 [cited by applicant]
US 20150340038A1 · Dzik · 2015 [cited by examiner]
US 20150356836A1 · Schlesinger et al. · 2015 [cited by applicant]
US 20150373477A1 · Norris et al. · 2015 [cited by applicant]
US 20160019895A1 · Poisner · 2016 [cited by applicant]
US 20160049146A1 · Chang · 2016 [cited by examiner]
US 20160064002A1 · Kim et al. · 2016 [cited by applicant]
US 20160098993A1 · Yamamoto et al. · 2016 [cited by applicant]
US 20160133257A1 · Namgoong et al. · 2016 [cited by applicant]
US 20160350664A1 · Devarajan et al. · 2016 [cited by applicant]
US 20160379638A1 · Basye et al. · 2016 [cited by applicant]
US 20170041699A1 · Mackellar et al. · 2017 [cited by applicant]
US 20170076000A1 · Ashoori et al. · 2017 [cited by applicant]
US 20170103748A1 · Weissberg et al. · 2017 [cited by applicant]
US 20170105662A1 · Silawan et al. · 2017 [cited by applicant]
US 20170111303A1 · Nesbitt · 2017 [cited by applicant]
US 20170208415A1 · Ojala · 2017 [cited by applicant]
US 20170221500A1 · Glasgow et al. · 2017 [cited by applicant]
US 20170364599A1 · Ohanyerenwa et al. · 2017 [cited by applicant]
US 20180018300A1 · Zass · 2018 [cited by examiner]
US 20180018963A1 · Zass · 2018 [cited by examiner]
US 20180018974A1 · Zass · 2018 [cited by examiner]
US 20180018975A1 · Zass · 2018 [cited by examiner]
US 20180018987A1 · Zass · 2018 [cited by examiner]
US 20180020285A1 · Zass · 2018 [cited by examiner]
US 20180077483A1 · Sahay · 2018 [cited by applicant]
US 20180151173A1 · Harris · 2018 [cited by examiner]
US 20180240458A1 · Zass · 2018 [cited by examiner]
US 20180285312A1 · Liu et al. · 2018 [cited by applicant]
US 20190279654A1 · Ganeshkumar · 2019 [cited by applicant]
US 20200050863A1 · Wexler · 2020 [cited by applicant]
US 20200066294A1 · Zass · 2020 [cited by examiner]
US 20210366505A1 · Zass · 2021 [cited by examiner]
US 20230048149A1 · Zass · 2023 [cited by examiner]
US 20240055014A1 · Zass · 2024 [cited by examiner]