IP Library Patent Application 19023598
Patent Application
App. No. 19/023,598

Visualizing Auditory Content

Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US None
App. No.
19/023,598
Abstract

Systems, methods and non-transitory computer readable media for processing audio and visually presenting information are provided. Audio data may be obtained. The audio data may be analyzed to obtain textual information. The audio data may be analyzed to associate different portions of the textual information with different speakers. Each portion of the textual information may be presented in a presentation region associated with the speaker associated with the portion of the textual information. The audio data may be analyzed to identify a nonverbal sound. A textual description of the nonverbal sound may be generated.

Claims (45)

1 . A non-transitory computer readable medium storing data and computer implementable instructions that when executed by at least one processor cause the at least one processor to perform a method for processing audio and visually presenting information, the method comprising:

obtaining audio data;

analyzing the audio data to obtain textual information;

analyzing the audio data to associate different portions of the textual information with different speakers;

presenting each portion of the textual information in a presentation region associated with the speaker associated with the portion of the textual information;

analyzing the audio data to identify a nonverbal sound; and

generating a textual description of the nonverbal sound.

2 . The non-transitory computer readable medium of claim 1 , wherein the method further comprises:

analyzing the audio data to identify different voice properties in different parts of the audio data;

associating different parts of the textual information with different voice properties;

determining information based on particular voice properties; and

presenting the information determined based on the particular voice properties along the part of the textual information associated with the particular voice properties.

3 . The non-transitory computer readable medium of claim 1 , wherein the method further comprises:

analyzing the audio data to identify a melody; and

generating textual description of the melody.

4 . The non-transitory computer readable medium of claim 1 , wherein the presentation is a presentation using a head mounted display system.

5 . The non-transitory computer readable medium of claim 1 , wherein the presentation is a presentation using an augmented reality display system, and the association of the presentation regions with the speakers is configured to overlay the portions of the textual information associated with a speaker over the speaker in the augmented reality display system.

6 . The non-transitory computer readable medium of claim 1 , wherein the method further comprises: analyzing the audio data to identify an item in the audio data; representing the item using a graphical symbol; and displaying the graphical symbol in conjunction with the textual information.

7 . The non-transitory computer readable medium of claim 1 , wherein the method further comprises displaying the different portions of the textual information using different sets of visual display parameters.

8 . The non-transitory computer readable medium of claim 1 , wherein the method further comprises: associating the different parts of the textual information with different textual formats based on the association of the different parts of the textual information with the different voice properties.

9 . The non-transitory computer readable medium of claim 1 , wherein the method further comprises associating the different portions of the textual information with different textual formats based on the association of the different portions of the textual information with the different speakers.

10 . The non-transitory computer readable medium of claim 1 , wherein the method further comprises: determining that a portion of a speech is said in a specific linguistic tone; selecting a visual display parameter for a part of the textual information associated with the portion of the speech based on the specific linguistic tone; and displaying the part of the textual information using the selected visual display parameter.

11 . The non-transitory computer readable medium of claim 1 , wherein the method further comprises visually presenting information associated with a speaker in conjunction with a portion of the textual information associated with the speaker.

12 . The non-transitory computer readable medium of claim 1 , wherein the association of the presentation regions with the speakers is based on spatial orientation of the speakers.

13 . The non-transitory computer readable medium of claim 1 , wherein the association of the presentation regions with the speakers is based on positions of the speakers.

14 . The non-transitory computer readable medium of claim 1 , wherein the method further comprises: associating a part of the textual information with speech produced by a user; and avoiding presenting the part of the textual information associated with the speech produced by the user.

15 . The non-transitory computer readable medium of claim 1 , wherein the method further comprises: determining that a part of the textual information is associated with speech that do not involve the user; and avoiding presenting the part of the textual information associated.

16 . The non-transitory computer readable medium of claim 15 , wherein the speech that do not involve the user is a conversation that do not involve the user.

17 . The non-transitory computer readable medium of claim 15 , wherein the speech that do not involve the user is a speech not directed at the user.

18 . The non-transitory computer readable medium of claim 1 , wherein the audio data is audio data captured by the one or more audio sensors from an environment of a wearer of a wearable apparatus.

19 . A system for processing audio and visually presenting information, the system comprising:

at least one processing unit configured to:

obtain audio data;

analyze the audio data to obtain textual information;

analyze the audio data to associate different portions of the textual information with different speakers;

present each portion of the textual information in a presentation region associated with the speaker associated with the portion of the textual information;

analyze the audio data to identify a nonverbal sound; and

generate a textual description of the nonverbal sound.

20 . A method for processing audio and visually presenting information, the method comprising:

obtaining audio data;

analyzing the audio data to obtain textual information;

analyzing the audio data to associate different portions of the textual information with different speakers;

presenting each portion of the textual information in a presentation region associated with the speaker associated with the portion of the textual information;

analyzing the audio data to identify a nonverbal sound; and

generating a textual description of the nonverbal sound.

Assignments (3)
SECURITY INTEREST Recorded Feb 19, 2026
From: RPX CORPORATION
To: BARINGS FINANCE LLC, AS COLLATERAL AGENT
Reel/Frame 073831/0310 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 12, 2026
From: ARGSQUARE, LTD
To: RPX CORPORATION
Reel/Frame 073432/0806 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 12, 2025
From: ZASS, RON, DR.
To: ARGSQUARE LTD
Reel/Frame 070822/0954 →