IP Library Granted Patent US 11,087,777
Granted Patent B1
US 11,087,777 · App. 16/853,449 · Granted Aug 10, 2021

Audio visual correspondence based signal augmentation

Inventors: Cesare Valerio Parise (Seattle, WA); William Owen Brimijoin, II (Kirkland, WA); Philip Robinson (Seattle, WA)
Assignee: Facebook Technologies, LLC
G10L21/0316G02B27/0093G06F3/167G10L21/0272G10L25/87H04R3/002H04R5/04
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,087,777
App. No.
16/853,449
Granted
Aug 10, 2021
Kind
B1
Abstract

A system includes a headset to capture sound and a visual signal of a local area including one or more sound sources. The system determines a strength of the audio signal and a portion of the visual signal associated with the audio signal, compares the strengths, selects the weaker signal, and augments the weaker signal. The headset accordingly presents augmented audio-visual content to a user, thereby enhancing the user's perception of the weak signal.

Claims (60)

1. A method comprising:

identifying an audio signal of a sound source based in part on a correspondence analysis of a visual signal describing a local area that includes the sound source and sound produced by a plurality of sound sources including the sound source within the local area;

determining an audio signal strength associated with the audio signal;

determining a visual signal strength associated with a portion of the visual signal corresponding to the audio signal;

selecting a weak signal from a group of signals including the audio signal and the portion of the visual signal, the selecting based in part on the visual signal strength and the audio signal strength; and

augmenting the weak signal, wherein the augmented weak signal is presented to a user in conjunction with other signals from the group of signals.

2. The method of claim 1 , further comprising:

responsive to selecting the audio signal as the weak signal, augmenting the audio signal.

3. The method of claim 2 , wherein augmenting the audio signal comprises amplifying the audio signal.

4. The method of claim 2 , wherein augmenting the audio signal comprises:

identifying frequencies in the audio signal that do not correspond to the sound source; and

applying a filter to the audio signal to attenuate the identified frequencies.

5. The method of claim 1 , further comprising:

responsive to selecting the visual signal as the weak signal, augmenting the visual signal.

6. The method of claim 5 , wherein augmenting the visual signal comprises:

modifying one or more properties of the visual signal, wherein the one or more properties include at least one of: brightness, contrast, color, and sharpness.

7. The method of claim 5 , wherein augmenting the visual signal comprises:

augmenting a portion of the visual signal corresponding to the audio signal.

8. The method of claim 5 , wherein augmenting the visual signal comprises:

modifying the visual signal to include a visual indicator proximate to a location of the sound source.

9. The method of claim 8 , wherein a movement of the visual indicator corresponds to the audio signal.

10. The method of claim 1 , wherein selecting the weak signal comprises:

receiving an input from the user indicating the weak signal; and

selecting the weak signal based on the received input.

11. The method of claim 1 , wherein selecting the weak signal comprises:

determining a signal-to-noise ratio of the audio signal;

determining a signal-to-noise ratio of the portion of the visual signal; and

based on a comparison of the signal-to-noise ratio of the audio signal and the portion of the visual signal, selecting the weak signal.

12. The method of claim 1 , wherein augmenting the weak signal comprises:

generating an instruction to present a haptic signal to the user in conjunction with the other signals.

13. A system comprising:

a transducer assembly configured to present audio to a user;

a display assembly configured to present visual content to the user; and

a controller configured to:

identify an audio signal of a sound source based on a correspondence analysis of a visual signal describing a local area that includes a sound source and sound produced by a plurality of sound sources including the sound source within the local area;

determine an audio signal strength associated with the audio signal;

determine a visual signal strength associated with a portion of the visual signal corresponding to the audio signal;

select a weak signal from a group of signals including the audio signal and the portion of the visual signal, the selecting based in part on the visual signal strength and the audio signal strength; and

augment the weak signal, wherein the augmented weak signal is presented to the user via at least one of the transducer assembly and the display assembly in conjunction with other signals from the group of signals.

14. The system of claim 13 , wherein the controller is configured to:

select the audio signal as the weak signal; and

responsive to selecting the audio signal, augment the audio signal.

15. The system of claim 14 , wherein the controller is further configured to:

amplify the audio signal.

16. The system of claim 14 , wherein the controller is further configured to:

identify frequencies in the audio signal that do not correspond to the sound source; and

apply a filter to the audio signal to attenuate the identified frequencies.

17. The system of claim 13 , wherein the controller is further configured to:

select the visual signal as the weak signal; and

responsive to selecting the visual signal, augment the visual signal.

18. The system of claim 17 , wherein the controller is further configured to:

modify one or more properties of a portion of the visual signal that corresponds to the audio signal, the one or more properties including at least one of: brightness, contrast, color, and sharpness.

19. The system of claim 17 , wherein the controller is further configured to:

modify the visual signal to include a visual indicator proximate to a location of the sound source.

20. A non-transitory computer readable medium configured to store program code instructions, when executed by a processor, cause the processor to perform steps comprising:

identifying an audio signal of a sound source based in part on a correspondence analysis of a visual signal describing a local area that includes the sound source and sound produced by a plurality of sound sources including the sound source within the local area;

determining an audio signal strength associated with the audio signal;

determining a visual signal strength associated with a portion of the visual signal corresponding to the audio signal;

selecting a weak signal from a group of signals including the audio signal and the portion of the visual signal, the selecting based in part on the visual signal strength and the audio signal strength; and

augmenting the weak signal, wherein the augmented weak signal is presented to a user in conjunction with other signals from the group of signals.

Assignments (2)
CHANGE OF NAME Recorded Jun 8, 2022
From: FACEBOOK TECHNOLOGIES, LLC
To: META PLATFORMS TECHNOLOGIES, LLC
Reel/Frame 060315/0224 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded May 13, 2020
From: VALERIO PARISE, CESARE; BRIMIJOIN, WILLIAM OWEN, II; ROBINSON, PHILIP
To: FACEBOOK TECHNOLOGIES, LLC
Reel/Frame 052654/0241 →
Continuity (1)
Provisional Application 62975096 · Feb 11, 2020