IP Library › Granted Patent US 11,211,076
Granted Patent B2
US 11,211,076 · App. 16/992,647 · Granted Dec 28, 2021

Key phrase detection with audio watermarking

Inventor: Ricardo Antonio Garcia (Mountain View, CA)
Assignee: Google LLC
G10L19/018G06F3/165G06F21/31G10L15/08G10L15/22G10L21/00G10L2015/088G10L2015/223H04N21/233H04N21/8358
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,211,076
App. No.
16/992,647
Granted
Dec 28, 2021
Kind
B2
Abstract

Methods, systems, and apparatus, including computer programs encoded on computer storage media, for using audio watermarks with key phrases. One of the methods includes receiving, by a playback device, an audio data stream; determining, before the audio data stream is output by the playback device, whether a portion of the audio data stream encodes a particular key phrase by analyzing the portion using an automated speech recognizer; in response to determining that the portion of the audio data stream encodes the particular key phrase, modifying the audio data stream to include an audio watermark; and providing the modified audio data stream for output.

Claims (40)

1. A method comprising:

receiving, at data processing hardware of a playback device, from a content provider, an audio data stream corresponding to one of music content or video content, wherein the playback device receives the audio data stream from the content provider through a wireless input connection other than a microphone;

creating, by the data processing hardware, a modified audio data stream by:

dynamically generating multiple audio watermarks encoding data that indicates the audio data stream originated from the content provider; and

inserting the dynamically generated multiple audio watermarks into the audio data stream to create the modified audio data stream; and

providing, by the data processing hardware, the modified audio data stream for output through a speaker in communication with the data processing hardware,

wherein after providing the modified audio data stream for output through the speaker, a listening device, while in an awake mode responsive to detecting a key phrase via a microphone, is configured to:

capture the modified audio data stream via the microphone; and

determine an action to perform using the multiple audio watermarks encoding the data that indicates the audio data stream originated from the content provider.

2. The method of claim 1 , wherein the playback device:

receives the audio data stream in a video stream from the content provider through the wireless input connection; and

connects to a display using a digital audio and video connection.

3. The method of claim 2 , further comprising, when providing the modified audio data stream for output through the speaker, providing, by the data processing hardware, using the digital audio and video connection, a video portion of the video for presentation by the display.

4. The method of claim 3 , wherein the playback device synchronizes presentation of the video portion of the video stream by the display with the modified audio data stream for output through the speaker.

5. The method of claim 2 , wherein the playback device connects to a television using the digital audio and video connection, the television comprising the display and the speaker.

6. The method of claim 1 , wherein the playback device comprises the speaker.

7. The method of claim 1 , wherein the listening device is located in a same room as the speaker.

8. The method of claim 1 , wherein a portion of the multiple audio watermarks in the modified audio data stream encode different data than the other multiple audio watermarks.

9. The method of claim 1 , wherein each of the multiple audio watermarks encode the same data.

10. A playback device comprising:

data processing hardware; and

memory hardware in communication with the data processing hardware and storing instructions that when executed on the data processing hardware cause the data processing hardware to perform operations comprising:

receiving, from a content provider, an audio data stream corresponding to one of music content or video content, wherein the playback device receives the audio data stream from the content provider through a wireless input connection other than a microphone;

creating a modified audio data stream by:

dynamically generating multiple audio watermarks encoding data that indicates the audio data stream originated from the content provider; and

inserting the dynamically generated multiple audio watermarks into the audio data stream to create the modified audio data stream; and

providing the modified audio data stream for output through a speaker in communication with the data processing hardware,

wherein after providing the modified audio data stream for output through the speaker, a listening device, while in an awake mode responsive to detecting a key phrase via a microphone, is configured to:

capture the modified audio data stream via the microphone; and

determine an action to perform using the multiple audio watermarks encoding the data that indicates the audio data stream originated from the content provider.

11. The playback device of claim 10 , wherein the playback device:

receives the audio data stream in a video stream from the content provider through the wireless input connection; and

connects to a display using a digital audio and video connection.

12. The playback device of claim 11 , wherein the operations further comprise, when providing the modified audio data stream for output through the speaker, providing, using the digital audio and video connection, a video portion of the video for presentation by the display.

13. The playback device of claim 12 , wherein the playback device synchronizes presentation of the video portion of the video stream by the display with the modified audio data stream for output through the speaker.

14. The playback device of claim 11 , wherein the playback device connects to a television using the digital audio and video connection, the television comprising the display and the speaker.

15. The playback device of claim 10 , wherein the playback device comprises the speaker.

16. The playback device of claim 10 , wherein the listening device is located in a same room as the speaker.

17. The playback device of claim 10 , wherein a portion of the multiple audio watermarks in the modified audio data stream encode different data than the other multiple audio watermarks.

18. The playback device of claim 10 , wherein each of the multiple audio watermarks encode the same data.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 19, 2020
From: GARCIA, RICARDO ANTONIO
To: GOOGLE LLC
Reel/Frame 053538/0025 →
Continuity (2)
Continuation 16358109 · Mar 19, 2019
Related Publication 20200372922A1 · Nov 26, 2020