IP Library Granted Patent US 10,777,210
Granted Patent B2
US 10,777,210 · App. 16/358,109 · Granted Sep 15, 2020

Key phrase detection with audio watermarking

Inventor: Ricardo Antonio Garcia (Mountain View, CA)
Assignee: Google LLC
G10L19/018G06F3/165G06F21/31G10L15/08G10L15/22G10L21/00G10L2015/088G10L2015/223H04N21/233H04N21/8358
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,777,210
App. No.
16/358,109
Granted
Sep 15, 2020
Kind
B2
Abstract

Methods, systems, and apparatus, including computer programs encoded on computer storage media, for using audio watermarks with key phrases. One of the methods includes receiving, by a playback device, an audio data stream; determining, before the audio data stream is output by the playback device, whether a portion of the audio data stream encodes a particular key phrase by analyzing the portion using an automated speech recognizer; in response to determining that the portion of the audio data stream encodes the particular key phrase, modifying the audio data stream to include an audio watermark; and providing the modified audio data stream for output.

Claims (34)

1. A computer-implemented method comprising:

receiving an audio data stream;

determining whether a portion of the audio data stream encodes a particular key phrase by analyzing the portion using an automated speech recognizer;

in response to determining that the portion of the audio data stream encodes the particular key phrase, modifying the audio data stream to include an audio watermark that includes data specifying that the particular key phrase is encoded in the portion of the audio stream; and

providing the modified audio data stream for output.

2. The method of claim 1 , wherein modifying the audio data stream to include the audio watermark comprises:

determining whether the received audio data stream includes a watermark for the particular key phrase; and

in response to determining that the received audio data stream does not include a watermark for the particular key phrase, modifying the audio data stream to include the audio watermark.

3. The method of claim 1 , wherein modifying the audio data stream to include the audio watermark comprises:

determining whether the received audio data stream includes a watermark for the particular key phrase;

in response to determining that the received audio data stream includes a watermark for the particular key phrase, determining whether specific data is encoded in the watermark by analyzing data encoded in the watermark; and

in response to determining that specific data is not encoded in the watermark, modifying the audio data stream to include the audio watermark that encodes the specific data.

4. The method of claim 3 , wherein modifying the audio data stream to include the audio watermark that encodes the specific data comprises modifying the watermark from the received audio data stream to encode the specific data.

5. The method of claim 3 , wherein the specific data comprises data for the particular key phrase.

6. The method of claim 3 , wherein the specific data comprises data for a source of the audio data stream.

7. The method of claim 3 , wherein the specific data comprises data about content encoded in the audio data stream.

8. The method of claim 1 , further comprising receiving another portion of the audio data stream concurrently with determining whether the portion of the audio data stream encodes the particular key phrase by analyzing the portion using the automated speech recognizer.

9. The method of claim 1 , wherein the particular key phrase is fixed.

10. The method of claim 1 , further comprising receiving input defining the particular key phrase prior to determining whether the portion of the audio data stream encodes the particular key phrase by analyzing the portion using the automated speech recognizer.

11. The method of claim 1 , wherein receiving the audio data stream comprises receiving the audio data stream through a wired or wireless input connection other than a microphone prior to providing the portion of the modified audio data stream for output.

12. The method of claim 1 , wherein modifying the audio data stream to include the audio watermark comprises modifying the audio data stream to include the audio watermark that identifies a source of the audio data stream.

13. A system comprising:

one or more computers and one or more storage devices storing instructions that are operable, when executed by the one or more computers, to cause the one or more computers to perform operations comprising:

receiving an audio data stream;

determining whether a portion of the audio data stream encodes a particular key phrase by analyzing the portion using an automated speech recognizer;

in response to determining that the portion of the audio data stream encodes the particular key phrase, modifying the audio data stream to include an audio watermark that includes data specifying that the particular key phrase is encoded in the portion of the audio stream; and

providing the modified audio data stream for output.

14. The system of claim 13 , wherein modifying the audio data stream to include the audio watermark comprises:

determining whether the received audio data stream includes a watermark for the particular key phrase; and

in response to determining that the received audio data stream does not include a watermark for the particular key phrase, modifying the audio data stream to include the audio watermark.

15. The system of claim 13 , wherein modifying the audio data stream to include the audio watermark comprises:

determining whether the received audio data stream includes a watermark for the particular key phrase;

in response to determining that the received audio data stream includes a watermark for the particular key phrase, determining whether specific data is encoded in the watermark by analyzing data encoded in the watermark; and

in response to determining that specific data is not encoded in the watermark, modifying the audio data stream to include the audio watermark that encodes the specific data.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 20, 2019
From: GARCIA, RICARDO ANTONIO
To: GOOGLE LLC
Reel/Frame 048651/0753 →
Continuity (2)
Continuation 15824183 · Nov 28, 2017
Related Publication 20190214030A1 · Jul 11, 2019
Cited By (3)
US 12,354,622 US 12,494,219 US 12,626,708