IP Library › Granted Patent US 10,930,266
Granted Patent B2
US 10,930,266 · App. 16/665,461 · Granted Feb 23, 2021

Methods and devices for selectively ignoring captured audio data

Inventors: James David Meyers (San Jose, CA); Kurt Wesley Piersol (San Jose, CA)
Assignee: Amazon Technologies, Inc.
G10L15/08G10L15/04G10L15/20G10L21/028G10L15/22G10L2015/088G10L2021/02082
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,930,266
App. No.
16/665,461
Filed
Oct 28, 2019
Granted
Feb 23, 2021
Kind
B2
Examiner
YEN, ERIC L
Art Unit
2658
USPC
704/251
Abstract

Systems and methods for selectively ignoring an occurrence of a wakeword within audio input data is provided herein. In some embodiments, a wakeword may be detected to have been uttered by an individual within a modified time window, which may account for hardware delays and echoing offsets. The detected wakeword that occurs during this modified time window may, in some embodiments, correspond to a word included within audio that is outputted by a voice activated electronic device. This may cause the voice activated electronic device to activate itself, stopping the audio from being outputted. By identifying when these occurrences of the wakeword within outputted audio are going to happen, the voice activated electronic device may selectively determine when to ignore the wakeword, and furthermore, when not to ignore the wakeword.

Claims (63)

1. A computer-implemented method comprising:

receiving audio data;

receiving, via a network from a first device external to a second device, metadata corresponding to detection of a wakeword represented in the audio data;

causing output of audio in response to the audio data; and

processing the metadata to disable audio processing to avoid a disruption of operation of the second device in response to the wakeword being detected in the audio by the second device.

2. The computer-implemented method of claim 1 , wherein the disruption of operation of the second device comprises a disruption of the output of the audio by the second device.

3. The computer-implemented method of claim 1 , wherein processing the metadata to avoid the disruption comprises:

using the metadata to generate a first command to alter operation of a component of the second device; and

sending the first command to the component.

4. The computer-implemented method of claim 3 , wherein:

the component comprises a hardware component of the second device; and

the first command causes the hardware component to be disabled.

5. The computer-implemented method of claim 3 , wherein:

the component comprises an audio input component of the second device; and

the first command causes the audio input component to be disabled.

6. The computer-implemented method of claim 3 , wherein:

the component comprises a speech processing component of the second device; and

the first command causes the speech processing component to be disabled.

7. The computer-implemented method of claim 3 , wherein:

the component comprises a wakeword detection component of the second device; and

the first command causes the wakeword detection component to be disabled.

8. The computer-implemented method of claim 3 , wherein:

the component comprises a wakeword detection component of the second device; and

the first command causes the wakeword detection component to disregard an indication of detection of the wakeword.

9. The computer-implemented method of claim 3 , further comprising:

determining, based at least in part on the metadata, an estimated time window when the wakeword will be represented in output audio; and

sending a second command to the component to restore operation of the component after the estimated time window.

10. A system, comprising:

at least one processor;

at least one memory comprising instructions that, when executed by the at least one processor, cause the system to:

receive audio data;

receive, via a network from a first device external to a second device, metadata corresponding to detection of a wakeword represented in the audio data;

cause output of audio in response to the audio data; and

process the metadata to disable audio processing to avoid a disruption of operation of the second device in response to the wakeword being detected in the audio by the second device.

11. The system of claim 10 , wherein the disruption of operation of the second device comprises a disruption of the output of the audio by the second device.

12. The system of claim 10 , wherein the instructions that cause the system to process the metadata to avoid the disruption comprise instructions that, when executed by the at least one processor, cause the system to:

use the metadata to generate a first command to alter operation of a component of the second device; and

send the first command to the component.

13. The system of claim 12 , wherein:

the component comprises a hardware component of the second device; and

the first command causes the hardware component to be disabled.

14. The system of claim 12 , wherein:

the component comprises an audio input component of the second device; and

the first command causes the audio input component to be disabled.

15. The system of claim 12 , wherein:

the component comprises a speech processing component of the second device; and

the first command causes the speech processing component to be disabled.

16. The system of claim 12 , wherein:

the component comprises a wakeword detection component of the second device; and

the first command causes the wakeword detection component to be disabled.

17. The system of claim 12 , wherein:

the component comprises a wakeword detection component of the second device; and

the first command causes the wakeword detection component to disregard an indication of detection of the wakeword.

18. The system of claim 12 , wherein the at least one memory further comprises instructions that, when executed by the at least one processor, further cause the system to:

determine, based at least in part on the metadata, an estimated time window when the wakeword will be represented in output audio; and

send a second command to the component to restore operation of the component after the estimated time window.

19. A computer-implemented method comprising:

receiving audio data;

receiving metadata indicating that a wakeword is represented in the audio data;

causing output of audio in response to the audio data;

processing the metadata to avoid a disruption of operation of a device in response to output of the audio corresponding to the wakeword;

using the metadata to generate a first command to alter operation of an audio input component of the device; and

sending the first command to the component.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 28, 2019
From: MEYERS, JAMES DAVID; PIERSOL, KURT WESLEY
To: AMAZON TECHNOLOGIES, INC.
Reel/Frame 050842/0911 →
Continuity (4)
Continuation 16036345 · Jul 16, 2018
Continuation 15633529 · Jun 26, 2017
Continuation 14934069 · Nov 5, 2015
Related Publication 20200066258A1 · Feb 27, 2020
Cited By (1)
US 12,249,331