IP Library Granted Patent US 12,009,007
Granted Patent B2
US 12,009,007 · App. 18/135,649 · Granted Jun 11, 2024

Voice trigger for a digital assistant

Inventors: Justin Binder (Oakland, CA); Samuel D. Post (Great Falls, MT); Onur Tackin (Saratoga, CA); Thomas R. Gruber (Santa Cruz, CA)
Assignee: Apple Inc.
G10L21/16G06F3/167G10L15/22G10L15/26G10L17/24G10L15/02G10L2015/223G10L15/30G10L25/51G10L25/84
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,009,007
App. No.
18/135,649
Granted
Jun 11, 2024
Kind
B2
Abstract

A method for operating a voice trigger is provided. In some implementations, the method is performed at an electronic device including one or more processors and memory storing instructions for execution by the one or more processors. The method includes receiving a sound input. The sound input may correspond to a spoken word or phrase, or a portion thereof. The method includes determining whether at least a portion of the sound input corresponds to a predetermined type of sound, such as a human voice. The method includes, upon a determination that at least a portion of the sound input corresponds to the predetermined type, determining whether the sound input includes predetermined content, such as a predetermined trigger word or phrase. The method also includes, upon a determination that the sound input includes the predetermined content, initiating a speech-based service, such as a voice-based digital assistant.

Claims (76)

1. A method for operating a voice trigger, the method performed at an electronic device including a first microphone, one or more processors, and memory storing instructions for execution by the one or more processors, the method comprising:

while the electronic device is coupled to an external electronic device with a second microphone:

in accordance with a determination that a predetermined condition is satisfied:

causing the second microphone to monitor audio input to determine whether the audio input includes predetermined content corresponding to a voice trigger operating on the electronic device; and

forgoing monitoring for the audio input using the first microphone; and

in accordance with a determination that the predetermined condition is not satisfied:

monitoring for the audio input using the first microphone; and

determining whether the audio input includes the predetermined content.

2. The method of claim 1 , wherein the predetermined condition is satisfied when the external electronic device is being worn by a user.

3. The method of claim 1 , wherein the external electronic device incudes a wearable device.

4. The method of claim 1 , wherein the predetermined condition is satisfied based on a determination of a difference between motion of the electronic device and motion of the external electronic device.

5. The method of claim 1 , further comprising:

while the electronic device is coupled to the external electronic device:

in accordance with a determination that the predetermined condition is satisfied and after a determination that the audio input includes the predetermined content:

receiving, from the external electronic device, a voice input that includes a command directed to a digital assistant operating on the electronic device.

6. The method of claim 1 , further comprising:

in accordance with a determination that the audio input includes the predetermined content:

initiating a speech-based service operating on the electronic device.

7. The method of claim 6 , wherein the speech-based service includes a digital assistant operating on the electronic device.

8. The method of claim 1 , further comprising:

in accordance with a determination that the audio input includes the predetermined content and in accordance with a determination that the audio input corresponds to a predetermined type of sound:

initiating a speech-based service operating on the electronic device.

9. The method of claim 8 , wherein the predetermined type of sound includes a human voice.

10. The method of claim 1 , wherein the predetermined content includes one or more words.

11. An electronic device, comprising:

one or more processors;

memory;

a first microphone; and

one or more programs, wherein the one or more programs are stored in the memory and configured to be executed by the one or more processors, the one or more programs including instructions for:

while the electronic device is coupled to an external electronic device with a second microphone:

in accordance with a determination that a predetermined condition is satisfied:

causing the second microphone to monitor audio input to determine whether the audio input includes predetermined content corresponding a voice trigger operating on the electronic device; and

forgoing monitoring for the audio input using the first microphone; and

in accordance with a determination that the predetermined condition is not satisfied:

monitoring for the audio input using the first microphone; and

determining whether the audio input includes the predetermined content.

12. The electronic device of claim 11 , wherein the predetermined condition is satisfied when the external electronic device is being worn by a user.

13. The electronic device of claim 11 , wherein the external electronic device incudes a wearable device.

14. The electronic device of claim 11 , wherein the predetermined condition is satisfied based on a determination of a difference between motion of the electronic device and motion of the external electronic device.

15. The electronic device of claim 11 , wherein the one or more programs further include instructions for:

while the electronic device is coupled to the external electronic device:

in accordance with a determination that the predetermined condition is satisfied and after a determination that the audio input includes the predetermined content:

receiving, from the external electronic device, a voice input that includes a command directed to a digital assistant operating on the electronic device.

16. The electronic device of claim 11 , wherein the one or more programs further include instructions for:

in accordance with a determination that the audio input includes the predetermined content:

initiating a speech-based service operating on the electronic device.

17. The electronic device of claim 16 , wherein the speech-based service includes a digital assistant operating on the electronic device.

18. The electronic device of claim 11 , wherein the one or more programs further include instructions for:

in accordance with a determination that the audio input includes the predetermined content and in accordance with a determination that the audio input corresponds to a predetermined type of sound:

initiating a speech-based service operating on the electronic device.

19. The electronic device of claim 18 , wherein the predetermined type of sound includes a human voice.

20. The electronic device of claim 11 , wherein the predetermined content includes one or more words.

21. A non-transitory computer-readable storage medium storing one or more programs for execution by one or more processors of an electronic device with a first microphone, the one or more programs including instructions for:

while the electronic device is coupled to an external electronic device with a second microphone:

in accordance with a determination that a predetermined condition is satisfied:

causing the second microphone to monitor audio input to determine whether the audio input includes predetermined content corresponding to a voice trigger operating on the electronic device; and

forgoing monitoring for the audio input using the first microphone; and

in accordance with a determination that the predetermined condition is not satisfied:

monitoring for the audio input using the first microphone; and

determining whether the audio input includes the predetermined content.

22. The non-transitory computer-readable storage medium of claim 21 , wherein the predetermined condition is satisfied when the external electronic device is being worn by a user.

23. The non-transitory computer-readable storage medium of claim 21 , wherein the external electronic device incudes a wearable device.

24. The non-transitory computer-readable storage medium of claim 21 , wherein the predetermined condition is satisfied based on a determination of a difference between motion of the electronic device and motion of the external electronic device.

25. The non-transitory computer-readable storage medium of claim 21 , wherein the one or more programs further include instructions for:

while the electronic device is coupled to the external electronic device:

in accordance with a determination that the predetermined condition is satisfied and after a determination that the audio input includes the predetermined content:

receiving, from the external electronic device, a voice input that includes a command directed to a digital assistant operating on the electronic device.

26. The non-transitory computer-readable storage medium of claim 21 , wherein the one or more programs further include instructions for:

in accordance with a determination that the audio input includes the predetermined content:

initiating a speech-based service operating on the electronic device.

27. The non-transitory computer readable storage medium of claim 26 , wherein the speech-based service includes a digital assistant operating on the electronic device.

28. The non-transitory computer-readable storage medium of claim 21 , wherein the one or more programs further include instructions for:

in accordance with a determination that the audio input includes the predetermined content and in accordance with a determination that the audio input corresponds to a predetermined type of sound:

initiating a speech-based service operating on the electronic device.

29. The non-transitory computer-readable storage medium of claim 28 , wherein the predetermined type of sound includes a human voice.

30. The non-transitory computer-readable storage medium of claim 21 , wherein the predetermined content includes one or more words.

Continuity (8)
Continuation 17962220 · Oct 7, 2022
Continuation 17713741 · Apr 5, 2022
Continuation 17150513 · Jan 15, 2021
Continuation 16879348 · May 20, 2020
Continuation 16222249 · Dec 17, 2018
Continuation 14175864 · Feb 7, 2014
Provisional Application 61762260 · Feb 7, 2013
Related Publication 20230253005A1 · Aug 10, 2023