IP Library › Granted Patent US 10,043,516
Granted Patent B2
US 10,043,516 · App. 15/385,606 · Granted Aug 7, 2018

Intelligent automated assistant

Inventors: Harry J. Saddler (Berkeley, CA); Aimee T. Piercy (Cupertino, CA); Garrett L. Weinberg (Cupertino, CA); Susan L. Booker (Cupertino, CA)
Assignee: Apple Inc.
G10L15/22G06F3/165G10L13/02G10L15/1815G06F3/0482G06F3/0488G06F3/04817G10L2015/088G10L2015/221G10L2015/223
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,043,516
App. No.
15/385,606
Granted
Aug 7, 2018
Kind
B2
Abstract

Systems and processes for operating an automated assistant are disclosed. In one example process, an electronic device provides an audio output via a speaker of the electronic device. While providing the audio output, the electronic device receives, via a microphone of the electronic device, a natural language speech input. The electronic device derives a representation of user intent based on the natural language speech input and the audio output, identifies a task based on the derived user intent; and performs the identified task.

Claims (98)

1. An electronic device for operating an automated assistant, the electronic device comprising:

one or more processors;

a memory;

a speaker;

a microphone; and

one or more programs, wherein the one or more programs are stored in the memory and configured to be executed by the one or more processors, the one or more programs including instructions for:

providing, via the speaker of the electronic device, an audio output;

while providing the audio output via the speaker of the electronic device, receiving, via the microphone of the electronic device, a natural language speech input;

in response to receiving the natural language speech input, determining a type of the audio output;

in response to a determination that the audio output is of a first type, adjusting the audio output;

in response to a determination that the audio output is of a second type different from the first type, ceasing to provide the audio output;

deriving a representation of user intent based on the natural language speech input and the audio output;

identifying a task based on the derived user intent; and

performing the identified task.

2. The electronic device of claim 1 , the one or more programs further including instructions for:

identifying one or more parameters associated with the task based on a portion of the audio output;

wherein performing the task includes performing the task based on the identified one or more parameters.

3. The electronic device of claim 2 , the one or more programs further including instructions for: in response to receipt of the natural language speech input, identifying the portion of the audio output.

4. The electronic device of claim 2 ,

wherein providing the audio output comprises providing a speech output indicative of a list of items, and

wherein the portion of the audio output is indicative of an item of the list of items.

5. The electronic device of claim 4 ,

wherein the item is a media item, and

wherein performing the task comprises performing playback, via the speaker, of the media item.

6. The electronic device of claim 4 ,

wherein the item is a location, and

wherein performing the task comprises providing, via the speaker, information associated with the location.

7. The electronic device of claim 1 ,

wherein providing the audio output comprises performing playback of media content, and

wherein performing the task comprises adjusting playback of the media content.

8. The electronic device of claim 7 , wherein adjusting playback of the media content comprises: adjusting a volume of the speaker of the electronic device.

9. The electronic device of claim 7 , wherein adjusting playback of the media content comprises: pausing playback of the media content.

10. The electronic device of claim 1

wherein adjusting the audio output comprises attenuating the audio output.

11. The electronic device of claim 1 , wherein the audio output is a first audio output, the one or more programs further including instructions for:

before providing the first audio output, providing a second audio output.

12. A method for operating an automated assistant, the method comprising:

at an electronic device with a speaker and a microphone,

providing, via the speaker of the electronic device, an audio output;

while providing the audio output via the speaker of the electronic device, receiving, via the microphone of the electronic device, a natural language speech input;

in response to receiving the natural language speech input, determining a type of the audio output;

in response to a determination that the audio output is of a first type, adjusting the audio output;

in response to a determination that the audio output is of a second type different from the first type, ceasing to provide the audio output;

deriving a representation of user intent based on the natural language speech input and the audio output;

identifying a task based on the derived user intent; and

performing the identified task.

13. The method of claim 12 , further comprising:

identifying one or more parameters associated with the task based on a portion of the audio output;

wherein performing the task includes performing the task based on the identified one or more parameters.

14. The method of claim 13 , further comprising: in response to receipt of the natural language speech input, identifying the portion of the audio output.

15. The method of claim 13 ,

wherein providing the audio output comprises providing a speech output indicative of a list of items, and

wherein the portion of the audio output is indicative of an item of the list of items.

16. The method of claim 15 ,

wherein the item is a media item, and

wherein performing the task comprises performing playback, via the speaker, of the media item.

17. The method of claim 15 ,

wherein the item is a location, and

wherein performing the task comprises providing, via the speaker, information associated with the location.

18. The method of claim 12 ,

wherein providing the audio output comprises performing playback of media content, and

wherein performing the task comprises adjusting playback of the media content.

19. The method of claim 18 , wherein adjusting playback of the media content comprises: adjusting a volume of the speaker of the electronic device.

20. The method of claim 18 , wherein adjusting playback of the media content comprises: pausing playback of the media content.

21. The method of claim 12 , wherein adjusting the audio output comprises attenuating the audio output.

22. The method of claim 12 , wherein the audio output is a first audio output, the method further comprising:

before providing the first audio output, providing a second audio output.

23. A non-transitory computer readable storage medium storing one or more programs, the one or more programs comprising instructions, which when executed by one or more processors of an electronic device, cause the device to:

provide, via a speaker of the electronic device, an audio output;

while providing the audio output via the speaker of the electronic device, receive, via a microphone of the electronic device, a natural language speech input;

in response to receiving the natural language speech input, determine a type of the audio output;

in response to a determination that the audio output is of a first type, adjust the audio output;

in response to a determination that the audio output is of a second type different from the first type, cease to provide the audio output;

derive a representation of user intent based on the natural language speech input and the audio output;

identify a task based on the derived user intent; and

perform the identified task.

24. The non-transitory computer readable storage medium of claim 23 , the one or more programs further comprising instructions, which when executed by one or more processors of the electronic device, cause the device to:

identify one or more parameters associated with the task based on a portion of the audio output;

wherein performing the task includes performing the task based on the identified one or more parameters.

25. The non-transitory computer readable storage medium of claim 24 , the one or more programs further comprising instructions, which when executed by one or more processors of the electronic device, cause the device to:

in response to receipt of the natural language speech input, identify he portion of the audio output.

26. The non-transitory computer readable storage medium of claim 24 ,

wherein providing the audio output comprises providing a speech output indicative of a list of items, and

wherein the portion of the audio output is indicative of an item of the list of items.

27. The non-transitory computer readable storage medium of claim 26 ,

wherein the item is a media item, and

wherein performing the task comprises performing playback, via the speaker, of the media item.

28. The non-transitory computer readable storage medium of claim 26 ,

wherein the item is a location, and

wherein performing the task comprises providing, via the speaker, information associated with the location.

29. The non-transitory computer readable storage medium of claim 23 ,

wherein providing the audio output comprises performing playback of media content, and

wherein performing the task comprises adjusting playback of the media content.

30. The non-transitory computer readable storage medium of claim 29 , wherein adjusting playback of the media content comprises: adjusting a volume of the speaker of the electronic device.

31. The non-transitory computer readable storage medium of claim 29 , wherein adjusting playback of the media content comprises: pausing playback of the media content.

32. The non-transitory computer readable storage medium of claim 23 , wherein adjusting the audio output comprises attenuating the audio output.

33. The non-transitory computer readable storage medium of claim 23 , wherein the audio output is a first audio output, the one or more programs further comprising instructions, which when executed by one or more processors of the electronic device, cause the device to:

before providing the first audio output, provide a second audio output.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Feb 10, 2017
From: SADDLER, HARRY J.; PIERCY, AIMEE T.; WEINBERG, GARRETT L.; BOOKER, SUSAN L.
To: APPLE INC.
Reel/Frame 041226/0426 →
Continuity (2)
Provisional Application 62399232 · Sep 23, 2016
Related Publication 20180090143A1 · Mar 29, 2018
Cited By (34)
US 12,197,712 US 12,197,817 US 12,200,297 US 12,204,932 US 12,211,502 US 12,216,894 US 12,216,963 US 12,219,314 US 12,223,282 US 12,236,952 US 12,254,887 US 12,260,234 US 12,277,954 US 12,293,763 US 12,301,635 US 12,327,558 US 12,327,559 US 12,333,404 US 12,334,060 US 12,340,809 US 12,361,943 US 12,367,879 US 12,380,876 US 12,386,434 US 12,386,491 US 12,431,128 US 12,469,033 US 12,475,883 US 12,477,470 US 12,556,890 US 12,608,171 US 12,613,730 US 12,619,452 US 12,651,020