IP Library Granted Patent US 10,089,072
Granted Patent B2
US 10,089,072 · App. 15/268,338 · Granted Oct 2, 2018

Intelligent device arbitration and control

Inventors: Kurt W. Piersol (Cupertino, CA); Ryan M. Orr (Cupertino, CA); Daniel J. Mandel (Tucson, AZ)
Assignee: Apple Inc.
G06F3/167G10L15/22G10L15/30G10L25/84G10L2015/228
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,089,072
App. No.
15/268,338
Granted
Oct 2, 2018
Kind
B2
Abstract

This relates to systems and processes for using a virtual assistant to arbitrate among and/or control electronic devices. In one example process, a first electronic device samples an audio input using a microphone. The first electronic device broadcasts a first set of one or more values based on the sampled audio input. Furthermore, the first electronic device receives a second set of one or more values, which are based on the audio input, from a second electronic device. Based on the first set of one or more values and the second set of one or more values, the first electronic device determines whether to respond to the audio input or forego responding to the audio input.

Claims (98)

1. An electronic device comprising:

a microphone;

one or more processors;

a memory; and

one or more programs, wherein the one or more programs are stored in the memory and configured to be executed by the one or more processors, the one or more programs including instructions for:

sampling, with the microphone at the electronic device, an audio input specifying a task, wherein the electronic device is a first electronic device;

identifying, with the first electronic device, a confidence value indicative of a likelihood that the audio input was provided by a particular user;

broadcasting a first set of one or more values based on the sampled audio input, wherein a first value of the first set of values is based on the confidence value;

receiving a second set of one or more values from a second electronic device, wherein the second set of one or more values is based on the audio input;

determining, with the first electronic device, whether a type of the first electronic device meets a requirement of the task; and

in accordance with a determination that the type of the first electronic device meets the requirement of the task:

determining whether the first electronic device is to respond to the audio input based on the first set of one or more values, the second set of one or more values, and the requirement of the task;

in accordance with a determination that the first electronic device is to respond to the audio input, responding to the audio input; and

in accordance with a determination that the first electronic device is not to respond to the audio input, foregoing responding to the audio input; and

in accordance with a determination that the type of the first electronic device does not meet the requirement of the task, foregoing responding to the audio input with the first electronic device.

2. The electronic device of claim 1 , wherein a value of the first set of values is based on a signal to noise ratio of speech of the audio input sampled with the first electronic device.

3. The electronic device of claim 1 , wherein a value of the first set of values is based on a sound pressure of the audio input sampled with the first electronic device.

4. The electronic device of claim 1 , wherein the one or more programs further include instructions for:

identifying, with the first electronic device, a state of the first electronic device, wherein a value of the first set of values is based on the identified state of the first electronic device.

5. The electronic device of claim 4 , wherein the state of the first electronic device is identified based on a user input received with the first electronic device.

6. The electronic device of claim 1 , wherein at least one value of the first set of one or more values is based on a type of the first electronic device.

7. The electronic device of claim 1 , wherein sampling the audio input comprises determining, with the first electronic device, whether the audio input comprises a spoken trigger and wherein the one or more programs further include instructions for:

in accordance with a determination that the audio input does not comprise the spoken trigger, foregoing broadcasting, with the first electronic device, the first set of one or more values.

8. The electronic device of claim 1 , wherein the one or more programs further include instructions for:

in accordance with the determination that the type of the first electronic device does not meet the requirement, determining, with the first electronic device, whether the second device is to respond to the audio input,

in accordance with a determination that the second device is to respond to the audio input, foregoing responding to the audio input with the first electronic device;

in accordance with a determination that the second device is not to respond to the audio input, providing, with the first electronic device, an output indicative of an error.

9. The electronic device of claim 1 , wherein the one or more programs further include instructions for:

receiving, with the first electronic device, data indicative of the requirement of the task from a server.

10. The electronic device of claim 1 , wherein foregoing responding to the audio input with the first electronic device comprises entering, with the first electronic device, an inactive mode.

11. The electronic device of claim 1 , wherein the first set of one or more values is broadcasted in accordance with a unidirectional broadcast communications protocol.

12. The electronic device of claim 1 , wherein the one or more programs further include instructions for:

in accordance with the determination that the first electronic device is to respond to the audio input, providing, with the first electronic device, a visual output, an auditory output, a haptic output, or a combination thereof.

13. The electronic device of claim 1 , wherein determining whether the first electronic device is to respond to the audio input comprises:

determining, with the first electronic device, whether a value of the first set of one or more values is higher than a corresponding value of the second set of one or more values.

14. A non-transitory computer readable storage medium storing one or more programs, the one or more programs comprising instructions, which when executed by one or more processors of a first electronic device with a microphone, cause the first electronic device to:

sample, with the microphone at the first electronic device, an audio input specifying a task;

identify, with the first electronic device, a confidence value indicative of a likelihood that the audio input was provided by a particular user;

broadcast a first set of one or more values based on the sampled audio input, wherein a first value of the first set of one or more values is based on the confidence value;

receive a second set of one or more values from a second electronic device, wherein the second set of one or more values is based on the audio input;

determine, with the first electronic device, whether a type of the first electronic device meets a requirement of the task; and

in accordance with a determination that the type of the first electronic device meets the requirement of the task:

determine whether the first electronic device is to respond to the audio input based on the first set of one or more values, the second set of one or more values, and the requirement of the task;

in accordance with a determination that the first electronic device is to respond to the audio input, respond to the audio input;

in accordance with a determination that the first electronic device is not to respond to the audio input, forego responding to the audio input; and

in accordance with a determination that the type of the first electronic device does not meet the requirement of the task, forego responding to the audio input with the first electronic device.

15. The non-transitory computer readable storage medium of claim 14 , wherein a value of the first set of values is based on a signal to noise ratio of speech of the audio input sampled with the first electronic device.

16. The non-transitory computer readable storage medium of claim 14 , wherein a value of the first set of values is based on a sound pressure of the audio input sampled with the first electronic device.

17. The non-transitory computer readable storage medium of claim 14 , wherein the instructions, which when executed by the electronic device, further cause the electronic device to:

identify, with the first electronic device, a state of the first electronic device, wherein a value of the first set of values is based on the identified state of the first electronic device.

18. The non-transitory computer readable storage medium of claim 17 , wherein the state of the first electronic device is identified based on a user input received with the first electronic device.

19. The non-transitory computer readable storage medium of claim 14 , wherein at least one value of the first set of one or more values is based on a type of the first electronic device.

20. The non-transitory computer readable storage medium of claim 14 , wherein sampling the audio input comprises determining, with the first electronic device, whether the audio input comprises a spoken trigger and wherein the instructions, which when executed by the electronic device, further case the electronic device to:

in accordance with a determination that the audio input does not comprise the spoken trigger, forego broadcasting, with the first electronic device, the first set of one or more values.

21. The non-transitory computer readable storage medium of claim 14 , wherein the instructions, which when executed by the electronic device, further cause the electronic device to:

in accordance with the determination that the type of the first electronic device does not meet the requirement, determine, with the first electronic device, whether the second device is to respond to the audio input;

in accordance with a determination that the second device is to respond to the audio input, forego responding to the audio input with the first electronic device;

in accordance with a determination that the second device is not to respond to the audio input, provide, with the first electronic device, an output indicative of an error.

22. The non-transitory computer readable storage medium of claim 14 , wherein the instructions, which when executed by the electronic device, further cause the electronic device to:

receive, with the first electronic device, data indicative of the requirement of the task from a server.

23. The non-transitory computer readable storage medium of claim 14 , wherein foregoing responding to the audio input with the first electronic device comprises entering, with the first electronic device, an inactive mode.

24. The non-transitory computer readable storage medium of claim 14 , wherein the first set of one or more values is broadcasted in accordance with a unidirectional broadcast communications protocol.

25. The non-transitory computer readable storage medium of claim 14 , wherein the instructions, which when executed by the electronic device, further cause the electronic device to:

in accordance with the determination that the first electronic device is to respond to the audio input, provide, with the first electronic device, a visual output, an auditory output, a haptic output, or a combination thereof.

26. The non-transitory computer readable storage medium of claim 14 , wherein determining whether the first electronic device is to respond to the audio input comprises:

determining, with the first electronic device, whether a value of the first set of one or more values is higher than a corresponding value of the second set of one or more values.

27. A method comprising:

at a first electronic device having a microphone:

sampling, with the microphone at the first electronic device, an audio input specifying a task;

identifying, with the first electronic device, a confidence value indicative of a likelihood that the audio input was provided by a particular user;

broadcasting, with the first electronic device, a first set of one or more values based on the sampled audio input, wherein a first value of the first set of values is based on the confidence value;

receiving, with the first electronic device, a second set of one or more values from a second electronic device, wherein the second set of one or more values is based on the audio input;

determining, with the first electronic device, whether a type of the first electronic device meets a requirement of the task; and

in accordance with a determination that the type of the first electronic device meets the requirement of the task;

determining, with the first electronic device, whether the first electronic device is to respond to the audio input based on the first set of one or more values, the second set of one or more values, and the requirement of the task;

in accordance with a determination that the first electronic device is to respond to the audio input, responding to the audio input;

in accordance with a determination that the first electronic device is not to respond to the audio input, foregoing responding to the audio input; and

in accordance with a determination that the type of the first electronic device does not meet the requirement of the task, foregoing responding to the audio input with the first electronic device.

28. The method of claim 27 , wherein a value of the first set of values is based on a signal to noise ratio of speech of the audio input sampled with the first electronic device.

29. The method of claim 27 , wherein a value of the first set of values is based on a sound pressure of the audio input sampled with the first electronic device.

30. The method of claim 27 , further comprising:

identifying, with the first electronic device, a state of the first electronic device, wherein a value of the first set of values is based on the identified state of the first electronic device.

31. The method of claim 30 , wherein the state of the first electronic device is identified based on a user input received with the first electronic device.

32. The method of claim 27 , wherein at least one value of the first set of one or more values is based on a type of the first electronic device.

33. The method of claim 27 , wherein sampling the audio input comprises determining, with the first electronic device, whether the audio input comprises a spoken trigger, the method further comprising:

in accordance with a determination that the audio input does not comprise the spoken trigger, foregoing broadcasting, with the first electronic device, the first set of one or more values.

34. The method of claim 27 , further comprising:

in accordance with the determination that the type of the first electronic device does not meet the requirement, determining, with the first electronic device, whether the second. device is to respond to the audio input,

in accordance with a determination that the second device is to respond to the audio input, foregoing responding to the audio input with the first electronic device;

in accordance with a determination that the second device is not to respond to the audio input, providing, with the first electronic device, an output indicative of an error.

35. The method of claim 27 , further comprising:

receiving, with the first electronic device, data indicative of the requirement of the task from a server.

36. The method of claim 27 , wherein foregoing responding to the audio input with the first electronic device comprises entering, with the first electronic device, an inactive mode.

37. The method of claim 27 , wherein the first set of one or more values is broadcasted in accordance with a unidirectional broadcast communications protocol.

38. The method of claim 27 , further comprising:

in accordance with the determination that the first electronic device is to respond to the audio input, providing, with the first electronic device, a visual output, an auditory output, a haptic output, or a combination thereof.

39. The method of claim 27 , wherein determining whether the first electronic device is to respond to the audio input comprises:

determining, with the first electronic device, whether a value of the first set of one or more values is higher than a corresponding value of the second set of one or more values.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Sep 29, 2016
From: PIERSOL, KURT W.; ORR, RYAN M.; MANDEL, DANIEL J.
To: APPLE INC.
Reel/Frame 039898/0148 →
Continuity (2)
Provisional Application 62348896 · Jun 11, 2016
Related Publication 20170357478A1 · Dec 14, 2017
Cited By (48)
US 12,197,712 US 12,197,817 US 12,200,297 US 12,204,932 US 12,211,502 US 12,216,894 US 12,219,314 US 12,223,282 US 12,230,257 US 12,236,164 US 12,236,952 US 12,254,047 US 12,254,887 US 12,260,234 US 12,271,658 US 12,277,954 US 12,293,203 US 12,293,763 US 12,299,755 US 12,301,635 US 12,333,404 US 12,340,025 US 12,361,943 US 12,367,879 US 12,379,894 US 12,380,281 US 12,380,876 US 12,386,434 US 12,386,491 US 12,431,128 US 12,470,503 US 12,475,883 US 12,477,470 US 12,518,323 US 12,524,771 US 12,556,890 US 12,567,415 US 12,567,435 US 12,574,627 US 12,585,966 US 12,608,171 US 12,613,621 US 12,613,730 US 12,619,452 US 12,620,179 US 12,640,151 US 12,670,639 US 12,675,839