IP Library › Granted Patent US 11,468,890
Granted Patent B2
US 11,468,890 · App. 16/888,450 · Granted Oct 11, 2022

Methods and user interfaces for voice-based control of electronic devices

Inventors: Jigar Vasant Gada (Sunnyvale, CA); Kevin Bartlett Aitken (Cupertino, CA)
Assignee: Apple Inc.
G10L15/22G06F3/167
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,468,890
App. No.
16/888,450
Filed
May 29, 2020
Granted
Oct 11, 2022
Kind
B2
Art Unit
2675
USPC
704/231
Abstract

The present disclosure generally relates to voice-control for electronic devices. In some embodiments, the method includes, in response to detecting a plurality of utterances, associating the plurality of operations with a first stored operation set and detecting a second set of one or more inputs corresponding to a request to perform the operations associated with the first stored operation set; and performing the plurality of operations associated with the first stored operation set, in the respective order.

Claims (44)

1. A non-transitory computer-readable storage medium storing one or more programs configured to be executed by one or more processors of an electronic device with a display device and a microphone, the one or more programs including instructions for:

displaying, via the display device, a first user interface;

while displaying the first user interface and at a first time, detecting, via the microphone, a first utterance; and

in response to detecting the first utterance:

in accordance with a determination that a set of performance criteria are met, the set of performance criteria including a criterion that is met when the first utterance corresponds to a first operation, performing the first operation; and

in accordance with a determination that the set of performance criteria are not met, displaying, in the first user interface, displaying a suggestion graphical object that includes a first text utterance suggestion corresponding to a second utterance that, when detected via the microphone, causes a second operation to be performed, wherein the first text utterance suggestion is selected for display based on a context of the first user interface at the first time and based on the first utterance, wherein the first text utterance suggestion corresponds to a valid voice command available for the current state of the first user interface.

2. The non-transitory computer-readable storage medium of claim 1 , wherein the first text utterance suggestion is further based on frequency of use of one or more utterances of a set of utterances that satisfy the set of performance criteria.

3. The non-transitory computer-readable storage medium of claim 1 , wherein the suggestion graphical object is displayed along an upper edge of the display.

4. The non-transitory computer-readable storage medium of claim 1 , wherein:

the suggestion graphical object includes a second text utterance suggestion corresponding to a third utterance that, when detected via the microphone, causes a third operation to be performed, wherein the second text utterance suggestion is selected based on the context of the first user interface at the first time and based on the utterance; and

an order of display of the first text utterance suggestion and the second text utterance suggestion is based on based on the context of the first user interface at the first time and based on the utterance.

5. The non-transitory computer-readable storage medium of claim 1 , wherein performing the first operation includes:

displaying a second suggestion graphical object that includes a third text utterance suggestion that is based on the context of the first user interface at the first time.

6. An electronic device, comprising:

a display device;

a microphone;

one or more processors; and

memory storing one or more programs configured to be executed by the one or more processors, the one or more programs including instructions for:

displaying, via the display device, a first user interface;

while displaying the first user interface and at a first time, detecting, via the microphone, a first utterance; and

in response to detecting the first utterance:

in accordance with a determination that a set of performance criteria are met, the set of performance criteria including a criterion that is met when the first utterance corresponds to a first operation, performing the first operation; and

in accordance with a determination that the set of performance criteria are not met, displaying, in the first user interface, displaying a suggestion graphical object that includes a first text utterance suggestion corresponding to a second utterance that, when detected via the microphone, causes a second operation to be performed, wherein the first text utterance suggestion is selected for display based on a context of the first user interface at the first time and based on the first utterance, wherein the first text utterance suggestion corresponds to a valid voice command available for the current state of the first user interface.

7. The electronic device of claim 6 , wherein the first text utterance suggestion is further based on frequency of use of one or more utterances of a set of utterances that satisfy the set of performance criteria.

8. The electronic device of claim 6 , wherein the suggestion graphical object is displayed along an upper edge of the display.

9. The electronic device of claim 6 , wherein:

the suggestion graphical object includes a second text utterance suggestion corresponding to a third utterance that, when detected via the microphone, causes a third operation to be performed, wherein the second text utterance suggestion is selected based on the context of the first user interface at the first time and based on the utterance; and

an order of display of the first text utterance suggestion and the second text utterance suggestion is based on based on the context of the first user interface at the first time and based on the utterance.

10. The electronic device of claim 6 , wherein performing the first operation includes:

displaying a second suggestion graphical object that includes a third text utterance suggestion that is based on the context of the first user interface at the first time.

11. A method comprising:

at an electronic device with a display device and a microphone:

displaying, via the display device, a first user interface;

while displaying the first user interface and at a first time, detecting, via the microphone, a first utterance; and

in response to detecting the first utterance:

in accordance with a determination that a set of performance criteria are met, the set of performance criteria including a criterion that is met when the first utterance corresponds to a first operation, performing the first operation; and

in accordance with a determination that the set of performance criteria are not met, displaying, in the first user interface, displaying a suggestion graphical object that includes a first text utterance suggestion corresponding to a second utterance that, when detected via the microphone, causes a second operation to be performed, wherein the first text utterance suggestion is selected for display based on a context of the first user interface at the first time and based on the first utterance, wherein the first text utterance suggestion corresponds to a valid voice command available for the current state of the first user interface.

12. The method of claim 11 , wherein the first text utterance suggestion is further based on frequency of use of one or more utterances of a set of utterances that satisfy the set of performance criteria.

13. The method of claim 11 , wherein the suggestion graphical object is displayed along an upper edge of the display.

14. The method of claim 11 , wherein:

the suggestion graphical object includes a second text utterance suggestion corresponding to a third utterance that, when detected via the microphone, causes a third operation to be performed, wherein the second text utterance suggestion is selected based on the context of the first user interface at the first time and based on the utterance; and

an order of display of the first text utterance suggestion and the second text utterance suggestion is based on based on the context of the first user interface at the first time and based on the utterance.

15. The method of claim 11 , wherein performing the first operation includes:

displaying a second suggestion graphical object that includes a third text utterance suggestion that is based on the context of the first user interface at the first time.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jun 7, 2021
From: GADA, JIGAR VASANT; AITKEN, KEVIN BARTLETT
To: APPLE INC.
Reel/Frame 056459/0961 →
Continuity (2)
Provisional Application 62856044 · Jun 1, 2019
Related Publication 20200380985A1 · Dec 3, 2020
Cited By (1)
US 12,676,151