IP Library Granted Patent US 10,643,615
Granted Patent B2
US 10,643,615 · App. 16/447,536 · Granted May 5, 2020

Voice function control method and apparatus

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,643,615
App. No.
16/447,536
Granted
May 5, 2020
Kind
B2
Abstract

A first recognition result of an input voice is generated, where the input voice is input by a user of a terminal, and the first recognition result is generated by a voice assistant of the terminal. An application of the terminal is determined based on the first recognition result, where the application provides a service, and the application is different from the voice assistant. The input voice is passed to the application, where the application performs voice recognition on the input voice to generate a second recognition result. The service is provided to the user based on the second recognition result.

Claims (67)

1. A computer-implemented method, comprising:

generating a first recognition result of an input voice, wherein the input voice is input by a user of a terminal, the first recognition result is generated by a voice assistant of the terminal, a mapping relationship between services and applications is maintained in the terminal, and the mapping relationship between services and applications includes a mapping relationship between function keywords and applications;

determining an application of the terminal based on the first recognition result, wherein the application provides a service, the application is different from the voice assistant, and determining the application of the terminal based on the first recognition result comprises:

extracting a function keyword from the first recognition result; and

determining the application based on the function keyword and the mapping relationship between function keywords and applications;

passing the input voice to the application, wherein the application performs voice recognition on the input voice to generate a second recognition result; and

providing the service to the user based on the second recognition result.

2. The computer-implemented method of claim 1 , wherein a mapping relationship between services and applications is maintained in the terminal, and determining the application of the terminal based on the first recognition result comprises:

determining the service that the user wants to use based on the first recognition result; and

determining the application based on the service and the mapping relationship between services and applications.

3. The computer-implemented method of claim 2 , wherein at least one of:

the mapping relationship between services and applications is set by the user;

in the mapping relationship between services and applications, for each particular service, a corresponding application is an application that is most frequently used by the user for the particular service; or

the mapping relationship between services and applications is submitted by a particular application.

4. The computer-implemented method of claim 1 , wherein determining the application of the terminal based on the first recognition result comprises:

determining the service that the user wants to use based on the first recognition result;

displaying a plurality of applications of the terminal, wherein each application of the plurality of applications provides the service and supports voice input; and

determining the application based on a user selection.

5. The computer-implemented method of claim 1 , wherein determining the application of the terminal based on the first recognition result comprises:

extracting an application name from the first recognition result; and

determining the application based on the application name.

6. The computer-implemented method of claim 1 , wherein passing the input voice to the application comprises passing the first recognition result and the input voice to the application, and wherein the application performs voice recognition on the first recognition result and the input voice to generate the second recognition result.

7. A non-transitory, computer-readable medium storing one or more instructions executable by a computer system to perform operations comprising:

generating a first recognition result of an input voice, wherein the input voice is input by a user of a terminal, the first recognition result is generated by a voice assistant of the terminal, a mapping relationship between services and applications is maintained in the terminal, and the mapping relationship between services and applications includes a mapping relationship between function keywords and applications;

determining an application of the terminal based on the first recognition result, wherein the application provides a service, the application is different from the voice assistant, and determining the application of the terminal based on the first recognition result comprises:

extracting a function keyword from the first recognition result; and

determining the application based on the function keyword and the mapping relationship between function keywords and applications;

passing the input voice to the application, wherein the application performs voice recognition on the input voice to generate a second recognition result; and

providing the service to the user based on the second recognition result.

8. The non-transitory, computer-readable medium of claim 7 , wherein a mapping relationship between services and applications is maintained in the terminal, and determining the application of the terminal based on the first recognition result comprises:

determining the service that the user wants to use based on the first recognition result; and

determining the application based on the service and the mapping relationship between services and applications.

9. The non-transitory, computer-readable medium of claim 8 , wherein at least one of:

the mapping relationship between services and applications is set by the user;

in the mapping relationship between services and applications, for each particular service, a corresponding application is an application that is most frequently used by the user for the particular service; or

the mapping relationship between services and applications is submitted by a particular application.

10. The non-transitory, computer-readable medium of claim 7 , wherein determining the application of the terminal based on the first recognition result comprises:

determining the service that the user wants to use based on the first recognition result;

displaying a plurality of applications of the terminal, wherein each application of the plurality of applications provides the service and supports voice input; and

determining the application based on a user selection.

11. The non-transitory, computer-readable medium of claim 7 , wherein determining the application of the terminal based on the first recognition result comprises:

extracting an application name from the first recognition result; and

determining the application based on the application name.

12. The non-transitory, computer-readable medium of claim 7 , wherein passing the input voice to the application comprises passing the first recognition result and the input voice to the application, and wherein the application performs voice recognition on the first recognition result and the input voice to generate the second recognition result.

13. A computer-implemented system, comprising:

one or more computers; and

one or more computer memory devices interoperably coupled with the one or more computers and having tangible, non-transitory, machine-readable media storing one or more instructions that, when executed by the one or more computers, perform one or more operations comprising:

generating a first recognition result of an input voice, wherein the input voice is input by a user of a terminal, the first recognition result is generated by a voice assistant of the terminal, a mapping relationship between services and applications is maintained in the terminal, and the mapping relationship between services and applications includes a mapping relationship between function keywords and applications;

determining an application of the terminal based on the first recognition result, wherein the application provides a service, the application is different from the voice assistant, and determining the application of the terminal based on the first recognition result comprises:

extracting a function keyword from the first recognition result; and

determining the application based on the function keyword and the mapping relationship between function keywords and applications;

passing the input voice to the application, wherein the application performs voice recognition on the input voice to generate a second recognition result; and

providing the service to the user based on the second recognition result.

14. The computer-implemented system of claim 13 , wherein a mapping relationship between services and applications is maintained in the terminal, and determining the application of the terminal based on the first recognition result comprises:

determining the service that the user wants to use based on the first recognition result; and

determining the application based on the service and the mapping relationship between services and applications.

15. The computer-implemented system of claim 14 , wherein at least one of:

the mapping relationship between services and applications is set by the user;

in the mapping relationship between services and applications, for each particular service, a corresponding application is an application that is most frequently used by the user for the particular service; or

the mapping relationship between services and applications is submitted by a particular application.

16. The computer-implemented system of claim 13 , wherein determining the application of the terminal based on the first recognition result comprises:

determining the service that the user wants to use based on the first recognition result;

displaying a plurality of applications of the terminal, wherein each application of the plurality of applications provides the service and supports voice input; and

determining the application based on a user selection.

17. The computer-implemented system of claim 13 , wherein determining the application of the terminal based on the first recognition result comprises:

extracting an application name from the first recognition result; and

determining the application based on the application name.

Assignments (3)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Sep 10, 2020
From: ADVANTAGEOUS NEW TECHNOLOGIES CO., LTD.
To: ADVANCED NEW TECHNOLOGIES CO., LTD.
Reel/Frame 053754/0625 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 31, 2020
From: ALIBABA GROUP HOLDING LIMITED
To: ADVANTAGEOUS NEW TECHNOLOGIES CO., LTD.
Reel/Frame 053743/0464 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 12, 2019
From: PAN, SHEN
To: ALIBABA GROUP HOLDING LIMITED
Reel/Frame 050982/0641 →