IP Library Granted Patent US 11,783,825
Granted Patent B2
US 11,783,825 · App. 17/178,009 · Granted Oct 10, 2023

Speech recognition method, speech wakeup apparatus, speech recognition apparatus, and terminal

Inventor: Junyang Zhou (Shenzhen, CN)
Assignee: Honor Device Co., Ltd.
G10L15/22G10L15/00G10L15/02G10L2015/223H04M1/725H04M1/72454
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,783,825
App. No.
17/178,009
Granted
Oct 10, 2023
Kind
B2
Abstract

Embodiments of the present invention provide a speech recognition method and a terminal. The method includes: listening, by a speech wakeup apparatus, to speech information in a surrounding environment; when determining that the speech information obtained by listening matches a speech wakeup model, buffering, by the speech wakeup apparatus, speech information, of first preset duration, obtained by listening, and sending a trigger signal for triggering enabling of a speech recognition apparatus, where the trigger signal is used to instruct the speech recognition apparatus to read and recognize the speech information buffered by the speech wakeup apparatus; and recognizing first speech information buffered by the speech wakeup apparatus and the second speech information obtained by listening, to obtain a recognition result.

Claims (59)

1. A speech control method, applied to a terminal that includes a speech wakeup apparatus and a speech recognition apparatus, the method comprising:

listening, by the speech wakeup apparatus, for first speech information in a surrounding environment, wherein the first speech information comprises wakeup information and a first portion of a command word, wherein the wakeup information is used to enable the speech recognition apparatus;

generating, by the speech wakeup apparatus, a trigger signal based on determining that the wakeup information matches a speech wakeup model;

sending, by the speech wakeup apparatus, the trigger signal to the speech recognition apparatus;

enabling, by the speech wakeup apparatus, the speech recognition apparatus according to the trigger signal;

listening, by the speech recognition apparatus, for second speech information, wherein the second speech information comprises a second portion of the command word; and

obtaining, by the speech recognition apparatus, speech instruction information according to the first speech information and the second speech information, wherein the speech instruction information matches the command word, and the command word comprises the first portion of the command word and the second portion of the command word.

2. The method according to claim 1 , wherein the determining that the wakeup information matches the speech wakeup model comprises:

in a case that the wakeup information matches predetermined wakeup speech information, determining that the wakeup information matches the speech wakeup model.

3. The method according to claim 1 , wherein the determining that the wakeup information matches the speech wakeup model comprises:

in a case that the wakeup information matches predetermined wakeup speech information, extracting a voiceprint feature in the wakeup information, and in a case that the extracted voiceprint feature matches a predetermined voiceprint feature, determining that the wakeup information matches the speech wakeup model.

4. The method according to claim 3 , wherein the voiceprint feature includes one or more of the following features:

a pitch contour, a linear prediction coefficient, a spectral envelope parameter, a harmonic energy ratio, a resonant peak frequency and its bandwidth, a cepstrum, or a Mel-frequency cepstrum coefficient.

5. The method according to claim 1 , wherein the obtaining, by the speech recognition apparatus, the speech instruction information according to the first speech information and the second speech information comprises:

obtaining, by the speech recognition apparatus, a recognition result according to the first speech information and the second speech information, wherein the recognition result comprises command word information; and

obtaining, by the speech recognition apparatus, the speech instruction information that matches the recognition result by matching between the obtained recognition result and pre-stored speech instruction information.

6. The method according to claim 1 ,

wherein the wakeup information is listened in a first period by the speech wakeup apparatus, and the first portion of the command word is listened in a second period by the speech wakeup apparatus; and

wherein the second speech information is listened in a third period by the speech recognition apparatus.

7. The method according to claim 1 , wherein the listening, by the speech wakeup apparatus, for the first speech information in the surrounding environment comprises:

listening for the first speech information in the surrounding environment in a standby state; or

listening for the first speech information in the surrounding environment in a non-standby state; or

listening for the first speech information in the surrounding environment in a screen-locked state.

8. The method according to claim 1 , further comprising:

controlling, by the speech recognition apparatus, execution of an operation corresponding to the speech instruction information.

9. The method according to claim 1 , wherein the first portion of the command word is listened by the speech wakeup apparatus before the speech recognition apparatus is enabled.

10. A terminal, comprising:

a speech wakeup apparatus; and

a speech recognition apparatus;

wherein the speech wakeup apparatus is configured to:

listen for first speech information in a surrounding environment, wherein the first speech information comprises a wakeup information and a first portion of a command word, wherein the wakeup information is used to enable the speech recognition apparatus;

generate a trigger signal based on determining that the wakeup information matches a speech wakeup model;

send the trigger signal to the speech recognition apparatus; and

enable the speech recognition apparatus according to the trigger signal; and

wherein the speech recognition apparatus is configured to:

listen for second speech information, wherein the second speech information comprises a second portion of the command word; and

obtain speech instruction information according to the first speech information and the second speech information, wherein the speech instruction information matches the command word, and the command word comprises the first portion of the command word and the second portion of the command word.

11. The terminal according to claim 10 , wherein the speech wakeup apparatus is configured to:

in a case that the wakeup information matches predetermined wakeup speech information, extract a voiceprint feature in the wakeup information, and in a case that the extracted voiceprint feature matches a predetermined voiceprint feature, determine that the wakeup information matches a speech wakeup model.

12. The terminal according to claim 11 , wherein the voiceprint feature includes one or more of the following features:

a pitch contour, a linear prediction coefficient, a spectral envelope parameter, a harmonic energy ratio, a resonant peak frequency and its bandwidth, a cepstrum, or a Mel-frequency cepstrum coefficient.

13. The terminal according to claim 10 , wherein the speech recognition apparatus is configured to:

obtain a recognition result according to the first speech information and the second speech information, wherein the recognition result comprises command word information; and

obtain the speech instruction information that matches the recognition result by matching between the obtained recognition result and pre-stored speech instruction information.

14. The terminal according to claim 10 ,

wherein the wakeup information is listened in a first period by the speech wakeup apparatus, and the first portion of the command word is listened in a second period by the speech wakeup apparatus; and

wherein the second speech information is listened in a third period by the speech recognition apparatus.

15. The terminal according to claim 10 , wherein the speech wakeup apparatus is configured to:

listen for the first speech information in the surrounding environment in a standby state; or

listen for the first speech information in the surrounding environment in a non-standby state; or

listen for the first speech information in the surrounding environment in a screen-locked state.

16. The terminal according to claim 10 , wherein the speech recognition apparatus is configured to:

automatically disable when determining that speech information is not received again within a preset duration after enabling the speech recognition apparatus.

17. The terminal according to claim 10 , further comprising:

a processor;

wherein the speech recognition apparatus is also configured to send an execution instruction that matches the speech instruction information to the processor; and

the processor is configured to execute an operation corresponding to the execution instruction.

18. The terminal according to claim 10 , wherein the speech wakeup apparatus is a digital signal processor (DSP), and the speech recognition apparatus is an application processor (AP).

19. The terminal according to claim 10 , wherein the first portion of the command word is listened by the speech wakeup apparatus before the speech recognition apparatus is enabled.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jun 3, 2021
From: ZHOU, JUNYANG
To: HUAWEI TECHNOLOGIES CO., LTD.
Reel/Frame 056425/0491 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 13, 2021
From: HUAWEI TECHNOLOGIES CO., LTD.
To: HONOR DEVICE CO., LTD.
Reel/Frame 055919/0344 →
Continuity (3)
Continuation 15729097 · Oct 10, 2017
Continuation PCTCN2015076342 · Apr 10, 2015
Related Publication 20210287671A1 · Sep 16, 2021