IP Library Granted Patent US 11,587,550
Granted Patent B2
US 11,587,550 · App. 17/117,786 · Granted Feb 21, 2023

Method and apparatus for outputting information

Inventors: Nengjun Ouyang (Beijing, CN); Ke Zhao (Beijing, CN); Rong Liu (Beijing, CN)
Assignee: Apollo Intelligent Connectivity (Beijing) Technology Co., Ltd.
G10L15/05G10L15/063G10L17/04G10L2015/088
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,587,550
App. No.
17/117,786
Granted
Feb 21, 2023
Kind
B2
Abstract

A method and an apparatus for outputting information are provided. The method includes acquiring voice information received within a preset time period before a device is awakened, where the device is provided with a wake-up model for outputting preset response information when a preset wake-up word is received; performing speech recognition on the voice information to obtain a recognition result; extracting feature information of the voice information in response to determining that the recognition result does not include a preset wake-up word; generating a counterexample training sample according to the feature information; and training the wake-up model using a counter-example training sample, and outputting the trained wake-up model.

Claims (51)

1. A method for outputting information, the method comprising:

acquiring voice information received within a preset time period before a device is awakened, wherein the device is provided with a wake-up model for outputting preset response information when a preset wake-up word is received;

performing speech recognition on the voice information to obtain a recognition result;

extracting feature information of the voice information in response to determining that the recognition result does not include the preset wake-up word;

generating a counterexample training sample based on the feature information; and

training the wake-up model using the counterexample training sample, and outputting a trained wake-up model, wherein the training is performed by using a new wake-up word included in the voice information in the counterexample training sample as an input, and using null information as an expected output.

2. The method according to claim 1 , wherein the training the wake-up model using the counterexample training sample and outputting the trained wake-up model comprises:

inputting the feature information of the voice information comprising the preset wake-up word into the trained wake-up model, and determining whether the preset response information is output; and

outputting the trained wake-up model in response to outputting the preset response information.

3. The method according to claim 2 , wherein the method further comprises:

in response to not outputting the preset response information, outputting alarm information.

4. The method according to claim 1 , wherein the training the wake-up model using the counterexample training sample, and outputting the trained wake-up model comprises:

determining a number of the counterexample training samples; and

training the wake-up model with the counterexample training samples in response to the number of the counterexample training samples being greater than or equal to a preset number threshold.

5. The method according to claim 4 , wherein the acquiring voice information received within a predetermined time period before a device is awakened comprises:

in response to the device being awakened by the preset wake-up word, acquiring the voice information received within the preset time period before the device is awakened.

6. A server comprising:

one or more processors;

a storage apparatus storing one or more programs,

wherein the one or more programs, when executed by the one or more processors, cause the one or more processors to perform operations, the operations comprising:

acquiring voice information received within a preset time period before a device is awakened, wherein the device is provided with a wake-up model for outputting preset response information when a preset wake-up word is received;

performing speech recognition on the voice information to obtain a recognition result;

extracting feature information of the voice information in response to determining that the recognition result does not include the preset wake-up word;

generating a counterexample training sample based on the feature information; and

training the wake-up model using the counterexample training sample, and outputting a trained wake-up model, wherein the training is performed by using a new wake-up word included in the voice information in the counterexample training sample as an input, and using null information as an expected output.

7. The server according to claim 6 , wherein the training the wake-up model using the counterexample training sample, and outputting the trained wake-up model comprises:

inputting the feature information of the voice information comprising the preset wake-up word into the trained wake-up model, and determining whether the preset response information is output; and

outputting the trained wake-up model in response to outputting the preset response information.

8. The server according to claim 7 , wherein the operations further comprises:

in response to not outputting the preset response information, outputting alarm information.

9. The server according to claim 6 , wherein the training the wake-up model using the counterexample training sample, and outputting the trained wake-up model comprises:

determining a number of the counterexample training samples; and

training the wake-up model with the counterexample training samples in response to the number of the counterexample training samples being greater than or equal to a preset number threshold.

10. The server according to claim 9 , wherein the acquiring voice information received within a predetermined time period before a device is awakened comprises:

in response to the device being awakened by the preset wake-up word, acquiring the voice information received within the preset time period before the device is awakened.

11. A non-transitory computer readable medium storing a computer program, wherein the computer program, when executed by a processor, causes the processor to perform operations, the operations comprising:

acquiring voice information received within a preset time period before a device is awakened, wherein the device is provided with a wake-up model for outputting preset response information when a preset wake-up word is received;

performing speech recognition on the voice information to obtain a recognition result;

extracting feature information of the voice information in response to determining that the recognition result does not include the preset wake-up word;

generating a counterexample training sample based on the feature information; and

training the wake-up model using the counterexample training sample, and outputting a trained wake-up model, wherein the training is performed by using a new wake-up word included in the voice information in the counterexample training sample as an input, and using null information as an expected output.

12. The non-transitory computer readable medium according to claim 11 , wherein the training the wake-up model using the counterexample training sample and outputting the trained wake-up model comprises:

inputting the feature information of the voice information comprising the preset wake-up word into the trained wake-up model, and determining whether the preset response information is output; and

outputting the trained wake-up model in response to outputting the preset response information.

13. The non-transitory computer readable medium according to claim 12 , wherein the operations further comprises:

in response to not outputting the preset response information, outputting alarm information.

14. The non-transitory computer readable medium according to claim 11 , wherein the training the wake-up model using the counterexample training sample, and outputting the trained wake-up model comprises:

determining a number of the counterexample training samples; and

training the wake-up model with the counterexample training samples in response to the number of the counterexample training samples being greater than or equal to a preset number threshold.

15. The non-transitory computer readable medium according to claim 14 , wherein the acquiring voice information received within a predetermined time period before a device is awakened comprises:

in response to the device being awakened by the preset wake-up word, acquiring the voice information received within the preset time period before the device is awakened.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 13, 2021
From: BEIJING BAIDU NETCOM SCIENCE AND TECHNOLOGY CO., LTD.
To: APOLLO INTELLIGENT CONNECTIVITY (BEIJING) TECHNOLOGY CO., LTD.
Reel/Frame 057789/0357 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 10, 2020
From: OUYANG, NENGJUN; ZHAO, KE; LIU, RONG
To: BEIJING BAIDU NETCOM SCIENCE AND TECHNOLOGY CO., LTD.
Reel/Frame 054607/0102 →
Priority Claims (1)
CN 202010522739.4 · Jun 10, 2020 · national
Continuity (1)
Related Publication 20210390947A1 · Dec 16, 2021