IP Library Granted Patent US 10,803,861
Granted Patent B2
US 10,803,861 · App. 15/857,008 · Granted Oct 13, 2020

Method and apparatus for identifying information

Inventors: Xiaojian Chen (Beijing, CN); Lifeng Zhao (Beijing, CN); Jun Li (Beijing, CN); Yushu Cao (Beijing, CN)
Assignee: BAIDU ONLINE NETWORK TECHNOLOGY (BEIJING) CO., LTD.
G10L15/22G10L15/08G10L15/30G10L25/51G10L2015/088
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,803,861
App. No.
15/857,008
Granted
Oct 13, 2020
Kind
B2
Abstract

Embodiments of the present disclosure disclose a method and apparatus for identifying information. One embodiment of the method includes: collecting to-be-processed audio in real-time; performing voice recognition on the to-be-processed audio; performing data-processing on the to-be-processed audio, when the audio is recognized as a wake-up word, the wake-up word is used for instructing performing data-processing on the to-be-processed audio. The embodiment can identify keywords from the to-be-processed audio obtained in real-time and then perform data-processing on the to-be-processed audio, which improves completeness in obtaining the to-be-processed audio and accuracy in performing data-processing on the to-be-processed audio.

Claims (42)

1. A method for identifying information, the method comprising:

collecting to-be-processed audio in real-time;

performing voice recognition on the to-be-processed audio based on preset wake-up word audio;

determining recognition accuracy for wake-up word audio recognized during the voice recognition, the determining recognition accuracy for the recognized wake-up word audio comprising:

performing an audio analysis on the recognized wake-up word audio to obtain audio information of the recognized wake-up word audio;

comparing the audio information of the recognized wake-up word audio with audio information of the preset wake-up word audio, to obtain the accuracy of the recognized wake-up word audio relative to the preset wake-up word audio; and

determining the recognized wake-up word audio to be accurate in response to determining that the accuracy is greater than a preset value, and determining the recognized wake-up word audio to be inaccurate in response to determining that the accuracy is less than or equal to the preset value; and

performing data-processing on the to-be-processed audio, when wake-up word audio is recognized from the to-be-processed audio, the recognized wake-up word being used for instructing the performing data-processing on the to-be-processed audio, comprising:

sending the to-be-processed audio to a server for the server to identify whether the recognized wake-up word audio is contained within the to-be-processed audio, in response to determining that the recognized wake-up word audio is inaccurate;

receiving identification information corresponding to the to-be-processed audio sent by the server; and

performing data-processing on the to-be-processed audio when the identification information indicates that the to-be-processed audio contains the recognized wake-up word.

2. The method according to claim 1 , wherein the performing data-processing on the to-be-processed audio further comprises:

performing data-processing on the to-be-processed audio when the recognized wake-up word audio is accurate, so as to obtain response information corresponding to the to-be-processed audio.

3. An apparatus for identifying information, the apparatus comprising:

at least one processor; and

a memory storing instructions, the instructions when executed by the at least one processor, cause the at least one processor to perform operations, the operations comprising:

collecting to-be-processed audio in real-time;

performing voice recognition on the to-be-processed audio based on preset wake-up word audio;

determining recognition accuracy for wake-up word audio recognized during the voice recognition, the determining recognition accuracy for the recognized wake-up word audio comprising:

performing an audio analysis on the recognized wake-up word audio to obtain audio information of the recognized wake-up word audio;

comparing the audio information of the recognized wake-up word audio with audio information of the preset wake-up word audio, to obtain the accuracy of the recognized wake-up word audio relative to the preset wake-up word audio; and

determining the recognized wake-up word audio to be accurate in response to determining that the accuracy is greater than a preset value, and determining the recognized wake-up word audio to be inaccurate in response to determining that the accuracy is less than or equal to the preset value; and

performing data-processing on the to-be-processed audio, when wake-up word audio is recognized from the to-be-processed audio, the recognized wake-up word being used for instructing the performing data-processing on the to-be-processed audio, comprising:

sending the to-be-processed audio to a server for the server to identify whether the recognized wake-up word audio is contained within the to-be-processed audio, in response to determining that the recognized wake-up word audio is inaccurate;

receiving identification information corresponding to the to-be-processed audio sent by the server; and

performing data-processing on the to-be-processed audio when the identification information indicates that the to-be-processed audio contains the recognized wake-up word.

4. The apparatus according to claim 3 , wherein the performing data-processing on the to-be-processed audio further comprises:

performing data-processing on the to-be-processed audio when the recognized wake-up word audio is accurate, so as to obtain response information corresponding to the to-be-processed audio.

5. A terminal device, comprising:

one or more processors;

a memory for storing one or more programs,

one or more processors execute the method according to claim 1 .

6. A non-transitory computer storage medium storing a computer program, the computer program when executed by one or more processors, causes the one or more processors to perform operations, the operations comprising: collecting to-be-processed audio in real-time;

performing voice recognition on the to-be-processed audio based on preset wake-up word audio;

determining recognition accuracy for wake-up word audio recognized during the voice recognition, the determining recognition accuracy for the recognized wake-up word audio comprising:

performing an audio analysis on the recognized wake-up word audio to obtain audio information of the recognized wake-up word audio;

comparing the audio information of the recognized wake-up word audio with audio information of the preset wake-up word audio, to obtain the accuracy of the recognized wake-up word audio relative to the preset wake-up word audio; and

determining the recognized wake-up word audio to be accurate in response to determining that the accuracy is greater than a preset value, and determining the recognized wake-up word audio to be inaccurate in response to determining that the accuracy is less than or equal to the preset value; and

performing data-processing on the to-be-processed audio, when wake-up word audio is recognized from the to-be-processed audio, the recognized wake-up word being used for instructing the performing data-processing on the to-be-processed audio, comprising:

sending the to-be-processed audio to a server for the server to identify whether the recognized wake-up word audio is contained within the to-be-processed audio, in response to determining that the recognized wake-up word audio is inaccurate;

receiving identification information corresponding to the to-be-processed audio sent by the server; and

performing data-processing on the to-be-processed audio when the identification information indicates that the to-be-processed audio contains the recognized wake-up word.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded May 30, 2021
From: BAIDU ONLINE NETWORK TECHNOLOGY (BEIJING) CO., LTD.
To: BAIDU ONLINE NETWORK TECHNOLOGY (BEIJING) CO., LTD.; SHANGHAI XIAODU TECHNOLOGY CO. LTD.
Reel/Frame 056811/0772 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 28, 2017
From: CHEN, XIAOJIAN; ZHAO, LIFENG; LI, JUN; CAO, YUSHU
To: BAIDU ONLINE NETWORK TECHNOLOGY (BEIJING) CO., LTD.
Reel/Frame 044995/0602 →
Priority Claims (1)
CN 2017 1 1128321 · Nov 15, 2017 · national
Continuity (1)
Related Publication 20190147860A1 · May 16, 2019