IP Library Granted Patent US 11,348,583
Granted Patent B2
US 11,348,583 · App. 16/907,269 · Granted May 31, 2022

Data processing method and apparatus for intelligent device, and storage medium

Inventors: Yang Liu (Beijing, CN); Xi Xi (Beijing, CN); Long Quan (Beijing, CN)
Assignees: BAIDU ONLINE NETWORK TECHNOLOGY (BEIJING) CO., LTD.; SHANGHAI XIAODU TECHNOLOGY CO. LTD.
G10L15/22G10L15/08G10L15/30G10L21/0232G10L2015/088G10L2015/223H04W88/02
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,348,583
App. No.
16/907,269
Granted
May 31, 2022
Kind
B2
Abstract

The present disclosure discloses a data processing method and apparatus for an intelligent device, and a storage medium, which relates to a field of artificial intelligence technologies. The method includes: extracting key voice information from collected user voice information; in a non-wireless fidelity (WiFi) network environment, transmitting the key voice information to a mobile terminal, so that the mobile terminal transmits the key voice information to a server, and receives a processing result fed back by the server after the server processes the key voice information; and obtaining the processing result from the mobile terminal to display the processing result.

Claims (58)

1. A data processing method for an intelligent device, comprising:

extracting key voice information from collected user voice information;

in a non-WiFi network environment, transmitting the key voice information to a mobile terminal, so that the mobile terminal transmits the key voice information to a server and receives a processing result fed back by the server after the server processes the key voice information; and

obtaining the processing result from the mobile terminal to display the processing result,

wherein transmitting the key voice information to the mobile terminal comprises:

controlling, using a first channel of a local Bluetooth module, a second channel of the local Bluetooth module to switch from an OFF state to an ON state; and

transmitting the key voice information to the mobile terminal via the second channel;

wherein power consumption of the first channel is lower than that of the second channel, and the first channel is in a normally ON state after the local Bluetooth module is enabled.

2. The method of claim 1 , wherein extracting the key voice information from the collected user voice information comprises:

in response to recognizing a wake-up word from the collected user voice information, extracting the key voice information from the collected user voice information.

3. The method of claim 1 , wherein extracting the key voice information from the collected user voice information comprises:

intercepting voice information containing a wake-up word from the collected user voice information as the key voice information.

4. The method of claim 2 , wherein extracting the key voice information from the collected user voice information comprises:

intercepting voice information containing the wake-up word from the collected user voice information as the key voice information.

5. The method of claim 1 , wherein extracting the key voice information from the collected user voice information comprises:

performing at least one of noise reduction processing and speech-to-text conversion processing on the collected user voice information to obtain the key voice information.

6. The method of claim 1 , after extracting the key voice information from the collected user voice information, further comprising:

in a WiFi network environment, transmitting the key voice information to the server based on the WiFi network, so that the server feeds back the processing result after processing the key voice information; and

obtaining the processing result from the server based on the WiFi network to display the processing result.

7. A data processing apparatus for an intelligent device, comprising:

one or more processors;

a memory storing instructions executable by the one or more processors;

wherein the one or more processors are configured to:

extract key voice information from collected user voice information;

in a non-WiFi network environment, transmit the key voice information to a mobile terminal, so that the mobile terminal transmits the key voice information to a server and receives a processing result fed back by the server after the server processes the key voice information; and

obtain the processing result from the mobile terminal to display the processing result,

wherein the one or more processors are configured to:

control, using a first channel of a local Bluetooth module, a second channel of the local Bluetooth module to switch from an OFF state to an ON state; and

transmit the key voice information to the mobile terminal via the second channel;

wherein power consumption of the first channel is lower than that of the second channel, and the first channel is in a normally ON state after the local Bluetooth module is enabled.

8. The apparatus of claim 7 , wherein the one or more processors are configured to:

in response to recognizing a wake-up word from the collected user voice information, extract the key voice information from the collected user voice information.

9. The apparatus of claim 7 , wherein the one or more processors are configured to:

intercept voice information containing a wake-up word from the collected user voice information as the key voice information.

10. The apparatus of claim 8 , wherein the one or more processors are configured to:

intercept voice information containing the wake-up word from the collected user voice information as the key voice information.

11. The apparatus of claim 7 , wherein the one or more processors are configured to:

perform at least one of noise reduction processing and speech-to-text conversion processing on the collected user voice information to obtain the key voice information.

12. The apparatus of claim 7 , after the key voice information is extracted from the collected user voice information, the one or more processors are configured to:

in a WiFi network environment, transmit the key voice information to the server based on the WiFi network, so that the server feeds back the processing result after processing the key voice information; and

obtain the processing result from the server based on the WiFi network to display the processing result.

13. A non-transitory computer-readable storage medium having computer instructions stored thereon, wherein when the computer instructions are executed on a computer, the computer is caused to execute a data processing method for an intelligent device, and the method comprises:

extracting key voice information from collected user voice information;

in a non-WiFi network environment, transmitting the key voice information to a mobile terminal, so that the mobile terminal transmits the key voice information to a server and receives a processing result fed back by the server after the server processes the key voice information; and

obtaining the processing result from the mobile terminal to display the processing result,

wherein transmitting the key voice information to the mobile terminal comprises:

controlling, using a first channel of a local Bluetooth module, a second channel of the local Bluetooth module to switch from an OFF state to an ON state; and

transmitting the key voice information to the mobile terminal via the second channel;

wherein power consumption of the first channel is lower than that of the second channel, and the first channel is in a normally ON state after the local Bluetooth module is enabled.

14. The non-transitory computer-readable storage medium of claim 13 , wherein extracting the key voice information from the collected user voice information comprises:

in response to recognizing a wake-up word from the collected user voice information, extracting the key voice information from the collected user voice information.

15. The non-transitory computer-readable storage medium of claim 13 , wherein extracting the key voice information from the collected user voice information comprises:

intercepting voice information containing a wake-up word from the collected user voice information as the key voice information.

16. The non-transitory computer-readable storage medium of claim 13 , wherein extracting the key voice information from the collected user voice information comprises:

performing at least one of noise reduction processing and speech-to-text conversion processing on the collected user voice information to obtain the key voice information.

17. The non-transitory computer-readable storage medium of claim 13 , after extracting the key voice information from the collected user voice information, further comprising:

in a WiFi network environment, transmitting the key voice information to the server based on the WiFi network, so that the server feeds back the processing result after processing the key voice information; and

obtaining the processing result from the server based on the WiFi network to display the processing result.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded May 30, 2021
From: BAIDU ONLINE NETWORK TECHNOLOGY (BEIJING) CO., LTD.
To: BAIDU ONLINE NETWORK TECHNOLOGY (BEIJING) CO., LTD.; SHANGHAI XIAODU TECHNOLOGY CO. LTD.
Reel/Frame 056811/0772 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jun 21, 2020
From: LIU, YANG; XI, XI; QUAN, LONG
To: BAIDU ONLINE NETWORK TECHNOLOGY (BEIJING) CO., LTD.
Reel/Frame 052995/0535 →
Priority Claims (1)
CN 201910935399.5 · Sep 29, 2019 · national
Continuity (1)
Related Publication 20210097994A1 · Apr 1, 2021