IP Library Granted Patent US 11,437,036
Granted Patent B2
US 11,437,036 · App. 16/886,935 · Granted Sep 6, 2022

Smart speaker wake-up method and device, smart speaker and storage medium

Inventors: Xiangdang Zhang (Beijing, CN); Xing Luo (Beijing, CN); Xiangdong Xue (Beijing, CN); Guohui Zhou (Beijing, CN); Wenjie Liao (Beijing, CN)
Assignees: Baidu Online Network Technology (Beijing) Co., Ltd.; Shanghai Xiaodu Technology Co., Ltd.
G10L15/22G10L15/08G10L2015/088G10L2015/223
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,437,036
App. No.
16/886,935
Granted
Sep 6, 2022
Kind
B2
Abstract

The present disclosure discloses a smart speaker wake-up method, a smart speaker wake-up device, a smart speaker and a storage medium, relates to the technical field of speech recognition. The method of the present disclosure is applied to a wireless network including two or more smart speakers, and a specific implementation thereof is: receiving, speech information including a wake-up word; performing a recognition processing to the speech information to obtain identification information corresponding to the wake-up word; and waking up one smart speaker in the wireless network to enter listening state according to the identification information. The present disclosure may be applied to a scenario where multiple smart speakers coexist, so as to quickly select one smart speaker that is most likely to be wakened, avoiding a chaotic speech interaction caused by multiple smart speakers being wakened simultaneously, improving efficiency and quality of speech interaction and achieving better user experience.

Claims (57)

1. A smart speaker wake-up method, which is applied to a wireless network comprising two or more smart speakers, comprising:

receiving speech information comprising a wake-up word;

performing a recognition processing to the speech information to obtain identification information corresponding to the wake-up word, wherein the identification information comprises time at which the wake-up word is recognized by any one smart speaker in the wireless network or a speech intensity corresponding to the wake-up word is recognized by any one smart speaker in the wireless network; and

waking up one smart speaker in the wireless network to enter listening state according to the time at which the wake-up word is recognized by any one smart speaker in the wireless network and time at which the wake-up word is recognized by other smart speakers in the wireless network or according to the speech intensity corresponding to the wake-up word is recognized by any one smart speaker in the wireless network and a speech intensity corresponding to the wake-up word is recognized by other smart speaker in the wireless network.

2. The method according to claim 1 , wherein the performing a recognition processing to the speech information to obtain identification information corresponding to the wake-up word, comprises:

recognizing the wake-up word from the speech information and recording a first timestamp corresponding to the recognition of the wake-up word.

3. The method according to claim 2 , wherein before the waking up one smart speaker in the wireless network to enter listening state according to the identification information, also comprising:

transmitting the first timestamp to other smart speakers in the wireless network by way of broadcasting; and

receiving a second timestamp transmitted by other smart speakers in the wireless network, wherein the second timestamp is referred to as time corresponding to a recognition of the wake-up word by other smart speakers.

4. The method according to claim 3 , wherein the waking up one smart speaker in the wireless network to enter listening state according to the identification information, comprises:

comparing the time corresponding to the first timestamp with the time corresponding to the second timestamp;

giving up to wake up if the time of at least one second timestamp is earlier than the time of the first timestamp; and

waking up to enter the listening state if none of the time of the second timestamp is earlier than the time of the first timestamp.

5. The method according to claim 1 , wherein the performing a recognition processing to the speech information to obtain identification information corresponding to the wake-up word, comprises:

recognizing the wake-up word from the speech information, and recording a first speech intensity corresponding to the wake-up word.

6. The method according to claim 5 , wherein before the waking up one smart speaker in the wireless network to enter listening state according to the identification information, also comprising:

transmitting the first speech intensity to other smart speakers in the wireless network by way of broadcasting; and

receiving a second speech intensity transmitted by other smart speakers in the wireless network, wherein the second speech intensity is referred to as a speech intensity corresponding to a recognized wake-up word by other smart speakers.

7. The method according to claim 6 , wherein the waking up one smart speaker in the wireless network to enter listening state according to the identification information, comprises:

comparing the first speech intensity with the second speech intensity;

giving up to wake up if at least one second speech intensity is larger than the first speech intensity; and

waking up to enter the listening state if none of the second speech intensity is larger than the first speech intensity.

8. The method according to claim 1 , wherein the smart speaker in the wireless network is located in a preset geographic range, and there exist at least two smart speakers with different account numbers in the wireless network.

9. A smart speaker system, comprising:

at least one processor; and

a memory in communication connection with the at least one processor; wherein,

the memory stores instruction executable by the at least one processor, the instruction is executed by the at least one processor, so as to make the at least one processor for executing the following steps while being applied to a wireless network comprising two or more smart speakers:

receiving speech information comprising a wake-up word;

performing a recognition processing to the speech information to obtain identification information corresponding to the wake-up word, wherein the identification information comprises time at which the wake-up word is recognized by any one smart speaker in the wireless network or a speech intensity corresponding to the wake-up word is recognized by any one smart speaker in the wireless network; and

determining whether to wake up the smart speaker itself to enter listening state according to the time at which the wake-up word is recognized by any one smart speaker in the wireless network and time at which the wake-up word is recognized by other smart speakers in the wireless network or according to the speech intensity corresponding to the wake-up word is recognized by any one smart speaker in the wireless network and a speech intensity corresponding to the wake-up word is recognized by other smart speaker in the wireless network.

10. The smart speaker according to claim 9 , wherein the performing a recognition processing to the speech information to obtain identification information corresponding to the wake-up word, comprises:

recognizing the wake-up word from the speech information and recording a first timestamp corresponding to the recognition of the wake-up word.

11. The smart speaker according to claim 10 , wherein before determining whether to wake up the smart speaker itself to enter listening state according to the identification information, also comprising:

transmitting the first timestamp to other smart speakers in the wireless network by way of broadcasting; and

receiving a second timestamp transmitted by other smart speakers in the wireless network, wherein the second timestamp is referred to as time corresponding to a recognition of the wake-up word by other smart speakers.

12. The smart speaker according to claim 11 , wherein determining whether to wake up the smart speaker itself to enter listening state according to the identification information, comprises:

comparing the time corresponding to the first timestamp with the time corresponding to the second timestamp;

giving up to wake up if the time of at least one second timestamp is earlier than the time of the first timestamp; and

waking up to enter the listening state if none of the time of the second timestamp is earlier than the time of the first timestamp.

13. The smart speaker according to claim 10 , wherein the smart speaker in the wireless network is located in a preset geographic range, and there exist at least two smart speakers with different account numbers in the wireless network.

14. The smart speaker according to claim 9 , wherein the performing a recognition processing to the speech information to obtain identification information corresponding to the wake-up word, comprises:

recognizing the wake-up word from the speech information, and recording a first speech intensity corresponding to the wake-up word.

15. The smart speaker according to claim 14 , wherein before determining whether to wake up the smart speaker itself to enter listening state according to the identification information, also comprising:

transmitting the first speech intensity to other smart speakers in the wireless network by way of broadcasting; and

receiving a second speech intensity transmitted by other smart speakers in the wireless network, wherein the second speech intensity is referred to as a speech intensity corresponding to a recognized wake-up word by other smart speakers.

16. The smart speaker according to claim 15 , wherein determining whether to wake up the smart speaker itself to enter listening state according to the identification information, comprises:

comparing the first speech intensity with the second speech intensity;

giving up to wake up if at least one second speech intensity is larger than the first speech intensity; and

waking up to enter the listening state if none of the second speech intensity is larger than the first speech intensity.

17. The smart speaker according to claim 14 , wherein the smart speaker in the wireless network is located in a preset geographic range, and there exist at least two smart speakers with different account numbers in the wireless network.

18. The smart speaker according to claim 9 , wherein the smart speaker in the wireless network is located in a preset geographic range, and there exist at least two smart speakers with different account numbers in the wireless network.

19. A non-transitory computer readable storage medium having a computer instruction stored therein, wherein, the computer instruction is configured to make the computer execute the method according to claim 1 .

20. A smart speaker wake-up method, comprising:

receiving speech information; and

waking up, if the speech information comprises a wake-up word, one smart speaker in a wireless network to enter listening state according to identification information corresponding to the wake-up word, wherein the identification information comprises time at which the wake-up word is recognized by any one smart speaker in the wireless network or a speech intensity corresponding to the wake-up word is recognized by any one smart speaker in the wireless network;

wherein the waking up one smart speaker in a wireless network to enter listening state according to identification information corresponding to the wake-up word, comprises:

waking up one smart speaker in the wireless network to enter listening state according to the time at which the wake-up word is recognized by any one smart speaker in the wireless network and time at which the wake-up word is recognized by other smart speakers in the wireless network or according to the speech intensity corresponding to the wake-up word is recognized by any one smart speaker in the wireless network and a speech intensity corresponding to the wake-up word is recognized by other smart speaker in the wireless network.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded May 30, 2021
From: BAIDU ONLINE NETWORK TECHNOLOGY (BEIJING) CO., LTD.
To: BAIDU ONLINE NETWORK TECHNOLOGY (BEIJING) CO., LTD.; SHANGHAI XIAODU TECHNOLOGY CO. LTD.
Reel/Frame 056811/0772 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded May 29, 2020
From: ZHANG, XIANGDANG; LUO, XING; XUE, XIANGDONG; ZHOU, GUOHUI; LIAO, WENJIE
To: BAIDU ONLINE NETWORK TECHNOLOGY (BEIJING) CO., LTD.
Reel/Frame 052783/0804 →
Priority Claims (1)
CN 201911128667.9 · Nov 18, 2019 · national
Continuity (1)
Related Publication 20210151048A1 · May 20, 2021
Cited By (1)
US 12,387,716