IP Library Granted Patent US 11,651,769
Granted Patent B2
US 11,651,769 · App. 17/018,234 · Granted May 16, 2023

Electronic device and operating method thereof

Inventor: Chanhee Choi (Suwon-si, KR)
Assignee: SAMSUNG ELECTRONICS CO., LTD.
G10L15/22G10L25/78G10L2015/223
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,651,769
App. No.
17/018,234
Granted
May 16, 2023
Kind
B2
Abstract

An electronic device includes a memory storing one or more instructions; and a processor configured to execute the one or more instructions stored in the memory to receive audio data corresponding to a user's utterance, to identify the user's utterance characteristics based on the received audio data, to determine a parameter for performing voice activity detection, by using the identified user's utterance characteristics, and to perform voice activity detection on the received audio data with respect to the user's utterance by using the determined parameter.

Claims (25)

1. An electronic device comprising:

a memory storing one or more instructions; and

a processor configured to execute the one or more instructions stored in the memory:

to receive audio data corresponding to a user utterance, the audio data including an utterance section comprising a trigger word that indicates a start of the utterance;

to identify whether the user utterance corresponds to the trigger word;

to analyze the identified user utterance corresponding to the trigger word to identify user utterance characteristics comprising at least one of an utterance rate, utterance energy level, and pitch of uttered sound;

to change a parameter for performing a function corresponding to voice detection by using the identified user utterance characteristics; and

to perform the function corresponding to voice detection for user utterance input after the trigger word is received by using the changed parameter.

2. The electronic device of claim 1 , wherein the processor is further configured to execute the one or more instructions stored in the memory to compare the identified user utterance characteristics with reference utterance characteristics with respect to the trigger word, and change the parameter based on a result of the comparison.

3. The electronic device of claim 1 , wherein the processor is further configured to execute the one or more instructions stored in the memory to detect a start point of the user utterance on the received audio data and an end point of the user utterance in the received audio data while the user is uttering to distinguish between the utterance section of the received audio data and a non-utterance section of the received audio data.

4. The electronic device of claim 1 , wherein the parameter for performing the function corresponding to voice detection comprises at least one of an energy threshold for identifying, in the received audio data, the utterance section, a hangover time, or automatic end time while the user is uttering.

5. The electronic device of claim 1 , further comprising a communicator configured to receive the audio data corresponding to the user utterance.

6. The electronic device of claim 1 , further comprising a microphone configured to receives the user utterance and to convert the received user utterance through the microphone into the audio data.

7. An operating method of an electronic device, the operating method comprising:

receiving audio data corresponding to a user utterance, the audio data comprising an utterance section comprising a trigger word that indicates a start of the utterance;

identifying whether the user utterance corresponds to the trigger word;

analyzing the identified user utterance corresponding to the trigger word to identify user utterance characteristics comprising at least one of an utterance rate, utterance energy level, and pitch of uttered sound;

changing a parameter for performing a function corresponding to voice detection by using the identified user utterance characteristics; and

performing the function corresponding to voice detection for the user utterance input after the trigger word is received by using the changed parameter.

8. The operating method of claim 7 , wherein the identifying of the parameter for performing the function corresponding to voice detection comprises comparing the identified user utterance characteristics with reference voice characteristics with respect to the trigger word and changing the parameter based on a result of the comparison.

9. The operating method of claim 7 , wherein the performing of the function corresponding to voice detection for the user utterance comprises detecting a start point of the user utterance and an end point of the user utterance while the user is uttering to distinguish between the utterance section of the received audio data and a non-utterance section of the received audio data.

10. The operating method of claim 7 , wherein the parameter for performing the function corresponding to voice detection comprises at least one of an energy threshold for identifying the utterance section, a hangover time, and an automatic end time while the user is uttering.

11. The operating method of claim 7 , wherein the receiving of the audio data of the user utterance comprises receiving the audio data corresponding to the user utterance through a communicator.

12. The operating method of claim 7 , wherein the receiving of the audio data of the user utterance comprises receiving the user utterance through a microphone and converting the received user utterance through the microphone into the audio data.

13. A non-transitory computer-readable recording media having recorded thereon a program for executing in a computer the operating method of claim 7 .

Priority Claims (1)
KR 10-2019-0113010 · Sep 11, 2019 · national
Continuity (1)
Related Publication 20210074290A1 · Mar 11, 2021