IP Library › Granted Patent US 11,094,323
Granted Patent B2
US 11,094,323 · App. 16/333,369 · Granted Aug 17, 2021

Electronic device and method for processing audio signal by electronic device

Inventors: Ki-hoon Shin (Yongin-si, KR); Myung-suk Song (Seoul, KR); Jong-uk Yoo (Suwon-si, KR)
Assignee: SAMSUNG ELECTRONICS CO., LTD.
G10L15/22G10L15/28G10L15/285G10L21/028G10L21/0216G10L21/0272G10L25/87G10L2015/223G10L2021/02166
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,094,323
App. No.
16/333,369
Granted
Aug 17, 2021
Kind
B2
Abstract

An electronic device is disclosed. The electronic device comprises: multiple microphones for receiving audio signals generated by multiple sound sources; a communication unit for communicating with a voice recognition server; and a processor for determining the direction in which each of the multiple sound sources is located with reference to the electronic device, on the basis of the multiple audio signals received through the multiple microphones, determining at least one target sound source among the multiple sound sources on the basis of the duration of the determined direction of each of the sound sources, and controlling the communication unit such that the communication unit transmits, to the voice recognition server, an audio signal of a target sound source from which a predetermined voice is generated among the at least one target sound source.

Claims (44)

1. An electronic device, comprising:

a plurality of microphones;

a communication unit; and

a processor configured to:

based on audio signals provided by a plurality of sound sources being received through the plurality of microphones, identify a direction in which each of the plurality of sound sources is located with respect to the electronic device,

identify at least one target sound source among the plurality of sound sources based on a duration of the direction of each of the plurality of sound sources,

control the communication unit to transmit, to a voice recognition server, an audio signal of a target sound source from which a predetermined voice is provided among the at least one target sound source,

wherein the processor is further configured to identify at least one sound source of which the duration is less than a predetermined time among the plurality of sound sources as the at least one target sound source.

2. The electronic device as claimed in claim 1 , wherein the processor is further configured to:

separate the audio signals received from each of the plurality of sound sources; and

perform voice recognition on the separated audio signals and identify a target sound source from which the predetermined voice is provided.

3. The electronic device as claimed in claim 2 , wherein the processor is further configured to control the communication unit to transmit, to the voice recognition server, only an audio signal of the target sound source identified to provide the predetermined voice among the audio signals received by the each of the plurality of microphones after the target sound source from which the predetermined voice is provided is identified.

4. The electronic device as claimed in claim 1 , wherein the processor is further configured to:

based on only one target sound source being present,

separate only an audio signal of the target sound source from audio signals received from each of the plurality of microphones; and

perform voice recognition on the audio signal of the separated target sound source and identify whether the predetermined voice is provided.

5. The electronic device as claimed in claim 1 , wherein the processor is further configured to:

identify a target sound source generating a voice among the at least one target sound source; and

perform voice recognition on only an audio signal of the target sound source to provide the voice, and identify the target sound source from which the predetermined voice is provided.

6. The electronic device as claimed in claim 1 , wherein the processor is further configured to identify a direction in which each of the plurality of sound sources is present, and

wherein a number of directions of the each of the plurality of sound sources is less than a number of the plurality of microphones.

7. The electronic device as claimed in claim 1 , further comprising:

a display configured to display a direction in which a target sound source from which the predetermined voice is provided is present with respect to the electronic device.

8. A method for processing an audio signal of an electronic device, the method comprising:

receiving by a plurality of microphones, audio signals provided by a plurality of sound sources;

identifying a direction in which each of the plurality of sound sources is located with respect to the electronic device based on the audio signals received through the plurality of microphones;

identifying at least one target sound source among the plurality of sound sources based on a duration of the direction of each of the plurality of sound sources; and

transmitting, to a voice recognition server, an audio signal of the at least one target sound source from which a predetermined voice is provided,

wherein the identifying is performed by identifying at least one sound source of which the duration is less than a predetermined time among the plurality of sound sources as the at least one target sound source.

9. The method as claimed in claim 8 , further comprising:

separating the audio signals received from each of the plurality of sound sources; and

performing voice recognition on the separated audio signals as a target sound source and identifying a target sound source from which the predetermined voice is provided.

10. The method as claimed in claim 9 , wherein the transmitting to the voice recognition server comprises:

transmitting, to the voice recognition server, only an audio signal of the target sound source identified to provide the predetermined voice among the audio signals received by the each of the plurality of microphones after the target sound source from which the predetermined voice is provided is identified.

11. The method as claimed in claim 8 , further comprising:

based on only one target sound source being present,

separating only an audio signal of the target sound source from audio signals received from each of the plurality of microphones; and

performing voice recognition on the audio signal of the separated target sound source and determine whether the predetermined voice is provided.

12. The method as claimed in claim 8 , further comprising:

identifying a target sound source generating a voice among the at least one target sound source; and

performing voice recognition on only an audio signal of the target sound source to provide the voice, and identifying the target sound source from which the predetermined voice is provided.

13. The method as claimed in claim 8 , wherein the determining the direction further comprises:

identifying a direction in which each of the plurality of sound sources is present,

wherein a number of directions of the each of the plurality of sound sources is less than a number of the plurality of microphones.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 14, 2019
From: SHIN, KI-HOON; SONG, MYUNG-SUK; YOO, JONG-UK
To: SAMSUNG ELECTRONICS CO., LTD.
Reel/Frame 048599/0821 →
Priority Claims (1)
KR 10-2016-0133386 · Oct 14, 2016 · national
Continuity (1)
Related Publication 20190214011A1 · Jul 11, 2019