IP Library Granted Patent US 12,640,140
Granted Patent B2
US 12,640,140 · App. 18/183,522 · Granted May 26, 2026

Electronic apparatus and controlling method thereof

Inventors: Dohyeong Hwang (Suwon-si, KR); Okhee Baek (Suwon-si, KR); Jongyeong Shin (Suwon-si, KR); Jeongwon Lee (Suwon-si, KR)
Assignee: Samsung Electronics Co., Ltd.
G10L15/16G06F3/167G10L15/063G10L15/22G10L15/30G10L2015/088G10L2015/223G10L2015/225
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,640,140
App. No.
18/183,522
Filed
Mar 14, 2023
Granted
May 26, 2026
Kind
B2
Art Unit
2659
USPC
704/232
Abstract

An electronic apparatus is provided. The electronic apparatus includes a communication interface with communication circuitry, a memory configured to store at least one instruction and a processor, and the processor is configured to receive a first audio recognized as a wake up word by an external device from the external device, determine whether the first audio corresponds to the wake up word by analyzing the first audio, based on determining that the first audio does not correspond to the wake up word, obtain a neural network model for detecting a wake up word misrecognition based on the first audio, and transmit information regarding the neural network model to the external device.

Claims (48)

1 . An apparatus comprising:

a communication interface with communication circuitry;

a memory configured to store at least one instruction; and

a processor,

wherein the processor is configured to:

receive, from an external device, a first audio, recognized as a wake up word by the external device,

receive, from the external device, second audio captured subsequent to the first audio, wherein the second audio includes a user voice subsequent to an operation performed by the external device based on recognition of the first audio as the wake up word,

analyze the first audio and the second audio so as to determine whether the first audio corresponds to the wake up word,

based on determining that the first audio does not correspond to the wake up word, obtain a neural network model trained to identify audio that corresponds to a misrecognized wake up word, and

transmit information regarding the neural network model to the external device so as to enable the external device to input the first audio to the neural network model and determine whether the first audio corresponds to the misrecognized wake up word based on output from the neural network model.

2 . The apparatus of claim 1 , wherein the processor is further configured to, based on a text corresponding to the first audio not being detected, determine that the first audio does not correspond to the wake up word.

3 . The apparatus of claim 1 , wherein the processor is further configured to:

obtain a text corresponding to the first audio; and

based on a similarity between the text corresponding to the first audio and the wake up word being less than a predetermined value, determine that the first audio does not correspond to the wake up word.

4 . The apparatus of claim 1 , wherein the processor is further configured to:

obtain a text corresponding to the second audio; and

based on the text corresponding to the second audio not having a predetermined sentence structure, determine that the first audio does not correspond to the wake up word.

5 . The apparatus of in claim 1 ,

wherein the second audio includes a user voice regarding an operation performed as the external device recognizes the first audio as the wake up word, and

wherein the processor is further configured to determine whether the first audio corresponds to the wake up word by analyzing the user voice.

6 . The apparatus of claim 1 , wherein the processor is further configured to determine whether the first audio corresponds to the wake up word based on a user feedback input through a user interface (UI) provided by the external device.

7 . The apparatus of claim 1 , wherein the processor is further configured to:

based on determining that the first audio does not correspond to the wake up word, store the first audio in the memory;

identify a plurality of third audios forming a cluster from among the first audio stored in the memory; and

train the neural network model based on the plurality of third audios.

8 . A method of controlling an electronic apparatus, the method comprising:

receiving, from an external device, a first audio recognized as a wake up word by the external device;

receiving, from the external device, a second audio captured subsequent to the first audio, wherein the second audio includes a user voice subsequent to an operation performed by the external device based on recognition of the first audio as the wake up word;

analyzing the first audio and the second audio to determine whether the first audio corresponds to the wake up word;

based on determining that the first audio does not correspond to the wake up word, obtaining a neural network model trained to identify audio that corresponds to a misrecognized wake up word; and

transmitting information regarding the neural network model to the external device.

9 . The method of claim 8 , wherein the determining of whether the first audio corresponds to the wake up word comprises determining, based on a text corresponding to the first audio not being detected, that the first audio does not correspond to the wake up word.

10 . The method of claim 8 , wherein the determining of whether the first audio corresponds to the wake up word comprises:

obtaining a text corresponding to the first audio; and

based on a similarity between the text corresponding to the first audio and the wake up word being less than a predetermined value, determining that the first audio does not correspond to the wake up word.

11 . The method of claim 8 , wherein the determining of whether the first audio corresponds to the wake up word comprises:

obtaining a text corresponding to the second audio; and

based on the text corresponding to the second audio not having a predetermined sentence structure, determining that the first audio does not correspond to the wake up word.

12 . The method of claim 8 ,

wherein the second audio includes a user voice regarding an operation performed as the external device recognizes the first audio as the wake up word, and

wherein the determining of whether the first audio corresponds to the wake up word comprises analyzing the user voice.

13 . The method of claim 8 , further comprising:

based on the first audio corresponding to the wake up word, obtaining a response corresponding to the second audio; and

transmitting information regarding the obtained response to the external device.

14 . The method of claim 8 , further comprising:

determining whether the first audio corresponds to the wake up word based on a user feedback input through a user interface (UI) provided by the external device.

15 . The method of claim 8 , further comprising determining whether the first audio corresponds to the wake up word by determining whether the first audio corresponds to a misrecognition word using the neural network model.

16 . The method of claim 8 , wherein the information regarding the neural network model comprises at least one of parameters regarding the neural network model or a message requesting to download the neural network model.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 14, 2023
From: HWANG, DOHYEONG; BAEK, OKHEE; SHIN, JONGYEONG; LEE, JEONGWON
To: SAMSUNG ELECTRONICS CO., LTD.
Reel/Frame 062977/0256 →
Priority Claims (1)
KR 10-2021-0154035 · Nov 10, 2021 · national
Continuity (2)
Continuation PCTKR2022017242 · Nov 4, 2022
Related Publication 20230223013A1 · Jul 13, 2023
References Cited (29)
US 10418027B2 · Ko et al. · 2019 [cited by applicant]
US 10699718B2 · Kim et al. · 2020 [cited by applicant]
US 10748524B2 · Wang et al. · 2020 [cited by applicant]
US 11417327B2 · Choi · 2022 [cited by applicant]
US 11514890B2 · Lee et al. · 2022 [cited by applicant]
US 11557292B1 · Wang · 2023 [cited by examiner]
US 20180102125A1 · Ko · 2018 [cited by examiner]
US 20190311719A1 · Adams · 2019 [cited by examiner]
US 20200013390A1 · Wang et al. · 2020 [cited by applicant]
US 20200027462A1 · Wang et al. · 2020 [cited by applicant]
US 20200090647A1 · Kurtz · 2020 [cited by examiner]
US 20200125603A1 · Ha et al. · 2020 [cited by applicant]
US 20200184966A1 · Yavagal · 2020 [cited by applicant]
US 20210151043A1 · Lee et al. · 2021 [cited by applicant]
US 20210210075A1 · Kim · 2021 [cited by examiner]
US 20210256965A1 · Kim et al. · 2021 [cited by applicant]
US 20210295833A1 · Rastrow et al. · 2021 [cited by applicant]
US 20220139377A1 · Lee et al. · 2022 [cited by applicant]
US 20220189481A1 · Ushakov · 2022 [cited by applicant]
CN 109872713A · 2019 [cited by applicant]
KR 1020160110085A · 2016 [cited by applicant]
KR 1020180040426A · 2018 [cited by applicant]
KR 1020200007530A · 2020 [cited by applicant]
KR 1020200025226A · 2020 [cited by applicant]
KR 1020200045851A · 2020 [cited by applicant]
KR 1020200063521A · 2020 [cited by applicant]
KR 1020210030160A · 2021 [cited by applicant]
European Search Report dated Oct. 10, 2024, issued in European Application No. 22893120.0. [cited by applicant]
International Search Report and written opinion dated Feb. 3, 2023, issued in International Application No. PCT/KR2022/017242. [cited by applicant]