IP Library › Granted Patent US 12,548,566
Granted Patent B2
US 12,548,566 · App. 18/221,676 · Granted Feb 10, 2026

Electronic device and method for controlling electronic device

Inventors: Jiyoun Hong (Suwon-si, KR); Kyenghun Lee (Suwon-si, KR); Hyeonmok Ko (Suwon-si, KR); Dayoung Kwon (Suwon-si, KR); Jonggu Kim (Suwon-si, KR); Seoha Song (Suwon-si, KR); Eunsik Lee (Suwon-si, KR); Pureum Jung (Suwon-si, KR); Changho Paeon (Suwon-si, KR)
Assignee: SAMSUNG ELECTRONICS CO., LTD.
G10L15/22G10L15/18G10L15/26
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,548,566
App. No.
18/221,676
Granted
Feb 10, 2026
Kind
B2
Abstract

An electronic device includes a microphone, a memory, and a processor configured to obtain a first natural language understanding result for a first user voice obtained through the microphone based on a first text corresponding to the first user voice, provide a first response corresponding to the first user voice based on the first natural language understanding result, identify whether the first user voice includes a tracking element based on the first natural language understanding result and a second text corresponding to the first response, based on identifying that the first user voice includes the tracking element, store the first text, the first natural language understanding result and the second text in the memory, and obtain a third text corresponding to the first response based on the first natural language understanding result.

Claims (67)

1 . An electronic device comprising:

a microphone;

a memory; and

a processor configured to:

obtain a first natural language understanding result for a first user voice obtained through the microphone based on a first text corresponding to the first user voice,

provide a first response corresponding to the first user voice based on the first natural language understanding result,

identify whether the first user voice comprises a tracking element based on the first natural language understanding result and a second text corresponding to the first response,

based on identifying that the first user voice comprises the tracking element, store the first text, the first natural language understanding result and the second text in the memory,

obtain a third text corresponding to the first response based on the first natural language understanding result,

identify whether a changed element from the second text is included in the third text by comparing the second text with the third text, and

based on identifying that the changed element is included in the third text, provide a second response corresponding to the first user voice, based on the third text.

2 . The electronic device of claim 1 , wherein the processor is further configured to identify that the first user voice comprises the tracking element based on a user intention included in the first user voice being to request information that is to be changed over time.

3 . The electronic device of claim 2 , wherein the processor is further configured to provide the first response by requesting a server to conduct a search corresponding to the first user voice, and

identify that the first user voice comprises the tracking element based on a search result obtained from the server.

4 . The electronic device of claim 1 , wherein the processor is further configured to:

identify whether the first user voice comprises time information based on the first natural language understanding result, and

based on identifying that the first user voice comprises the time information, obtain the third text until a time point corresponding to the time information after the first response is provided.

5 . The electronic device of claim 1 , wherein the processor further is configured to:

obtain a second natural language understanding result for the second text and a third natural language understanding result for the third text by inputting the second text and the third text to a natural language understanding model, and

identify whether the changed element is included in the third text based on the obtained second natural language understanding result and the third natural language understanding result.

6 . The electronic device of claim 1 , wherein the processor is further configured to identify whether a context of a user that produces the first user voice corresponds to the third text at a time point when the first response is provided, and

based on identifying that the context of the user corresponds to the third text at the time point when the first response is provided, provide the second response based on the third text.

7 . The electronic device of claim 6 , wherein the processor is further configured to:

based on identifying that the context of the user that produces the first user voice does not correspond to the third text, identify whether a second user voice of the user related to the first user voice is included in history information, the history information comprising information on the first user voice, information on the second user voice, and information on responses respectively corresponding to the first user voice and the second user voice, and

based on identifying that the second user voice is included in the history information, provide the second response based on the third text.

8 . A method for controlling an electronic device, the method comprising:

obtaining a first natural language understanding result for a first user voice based on a first text corresponding to the first user voice,

providing a first response corresponding to the first user voice based on the first natural language understanding result;

identifying whether the first user voice comprises a tracking element based on the first natural language understanding result and a second text corresponding to the first response;

based on identifying that the first user voice comprises the tracking element, storing the first text, the first natural language understanding result and the second text;

obtaining a third text corresponding to the first response based on the first natural language understanding result;

identifying whether a changed element from the second text is included in the third text by comparing the second text with the third text; and

based on identifying that the changed element is included in the third text, providing a second response corresponding to the first user voice and based on the third text.

9 . The method of claim 8 , wherein identifying whether the first user voice comprises the tracking element comprises identifying that a user intention included in the first user voice is to request information that is to be changed over time.

10 . The method of claim 9 , wherein identifying whether the first user voice comprises the tracking element further comprises:

requesting a server to search for the first user voice, and

identifying that the first user voice comprises the tracking element based a search result obtained from the server.

11 . The method of claim 8 , wherein obtaining the third text comprises:

identifying whether the first user voice comprises time information based on the first natural language understanding result; and

based on identifying that the first user voice comprises the time information, obtaining the third text until a time point corresponding to the time information after the first response is provided.

12 . The method of claim 8 , wherein identifying whether the changed element from the second text is included in the third text further comprises:

obtaining, a second natural language understanding result for the second text and a third natural language understanding result for the third text, and

identifying whether the changed element is included in the third text based on the second natural language understanding result and the third natural language understanding result.

13 . The method of claim 8 , wherein providing the second response based on the third text comprises identifying whether a context of a user that produces the first user voice corresponds to the third text at a time point when the first response is provided, and

wherein the method further comprises, based on identifying that the context of the user corresponds to the third text at the time point when the first response is provided, providing the second response based on the third text.

14 . The method of claim 13 , wherein providing the second response based on the third text comprises:

based on identifying that the context of the user that produces the first user voice does not correspond to the third text, identifying whether a second user voice of the user related to the first user voice is included in history information, the history information comprising information on the first user voice, information on the second user voice, and information on responses respectively corresponding to the first user voice and the second user voice, and

based on identifying that the second user voice is included in the history information, providing the second response based on the third text.

15 . A non-transitory computer-readable storage medium storing instructions that, when executed by at least one processor, cause the at least one processor to:

obtain a first natural language understanding result for a first user voice obtained through a microphone based on a first text corresponding to the first user voice,

provide a first response corresponding to the first user voice based on the first natural language understanding result,

obtain a third text corresponding to the first response based on the first natural language understanding result,

identify whether a changed element from a second text is included in the third text by comparing the second text with the third text, and

based on identifying that the changed element is included in the third text, provide a second response corresponding to the first user voice and based on the third text.

16 . The storage medium of claim 15 , wherein the instructions, when executed, further cause the at least one processor to identify that the first user voice comprises a tracking element based on a user intention included in the first user voice being to request information that is to be changed over time.

17 . The storage medium of claim 16 , wherein the instructions, when executed, further cause the at least one processor to:

provide the first response by requesting a server to conduct a search corresponding to the first user voice, and

identify that the first user voice comprises the tracking element based on a search result obtained from the server.

18 . The storage medium of claim 15 , wherein the instructions, when executed, further cause the at least one processor to:

identify whether the first user voice comprises time information based on the first natural language understanding result, and

based on identifying that the first user voice comprises the time information, obtain the third text until a time point corresponding to the time information after the first response is provided.

19 . The storage medium of claim 15 , wherein the instructions, when executed, further cause the at least one processor to:

obtain a second natural language understanding result for the second text and a third natural language understanding result for the third text by inputting the second text and the third text to a natural language understanding model, and

identify whether the changed element is included in the third text based on the obtained second natural language understanding result and the third natural language understanding result.

20 . The storage medium of claim 15 , wherein the instructions, when executed, further cause the at least one processor to:

identify whether a context of a user that produces the first user voice corresponds to the third text at a time point when the first response is provided, and

based on identifying that the context of the user corresponds to the third text at the time point when the first response is provided, provide the second response based on the third text.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jul 13, 2023
From: HONG, JIYOUN; LEE, KYENGHUN; KO, HYEONMOK; KWON, DAYOUNG; KIM, JONGGU; SONG, SEOHA; LEE, EUNSIK; JUNG, PUREUM; PAEON, CHANGHO
To: SAMSUNG ELECTRONICS CO., LTD.
Reel/Frame 064248/0660 →
Priority Claims (1)
KR 10-2022-0000559 · Jan 3, 2022 · national
Continuity (2)
Continuation PCTKR2023000026 · Jan 2, 2023
Related Publication 20230360648A1 · Nov 9, 2023
References Cited (44)
US 8126456B2 · Lotter et al. · 2012 [cited by applicant]
US 9858925B2 · Gruber et al. · 2018 [cited by applicant]
US 9959129B2 · Kannan et al. · 2018 [cited by applicant]
US 10102844B1 · Mois · 2018 [cited by examiner]
US 10880378B2 · VanBlon et al. · 2020 [cited by applicant]
US 11302319B2 · Jang et al. · 2022 [cited by applicant]
US 11314481B2 · Wang et al. · 2022 [cited by applicant]
US 11501755B2 · Hwang · 2022 [cited by applicant]
US 11837215B1 · Srivatsa · 2023 [cited by examiner]
US 11966701B2 · Pu · 2024 [cited by examiner]
US 20120265528A1 · Gruber et al. · 2012 [cited by applicant]
US 20130346396A1 · Stamm · 2013 [cited by examiner]
US 20160203002A1 · Kannan et al. · 2016 [cited by applicant]
US 20180144055A1 · Wu · 2018 [cited by applicant]
US 20190180747A1 · Back et al. · 2019 [cited by applicant]
US 20200005784A1 · Vadackupurath Mani et al. · 2020 [cited by applicant]
US 20200111490A1 · Jang et al. · 2020 [cited by applicant]
US 20210026593A1 · Wang et al. · 2021 [cited by applicant]
US 20210050006A1 · Andreas et al. · 2021 [cited by applicant]
US 20210065685A1 · Hwang · 2021 [cited by applicant]
US 20210097999A1 · Gao et al. · 2021 [cited by applicant]
US 20210166678A1 · Lee et al. · 2021 [cited by applicant]
US 20210375286A1 · Choudhury et al. · 2021 [cited by applicant]
US 20220244910A1 · Wang et al. · 2022 [cited by applicant]
US 20220270605A1 · Jang et al. · 2022 [cited by applicant]
US 20220375467A1 · Jin et al. · 2022 [cited by applicant]
JP 2016126294A · 2016 [cited by applicant]
JP 2018180409A · 2018 [cited by applicant]
KR 1020200044175A · 2009 [cited by applicant]
KR 1020130012240A · 2013 [cited by applicant]
KR 1020190067638A · 2019 [cited by applicant]
KR 1020190142219A · 2019 [cited by applicant]
KR 1020210026962A · 2021 [cited by applicant]
KR 1020210066651A · 2021 [cited by applicant]
KR 102447546B1 · 2022 [cited by applicant]
KR 102490776B1 · 2023 [cited by applicant]
KR 102520068B1 · 2023 [cited by applicant]
WO WO2013192584A1 · 2013 [cited by examiner]
WO 2022234930A1 · 2022 [cited by applicant]
Written Opinion (PCT/ISA/237) issued Apr. 11, 2023 from the International Searching Authority in International Application No. PCT/KR2023/000026. [cited by applicant]
International Search Report (PCT/ISA/210) issued Apr. 11, 2023 from the International Searching Authority in International Application No. PCT/KR2023/000026. [cited by applicant]
Extended European Search Report issued on Nov. 11, 2024 by the European Patent Office for European Patent Application No. 23735153.1. [cited by applicant]
Communication dated Jul. 2, 2025, issued by the European Patent Office in counterpart European Application No. 23735153.1. [cited by applicant]
Communication issued Nov. 21, 2025 by the European Patent Office in European Patent Application No. 25204456.5. [cited by applicant]