IP Library › Granted Patent US 12,749,487
Granted Patent B2
US 12,749,487 · App. 18/790,904 · Granted Sep 29, 2026

Method and apparatus for providing voice recognition service

Inventor: Seong Soo Yae (Hwaseong-si, KR)
Assignees: Hyundai Motor Company; Kia Corporation
G10L15/22G10L15/1815G10L15/20G10L15/26
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,749,487
App. No.
18/790,904
Granted
Sep 29, 2026
Kind
B2
Abstract

A method and apparatus for providing a voice recognition service can include separating a user utterance from noise and converting the user utterance into a text to generate content of the user utterance, extracting, from the content of the user utterance, call words information, domain information, end service name information, and operations information, generating a corrected user command by correcting a user command, and generating response information by using the corrected user command.

Claims (57)

1 . An apparatus for providing a voice recognition service, comprising:

at least one processor; and

a memory storing instructions that, when executed by the at least one processor, cause the apparatus to:

separate a user utterance, in which call word information for specifying one of a plurality of voice assistants is omitted, from noise;

convert the user utterance into text to generate content of the user utterance;

extract, from the content of the user utterance, extracted information including at least one of domain information, end service name information, and operations information;

generate a corrected user command including the call word information, the domain information, the end service name information, and the operations information, based on usage counts of combinations of each voice assistant of the plurality of voice assistants and the extracted information;

convert the corrected user command into a voice command by generating speech through a text-to-speech (TTS) process;

generate a control signal for a vehicle in response to the voice command; and

control the vehicle based on the control signal.

2 . The apparatus of claim 1 , wherein execution of the instructions by the at least one processor further causes the apparatus to:

classify an intention of the user utterance, by using a natural language understanding engine.

3 . The apparatus of claim 1 , wherein execution of the instructions by the at least one processor further causes the apparatus to:

determine the call word information by using a voice assistant usage history of a voice assistant, in response to the content of the user utterance not including the call word information but including the end service name information.

4 . The apparatus of claim 1 , wherein execution of the instructions by the at least one processor further causes the apparatus to:

determine the call word information and the end service name information by using a voice assistant usage history of a voice assistant, an end service usage history of an end service, and a recently utilized service usage history of a recently utilized service, in response to the content of the user utterance not including the call word information and the end service name information.

5 . The apparatus of claim 1 , wherein execution of the instructions by the at least one processor further causes the apparatus to:

determine the call word information and the end service name information by using an end service based on the domain information, in response to the content of the user utterance including the domain information not being supported by an invoked voice assistant.

6 . The apparatus of claim 1 , wherein execution of the instructions by the at least one processor further causes the apparatus to:

determine the call word information and the end service name information by using a voice assistant activated earliest among the plurality of voice assistants, in response to the content of the user utterance including the domain information that has no usage history.

7 . The apparatus of claim 1 , wherein:

execution of the instructions by the at least one processor further causes the apparatus to store the usage counts by using a source database and a virtual database copy,

the source database stores all existing usage histories, and

the virtual database copy is a copy of the source database.

8 . A method of voice recognition service, the method comprising:

separating a user utterance, in which call word information for specifying one of a plurality of voice assistants is omitted, from noise;

converting the user utterance into text to generate content of the user utterance;

extracting, from the content of the user utterance, extracted information including at least one of domain information, end service name information, and operations information;

generating a corrected user command including the call word information, the domain information, the end service name information, and the operations information, based on usage counts of combinations of each voice assistant of the plurality of voice assistants and the extracted information;

converting the corrected user command into a voice command by generating speech through a text-to-speech process;

generating a control signal for a vehicle in response to the voice command; and

controlling the vehicle based on the control signal.

9 . The method of claim 8 , wherein extracting, from the content of the user utterance, the extracted information including at least one of the domain information, the end service name information, and the operations information comprises:

classifying an intention of the user utterance by using a natural language understanding engine to extract the domain information, the end service name information, and the operations information.

10 . The method of claim 8 , wherein generating the corrected user command comprises;

determining the call word information by using a voice assistant usage history of a voice assistant to generate the corrected user command, in response to the content of the user utterance not including the call word information but including the end service name information.

11 . The method of claim 8 , wherein generating the corrected user command comprises:

determining the call word information and the end service name information by using a voice assistant usage history of a voice assistant, an end service usage history of an end service, and a recently utilized service usage history of a recently utilized service to generate the corrected user command, in response to the content of the user utterance not including the call word information and the end service name information.

12 . The method of claim 8 , wherein generating the corrected user command comprises:

determining the call word information and the end service name information by using an end service based on the domain information to generate the corrected user command, in response to the content of the user utterance including the domain information not being supported by an invoked voice assistant.

13 . The method of claim 8 , wherein generating the corrected user command comprises:

determining the call word information and the end service name information by using a voice assistant activated earliest among the plurality of voice assistants to generate the corrected user command, in response to the content of the user utterance including the domain information having no usage history.

14 . A non-transitory computer-readable medium storing a computer program including computer-executable instructions for causing, when executed by a computer, the computer to perform steps of:

separating a user utterance, in which call word information for specifying one of a plurality of voice assistants is omitted, from noise;

converting the user utterance into text to generate content of the user utterance;

extracting, from the content of the user utterance, extracted information including at least one of domain information, end service name information, and operations information;

generating a corrected user command including the call word information, the domain information, the end service name information, and the operations information, based on usage counts of combinations of each voice assistant plurality of voice assistants and the extracted information;

converting the corrected user command into a voice command by generating speech through a text-to-speech process;

generating a control signal for a vehicle in response to the voice command; and

controlling the vehicle based on the control signal.

15 . The apparatus of claim 1 , wherein execution of the instructions by the at least one processor further causes the apparatus to:

operate a source database and a virtual database copy to manage the usage counts of combinations,

wherein:

the virtual database copy is generated by copying the source database when the vehicle is started, and

the source database is updated by using the virtual database copy when the vehicle is turned off.

16 . The apparatus of claim 1 , wherein execution of the instructions by the at least one processor further causes the apparatus to:

update the usage counts of combinations when a predetermined amount of time has passed after a service corresponding to the corrected user command is provided to a user.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jul 31, 2024
From: YAE, SEONG SOO
To: HYUNDAI MOTOR COMPANY; KIA CORPORATION
Reel/Frame 068143/0614 →
Priority Claims (1)
KR 10-2023-0140353 · Oct 19, 2023 · national
Continuity (1)
Related Publication 20250131922A1 · Apr 24, 2025
References Cited (13)
US 6064646A · Shal · 2000 [cited by examiner]
US 10515637B1 · Devries · 2019 [cited by examiner]
US 11887603B2 · Casado · 2024 [cited by examiner]
US 20050015197A1 · Ohtsuji · 2005 [cited by examiner]
US 20060287866A1 · Cross · 2006 [cited by examiner]
US 20110054899A1 · Phillips · 2011 [cited by examiner]
US 20120035924A1 · Jitkoff · 2012 [cited by examiner]
US 20170371861A1 · Barborak · 2017 [cited by examiner]
US 20180047387A1 · Nir · 2018 [cited by examiner]
US 20190095428A1 · Asano · 2019 [cited by examiner]
US 20220157307A1 · Hartung · 2022 [cited by examiner]
US 20240021199A1 · Ishimaru · 2024 [cited by examiner]
US 20240180482A1 · Stegmann · 2024 [cited by examiner]