IP Library › Granted Patent US 11,900,932
Granted Patent B2
US 11,900,932 · App. 17/366,270 · Granted Feb 13, 2024

Determining a system utterance with connective and content portions from a user utterance

Inventors: Atsushi Ikeno (Kyoto, JP); Yusuke Jinguji (Hiroo-gun, JP); Toshifumi Nishijima (Kasugai, JP); Fuminori Kataoka (Nisshin, JP); Hiromi Tonegawa (Okazaki, JP); Norihide Umeyama (Nisshin, JP)
Assignee: TOYOTA JIDOSHA KABUSHIKI KAISHA
G10L15/22G10L13/08G10L15/1815G10L15/30G10L2015/223
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,900,932
App. No.
17/366,270
Granted
Feb 13, 2024
Kind
B2
Abstract

A voice dialogue system includes a voice input unit which acquires a user utterance, an intention understanding unit which interprets an intention of utterance of a voice acquired by the voice input unit, a dialogue text creator which creates a text of a system utterance, and a voice output unit which outputs the system utterance as voice data. When creating a text of a system utterance, the dialogue text creator creates the text by inserting a tag in a position in the system utterance, and the intention understanding unit interprets an utterance intention of a user in accordance with whether a timing at which the user utterance is made is before or after an output of a system utterance at a position corresponding to the tag from the voice output unit.

Claims (17)

1. A voice dialogue system, comprising:

a voice input unit configured to acquire user utterances of a user;

a dialogue text creator configured to create system utterances, wherein the dialogue text creator creates the system utterances based upon a stored history of dialogue performed in a past between the system and the user stored in a dialogue manager, the dialogue manager storing a time and date or location of the dialogue and enabling what kind of conversation had taken place with the user to be discerned and a response using previous dialogue as a reference to be generated;

a voice output unit configured to output the system utterances as voice data; and

a determiner configured to determine whether or not that the user utterance acquired by the voice input unit is a response to the system utterance currently being output as voice data, wherein

in response to a first system utterance output by the voice output unit and a second system utterance output after the first system utterance without having acquired a user utterance, the second system utterance comprising a connective portion for connecting following sentences and a content portion that is a subject of the second system utterance, the content portion including (a) a first content portion that is a first system utterance that is output before the connective portion and (b) a second content portion that is a second system utterance that is output after the connective portion, the first content portion is a first question and the second content portion is a second question different from the first question, the determiner determines:

that the user utterance is a response to the first system utterance when the user utterance is acquired during output of the connective portion of the second system utterance by the voice output unit, and

that the user utterance is a response to the second system utterance when the user utterance is acquired during output of the content portion of the second system utterance by the voice output unit.

2. The voice dialogue system according to claim 1 , wherein

the connective portion comprises one of an interjection, a gambit, or a repetition of a part of a previously acquired user utterance.

3. The voice dialogue system according to claim 2 , wherein

the dialogue text creator is further configured to, when creating the system utterances, to insert an unvoiced tag between the connective portion and the content portion of the system utterances, and

the determiner is further configured to determine whether the user utterance is acquired during the output of the content portion or the connective portion of the second system utterance based at least on a position of the unvoiced tag in the second system utterance.

4. The voice dialogue system according to claim 2 , wherein the determiner is further configured to:

calculate a first period of time that is a period of time that it will take to output the connective portion of the second system utterance as voice data;

acquire a second period of time that is a period of time from a start of output of the second system utterance as voice data to a start of the user utterance; and

compare the first period of time and the second period of time with each other to determine whether the user utterance is acquired during output of the content portion of the second system utterance or during output of the connective portion of the second system utterance.

Priority Claims (1)
JP 2016-189406 · Sep 28, 2016 · national
Continuity (3)
Division 16390261 · Apr 22, 2019
Continuation 15704691 · Sep 14, 2017
Related Publication 20210335362A1 · Oct 28, 2021
Cited By (1)
US 12,340,803