IP Library Granted Patent US 11,355,117
Granted Patent B2
US 11,355,117 · App. 16/990,525 · Granted Jun 7, 2022

Dialog system with automatic reactivation of speech acquiring mode

Inventors: Ilya Gennadyevich Gelfenbeyn (Sunnyvale, CA); Artem Goncharuk (Mountain View, CA); Pavel Aleksandrovich Sirotin (Sunnyvale, CA)
Assignee: GOOGLE LLC
G10L15/22G06F3/167G06F16/3329G10L15/083G10L15/1815G10L15/1822G10L15/26G10L15/30G10L25/48G10L15/18G10L2015/223G10L2015/228
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,355,117
App. No.
16/990,525
Granted
Jun 7, 2022
Kind
B2
Abstract

Embodiments of the disclosure generally relate to a dialog system allowing for automatically reactivating a speech acquiring mode after the dialog system delivers a response to a user request. The reactivation parameters, such as a delay, depend on a number of predetermined factors and conversation scenarios. The embodiments further provide for a method of operating of the dialog system. An exemplary method comprises the steps of: activating a speech acquiring mode, receiving a first input of a user, deactivating the speech acquiring mode, obtaining a first response associated with the first input, delivering the first response to the user, determining that a conversation mode is activated, and, based on the determination, automatically re-activating the speech acquiring mode within a first predetermined time period after delivery of the first response to the user.

Claims (46)

1. A method for operating a dialog system, the method implemented by one or more processors and comprising:

providing, at a graphical interface associated with the dialog system, an actionable graphical button that is selectable through user interface input to activate or deactivate a conversation mode of the dialog system;

activating the conversation mode of the dialog system in response to a selection of the provided actionable graphical button;

after activating the conversation mode of the dialog system:

activating a speech acquiring mode of the dialog system in response to a user speaking an activation phrase;

receiving, via the speech acquiring mode after activating the speech acquiring mode, a first spoken input of a user;

deactivating, after receiving the first spoken input of the user, the speech acquiring mode;

obtaining a response associated with the first spoken input;

delivering the response in response to receiving the first spoken input;

selecting, based on one or more properties of the response, a time period for the speech acquiring mode to last when the speech acquiring mode is automatically reactivated responsive to the first spoken input;

determining, in response to the conversation mode being activated, to automatically re-activate the speech acquiring mode responsive to the first spoken input;

automatically re-activating the speech acquiring mode responsive to the first spoken input and in response to determining to automatically re-activate the speech acquiring mode,

wherein automatically re-activating the speech acquiring mode comprises causing the speech acquiring mode to last for the selected time period.

2. The method of claim 1 , wherein the one or more properties of the response include a length of the response and wherein selecting, based on one or more properties of the response, the time period for the speech acquiring mode to last when the speech acquiring mode is automatically reactivated responsive to the first spoken input, comprises:

selecting the time period based on the length of the response.

3. The method of claim 1 , wherein the one or more properties of the response include a type of the response and wherein selecting, based on one or more properties of the response, the time period for the speech acquiring mode to last when the speech acquiring mode is automatically reactivated responsive to the first spoken input, comprises:

selecting the time period based on the type of the response.

4. The method of claim 1 , wherein obtaining the response associated with the first spoken input comprises:

processing the first spoken input, using an automatic speech recognizer, to generate a recognized input; and

generating the response based on the recognized input.

5. The method of claim 4 , wherein generating the response based on the recognized input comprises:

retrieving text based on the recognized input;

transforming the text into a machine-generated audio signal; and

using the machine-generated audio signal as the response.

6. The method of claim 5 , wherein automatically re-activating the speech acquiring mode occurs immediately after the machine-generated audio signal has been provided as output to the user.

7. A user device, the user device comprising:

instructions stored in memory,

one or more processors executing the stored instructions to cause the one or more processors to:

activate a speech acquiring mode in response to a user speaking an activation phrase;

receive, via the speech acquiring mode after activating the speech acquiring mode, a first spoken input of a user;

deactivate, after receiving the first spoken input of the user, the speech acquiring mode;

obtain, in response to the first spoken input:

a response that includes content to be rendered responsive to the first spoken input, and

metadata associated with the response, wherein the metadata dictates that the speech acquiring mode is to be automatically reactivated after rendering of the content of the response;

render the content in response to receiving the response;

determine, in response to the metadata dictating that the speech acquiring mode is to be automatically reactivated after rendering of the content of the response, to automatically re-activate the speech acquiring mode after rendering of the content of the response; and

automatically re-activate the speech acquiring mode after rendering of the content of the response in response to determining to automatically re-activate the speech acquiring mode after delivery of the response.

8. The user device of claim 7 ,

wherein the metadata further dictates a time period for the speech acquiring mode to last when the speech acquiring mode is automatically reactivated after rendering of the content of the response; and

wherein in automatically re-activating the speech acquiring mode after rendering of the content of the response, one or more of the processors are to re-activate the speech acquiring mode for the time period in response to the metadata dictating the time period for the speech acquiring mode to last when the speech acquiring mode is automatically reactivated after rendering of the content of the response.

9. The user device of claim 7 , wherein the user device further comprises a speaker and wherein the content is an audio message and is rendered via the speaker.

10. The user device of claim 7 , wherein in obtaining the response one or more of the processors are to obtain the response via a communication network and from a dialog system.

11. The user device of claim 7 , wherein the metadata, that dictates that the speech acquiring mode is to be automatically reactivated after rendering of the content of the response, comprises a flag with a true value.

12. The user device of claim 7 , wherein the user device further comprises a speaker, wherein the content is an audio message and is rendered via the speaker, wherein the metadata further dictates one or more parameters of the audio message, and wherein in rendering the audio message one or more of the processors are to render the audio message with the one or more parameters.

13. The user device of claim 12 , wherein the one or more parameters include a volume of the response.

14. The user device of claim 7 , wherein the response includes the metadata.

Assignments (3)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 19, 2020
From: GELFENBEYN, ILYA GENNADYEVICH; GONCHARUK, ARTEM; SIROTIN, PAVEL ALEKSANDROVICH
To: SPEAKTOIT, INC.
Reel/Frame 053537/0455 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 19, 2020
From: SPEAKTOIT, INC.
To: GOOGLE INC.
Reel/Frame 053537/0580 →
CHANGE OF NAME Recorded Aug 19, 2020
From: GOOGLE INC.
To: GOOGLE LLC
Reel/Frame 053540/0879 →
Priority Claims (1)
RU 2012150996 · Nov 28, 2012 · national
Continuity (10)
Continuation 16137069 · Sep 20, 2018
Continuation 15395476 · Dec 30, 2016
Continuation 15169926 · Jun 1, 2016
Continuation In Part 14721012 · May 26, 2015
Continuation In Part PCTIB2012056955 · Dec 4, 2012
Continuation In Part 14721044 · May 26, 2015
Continuation In Part PCTIB2012056973 · Dec 5, 2012
Continuation In Part 14775729
Continuation In Part 14901026
Related Publication 20200372914A1 · Nov 26, 2020