IP Library Granted Patent US 12,183,342
Granted Patent B2
US 12,183,342 · App. 18/230,581 · Granted Dec 31, 2024

Proactive incorporation of unsolicited content into human-to-computer dialogs

Inventors: Vladimir Vuskovic (Zollikerberg, CH); Stephan Wenger (Zurich, CH); Zineb Ait Bahajji (Zurich, CH); Martin Baeuml (Wollerau, CH); Alexandru Dovlecel (Zurich, CH); Gleb Skobeltsyn (Kilchberg, CH)
Assignee: GOOGLE LLC
G10L15/22G06F40/295G06F40/35G06F40/56G10L15/1815G10L15/222G10L2015/227
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,183,342
App. No.
18/230,581
Granted
Dec 31, 2024
Kind
B2
Abstract

Methods, apparatus, and computer readable media are described related to automated assistants that proactively incorporate, into human-to-computer dialog sessions, unsolicited content of potential interest to a user. In various implementations, based on content of an existing human-to-computer dialog session between a user and an automated assistant, an entity mentioned by the user or automated assistant may be identified. Fact(s)s related to the entity or to another entity that is related to the entity may be identified based on entity data contained in database(s). For each of the fact(s), a corresponding measure of potential interest to the user may be determined. Unsolicited natural language content may then be generated that includes one or more of the facts selected based on the corresponding measure(s) of potential interest. The automated assistant may then incorporate the unsolicited content into the existing human-to-computer dialog session or a subsequent human-to-computer dialog session.

Claims (41)

1. A method implemented using one or more processors, the method comprising:

processing a voice input provided by a user as part of a dialog session involving the user and an automated assistant executed by one or more of the processors;

generating solicited natural language content, wherein the solicited natural language content is responsive to a request identified in the voice input based on the processing;

incorporating, by the automated assistant into the dialog session involving the user and the automated assistant, the solicited natural language content;

identifying additional content that is tangential to the request identified in the voice input or to the solicited natural language content, wherein the additional content includes one or more facts;

determining a measure of potential interest of the user to receive one or more of the facts, wherein the measure of potential interest reflects whether one or more of the same facts has been previously presented to the user in the existing human-to-computer dialog session between the user and the automated assistant or in a previous human-to-computer dialog session between the user and the automated assistant;

in response to determining that the measure of potential interest satisfies a threshold, causing unsolicited natural language content to be automatically output to the user without the user specifically requesting the unsolicited natural language content, wherein the unsolicited natural language output incorporates the additional content; and

in response to determining that the measure of potential interest fails to satisfy the threshold, refraining from causing unsolicited natural language content to be automatically output to the user.

2. The method of claim 1 , wherein the measure of potential interest is further determined based on contextual information associated with the user, wherein the contextual information includes traffic detected near a current location of the user or an accelerometer signal generated by a computing device carried by the user.

3. The method of claim 1 , wherein the measure of potential interest is further determined based on contextual information associated with the user, wherein the contextual information includes past human-to-computer dialogs between the user and the automated assistant or sentiment analysis of speech recognition output of the voice input.

4. The method of claim 1 , wherein the measure of potential interest is further determined based on contextual information associated with the user, wherein the contextual information includes one or more applications currently being interacted with by the user or a state of an application operating on a computing device controlled by the user.

5. The method of claim 1 , wherein one or more of the facts are selected based on one or more entities mentioned in the request identified in the voice input.

6. The method of claim 1 , wherein one or more of the facts are selected based on one or more entities mentioned in the solicited natural language content.

7. The method of claim 1 , wherein the additional content comprises a query that is tangential to the request identified in the voice input.

8. The method of claim 1 , wherein the additional content comprises a query that is tangential to the solicited natural language content.

9. A system comprising one or more processors and memory storing instructions that, in response to execution by the one or more processors, cause the one or more processors to:

process a voice input provided by a user as part of a dialog session involving the user and an automated assistant executed by one or more of the processors;

generate solicited natural language content, wherein the solicited natural language content is responsive to a request identified in the voice input based on the processing;

incorporate, by the automated assistant into the dialog session involving the user and the automated assistant, the solicited natural language content;

identify additional content that is tangential to the request identified in the voice input or to the solicited natural language content, wherein the additional content includes one or more facts;

determine a measure of potential interest desirability of the user to receive one or more of the facts, wherein the measure of potential interest reflects whether one or more of the same facts has been previously presented to the user in the existing human-to-computer dialog session between the user and the automated assistant or in a previous human-to-computer dialog session between the user and the automated assistant;

in response to a determination that the measure of potential interest satisfies a threshold, cause unsolicited natural language content to be automatically output to the user without the user specifically requesting the unsolicited natural language content, wherein the unsolicited natural language output incorporates the additional content; and

in response to a determination that the measure of potential interest fails to satisfy the threshold, refrain from causing unsolicited natural language content to be automatically output to the user.

10. The system of claim 9 , wherein the measure of potential interest is further determined based on contextual information associated with the user, wherein the contextual information includes traffic detected near a current location of the user or an accelerometer signal generated by a computing device carried by the user.

11. The system of claim 9 , wherein the measure of potential interest is further determined based on contextual information associated with the user, wherein the contextual information includes past human-to-computer dialogs between the user and the automated assistant or sentiment analysis of speech recognition output of the voice input.

12. The system of claim 9 , wherein the measure of potential interest is further determined based on contextual information associated with the user, wherein the contextual information includes one or more applications currently being interacted with by the user or a state of an application operating on a computing device controlled by the user.

13. The system of claim 9 , wherein one or more of the facts are selected based on one or more entities mentioned in the request identified in the voice input.

14. The system of claim 9 , wherein one or more of the facts are selected based on one or more entities mentioned in the solicited natural language content.

15. The system of claim 9 , wherein the additional content comprises a query that is tangential to the request identified in the voice input.

16. The system of claim 9 , wherein the additional content comprises a query that is tangential to the solicited natural language content.

17. At least one non-transitory computer-readable medium comprising instructions that, in response to execution by one or more processors, cause the one or more processors to:

process a voice input provided by a user as part of a dialog session involving the user and an automated assistant executed by one or more of the processors;

generate solicited natural language content, wherein the solicited natural language content is responsive to a request identified in the voice input based on the processing;

incorporate, by the automated assistant into the dialog session involving the user and the automated assistant, the solicited natural language content;

identify additional content that is tangential to the request identified in the voice input or to the solicited natural language content, wherein the additional content includes one or more facts;

determine a measure of potential interest of the user to receive one or more of the facts, wherein the measure of potential interest reflects whether one or more of the same facts has been previously presented to the user in the existing human-to-computer dialog session between the user and the automated assistant or in a previous human-to-computer dialog session between the user and the automated assistant;

in response to a determination that the measure of potential interest satisfies a threshold, cause unsolicited natural language content to be automatically output to the user without the user specifically requesting the unsolicited natural language content, wherein the unsolicited natural language output incorporates the additional content; and

in response to a determination that the measure of potential interest fails to satisfy the threshold, refrain from causing unsolicited natural language content to be automatically output to the user.

18. The at least one non-transitory computer-readable medium of claim 17 , wherein the measure of potential interest is further determined based on contextual information associated with the user, wherein the contextual information includes traffic detected near a current location of the user or an accelerometer signal generated by a computing device carried by the user.

19. The at least one non-transitory computer-readable medium of claim 17 , wherein the measure of potential interest is further determined based on contextual information associated with the user, wherein the contextual information includes past human-to-computer dialogs between the user and the automated assistant or sentiment analysis of speech recognition output of the voice input.

20. The at least one non-transitory computer-readable medium of claim 17 , wherein the measure of potential interest is further determined based on contextual information associated with the user, wherein the contextual information includes one or more applications currently being interacted with by the user or a state of an application operating on a computing device controlled by the user.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 14, 2023
From: VUSKOVIC, VLADIMIR; WENGER, STEPHAN; BAHAJJI, ZINEB AIT; BAEUML, MARTIN; DOVLECEL, ALEXANDRU; SKOBELTSYN, GLEB
To: GOOGLE INC.
Reel/Frame 064580/0205 →
CHANGE OF NAME Recorded Aug 14, 2023
From: GOOGLE INC.
To: GOOGLE LLC
Reel/Frame 064585/0412 →
Continuity (5)
Continuation 17411532 · Aug 25, 2021
Continuation 16549403 · Aug 23, 2019
Continuation 15825919 · Nov 29, 2017
Continuation 15585363 · May 3, 2017
Related Publication 20230377571A1 · Nov 23, 2023