IP Library Granted Patent US 12,243,529
Granted Patent B2
US 12,243,529 · App. 18/404,452 · Granted Mar 4, 2025

Conversational agent response determined using a sentiment

Inventors: Johnny Chen (Sunnyvale, CA); Thomas L. Dean (Los Altos Hills, CA); Qiangfeng Peter Lau (Mountain View, CA); Sudeep Gandhe (Sunnyvale, CA); Gabriel Schine (Los Banos, CA)
Assignee: GOOGLE LLC
G10L15/22G06F16/3329G06F21/6245G10L13/00G10L13/033G10L13/08G10L15/1815G10L15/1822G10L15/26G10L15/30G10L17/22H04L67/104G10L2015/223G10L2015/228
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,243,529
App. No.
18/404,452
Granted
Mar 4, 2025
Kind
B2
Abstract

Methods, systems, and apparatus, including computer programs encoded on computer storage media, for handing off a user conversation between computer-implemented agents. One of the methods includes receiving, by a computer-implemented agent specific to a user device, a digital representation of speech encoding an utterance, determining, by the computer-implemented agent, that the utterance specifies a requirement to establish a communication with another computer-implemented agent, and establishing, by the computer-implemented agent, a communication between the other computer-implemented agent and the user device.

Claims (35)

1. A computer-implemented method comprising:

receiving, by a first computer-implemented agent for a user device, a text representation of an utterance that includes a command, wherein the text representation of the utterance is determined based on a speech encoding of the utterance; and

in response to processing the text representation of the utterance to determine words included in the utterance:

determining, from among a plurality of different demographics, a particular demographic of a speaker of the utterance;

in response to determining the particular demographic, selecting, from a plurality of different computer-implemented agents, a particular computer- implemented agent based on the particular demographic associated with the particular computer-implemented agent matching the particular demographic, wherein each agent is associated with a respective one of the plurality of different demographics; and

causing the particular computer-implemented agent to provide an interface for processing subsequent utterances spoken by the speaker.

2. The computer-implemented method of claim 1 , wherein the command comprises a request for results responsive to the first computer-implemented agent that is different from the particular computer-implemented agent.

3. The computer-implemented method of claim 1 , wherein determining, from among the plurality of different demographics, the particular demographic of the speaker of the utterance comprises determining the particular demographic using the text representation of the utterance.

4. The computer-implemented method of claim 1 , wherein determining, from among the plurality of different demographics, the particular demographic of the speaker of the utterance comprises determining the particular demographic based on profile data that includes demographic information for the speaker.

5. The computer-implemented method of claim 3 , wherein determining, from among the plurality of different demographics, the particular demographic of the speaker of the utterance comprises determining the particular demographic based on a style of speech determined from the speech encoding of the utterance.

6. The computer-implemented method of claim 3 , wherein determining, from among the plurality of different demographics, the particular demographic of the speaker of the utterance comprises receiving data specifying the particular demographic based on a response from the speaker.

7. A user device comprising:

one or more processors; and

memory storing instructions that, when executed by the one or more processors, are operable to:

receive, by a first computer-implemented agent for the user device, a text representation of an utterance that includes a command, wherein the text representation of the utterance is determined based on a speech encoding of the utterance; and

in response to processing the text representation of the utterance to determine words included in the utterance:

determine, from among a plurality of different demographics, a particular demographic of a speaker of the utterance;

in response to determining the particular demographic, select, from a plurality of different computer-implemented agents, a particular computer-implemented agent based on the particular demographic associated with the particular computer-implemented agent matching the particular demographic, wherein each agent is associated with a respective one of the plurality of different demographics; and

cause the particular computer-implemented agent to provide an interface for processing subsequent utterances spoken by the speaker.

8. The user device of claim 7 , wherein the command comprises a request for results responsive to the first computer-implemented agent that is different from the particular computer-implemented agent.

9. The user device of claim 7 , wherein determining, from among the plurality of different demographics, the particular demographic of the speaker of the utterance comprises determining the particular demographic using the text representation of the utterance.

10. The user device of claim 7 , wherein determining, from among the plurality of different demographics, the particular demographic of the speaker of the utterance comprises determining the particular demographic based on profile data that includes demographic information for the speaker.

11. The user device of claim 9 , wherein determining, from among the plurality of different demographics, the particular demographic of the speaker of the utterance comprises determining the particular demographic based on a style of speech determined from the speech encoding of the utterance.

12. The user device of claim 9 , wherein determining, from among the plurality of different demographics, the particular demographic of the speaker of the utterance comprises receiving data specifying the particular demographic based on a response from the speaker.

13. A non-transitory computer readable storage medium storing instructions executable by a data processing apparatus and, upon such execution, causes the data processing apparatus to perform operations comprising:

receiving, by a first computer-implemented agent for a user device, a text representation of an utterance that includes a command, wherein the text representation of the utterance is determined based on a speech encoding of the utterance; and

in response to processing the text representation of the utterance to determine words included in the utterance:

determining, from among a plurality of different demographics, a particular demographic of a speaker of the utterance;

in response to determining the particular demographic, selecting, from a plurality of different computer-implemented agents, a particular computer-implemented agent based on the particular demographic associated with the particular computer-implemented agent matching the particular demographic, wherein each agent is associated with a respective one of the plurality of different demographics; and

cause the particular computer-implemented agent to provide an interface for processing subsequent utterances spoken by the speaker.

14. The non-transitory computer readable storage medium of claim 13 , wherein the command comprises a request for results responsive to the first computer-implemented agent that is different from the particular computer-implemented agent.

15. The non-transitory computer readable storage medium of claim 13 , wherein determining, from among the plurality of different demographics, the particular demographic of the speaker of the utterance comprises determining the particular demographic using the text representation of the utterance.

16. The non-transitory computer readable storage medium of claim 13 , wherein determining, from among the plurality of different demographics, the particular demographic of the speaker of the utterance comprises determining the particular demographic based on profile data that includes demographic information for the speaker.

17. The non-transitory computer readable storage medium of claim 16 , wherein determining, from among the plurality of different demographics, the particular demographic of the speaker of the utterance comprises determining the particular demographic based on a style of speech determined from the speech encoding of the utterance.

18. The non-transitory computer readable storage medium of claim 16 , wherein determining, from among the plurality of different demographics, the particular demographic of the speaker of the utterance comprises receiving data specifying the particular demographic based on a response from the speaker.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 9, 2024
From: CHEN, JOHNNY; DEAN, THOMAS L.; LAU, QIANGFENG PETER; GANDHE, SUDEEP; SCHINE, GABRIEL
To: GOOGLE INC.
Reel/Frame 066068/0074 →
CHANGE OF NAME Recorded Jan 9, 2024
From: GOOGLE INC.
To: GOOGLE LLC
Reel/Frame 066242/0464 →
Continuity (8)
Continuation 17867161 · Jul 18, 2022
Continuation 16939298 · Jul 27, 2020
Continuation 16395533 · Apr 26, 2019
Continuation 15966975 · Apr 30, 2018
Continuation 15464935 · Mar 21, 2017
Continuation 15228488 · Aug 4, 2016
Continuation 14447737 · Jul 31, 2014
Related Publication 20240135928A1 · Apr 25, 2024
References Cited (49)
US 8195460B2 · Degani et al. · 2012 [cited by applicant]
US 8682666B2 · Degani et al. · 2014 [cited by applicant]
US 8685666B2 · Cao · 2014 [cited by applicant]
US 9418663B2 · Chen et al. · 2016 [cited by applicant]
US 9601115B2 · Chen et al. · 2017 [cited by applicant]
US 9640180B2 · Chen et al. · 2017 [cited by applicant]
US 9666185B2 · Goussard · 2017 [cited by examiner]
US 9858923B2 · Wasserblat · 2018 [cited by examiner]
US 9997158B2 · Chen et al. · 2018 [cited by applicant]
US 10276170B2 · Gruber · 2019 [cited by examiner]
US 10325595B2 · Chen et al. · 2019 [cited by applicant]
US 10403273B2 · Lee · 2019 [cited by examiner]
US 10672397B2 · Lee · 2020 [cited by examiner]
US 10726840B2 · Chen · 2020 [cited by examiner]
US 10741185B2 · Gruber · 2020 [cited by examiner]
US 11423902B2 · Chen · 2022 [cited by applicant]
US 20040176958A1 · Salmenkaita et al. · 2004 [cited by applicant]
US 20040193420A1 · Kennewick et al. · 2004 [cited by applicant]
US 20070265830A1 · Sidhu et al. · 2007 [cited by applicant]
US 20090204711A1 · Binyamin · 2009 [cited by applicant]
US 20100114944A1 · Adler et al. · 2010 [cited by applicant]
US 20110282669A1 · Michaelis · 2011 [cited by applicant]
US 20120143598A1 · Bandara · 2012 [cited by applicant]
US 20120310850A1 · Zeng et al. · 2012 [cited by applicant]
US 20130086652A1 · Kavantzas et al. · 2013 [cited by applicant]
US 20130226580A1 · Witt-Ehsani · 2013 [cited by applicant]
US 20130266925A1 · Nunamaker et al. · 2013 [cited by applicant]
US 20130304457A1 · Kang et al. · 2013 [cited by applicant]
US 20130325447A1 · Levien et al. · 2013 [cited by applicant]
US 20140149121A1 · Di Fabbrizio et al. · 2014 [cited by applicant]
US 20140277735A1 · Breazeal · 2014 [cited by applicant]
US 20150169284A1 · Quast et al. · 2015 [cited by applicant]
US 20180247649A1 · Chen et al. · 2018 [cited by applicant]
US 20200357403A1 · Chen et al. · 2020 [cited by applicant]
US 20220351731A1 · Chen et al. · 2022 [cited by applicant]
DE 10127558 · 2002 [cited by applicant]
EP 2273491 · 2011 [cited by applicant]
WO 200065773 · 2000 [cited by applicant]
WO 2005038775 · 2005 [cited by applicant]
WO 2010049582 · 2010 [cited by applicant]
WO 2013155619 · 2013 [cited by applicant]
Deutsches Patent Office; Office Action issued in Application No. 112015003521.4, 17 pages; dated Dec. 20, 2023. [cited by applicant]
German Patent and Trademark Office; Office Action issued in DE112015003521.4 on Mar. 30, 2020. [cited by applicant]
“Google Now”, Wikipedia, the free encyclopedia, downloaded from the internet on Jun. 9, 2014, 6 pages, http://en.wikipedia.org/wiki/Google_Now Jun. 9, 2014. [cited by applicant]
Cassell, Justine et al. “Designing and Evaluating Conversational Interfaces With Animated Characters”, Embodied Conversational Agents, MIT Press, 2000, 5 pages 2000. [cited by applicant]
Cassell, Justine et al., “Nudge Nudge Wink Wink: Elements of Face to Face Conversation for Embodied Conversational Embodied Conversational Agents,”, Embodied Conversational Agents, MIT 2000, 29 pages 2000. [cited by applicant]
McBreen, Helen et al., “Evaluating 3D Embodied Conversational Agents In Contrasting VRML Retail Applications”, Workshop on Multimodal Communication and Context in Embodied Agents, Autonomous Agents 2001, held in conjunc… [cited by applicant]
International Preliminary Report on Patentability issued in International Application No. PCT/US2015/042048 mailed on Jan. 31, 2017, 10 pages. [cited by applicant]
International Search Report and Written Opinion in International Application No. PCT/US2015/042048, mailed Oct. 26, 2015, 12 pages. [cited by applicant]