IP Library › Granted Patent US 12,614,090
Granted Patent B2
US 12,614,090 · App. 17/872,937 · Granted Apr 28, 2026

Modeling characters that interact with users as part of a character-as-a-service implementation

Inventors: Michael Abrams (Los Angeles, CA); Eric Haseltine (Rancho Palos Verdes, CA)
Assignee: DISNEY ENTERPRISES, INC.
G06N5/04G06N3/006G06N20/00
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,614,090
App. No.
17/872,937
Granted
Apr 28, 2026
Kind
B2
Abstract

In various embodiments, a character engine models a character that interacts with users The modeling techniques include evaluating user input data that is associated with a user device to identify a user intent and an assessment domain, selecting a first set of inference algorithms from a plurality of inference algorithms based, at least in part, on the user intent and the assessment domain, and applying the user intent and the assessment domain to the first set of inference algorithms to generate a plurality of inferences. The modeling techniques further include consolidating the plurality of inferences into a consolidated inference, applying a second set of inference algorithms from the plurality of inference algorithms to the consolidated inference to generate a context-specific inference, computing the character response to the user input data based on the user input data, the context-specific inference, and data representing knowledge associated with a character, and causing the user device to output the character response to the user.

Claims (78)

1 . A computer-implemented method for generating a character response during an interaction with a user, the method comprising:

generating, by a processor, recognition data from user input data that is generated by one or more sensors associated with a user device of a user;

evaluating, by the processor, the recognition data to identify a user intent of the user and an assessment domain associated with the user intent;

selecting, from a plurality of inference algorithms based, at least in part, on the user intent and the assessment domain, a first set of at least two inference algorithms;

applying the user intent and the assessment domain to each inference algorithm in the first set of inference algorithms to generate a plurality of inferences;

consolidating the plurality of inferences into a consolidated inference by weighing the plurality of inferences based on a plurality of confidence factors;

applying a second set of inference algorithms from the plurality of inference algorithms to the consolidated inference to generate a context-specific inference, wherein the first set of inference algorithms and the second set of inference algorithms are mutually exclusive;

computing, based on the user input data, the context-specific inference, and data representing knowledge associated with a character, a content of the character response to the user input data;

determining an expression of the character response based on one or more capabilities of the user device;

configuring, by the processor, the user device to output the content of the character response and the expression of the user device to the user;

generating training data based on at least one of the recognition data, the user intent, the context-specific inference, the character response, or one or more additional interactions between the user and the character;

identifying, for an item of data included in the training data, one or more inference algorithms included in the plurality of inference algorithms and associated with (i) a data type included in the item of data or (ii) a type of inference operation performed on the item of data;

training, in an offline training mode and using at least the item of data, at least one of the one or more identified inference algorithms included in the plurality of inference algorithms to generate at least one updated inference algorithm;

generating, by the at least one updated inference algorithm, an additional content of an additional character response to additional user input data associated with the user, wherein the additional character response is generated based on at least one updated inference outputted by the at least one updated inference algorithm and at least one updated confidence factor associated with the at least one updated inference;

determining an additional expression of the additional character response based on one or more additional capabilities of an additional user device, wherein the additional expression of the additional character response differs from the expression of the character response; and

configuring the additional user device to output the additional content of the additional character response and the additional expression of the additional character response.

2 . The computer-implemented method of claim 1 , further comprising:

for each inference in the plurality of inferences, generating a corresponding merit rating,

wherein consolidating the plurality of inferences comprises selecting an inference corresponding to a highest merit rating in a plurality of merit ratings corresponding to the plurality of inferences.

3 . The computer-implemented method of claim 1 , wherein the plurality of inference algorithms includes at least one of: a theory-of-mind (TOM) system, a Markov model, a neural network, a computer vision system, or a support vector machine (SVM).

4 . The computer-implemented method of claim 1 , wherein the plurality of inferences includes a predicted user intent based on the user input data.

5 . The computer-implemented method of claim 1 , wherein the user device comprises a robot, a walk around character, or a toy.

6 . The computer-implemented method of claim 1 , wherein the data representing knowledge associated with the character includes information obtained from at least one of a World Wide Web, a script, a book, or a user-specific history.

7 . The computer-implemented method of claim 1 , wherein the expression of the character response and the additional expression of the additional character response comprise at least one of: text, speech, a gesture, a facial expression, a movement, a physical action, a sound, or an image.

8 . The computer-implemented method of claim 1 , wherein:

computing the content of the character response comprises selecting, based at least on the context-specific inference and from a plurality of sets of personality characteristics, a first set of personality characteristics for generating the content of the character response; and

the first set of personality characteristics comprises a first plurality of parameters, each parameter being associated with a personality dimension.

9 . The computer-implemented method of claim 8 , wherein:

determining the expression of the character response comprises selecting, based at least on the context-specific inference and from the plurality of sets of personality characteristics, a second set of personality characteristics for generating the expression of the character response, and

the second set of personality characteristics comprises a second plurality of parameters that is different from the first plurality of parameters.

10 . The computer-implemented method of claim 1 , wherein the assessment domain is further generated from a mode of interaction with the user associated with the user input data.

11 . The computer-implemented method of claim 1 , further comprising:

combining a set of historical data associated with the user with at least one of the consolidated inference, the context-specific inference, or the character response; and

generating a set of user-specific consolidated data.

12 . One or more non-transitory computer-readable media storing instructions that, when executed by one or more processors, cause the one or more processors to generate a character response during an interaction with a user by performing the steps of:

generating recognition data from user input data that is generated by one or more sensors associated with a user device of a user;

evaluating the recognition data to identify a user intent of the user and an assessment domain associated with the user intent;

selecting, from a plurality of inference algorithms based, at least in part, on the user intent and the assessment domain, a first set of at least two inference algorithms;

applying the user intent and the assessment domain to each inference algorithm in the first set of inference algorithms to generate a plurality of inferences;

consolidating the plurality of inferences into a consolidated inference by weighing the plurality of inferences based on a plurality of confidence factors;

applying a second set of inference algorithms from the plurality of inference algorithms to the consolidated inference to generate a context-specific inference, wherein the first set of inference algorithms and the second set of inference algorithms are mutually exclusive;

computing, based on the user input data, the context-specific inference, and data representing knowledge associated with a character, the character response to the user input data;

receiving, based on a plurality of interactions with a plurality of users, structured data that includes (i) a set of named entities and (ii) a set of relationships associated with the set of named entities;

generating training data that includes the structured data, the context-specific inference, the character response, and one or more additional interactions between the user and the character;

identifying, for an item of data included in the training data, one or more inference algorithms included in the plurality of inference algorithms and associated with (i) a data type included in the item of data or (ii) a type of inference operation performed on the item of data;

training, in an offline training mode and using at least the item of data, at least one of the one or more identified inference algorithms included in the plurality of inference algorithms to generate at least one updated inference algorithm; and

generating, by the at least one updated inference algorithm, an additional character response to additional user input data associated with the user, wherein the additional character response is generated based on at least one updated inference outputted by the at least one updated inference algorithm and at least one updated confidence factor associated with the at least one updated inference.

13 . The one or more non-transitory computer-readable media of claim 12 , further storing instructions that, when executed by the one or more processors, cause the one or more processors to perform the step of:

determining an expression of the character response based on one or more capabilities of a robot corresponding to the user device; and

configuring the robot to output the expression of the character response to the user,

wherein the expression of the character response comprises at least one of a text output, a voice output, a movement, a gesture, or a facial expression.

14 . The one or more non-transitory computer-readable media of claim 12 , wherein the plurality of inference algorithms includes at least one of: a theory-of-mind (TOM) system, a Markov model, a neural network, a computer vision system, or a support vector machine (SVM).

15 . The one or more non-transitory computer-readable media of claim 12 , wherein the plurality of inferences includes a predicted user intent based on the user input data.

16 . The one or more non-transitory computer-readable media of claim 12 , wherein the data representing knowledge associated with the character includes information obtained from at least one of a World Wide Web, a script, a book, or a user-specific history.

17 . The one or more non-transitory computer-readable media of claim 12 , wherein:

computing the character response comprises:

selecting, based at least on the context-specific inference and from a plurality of sets of personality characteristics, a first set of personality characteristics for generating content of the character response, and

selecting, based at least on the context-specific inference and from the plurality of sets of personality characteristics, a second set of personality characteristics for generating an expression of the content of the character response,

the first set of personality characteristics comprises a first plurality of parameters, each parameter being associated with a personality dimension; and

the second set of personality characteristics comprises a second plurality of parameters that is different from the first plurality of parameters.

18 . The one or more non-transitory computer-readable media of claim 12 , wherein the assessment domain is further generated from a mode of interaction with the user associated with the user input data.

19 . The one or more non-transitory computer-readable media of claim 12 , further storing instructions that, when executed by the one or more processors, cause the one or more processors to perform the steps of:

combining a set of historical data associated with the user with at least one of the consolidated inference, the context-specific inference, or the character response; and

generating a set of user-specific consolidated data.

20 . A computer-implemented method for generating a character response during an interaction with a user, the method comprising:

generating, by a processor, recognition data from user input data that is generated by one or more sensors associated with a user device of a user;

evaluating, by the processor, the recognition data to identify a user intent of the user and an assessment domain associated with the user intent;

selecting at least two of a plurality of inference algorithms based, at least in part, on the user intent and the assessment domain;

applying the user intent and the assessment domain to each of the two or more selected inference algorithms to generate a plurality of inferences;

consolidating the plurality of inferences into a consolidated inference by weighing the plurality of inferences based on a plurality of confidence factors;

determining one or more expression capabilities of the user device, wherein the one or more expression capabilities comprise at least one of a graphical processing power, a facial expression, a speech output, a movement, or a gesture;

applying, based on the one or more expression capabilities of the user device, at least one of the plurality of inference algorithms to the consolidated inference to generate a context-specific inference;

computing, based on the user input data, the context-specific inference, and data representing knowledge associated with a character, the character response to the user input data;

configuring the user device to output the character response to the user based on the one or more expression capabilities of the user device;

receiving, based on a plurality of interactions with a plurality of users, structured data that includes (i) a set of named entities and (ii) a set of relationships associated with the set of named entities;

training, in an offline training mode, at least one inference algorithm included in the plurality of inference algorithms based on training data that includes the structured data, the user input data, the character response, and one or more additional interactions between the user and the character to generate at least one updated inference algorithm, wherein the at least one inference algorithm is identified for training on a data item included in the training data, based at least on a type of inference operation performed by the at least one inference algorithm;

generating, by the at least one updated inference algorithm, an additional character response to additional user input data associated with the user, wherein the additional character response is generated based on at least one updated inference outputted by the at least one updated inference algorithm and at least one updated confidence factor associated with the at least one updated inference; and

configuring an additional user device to output the additional character response based on one or more additional expression capabilities of the additional user device, wherein an expression of the additional character response by the additional user device differs from an expression of the character response by the user device.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 5, 2022
From: ABRAMS, MICHAEL; HASELTINE, ERIC
To: DISNEY ENTERPRISES, INC.
Reel/Frame 060732/0910 →
Continuity (2)
Continuation 15373352 · Dec 8, 2016
Related Publication 20220366281A1 · Nov 17, 2022
References Cited (61)
US 6463585B1 · Hendricks et al. · 2002 [cited by applicant]
US 7146627B1 · Ismail et al. · 2006 [cited by applicant]
US 8171509B1 · Girouard et al. · 2012 [cited by applicant]
US 9038107B2 · Khoo et al. · 2015 [cited by applicant]
US 9602884B1 · Eldering et al. · 2017 [cited by applicant]
US 9785684B2 · Allen et al. · 2017 [cited by applicant]
US 10078917B1 · Gaeta et al. · 2018 [cited by applicant]
US 10357881B2 · Faridi · 2019 [cited by examiner]
US 20020059584A1 · Ferman et al. · 2002 [cited by applicant]
US 20040098747A1 · Kay et al. · 2004 [cited by applicant]
US 20050149987A1 · Boccon-Gibod et al. · 2005 [cited by applicant]
US 20060168616A1 · Candelore · 2006 [cited by applicant]
US 20070094041A1 · Coale et al. · 2007 [cited by applicant]
US 20100306655A1 · Mattingly et al. · 2010 [cited by applicant]
US 20110069940A1 · Shimy et al. · 2011 [cited by applicant]
US 20120127284A1 · Bar-Zeev et al. · 2012 [cited by applicant]
US 20120185886A1 · Charania et al. · 2012 [cited by applicant]
US 20120324492A1 · Treadwell, III et al. · 2012 [cited by applicant]
US 20140007157A1 · Harrison et al. · 2014 [cited by applicant]
US 20140143806A1 · Steinberg et al. · 2014 [cited by applicant]
US 20140279050A1 · Makar · 2014 [cited by examiner]
US 20150032766A1 · Greenbaum · 2015 [cited by applicant]
US 20150095277A1 · Tristan et al. · 2015 [cited by applicant]
US 20150113570A1 · Klarfeld et al. · 2015 [cited by applicant]
US 20150302639A1 · Malekian et al. · 2015 [cited by applicant]
US 20150328550A1 · Herzig · 2015 [cited by examiner]
US 20160151917A1 · Faridi · 2016 [cited by examiner]
US 20160294739A1 · Stoehr et al. · 2016 [cited by applicant]
US 20160378861A1 · Eledath · 2016 [cited by examiner]
US 20170076498A1 · Dakss et al. · 2017 [cited by applicant]
US 20170344532A1 · Zhou · 2017 [cited by examiner]
US 20180052885A1 · Gaskill · 2018 [cited by examiner]
US 20180150749A1 · Wu · 2018 [cited by examiner]
US 20180165596A1 · Abrams et al. · 2018 [cited by applicant]
US 20190174172A1 · Eubanks · 2019 [cited by applicant]
US 20190230387A1 · Gersten · 2019 [cited by applicant]
US 20200053416A1 · Maltar et al. · 2020 [cited by applicant]
US 20210006870A1 · Navin et al. · 2021 [cited by applicant]
US 20210076106A1 · Marten · 2021 [cited by applicant]
Bosser, “Dialogs Taking into Account Experiences Emotions and personality”, (Year: 2007). [cited by examiner]
Yu Wu, “Automatic Chatbot Knowledge Acquisition from Online Forum via Rough Set and Ensemble Learning”, IEEE, 2008 (Year: 2008). [cited by examiner]
Anne-Gwenn Bosser, “Dialogs Taking into Account Experience, Emotions and Personality”, AMC, 2007 (Year: 2007). [cited by examiner]
Wu et al., “Automatic Chatbot Knowledge Acquisition from Online Forum via Rough Set and Ensemble Learning”, FIP International Conference on Network and Parallel Computing, 2008, 5 pages. [cited by applicant]
Santangelo et al., “A Virtual Shopper Customer Assistant In Pervasive Environments”, 2007, 10 pages. [cited by applicant]
Bosser et al., “Dialogs Taking into Account Experience, Emotions and Personality”, ACM, 2007, pp. 9-12. [cited by applicant]
Galvao et al., “Persona-AIML: An Architecture Developing Chatterbots with Personality”, AAMAS, Proceedings of thE Third International Joint Conference on Autonomous Agents and Multiagent Systems, vol. 3, Jul. 2004, pp. … [cited by applicant]
Non Final Office Action received for U.S. Appl. No. 16/853,451 dated Aug. 17, 2022, 14 pages. [cited by applicant]
Galvao et al., “Persona-AIML: An Architecture Developing Chatterbots with Personality”, AAMAS, Proceedings of hE Third International Joint Conference on Autonomous Agents and Multiagent Systems, vol. 3, Jul. 2004, pp. 1… [cited by applicant]
Non Final Office Action received for U.S. Appl. No. 16/049,745 dated Jun. 28, 2019, 19 pages. [cited by applicant]
Final Office Action received for U.S. Appl. No. 16/049,745 dated Jan. 10, 2020, 22 pages. [cited by applicant]
Non Final Office Action received for U.S. Appl. No. 16/049,745 dated Jun. 26, 2020, 22 pages. [cited by applicant]
Final Office Action received for U.S. Appl. No. 16/049,745 dated Jan. 8, 2021, 22 pages. [cited by applicant]
Non Final Office Action received for U.S. Appl. No. 16/853,451 dated Mar. 4, 2021, 11 pages. [cited by applicant]
Notice of Allowance received for U.S. Appl. No. 16/049,745 dated May 18, 2021, 14 pages. [cited by applicant]
Final Office Action received for U.S. Appl. No. 16/853,451 dated Aug. 19, 2021, 19 pages. [cited by applicant]
Non Final Office Action received for U.S. Appl. No. 16/853,451 dated Dec. 8, 2021, 10 pages. [cited by applicant]
Final Office Action received for U.S. Appl. No. 16/853,451 dated Mar. 17, 2022, 19 pages. [cited by applicant]
Final Office Action received for U.S. Appl. No. 16/853,451 dated Dec. 8, 2022, 12 pages. [cited by applicant]
Final Office Action received for U.S. Appl. No. 16/853,451 dated Aug. 17, 2023, 11 pages. [cited by applicant]
Non Final Office Action for U.S. Appl. No. 16/853,451 dated Mar. 30, 2023, 11 pages. [cited by applicant]
Non Final Office Action received for U.S. Appl. No. 16/853,451 dated Dec. 7, 2023, 16 pages. [cited by applicant]