IP Library › Granted Patent US 12,118,320
Granted Patent B2
US 12,118,320 · App. 18/140,136 · Granted Oct 15, 2024

Controlling generative language models for artificial intelligence characters

Inventors: Ilya Gelfenbeyn (Palo Alto, CA); Mikhail Ermolenko (Mountain View, CA); Kylan Gibbs (San Francisco, CA)
Assignee: Theai, Inc.
G06F40/35G06F40/20
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,118,320
App. No.
18/140,136
Granted
Oct 15, 2024
Kind
B2
Abstract

Systems and methods for conducting communications between a user and an Artificial Intelligence (AI) character model are provided. An example method includes determining a context of a dialog between the AI character model and the user, the context being determined based on a data stream received from a client-side computing device associated with the user; receiving a message of the user in the dialog; and generating, based on the context and the message, an input to a language model configured to predict a response to the message; providing the input to the language model to obtain the response; and transmitting the response to the client-side computing device, where the client-side computing device presents the response to the user. The input to the language model includes the message expanded by a keyword associated with the context. The context includes an intent of the user and an emotional state of the user.

Claims (52)

1. A method for conducting communications between a user and an Artificial Intelligence (AI) character model, the method being implemented with a processor of a computing platform providing the AI character model, the method comprising:

determining, by the processor, based on a data stream received from a client-side computing device associated with the user, a context of a dialog between the AI character model and the user in a virtual environment, the AI character model and the virtual environment being displayed on the client-side computing device by an application running on the client-side computing device, the application communicating with the computing platform using a pre-determined communication protocol associated with the application;

receiving, by the processor, a message from the user in the dialog;

passing, by the processor, the context and the message through a plurality of heuristics machine learning models to produce intermediate outputs of the plurality of heuristics machine learning models;

upon the producing the intermediate outputs of the plurality of heuristics machine learning models, composing, by the processor and based on the intermediate outputs of the plurality of heuristics machine learning models and a predetermined templated format ingestible by a language model, dialogue prompts being an input to the language model configured to predict a response to the message, the context including parameters of an event occurring in the virtual environment independently of the dialog between the AI character model and the user;

upon the composing the dialogue prompts, providing, by the processor, the input to the language model to obtain the response, wherein the language model is configured to:

process the input to form a request for a large language model, the processing including at least one of classifying and adjusting the message based on the context;

provide the request to the large language model;

receive, from the large language model, a model response to the request; and

modify the model response from the large language model based on predetermined criteria to form the response, wherein the modification includes adding, to the model response, a specific word corresponding to a current emotional state of the AI character model; and

transmitting, by the processor, the response to the client-side computing device, wherein the client-side computing device presents the response to the user.

2. The method of claim 1 , wherein the language model includes a generative language model configured to predict, based on a sequence of words, at least one further word to follow the sequence of words.

3. The method of claim 1 , wherein the input to the language model includes the message expanded by at least one keyword associated with the context or structured data.

4. The method of claim 1 , wherein the input to the language model includes the message and a parameter associated with the language model, the parameter being determined based on the context.

5. The method of claim 1 , wherein the input includes third party data retrieved, based on the context, from one or more online web resources.

6. The method of claim 1 , wherein the context includes parameters of a scene associated with the AI character model in the virtual environment generated by the client-side computing device.

7. The method of claim 1 , wherein the context includes an intent of the user during the dialog.

8. The method of claim 1 , wherein the context includes an intent of the AI character model during the dialog.

9. The method of claim 1 , wherein the context includes an emotional state of the user.

10. The method of claim 1 , wherein the context includes the current emotional state of the AI character model.

11. A computing platform for providing an Artificial Intelligence (AI) character model, the computing platform comprising:

a processor; and

a memory storing instructions that, when executed by the processor, configure the computing platform to:

determine, based on a data stream received from a client-side computing device associated with a user, a context of a dialog between the AI character model and the user in a virtual environment, the AI character model and the virtual environment being displayed on the client-side computing device by an application running on the client-side computing device, the application communicating with the computing platform using a pre-determined communication protocol associated with the application;

receive a message of the user in the dialog;

pass the context and the message through a plurality of heuristics machine learning models to produce intermediate outputs of the plurality of heuristics machine learning models;

upon the producing the intermediate outputs of the plurality of heuristics machine learning models, compose, based on the intermediate outputs of the plurality of heuristics machine learning models and a predetermined templated format ingestible by a language model, dialogue prompts being an input to the language model configured to predict a response to the message, the context including parameters of an event occurring in the virtual environment independently of the dialog between the AI character model and the user;

upon the composing the dialogue prompts, provide the input to the language model to obtain the response, wherein the language model is configured to:

process the input to form a request for a large language model, the processing including at least one of classifying and adjusting the message based on the context;

provide the request to the large language model;

receive, from the large language model, a model response to the request; and

modify the model response from the large language model based on predetermined criteria to form the response, wherein the modification includes adding, to the model response, a specific word corresponding to a current emotional state of the AI character model; and

transmit the response to the client-side computing device, wherein the client-side computing device presents the response to the user.

12. The computing platform of claim 11 , wherein the language model includes a generative language model configured to predict, based on a sequence of words, at least one further word to follow the sequence of words.

13. The computing platform of claim 11 , wherein the input to the language model includes the message expanded by at least one keyword associated with the context.

14. The computing platform of claim 11 , wherein the input to the language model includes the message and a parameter associated with the language model, the parameter being determined based on the context.

15. The computing platform of claim 11 , wherein the input includes third party data retrieved, based on the context, from one or more online web resources.

16. The computing platform of claim 11 , wherein the context includes parameters of a scene associated with the AI character model in the virtual environment generated by the client-side computing device.

17. The computing platform of claim 11 , wherein the context includes an intent of the user during the dialog.

18. The computing platform of claim 11 , wherein the context includes an intent of the AI character model during the dialog.

19. The computing platform of claim 11 , wherein the context includes an emotional state of the user.

20. A non-transitory computer-readable storage medium, the computer-readable storage medium including instructions that, when executed by a processor of a computing platform providing an Artificial Intelligence (AI) character model, cause the computing platform to:

determine, based on a data stream received from a client-side computing device associated with a user, a context of a dialog between the AI character model and the user in a virtual environment, the AI character model and the virtual environment being displayed on the client-side computing device by an application running on the client-side computing device, the application communicating with the computing platform using a pre-determined communication protocol associated with the application;

receive a message of the user in the dialog;

pass the context and the message through a plurality of heuristics machine learning models to produce intermediate outputs of the plurality of heuristics machine learning models;

upon the producing the intermediate outputs of the plurality of heuristics machine learning models, compose, based on the intermediate outputs of the plurality of heuristics machine learning models and a predetermined templated format ingestible by a language model, dialogue prompts being an input to the language model configured to predict a response to the message, the context including parameters of an event occurring in the virtual environment independently of the dialog between the AI character model and the user;

upon the composing the dialogue prompts, provide the input to the language model to obtain the response, wherein the language model is configured to:

process the input to form a request for a large language model, the processing including at least one of classifying and adjusting the message based on the context;

provide the request to the large language model;

receive, from the large language model, a model response to the request; and

modify the model response from the large language model based on predetermined criteria to form the response, wherein the modification includes adding, to the model response, a specific word corresponding to a current emotional state of the AI character model; and

transmit the response to the client-side computing device, wherein the client-side computing device presents the response to the user.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 27, 2023
From: GELFENBEYN, ILYA; ERMOLENKO, MIKHAIL; GIBBS, KYLAN
To: THEAI, INC.
Reel/Frame 063462/0786 →
Continuity (2)
Provisional Application 63335856 · Apr 28, 2022
Related Publication 20230351118A1 · Nov 2, 2023
Cited By (7)
US 12,405,580 US 12,499,882 US 12,632,913 US 12,657,643 US 12,731,201 US 12,731,202 US 12,743,735