IP Library Granted Patent US 11,893,669
Granted Patent B2
US 11,893,669 · App. 17/571,099 · Granted Feb 6, 2024

Development platform for digital humans

Inventors: Abhijit Z. Bendale (Campbell, CA); Pranav K. Mistry (Saratoga, CA); Bola Yoo (Seoul, KR); Kijeong Kwon (Seoul, KR); Simon Gibbs (San Jose, CA); Anil Unnikrishnan (Los Gatos, CA); Link Huang (Mountain View, CA)
Assignee: SAMSUNG ELECTRONICS CO., LTD.
G06T13/00G06F3/011G06F3/147G06F3/16G06T13/40G06V40/174G06V40/20G09F13/049G10L13/02G10L15/063G10L15/16G10L15/22G10L25/63H05B47/12H05B47/125G06F3/012G06F3/017G06F2203/011G10L2015/0638G10L2015/227
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,893,669
App. No.
17/571,099
Granted
Feb 6, 2024
Kind
B2
Abstract

A digital human development platform can enable a user to generate a digital human. The digital human development platform can receive user input specifying a dialogue for the digital human and one or more behaviors for the digital human, the one or more specified behaviors corresponding with one or more portions of the dialog on a common timeline. Scene data can be generated with the digital human development platform by merging the one or more behaviors with one or more portions of the dialogue based on times of the one or more behaviors and the one or more portions of the dialog on the common timeline.

Claims (54)

1. A method, comprising:

selecting, with a development platform, a digital human;

receiving, with the development platform, user input specifying a dialog for the digital human and one or more behaviors for the digital human corresponding with one or more portions of the dialog on a common timeline;

wherein the dialog includes words to be spoken by the digital human in response to one or more predetermined cues received during an interactive dialog with an individual; and

generating scene data, with the development platform, by merging the one or more behaviors with the one or more portions of the dialog based on times of the one or more behaviors and the one or more portions of the dialog on the common timeline;

wherein the scene data is executable by a device to render the digital human and engage in the interactive dialog with the individual based on the one or more predetermined cues from the individual as received by the device during the interactive dialog.

2. The method of claim 1 , wherein

the selecting comprises selecting a visual representation of the digital human from a plurality of visual representations corresponding to different digital humans electronically stored in a database communicatively coupled with the development platform.

3. The method of claim 1 , wherein

the selecting comprises selecting a voice type corresponding to the digital human for audibly rendering the dialog.

4. The method of claim 1 , wherein

the selecting comprises selecting a language corresponding to the digital human for audibly rendering the dialog.

5. The method of claim 1 , wherein

the merging comprises determining individual time segments of the common timeline during which distinct portions of the dialog are rendered and editing at least one of the one or more behaviors within each of the individual time segments.

6. The method of claim 1 , wherein

the one or more behaviors comprise one or more facial expressions rendered by the digital human, each facial expression of the one or more facial expressions is timed based on the common timeline to correspond to a time interval including a select portion of the dialog or an absence of dialog between two consecutive portions of the dialog.

7. The method of claim 1 , wherein

the one or more behaviors comprise one or more gestures made by the digital human, each gesture of the one or more gestures corresponding to a select portion of the dialog or an absence of dialog between two consecutive portions of the dialog.

8. The method of claim 1 , further comprising:

editing the dialog, wherein the editing provides at least one of an inflection, a stress, a speech rate, or a tone of one or more words of the dialog.

9. The method of claim 1 , further comprising:

generating a background against which the digital human is visually rendered.

10. The method of claim 9 , further comprising:

positioning the digital human against the background.

11. A system, comprising:

a processor configured to initiate operations including:

selecting a digital human;

receiving user input specifying a dialog for the digital human and one or more behaviors for the digital human corresponding with one or more portions of the dialog on a common timeline;

wherein the dialog includes words to be spoken by the digital human in response to one or more predetermined cues received during an interactive dialog with an individual; and

generating scene data by merging the one or more behaviors with the one or more portions of the dialog based on times of the one or more behaviors and the one or more portions of the dialog on the common timeline;

wherein the scene data is executable by a device to render the digital human and engage in the interactive dialog with the individual based on the one or more predetermined cues from the individual as received by the device during the interactive dialog.

12. The system of claim 11 , wherein

the selecting comprises selecting a visual representation of the digital human from a plurality of visual representations corresponding to different digital humans electronically stored in a database communicatively coupled with the system.

13. The system of claim 11 , wherein

the selecting comprises selecting a voice type corresponding to the digital human for audibly rendering the dialog.

14. The system of claim 11 , wherein

the selecting comprises selecting a language corresponding to the digital human for audibly rendering the dialog.

15. The system of claim 11 , wherein

the merging comprises determining individual time segments of the common timeline during which distinct portions of the dialog are rendered and editing at least one of the one or more behaviors within each of the individual time segments.

16. The system of claim 11 , wherein

the one or more behaviors comprise one or more facial expressions rendered by the digital human, each facial expression of the one or more facial expressions is timed based on the common timeline to correspond to a time interval including a select portion of the dialog or an absence of dialog between two consecutive portions of the dialog.

17. The system of claim 11 , wherein

the one or more behaviors comprise one or more gestures made by the digital human, each gesture of the one or more gestures corresponding to a select portion of the dialog or an absence of dialog between two consecutive portions of the dialog.

18. The system of claim 11 , wherein the processor is configured to initiate operations further including:

editing the dialog, wherein the editing provides at least one of an inflection, a stress, a speech rate, or a tone of one or more words of the dialog.

19. The system of claim 11 , wherein the processor is configured to initiate operations further including:

generating a background against which the digital human is visually rendered.

20. A computer program product, the computer program product comprising:

one or more computer-readable storage media and program instructions collectively stored on the one or more computer-readable storage media, the program instructions executable by a processor to cause the processor to initiate operations including:

selecting a digital human;

receiving user input specifying a dialog for the digital human and one or more behaviors for the digital human corresponding with one or more portions of the dialog on a common timeline;

wherein the dialog includes words to be spoken by the digital human in response to one or more predetermined cues received during an interactive dialog with an individual; and

generating scene data by merging the one or more behaviors with one or more portions of the dialog based on times of the one or more behaviors and the one or more portions of the dialog on the common timeline;

wherein the scene data is executable by a device to render the digital human and engage in the interactive dialog with the individual based on the one or more predetermined cues from the individual as received by the device during the interactive dialog.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 8, 2023
From: BENDALE, ABHIJIT Z.; MISTRY, PRANAV K.; YOO, BOLA; KWON, KIJEONG; GIBBS, SIMON; UNNIKRISHNAN, ANIL; HUANG, LINK
To: SAMSUNG ELECTRONICS CO., LTD.
Reel/Frame 065496/0721 →
Continuity (5)
Provisional Application 63135855 · Jan 11, 2021
Provisional Application 63135526 · Jan 8, 2021
Provisional Application 63135516 · Jan 8, 2021
Provisional Application 63135505 · Jan 8, 2021
Related Publication 20220222883A1 · Jul 14, 2022
Cited By (1)
US 12,217,341