IP Library Granted Patent US 10,475,225
Granted Patent B2
US 10,475,225 · App. 15/124,811 · Granted Nov 12, 2019

Avatar animation system

Inventors: Minje Park (Seoul, KR); Tae-Hoon Kim (Seoul, KR); Myung-Ho Ju (Seoul, KR); Jihyeon Yi (Yongin-si, KR); Xiaolu Shen (Beijing, CN); Lidan Zhang (Beijing, CN); Qiang Li (Beijing, CN)
Assignee: Intel Corporation
G06T13/40G06T7/73G06T17/20G06T2207/20084G06T2207/30201G06T2210/44
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,475,225
App. No.
15/124,811
Filed
Sep 9, 2016
Granted
Nov 12, 2019
Kind
B2
Art Unit
2612
USPC
345/473
Abstract

Avatar animation systems disclosed herein provide high quality, real-time avatar animation that is based on the varying countenance of a human face. In some example embodiments, the real-time provision of high quality avatar animation is enabled at least in part, by a multi-frame regressor that is configured to map information descriptive of facial expressions depicted in two or more images to information descriptive of a single avatar blend shape. The two or more images may be temporally sequential images. This multi-frame regressor implements a machine learning component that generates the high quality avatar animation from information descriptive of a subject's face and/or information descriptive of avatar animation frames previously generated by the multi-frame regressor. The machine learning component may be trained using a set of training images that depict human facial expressions and avatar animation authored by professional animators to reflect facial expressions depicted in the set of training images.

Claims (25)

1. An avatar animation system comprising: a memory; and at least one processor coupled to the memory and configured to: receive image data comprising three or more temporally sequential frames depicting three or more temporally sequential of facial expressions of a subject; identify a set of landmark points within each facial expression of the three or more temporally sequential facial expressions; generate a mesh based on each set of landmark points, thereby creating three or more temporally sequential meshes corresponding to the three or more temporally sequential facial expressions; generate a set of blend shape weights for a single frame of avatar animation at least in part by mapping, via a multi-frame regressor, the three or more temporally sequential meshes to the set of blend shape weights; and provide the set of blend shape weights to a user interface component, thereby causing the user interface component to render the single frame of avatar animation.

2. The avatar animation system of claim 1 , wherein the image data includes at least one of two-dimensional image data and three-dimensional image data.

3. The avatar animation system of claim 1 , wherein the at least one processor is configured to generate the set of blend shape weights at least in part by mapping, via the multi-frame regressor, the three or more temporally sequential meshes and at least one previously generated blend shape weight to the set of blend shape weights.

4. The avatar animation system of claim 1 , further comprising an avatar client component including the user interface component and configured to acquire the image data.

5. The avatar animation system of claim 1 , wherein the multi-frame regressor comprises an artificial neural network.

6. The avatar animation system of claim 5 , wherein the artificial neural network is configured to: process the three or more temporally sequential meshes via a plurality of input nodes; and generate the set of blend shape weights via a plurality, of output nodes.

7. The avatar animation system of claim 6 , wherein each input node of the plurality of input nodes is configured to receive either a coordinate value or a blend shape weight.

8. The avatar animation system of claim 6 , wherein each output node of the plurality of output nodes is configured to identify a blend shape weight.

9. A method of generating avatar animation using a system, the method comprising: receiving image data comprising three or more temporally sequential frames depicting three or more temporally sequential facial expressions of a subject; identifying a set of landmark points within each facial expression of the three or more temporally sequential facial expressions; generating a mesh based on each set of landmark points, thereby creating three or more temporally sequential meshes corresponding to the three or more temporally sequential facial expressions;

generating a set of blend shape weights for a single frame of avatar animation at least in part by mapping, via a multi-frame regressor, the three or more temporally sequential meshes to the set of blend shape weights; and providing the set of blend shape weights to a user interface component, thereby causing the user interface component to render the single frame of avatar animation.

10. The method of claim 9 , wherein generating the set of blend shape weights comprises mapping the three or more temporally sequential meshes and at least one previously generated blend shape weight to the set of blend shape weights.

11. The method of claim 9 , further comprising acquiring the image data.

12. The method of claim 9 , wherein mapping the three or more temporally sequential meshes includes mapping the three or more temporally sequential of meshes via an artificial neural network.

13. The method of claim 12 , wherein the artificial neural network includes a plurality of input nodes and a plurality of output nodes and the method further comprises: processing the three or more temporally sequential meshes via the plurality of input nodes; and generating the set of blend shape weights via the plurality of output nodes.

14. The method of claim 13 , further comprising receiving, at each input node of the plurality of input nodes, either a coordinate value or a blend shape weight.

15. The method of claim 13 , further comprising identifying, at each output node of the plurality of output nodes, a blend shape weight.

16. The method of claim 9 , wherein receiving the image data comprises receiving at least one of two-dimensional image data and three-dimensional image data.

17. A non-transient computer program product encoded with instructions that when executed by one or more processors cause a process of animating avatars to be carried out, the process comprising: receiving image data comprising three or more temporally sequential frames depicting three or more temporally sequential facial expressions of a subject; identifying a set of landmark points within each facial expression of the three or more temporally sequential facial expressions; generating a mesh based on each set of landmark points, thereby creating three or more temporally sequential meshes corresponding to the three or more temporally sequential facial expressions; generating a set of blend shape weights for a single frame of avatar animation at least in part by mapping, via a multi-frame regressor, the three or more temporally sequential meshes to the set of blend shape weights; and providing the set of blend shape weights to a user interface component, thereby causing the user interface component to render the single frame of avatar animation.

18. The computer program product of claim 17 , the process further comprising acquiring the image data.

19. The computer program product of claim 17 , wherein mapping the three or more temporally sequential meshes includes mapping the three or more temporally sequential meshes via an artificial neural network.

20. The computer program product of claim 19 , wherein the artificial neural network includes a plurality of input nodes and a plurality of output nodes and the process further comprises: processing the three or more temporally sequential meshes via the plurality of input nodes; and generating the set of blend shape weights via the plurality of output nodes.

21. The computer program product of claim 20 , the process further comprising receiving, at each input node of the plurality of input nodes, either a coordinate value or a blend shape weight.

22. The computer program product of claim 20 , the process further comprising identifying, at each output node of the plurality of output nodes, a blend shape weight.

23. The computer program product of claim 17 , wherein generating the set of blend shape weights comprises mapping the three or more temporally sequential meshes and at least one previously generated blend shape weight to the set of blend shape weights.

24. The computer program product of claim 17 , wherein the image data includes at least one of two-dimensional image data and three-dimensional image data.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 15, 2022
From: INTEL CORPORATION
To: TAHOE RESEARCH, LTD.
Reel/Frame 061175/0176 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Sep 9, 2016
From: PARK, MINJE; KIM, TAE-HOON; JU, MYUNG-HO; YI, JIHYEON; SHEN, XIAOLU; ZHANG, LIDAN; LI, QIANG
To: INTEL CORPORATION
Reel/Frame 039686/0382 →
Continuity (1)
Related Publication 20170256086A1 · Sep 7, 2017
Cited By (153)
US 1,089,291 US 12,192,854 US 12,198,281 US 12,198,287 US 12,198,372 US 12,198,398 US 12,198,664 US 12,205,295 US 12,206,635 US 12,211,159 US 12,212,614 US 12,213,028 US 12,217,374 US 12,217,453 US 12,218,893 US 12,223,156 US 12,223,672 US 12,226,001 US 12,229,860 US 12,229,901 US 12,231,709 US 12,235,991 US 12,236,512 US 12,242,708 US 12,242,979 US 12,243,173 US 12,243,266 US 12,254,577 US 12,260,450 US 12,265,692 US 12,271,536 US 12,273,308 US 12,277,632 US 12,277,638 US 12,284,146 US 12,284,698 US 12,288,273 US 12,293,009 US 12,293,433 US 12,299,004 US 12,299,256 US 12,299,775 US 12,299,830 US 12,299,832 US 12,299,905 US 12,307,564 US 12,315,495 US 12,316,589 US 12,316,597 US 12,321,577 US 12,322,021 US 12,327,277 US 12,335,213 US 12,335,575 US 12,340,064 US 12,340,453 US 12,341,736 US 12,347,013 US 12,347,045 US 12,348,467 US 12,354,353 US 12,354,355 US 12,361,627 US 12,361,652 US 12,361,934 US 12,363,056 US 12,367,616 US 12,380,618 US 12,380,649 US 12,386,485 US 12,387,405 US 12,387,436 US 12,387,444 US 12,387,447 US 12,393,318 US 12,394,154 US 12,395,456 US 12,400,101 US 12,400,389 US 12,406,416 US 12,412,205 US 12,412,347 US 12,417,562 US 12,418,504 US 12,429,953 US 12,436,598 US 12,438,837 US 12,439,223 US 12,443,325 US 12,444,138 US 12,462,507 US 12,469,090 US 12,469,196 US 12,469,273 US 12,472,435 US 12,475,621 US 12,475,658 US 12,482,131 US 12,488,548 US 12,488,551 US 12,499,483 US 12,499,626 US 12,499,638 US 12,501,233 US 12,504,287 US 12,504,866 US 12,513,098 US 12,517,626 US 12,518,437 US 12,518,738 US 12,520,101 US 12,524,128 US 12,530,408 US 12,530,847 US 12,530,852 US 12,536,751 US 12,541,929 US 12,541,930 US 12,548,267 US 12,555,274 US 12,555,310 US 12,556,659 US 12,561,914 US 12,567,102 US 12,579,204 US 12,580,784 US 12,580,980 US 12,586,562 US 12,602,842 US 12,614,354 US 12,614,359 US 12,616,903 US 12,620,188 US 12,620,216 US 12,625,026 US 12,627,627 US 12,632,890 US 12,633,073 US 12,646,266 US 12,646,268 US 12,651,292 US 12,651,409 US 12,670,643 US 12,670,674 US 12,682,508 US 12,695,847 US 12,699,802 US 12,700,187 US 12,705,009 US 12,705,840 US 12,705,875 US 12,711,380 US 12,711,538