IP Library Granted Patent US 12,254,561
Granted Patent B2
US 12,254,561 · App. 17/706,420 · Granted Mar 18, 2025

Producing a digital image representation of a body

Inventors: Evan Smyth (La Crescenta, CA); Gil Spencer (Incline Village, NV)
Assignee: SPREEAI CORPORATION
G06T15/205G06T7/70G06T13/40G06T2207/20084
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,254,561
App. No.
17/706,420
Granted
Mar 18, 2025
Kind
B2
Abstract

Methods and apparati for using an avatar ( 2 ) to help produce a digital image representation ( 17 ) of a body ( 1 ). A method embodiment of the present invention comprises the steps of capturing ( 21 ) measurements of the body ( 1 ) and combining the measurements with an avatar ( 2 ) to produce a set of avatar images ( 11 ); invoking ( 22 ) a neural network ( 3 ) to generate a set of photorealistic face and hair images ( 12 ); combining ( 26 ) the avatar images ( 11 ) and the face and hair images ( 12 ) to produce a set of composite images ( 16 ); and converting ( 27 ) the set of composite images ( 16 ) into a set of final images ( 17 ) representing the body ( 1 ).

Claims (32)

1. A computer-implemented method for using an avatar to produce a digital image representation of a body of a user, said method comprising the steps of:

capturing, from one or more input images of the body from a video, measurements of the body and combining the measurements with an avatar to produce a set of avatar images;

invoking a neural network to generate a set of photorealistic face and hair images of the user;

combining the avatar images and the face and hair images to produce a set of composite images; and

converting the set of composite images into a set of final images representing the body by inputting, by the neural network, data describing an image sequence and matte representing a boundary between a head and a neck to facilitate accurate blending of image inputs.

2. The method of claim 1 wherein each image is a single still frame.

3. The method of claim 1 wherein the combining step further combines foreground images in addition to combining the avatar images and the face and hair images.

4. The method of claim 1 wherein the user is a human person.

5. The method of claim 4 wherein the digital image representation is a representation of the person's face, hair, body, garments, and shoes.

6. The method of claim 1 wherein the avatar images comprise at least one of garments and shoes.

7. The method of claim 1 wherein the step of invoking a neural network takes into account skin tone of the user, as well as illumination of the user's face and hair.

8. The method of claim 1 wherein the neural network inputs data describing an image sequence representing successive positions of an avatar head during playing of an avatar video.

9. The method of claim 1 wherein the neural network inputs data describing a sequence representing desired hair motion during playing of an avatar video.

10. The method of claim 1 wherein the neural network inputs data describing an image sequence representing illumination on a head area during playing of an avatar video.

11. The method of claim 1 wherein inputs to the neural network comprise a set of face and hair images.

12. The method of claim 1 wherein the neural network produces, in addition to face and hair images, an entire representation of the body, including the head of the body.

13. The method of claim 12 wherein the neural network produces measurements of the body.

14. Apparatus for producing a digital image representation of a body, said apparatus comprising one or more processors configured to:

capture, from one or more input images of the body from a video, measurements of the body and combine the measurements with an avatar to produce a set of avatar images;

generate, using a neural network, a set of photorealistic face and hair images;

combine the avatar images and the face and hair images to produce a set of composite images; and

convert the set of composite images into a set of final images representing the body by inputting, by the neural network, data describing an image sequence and matte representing a boundary between a head and a neck to facilitate accurate blending of image inputs.

15. The apparatus of claim 14 , wherein the one or more processors are further configured to convert a background scene into a set of background images and combine the set of background images along with the avatar images and the face and hair images.

16. The apparatus of claim 14 wherein the one or more processors are further configured to combine at least one of garments, shoes, accessories, background, jewelry, hats, eyeglasses, sunglasses, belts, purses, handbags, and backpacks associated with the body.

17. The apparatus of claim 14 wherein the one or more processors are further configured to add an audio track to the set of composite images.

18. One or more non-transitory computer-readable storage media storing instructions that, responsive to execution by a processing device, causes the processing device to:

capture, from one or more input images of the body from a video, measurements of the body and combine the measurements with an avatar to produce a set of avatar images;

invoke a neural network to generate a set of photorealistic face and hair images of the user;

combine the avatar images and the face and hair images to produce a set of composite images; and

convert the set of composite images into a set of final images representing the body by inputting, by the neural network, data describing an image sequence and matte representing a boundary between a head and a neck to facilitate accurate blending of image inputs.

19. The one or more non-transitory computer-readable storage media of claim 18 , wherein the instructions, responsive to execution by the processing device, further cause the processing device to convert a background scene into a set of background images and combine the set of background images along with the avatar images and the face and hair images.

20. The one or more non-transitory computer-readable storage media of claim 18 , wherein the instructions, responsive to execution by the processing device, further cause the processing device to combine at least one of garments, shoes, accessories, background, jewelry, hats, eyeglasses, sunglasses, belts, purses, handbags, and backpacks associated with the body.

Assignments (2)
CHANGE OF NAME Recorded Nov 29, 2023
From: SPREE3D CORPORATION
To: SPREEAI CORPORATION
Reel/Frame 065714/0419 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 31, 2022
From: SMYTH, EVAN; SPENCER, GIL
To: SPREE3D CORPORATION
Reel/Frame 059460/0728 →
Continuity (3)
Continuation In Part 17231325 · Apr 15, 2021
Provisional Application 63142294 · Jan 27, 2021
Related Publication 20220237857A1 · Jul 28, 2022
References Cited (68)
US 6466215B1 · Matsuda et al. · 2002 [cited by applicant]
US 6546309B1 · Gazzuolo · 2003 [cited by applicant]
US 6731287B1 · Erdem · 2004 [cited by applicant]
US 10255681B2 · Price et al. · 2019 [cited by applicant]
US 10936853B1 · Sethi · 2021 [cited by examiner]
US 20070091085A1 · Wang et al. · 2007 [cited by applicant]
US 20070188502A1 · Bishop · 2007 [cited by applicant]
US 20090066700A1 · Harding et al. · 2009 [cited by applicant]
US 20130100140A1 · Ye et al. · 2013 [cited by applicant]
US 20130314412A1 · Gravois et al. · 2013 [cited by applicant]
US 20160086500A1 · Kaleal, III · 2016 [cited by examiner]
US 20160163084A1 · Corazza et al. · 2016 [cited by applicant]
US 20160247017A1 · Sareen et al. · 2016 [cited by applicant]
US 20160284018A1 · Adeyoola et al. · 2016 [cited by applicant]
US 20170004657A1 · Zagel et al. · 2017 [cited by applicant]
US 20170080346A1 · Abbas · 2017 [cited by applicant]
US 20180047200A1 · O'Hara et al. · 2018 [cited by applicant]
US 20180197347A1 · Tomizuka · 2018 [cited by applicant]
US 20180240280A1 · Chen et al. · 2018 [cited by applicant]
US 20180240281A1 · Vincelette · 2018 [cited by applicant]
US 20190035149A1 · Chen · 2019 [cited by examiner]
US 20190156541A1 · Isgar · 2019 [cited by applicant]
US 20190265945A1 · Newell · 2019 [cited by examiner]
US 20190287301A1 · Colbert · 2019 [cited by applicant]
US 20190371032A1 · Scapel et al. · 2019 [cited by applicant]
US 20200066029A1 · Chen · 2020 [cited by examiner]
US 20200126316A1 · Sharma et al. · 2020 [cited by applicant]
US 20200234508A1 · Shaburov et al. · 2020 [cited by applicant]
US 20200258280A1 · Park et al. · 2020 [cited by applicant]
US 20200294294A1 · Petriv et al. · 2020 [cited by applicant]
US 20200306640A1 · Kolen et al. · 2020 [cited by applicant]
US 20200320769A1 · Chen et al. · 2020 [cited by applicant]
US 20200334867A1 · Chen et al. · 2020 [cited by applicant]
US 20200346420A1 · Friedrich · 2020 [cited by applicant]
US 20200364533A1 · Sareen et al. · 2020 [cited by applicant]
US 20200402307A1 · Tanwer et al. · 2020 [cited by applicant]
US 20210049811A1 · Fedyukov et al. · 2021 [cited by applicant]
US 20210074005A1 · Xie · 2021 [cited by examiner]
US 20210150187A1 · Karras et al. · 2021 [cited by applicant]
US 20210303919A1 · Niu · 2021 [cited by applicant]
US 20210398337A1 · McDuff · 2021 [cited by examiner]
US 20220122344A1 · Liu · 2022 [cited by examiner]
CN 110930500A · 2020 [cited by applicant]
WO 2014161429A1 · 2014 [cited by applicant]
WO 2017029488A2 · 2017 [cited by applicant]
WO 2017143392A1 · 2017 [cited by applicant]
WO 2018089039A1 · 2018 [cited by applicant]
WO 2018154331A1 · 2018 [cited by applicant]
WO 2019050808A1 · 2019 [cited by applicant]
WO 2019164266A1 · 2019 [cited by applicant]
WO 2020038254A1 · 2020 [cited by applicant]
Kim et al., “Deep Video Portraits”, ACM Trans. Graph, vol. 37, No. 4, Article 163; published online May 29, 2018, pp. 163:1-14; Association for Computing Machinery, U.S.A. https://arxiv.org/pdf/1805.11714.pdf. [cited by applicant]
Lewis et al., “Pose Space Deformation: A Unified Approach to Shape Interpolation and Skeleton-Driven Deformation”, SIGGRAPH 2000, New Orleans, Louisiana, USA, pp. 165-172. [cited by applicant]
Neophytou et al., “Shape and Pose Space Deformation for Subject Specific Animation”, Centre for Vision Speech and Signal Processing (CVSSP), University of Surrey, Guildford, United Kingdom; IEEE Conference Publication, … [cited by applicant]
“Pose space deformation,” article in Wikipedia, downloaded Jul. 29, 2022, 2 pages. [cited by applicant]
Burkov, “Neural Head Reenactment with Latent Pose Descriptors”, Procedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2020, pp. 13766-13795; Oct. 30, 2020. https://arxiv.org/pdf/2004.12… [cited by applicant]
Deng, “Disentangled and Controllable Face Image Generation via 3D Imitative-Contrastive Learning”, Computer Vision Foundation Conference, Procedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition … [cited by applicant]
Tripathy, “ICface: Interpretable and Controllable Face Reenactment Using GANs”, Procedings of the IEEE/CVF Winter Conference on Applications of Computer Vision (WACV) 2020, pp. 3385-3394; Jan. 17, 2020. https://arxiv.or… [cited by applicant]
Huang, “Learning Identity-Invariant Motion Representations for Cross-ID Face Reenactment”, Procedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2020, pp. 7084-7092, 2020, open access v… [cited by applicant]
Thies, “Face2Face: Real-Time Face Capture and Reenactment of RGB Videos”, Jul. 29, 2020; abstract of this paper published in 2016 by IEEE at https://ieeexplore.ieee.org/document/7780631. [cited by applicant]
Zhao, “Joint face alignment and segmentation via deep multi-task learning”, published in Multimedia Tools and Applications 78, 13131-13148 (2019), published by Springer Nature, 1 New York Plaza, Suite 4600, New York, NY… [cited by applicant]
Li, “FaceShifter: Towards High Fidelity and Occlusion Aware Face Swapping”, Peking University and Microsoft Research, Sep. 15, 2020. [email protected] and [email protected]; pdf version availab… [cited by applicant]
Nirkin, “FSGAN: Subject Agnostic Face Swapping and Reenactment”, Computer Vision Foundation, ICCV 2019, open access version, Aug. 2019, https://arxiv.org/pdf/1908.05932.pdf ; also published in Proceedings of the IEEE In… [cited by applicant]
Naruniec, “High-Resolution Neural Face Swapping for Visual Effects”, Eurographics Symposium on Rendering 2020, vol. 39 (2020), No. 4. https://studios.disneyresearch.com/wp-content/uploads/2020/06/High-Resolution-Neural-… [cited by applicant]
Wawrzonowski et al. “Mobile devices' GPUs in cloth dynamics simulation”, Proceedings of the Federated Conference on Computer Science and Information Systems, Prague, Czech Republic, 2017, pp. 1283-1290. Retrieved on Feb… [cited by applicant]
International Preliminary Report on Patentability (issued by the USPTO/RO as the designated IPEA, after the filing of an Article 34 Amendment) mailed Oct. 11, 2023 for PCT/US2022/022180 with an international filing date… [cited by applicant]
Goes, F. et al., “Garment Refitting for Digital Characters”, SIGGRAPH '20 Talks, Aug. 17, 2020, Virtual Event, USA; 2 pages. ACM ISBN 978-1-4503-7971-7/20/08. https://doi.org/10.1145/3388767.3407348. [cited by applicant]
22788638.9 , “EP Search Report”, EP Application No. 22788638.9, Dec. 13, 2024, 7 pages. [cited by applicant]