IP Library › Granted Patent US 12,260,485
Granted Patent B2
US 12,260,485 · App. 18/046,077 · Granted Mar 25, 2025

Cascaded domain bridging for image generation

Inventors: Guoxian Song (Los Angeles, CA); Shen Sang (Los Angeles, CA); Tiancheng Zhi (Los Angeles, CA); Jing Liu (Los Angeles, CA); Linjie Luo (Los Angeles, CA)
Assignee: Lemon Inc.
G06T15/02G06T7/11G06T2207/20081G06T2207/20084G06T2207/30201
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,260,485
App. No.
18/046,077
Granted
Mar 25, 2025
Kind
B2
Abstract

A method of generating a style image is described. The method includes receiving an input image of a subject. The method further includes encoding the input image using a first encoder of a generative adversarial network (GAN) to obtain a first latent code. The method further includes decoding the first latent code using a first decoder of the GAN to obtain a normalized style image of the subject, wherein the GAN is trained using a loss function according to semantic regions of the input image and the normalized style image.

Claims (32)

1. A method of generating a style image, the method comprising:

receiving an input image of a subject;

encoding the input image using a first encoder of a generative adversarial network (GAN) to obtain a first latent code;

decoding the first latent code using a first decoder of the GAN to obtain a normalized style image of the subject, wherein:

the GAN is trained using a loss function according to semantic regions of the input image and the normalized style image, and

a distribution prior of a W+ space is modeled for training the GAN by inverting a dataset of real face images using a second encoder that is pre-trained.

2. The method of claim 1 , further comprising training the GAN by inverting the dataset of real face images to obtain a plurality of latent codes.

3. The method of claim 1 , the second encoder is different from the first decoder.

4. The method of claim 1 , wherein the second encoder is a pre-trained StyleGAN encoder.

5. The method of claim 1 , wherein training the GAN further comprises performing a W+ space transfer learning from the second encoder to the first encoder.

6. The method of claim 5 , wherein performing the W+ space transfer learning comprises using a normalized exemplar set with only neutral expressions of the subject.

7. The method of claim 5 , wherein performing the W+ space transfer learning comprises using a normalized exemplar set with only neutral poses of the subject.

8. The method of claim 5 , wherein performing the W+ space transfer learning comprises using a normalized exemplar set with only neutral lighting of the subject.

9. The method of claim 1 , further comprising training the GAN using a difference between a first face segmentation model trained using real face images and a second face segmentation model using style exemplars as the loss function.

10. The method of claim 9 , wherein the semantic regions include one or more of hair regions of the subject or skin regions of the subject.

11. A system for generating a style image, the system comprising:

a processor; and

memory storing instructions that, when executed by the processor, cause the system to perform a set of operations, the set of operations comprising:

receiving an input image of a subject;

encoding the input image using a first encoder of a generative adversarial network (GAN) to obtain a first latent code;

decoding the first latent code using a first decoder of the GAN to obtain a normalized style image of the subject, wherein:

the GAN is trained using a loss function according to semantic regions of the input image and the normalized style image, and

a distribution prior of a W+ space is modeled for training the GAN by inverting a dataset of real face images using a second encoder that is pre-trained.

12. The method of claim 11 , wherein the set of operations further comprise training the GAN by inverting the dataset of real face images to obtain a plurality of latent codes.

13. The method of claim 11 , wherein the second encoder is different from the first decoder.

14. The method of claim 11 , wherein the second encoder is a pre-trained StyleGAN encoder.

15. The method of claim 11 , wherein the set of operations further comprise performing a W+ space transfer learning from the second encoder to the first encoder.

16. The method of claim 15 , wherein the set of operations further comprise using a normalized exemplar set with only neutral expressions of the subject.

17. The method of claim 15 , wherein the set of operations further comprise using a normalized exemplar set with only neutral poses of the subject.

18. The method of claim 15 , wherein the set of operations further comprise using a normalized exemplar set with only neutral lighting of the subject.

19. The method of claim 11 , wherein the set of operations further comprise training the GAN using a difference between a first face segmentation model trained using real face images and a second face segmentation model using style exemplars as the loss function.

20. The method of claim 19 , wherein the semantic regions include one or more of hair regions of the subject or skin regions of the subject.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 11, 2024
From: SONG, GUOXIAN; SANG, SHEN; ZHI, TIANCHENG; LIU, JING; LUO, LINJIE
To: BYTEDANCE INC.
Reel/Frame 069548/0927 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 11, 2024
From: BYTEDANCE INC.
To: LEMON INC.
Reel/Frame 069549/0084 →
Continuity (1)
Related Publication 20240135627A1 · Apr 25, 2024
References Cited (7)
US 20190295302A1 · Fu · 2019 [cited by examiner]
US 20210280322A1 · Frank · 2021 [cited by applicant]
US 20210358164A1 · Liu · 2021 [cited by examiner]
US 20210406682A1 · Meeyakhan Rawther · 2021 [cited by examiner]
US 20230130535A1 · Ma · 2023 [cited by applicant]
US 20230316606A1 · Qu · 2023 [cited by examiner]
Office Action dated Jun. 5, 2024 in U.S. Appl. No. 18/046,073 (39 pages). [cited by applicant]