IP Library › Granted Patent US 12,361,614
Granted Patent B2
US 12,361,614 · App. 17/804,268 · Granted Jul 15, 2025

Protecting image features in stylized representations of a source image

Inventors: Konstantin Gudkov (London, GB); Sergey Demyanov (Santa Monica, CA); Andrey Alejandrovich Gomez Zharkov (London, GB); Fedor Zhdanov (London, GB); Vadim Velicodnii (London, GB)
Assignee: Snap Inc.
G06T11/60G06N20/20G06T2210/44
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,361,614
App. No.
17/804,268
Granted
Jul 15, 2025
Kind
B2
Abstract

Systems and methods herein describe an image stylization system. The image stylization system accesses a set of images corresponding to a target domain style, generates a set of paired images using a first machine learning model, analyze the generated set of paired images using a second machine learning model trained to analyze the generated set of paired images based on a plurality of protected feature criteria, determines a set of image transformations for the generated set of pairs, generates a transformed set of paired images by performing the set of image transformations on the set of paired images, and generates stylized images corresponding to the target domain style using a supervised image translation model trained on the transformed set of paired images.

Claims (43)

1. A method comprising:

generating a set of paired images using a first machine learning model, the set of paired images comprising a first image corresponding to a source image and a second image corresponding to a target domain style;

analyzing the generated set of paired images using a second machine learning model trained to analyze the generated set of paired images based on a plurality of protected feature criteria;

based on the analyzing, determining a set of image transformations for the generated set of paired images;

generating a transformed set of paired images, the generating comprising performing the set of image transformations on the generated set of paired images; and

generating one or more stylized images corresponding to the target domain style, the one or more stylized images generated using a supervised image translation model trained on the transformed set of paired images.

2. The method of claim 1 , wherein the source image comprises a human face.

3. The method of claim 1 , wherein the set of image transformations comprises: image processing and image-to-image translation.

4. The method of claim 1 , wherein each protected feature criterion of the plurality of protected feature criteria is an unbiased classifier for a protected feature.

5. The method of claim 4 , wherein the protected feature is a feature of the first image that is preserved during transformation of the first image to the second image.

6. The method of claim 1 , wherein the target domain style is an anime image style.

7. The method of claim 1 , wherein analyzing the generated set of paired images further comprises:

accessing a set of images corresponding to the target domain style, the accessed set of images different from the set of paired images; and

generating an adjusted set of images, the generating comprising, augmenting the accessed set of images corresponding to the target domain style with a set of new images corresponding to the target domain style.

8. The method of claim 7 , wherein generating the adjusted set of images further comprises:

removing paired images from the accessed set of images, wherein the removed paired images fail to meet at least one protected feature criteria of the plurality of protected feature criteria.

9. The method of claim 7 , wherein the first machine learning model is retrained on the adjusted set of images.

10. A system comprising:

a processor; and

a memory storing instructions that, when executed by the processor, configure the system to perform operations comprising:

generating a set of paired images using a first machine learning model, the set of paired images comprising a first image corresponding to a source image and a second image corresponding to a target domain style;

analyzing the generated set of paired images using a second machine learning model trained to analyze the generated set of paired images based on a plurality of protected feature criteria;

based on the analyzing, determining a set of image transformations for the generated set of paired images;

generating a transformed set of paired images, the generating comprising performing the set of image transformations on the generated set of paired images; and

generating one or more stylized images corresponding to the target domain style, the one or more stylized images generated using a supervised image translation model trained on the transformed set of paired images.

11. The system of claim 10 , wherein the first image is a source image comprising a human face.

12. The system of claim 10 , wherein the set of image transformations comprises: image processing and image-to-image translation.

13. The system of claim 10 , wherein each protected feature criterion of the plurality of protected feature criteria is an unbiased classifier for a protected feature.

14. The system of claim 13 , wherein the protected feature is a feature of the first image that is preserved during transformation of the first image to the second image.

15. The system of claim 13 , wherein the protected feature is provided as input to the first machine learning model.

16. The system of claim 10 , wherein the target domain style is an anime image style.

17. The system of claim 10 , wherein analyzing the generated set of paired images further comprises:

accessing a set of images corresponding to the target domain style, the accessed set of images different from the set of paired images; and

generating an adjusted set of images, the generating comprising, augmenting the accessed set of images corresponding to the target domain style with a set of new images corresponding to the target domain style.

18. The system of claim 17 , wherein generating the adjusted set of paired images further comprises:

removing paired images from the accessed set of images, wherein the removed paired images fail to meet at least one protected feature criteria of the plurality of protected feature criteria.

19. The system of claim 17 , wherein the first machine learning model is retrained on the adjusted set of images.

20. A non-transitory processor-readable storage medium storing processor executable instructions that, when executed by one or more processors of a machine, cause the machine to perform operations comprising:

generating a set of paired images using a first machine learning model, the set of paired images comprising a first image corresponding to a source image and a second image corresponding to a target domain style;

analyzing the generated set of paired images using a second machine learning model trained to analyze the generated set of paired images based on a plurality of protected feature criteria;

based on the analyzing, determining a set of image transformations for the generated set of paired images;

generating a transformed set of paired images, the generating comprising performing the set of image transformations on the generated set of paired images; and

generating one or more stylized images corresponding to the target domain style, the one or more stylized images generated using a supervised image translation model trained on the transformed set of paired images.

Assignments (2)
CORRECTIVE ASSIGNMENT TO CORRECT THE CORRECT THE THIRD INVENTOR'S NAME FROM "ANDREI ZHARKOV" TO "ANDREY ALEJANDROVICH GOMEZ ZHARKOV" PREVIOUSLY RECORDED AT REEL: 64887 FRAME: 206. ASSIGNOR(S) HEREBY CONFIRMS THE ASSIGNMENT. Recorded Mar 26, 2025
From: GUDKOV, KONSTANTIN; DEMYANOV, SERGEY; GOMEZ ZHARKOV, ANDREY ALEJANDROVICH; ZHDANOV, FEDOR; VELICODNII, VADIM
To: SNAP INC.
Reel/Frame 070662/0290 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Sep 13, 2023
From: GUDKOV, KONSTANTIN; DEMYANOV, SERGEY; ZHARKOV, ANDREI; ZHDANOV, FEDOR; VELICODNII, VADIM
To: SNAP INC.
Reel/Frame 064887/0206 →
Continuity (2)
Provisional Application 63295394 · Dec 30, 2021
Related Publication 20230215062A1 · Jul 6, 2023
References Cited (79)
US 10389676B2 · Rubinstein et al. · 2019 [cited by applicant]
US 11120526B1 · Demyanov et al. · 2021 [cited by applicant]
US 11205086B2 · Sriram et al. · 2021 [cited by applicant]
US 11657479B2 · Demyanov et al. · 2023 [cited by applicant]
US 11900565B2 · Demyanov et al. · 2024 [cited by applicant]
US 20130163854A1 · Cheng · 2013 [cited by applicant]
US 20180075581A1 · Shi et al. · 2018 [cited by applicant]
US 20180174052A1 · Rippel et al. · 2018 [cited by applicant]
US 20180176576A1 · Rippel et al. · 2018 [cited by applicant]
US 20180293712A1 · Vogels et al. · 2018 [cited by applicant]
US 20180308243A1 · Justice · 2018 [cited by applicant]
US 20180314716A1 · Kim et al. · 2018 [cited by applicant]
US 20180373979A1 · Wang et al. · 2018 [cited by applicant]
US 20190087660A1 · Hare et al. · 2019 [cited by applicant]
US 20190171929A1 · Abadi et al. · 2019 [cited by applicant]
US 20190180136A1 · Bousmalis et al. · 2019 [cited by applicant]
US 20190213439A1 · Liu · 2019 [cited by examiner]
US 20190244060A1 · Dundar · 2019 [cited by examiner]
US 20190258878A1 · Koivisto et al. · 2019 [cited by applicant]
US 20190304065A1 · Bousmalis et al. · 2019 [cited by applicant]
US 20190318261A1 · Deng et al. · 2019 [cited by applicant]
US 20190333198A1 · Wang · 2019 [cited by examiner]
US 20190379630A1 · Rubinstein et al. · 2019 [cited by applicant]
US 20190392624A1 · Elgammal · 2019 [cited by applicant]
US 20200053034A1 · Kozhemiak et al. · 2020 [cited by applicant]
US 20200112531A1 · Tang · 2020 [cited by applicant]
US 20200160113A1 · Zhang et al. · 2020 [cited by applicant]
US 20200193717A1 · Daly · 2020 [cited by applicant]
US 20200257985A1 · West et al. · 2020 [cited by applicant]
US 20210118112A1 · Huang · 2021 [cited by examiner]
US 20210150684A1 · Elmoznino · 2021 [cited by examiner]
US 20210232932A1 · Liu · 2021 [cited by examiner]
US 20210241498A1 · Sun · 2021 [cited by examiner]
US 20210264236A1 · Xu · 2021 [cited by examiner]
US 20210295045A1 · Li · 2021 [cited by examiner]
US 20210334595A1 · Berlin · 2021 [cited by examiner]
US 20210358164A1 · Liu · 2021 [cited by examiner]
US 20210383509A1 · Demyanov et al. · 2021 [cited by applicant]
US 20220044352A1 · Liao · 2022 [cited by examiner]
US 20220101635A1 · Koivisto et al. · 2022 [cited by applicant]
US 20230124252A1 · Liu et al. · 2023 [cited by applicant]
US 20230206398A1 · Demyanov et al. · 2023 [cited by applicant]
US 20230360294A1 · Aggarwal et al. · 2023 [cited by applicant]
US 20230410249A1 · Noh · 2023 [cited by examiner]
US 20240144515A1 · Wang · 2024 [cited by examiner]
CA 3152644 · 2022 [cited by applicant]
CN 109816589B · 2020 [cited by examiner]
CN 118525301A · 2024 [cited by applicant]
WO 2023005358 · 2023 [cited by applicant]
WO 2023129391 · 2023 [cited by applicant]
Roger et al., XGAN: Unsupervised Image-to-Image Translation for Many-to-Many Mappings, Domain Adaptation for Visual Understanding at ICML'18, 2018, pp. 1-19, arXiv:1711.05139, doi.org/10.48550/arXiv.1711.0513, AAPA furn… [cited by examiner]
Dundar et al., Domain Stylization: A Strong, Simple Baseline for Synthetic to Real Image Domain Adaptation, 2018, pp. 1-10, arXiv:1807.09384. [cited by examiner]
Dundar et al., Domain Stylization: A Fast Covariance Matching Framework Towards Domain Adaptation, in IEEE Transactions on Pattern Analysis and Machine Intelligence, vol. 43, No. 7, pp. 2360-2372, Jul. 1, 2021, doi: 10.… [cited by examiner]
Gatys et al., Image Style Transfer Using Convolutional Neural Networks, 2016 IEEE Conference on Computer Vision and Pattern Recognition, pp. 2414-2423, doi: 10.1109/CVPR.2016.265. [cited by examiner]
“International Application Serial No. PCT US2022 052980, International Search Report mailed May 2, 2023”, 4 pgs. [cited by applicant]
“International Application Serial No. PCT US2022 052980, Written Opinion mailed May 2, 2023”, 5 pgs. [cited by applicant]
Royer, Amelie, “XGAN: Unsupervised Image-to-Image Translation for Many-to-Many Mappings”, [Online] Retrieved from the internet:https: arxiv.org pdf 1711.05139.pdf, (Jul. 10, 2018), 19 pgs. [cited by applicant]
Wang, Tianying, “RoboCoDraw: Robotic Avatar Drawing with GAN-based Style Transfer and Time-efficient Path Optimization”, arxiv.org, Cornell University Library, 201 Olin Library Cornell University Ithaca, NY 14853, (Dec.… [cited by applicant]
U.S. Appl. No. 16/376,564 U.S. Pat. No. 11,120,526, Apr. 5, 2019, Deep Feature Generative Adversarial Neural Networks. [cited by applicant]
U.S. Appl. No. 17/445,362 U.S. Pat. No. 11,657,479, Aug. 18, 2021, Deep Feature Generative Adversarial Neural Networks. [cited by applicant]
U.S. Appl. No. 18/116,682 U.S. Pat. No. 11,900,565, Mar. 2, 2023, Deep Feature Generative Adversarial Neural Networks. [cited by applicant]
U.S. Appl. No. 18/478,783, filed Sep. 29, 2023, Generalizing Image Stylization Effects. [cited by applicant]
“U.S. Appl. No. 16/376,564, Non Final Office Action mailed Oct. 6, 2020”, 13 pgs. [cited by applicant]
“U.S. Appl. No. 16/376,564, Notice of Allowance mailed May 12, 2021”, 8 pgs. [cited by applicant]
“U.S. Appl. No. 16/376,564, Response filed Feb. 8, 2021 to Non Final Office Action mailed Oct. 6, 2020”, 9 pgs. [cited by applicant]
“U.S. Appl. No. 17/445,362, Non Final Office Action mailed Sep. 21, 2022”, 16 pgs. [cited by applicant]
“U.S. Appl. No. 17/445,362, Notice of Allowance mailed Jan. 17, 2023”, 8 pgs. [cited by applicant]
“U.S. Appl. No. 18/116,682, Non Final Office Action mailed Jul. 6, 2023”, 18 pgs. [cited by applicant]
“U.S. Appl. No. 18/116,682, Notice of Allowance mailed Oct. 2, 2023”, 8 pgs. [cited by applicant]
“U.S. Appl. No. 18/116,682, Response filed Sep. 22, 2023 to Non Final Office Action mailed Jul. 6, 2023”, 9 pgs. [cited by applicant]
“International Application Serial No. PCT/US2022/052980, International Preliminary Report on Patentability mailed Jul. 11, 2024”, 7 pgs. [cited by applicant]
Brunner, et al., “Symbolic Music Genre Transfer with CycleGAN”, arXiv:1809.07575v1 [cs.SD], (2018), 8 pgs. [cited by applicant]
Chen, Ying-Cong, et al., “Facelet-Bank for Fast Portrait Manipulation”, arXiv:1803.05576v3 [cs.CV], (Mar. 30, 2018), 9 pgs. [cited by applicant]
Upchurch, Paul, et al., “Deep Feature Interpolation for Image Content Changes”, arXiv:1611.05507v2 [cs.CV], (Jun. 19, 2017), 10 pgs. [cited by applicant]
“International Application Serial No. PCT US2024 048696, International Search Report mailed Dec. 9, 2024”, 4 pgs. [cited by applicant]
“International Application Serial No. PCT US2024 048696, Written Opinion mailed Dec. 9, 2024”, 11 pgs. [cited by applicant]
Aguinaldo, Angeline, “Compressing GANs using Knowledge Distillation”, arxiv.org, Cornell University Library, 201 Olin Library Cornell University Ithaca, NY 14853, (Feb. 28, 2019), 10 pgs. [cited by applicant]
“European Application Serial No. 22850809.9, Response to Communication Pursuant to Rules 161 and 162 EPC Filed Jan. 20, 2025”, 14 pages. [cited by applicant]
“U.S. Appl. No. 18/478,783, Non Final Office Action mailed May 15, 2025”, 19 pgs. [cited by applicant]