IP Library › Granted Patent US 12,456,331
Granted Patent B2
US 12,456,331 · App. 18/053,641 · Granted Oct 28, 2025

Guided CoModGaN optimization

Inventors: Zohreh Azizi (Los Angeles, CA); Surabhi Sinha (San Jose, CA); Siavash Khodadadeh (Mountain View, CA)
Assignee: ADOBE INC.
G06V40/172G06T5/77
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,456,331
App. No.
18/053,641
Granted
Oct 28, 2025
Kind
B2
Abstract

Methods for image processing are described. Embodiments of the present disclosure identifies an image generation network that includes an encoder and a decoder; prunes channels of a block of the encoder; prunes channels of a block of the decoder that is connected to the block of the encoder by a skip connection, wherein the channels of the block of the decoder are pruned based on the pruned channels of the block of the encoder; and generates an image using the image generation network based on the pruned channels of the block of the encoder and the pruned channels of the block of the decoder.

Claims (59)

1. A method for image generation comprising:

identifying, by a machine learning model, an image generation network that includes an encoder and a decoder;

pruning channels of a first layer of a block of the encoder;

pruning channels of a second layer of the block of the encoder based on the pruned channels of the first layer of the block of the encoder;

pruning channels of a block of the decoder that is connected to the block of the encoder by a skip connection, wherein the channels of the block of the decoder are pruned based on the pruned channels of the block of the encoder;

wherein pruning channels of the block of the decoder comprises:

pruning channels of a first layer of the block of the decoder based on the pruned channels of the first layer of the block of the encoder; and

pruning channels of a second layer of the block of the decoder based on the pruned channels of the second layer of the block of the encoder; and

generating an image using the image generation network based on the pruned channels of the block of the encoder and the pruned channels of the block of the decoder.

2. The method of claim 1 , further comprising:

identifying an input image and a portion of the input image for inpainting; and

inpainting the portion of the input image using the image generation network to obtain an inpainted image.

3. The method of claim 1 , further comprising:

identifying an image of a face; and

generating an anonymized image of the face using the image generation network.

4. The method of claim 1 , further comprising:

fine-tuning the image generation network based on the pruned channels of the block of the encoder and the pruned channels of the block of the decoder.

5. The method of claim 1 , further comprising:

refraining from pruning a mapping network of the image generation network, wherein the encoder and the decoder are components of a synthesis network of the image generation network.

6. The method of claim 1 , further comprising:

refraining from pruning a global encoder block of the encoder and a global decoder block of the decoder.

7. The method of claim 1 , wherein:

the block of the encoder and the block of the decoder comprise convolutional layers.

8. A non-transitory computer readable medium storing code for image processing, the code comprising instructions that, when executed by at least one processor, cause the at least one processor to perform operations for an image generation method comprising:

identifying, by a machine learning model, an image generation network that includes an encoder and a decoder;

pruning channels of a first layer of a block of the encoder;

pruning channels of a second layer of the block of the encoder based on the pruned channels of the first layer of the block of the encoder;

pruning channels of a block of the decoder that is connected to the block of the encoder by a skip connection, wherein the channels of the block of the decoder are pruned based on the pruned channels of the block of the encoder;

wherein pruning channels of the block of the decoder comprises:

pruning channels of a first layer of the block of the decoder based on the pruned channels of the first layer of the block of the encoder; and

pruning channels of a second layer of the block of the decoder based on the pruned channels of the second layer of the block of the encoder; and

generating an image using the image generation network based on the pruned channels of the block of the encoder and the pruned channels of the block of the decoder.

9. The non-transitory computer readable medium of claim 8 , the code further comprising instructions executable by the at least one processor to perform operations comprising:

identifying an input image and a portion of the input image for inpainting; and

inpainting the portion of the input image using the image generation network to obtain an inpainted image.

10. The non-transitory computer readable medium of claim 8 , the code further comprising instructions executable by the at least one processor to perform operations comprising:

identifying an image of a face; and

generating an anonymized image of the face using the image generation network.

11. The non-transitory computer readable medium of claim 8 , the code further comprising instructions executable by the at least one processor to perform operations comprising:

fine-tuning the image generation network based on the pruned channels of the block of the encoder and the pruned channels of the block of the decoder.

12. The non-transitory computer readable medium of claim 8 , the code further comprising instructions executable by the at least one processor to perform operations comprising:

refraining from pruning a mapping network of the image generation network, wherein the encoder and the decoder are components of a synthesis network of the image generation network.

13. The non-transitory computer readable medium of claim 8 , the code further comprising instructions executable by the at least one processor to perform operations comprising:

refraining from pruning a global encoder block of the encoder and a global decoder block of the decoder.

14. A system for image generation comprising:

a memory component; and

a processing device coupled to the memory component, the processing device configured to perform operations comprising:

identifying, by a machine learning model, an image generation network that includes an encoder and a decoder;

pruning channels of a first layer of a block of the encoder;

pruning channels of a second layer of the block of the encoder based on the pruned channels of the first layer of the block of the encoder;

pruning channels of a block of the decoder that is connected to the block of the encoder by a skip connection, wherein the channels of the block of the decoder are pruned based on the pruned channels of the block of the encoder;

wherein pruning channels of the block of the decoder comprises:

pruning channels of a first layer of the block of the decoder based on the pruned channels of the first layer of the block of the encoder; and

pruning channels of a second layer of the block of the decoder based on the pruned channels of the second layer of the block of the encoder; and

generating an image using the image generation network based on the pruned channels of the block of the encoder and the pruned channels of the block of the decoder.

15. The system of claim 14 , wherein:

the image generation network includes a synthesis network and a mapping network, and wherein the synthesis network includes the encoder and the decoder.

16. The system of claim 14 , further comprising:

a decomposition component configured to perform tensor decomposition on a layer of the image generation network and to compress the layer of the image generation network based on the tensor decomposition.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 8, 2022
From: AZIZI, ZOHREH; SINHA, SURABHI; KHODADADEH, SIAVASH
To: ADOBE INC.
Reel/Frame 061696/0272 →
Continuity (1)
Related Publication 20240152757A1 · May 9, 2024
References Cited (23)
US 20180260984A1 · Severenuk et al. · 2018 [cited by applicant]
US 20210142032A1 · Kim · 2021 [cited by examiner]
US 20220004875A1 · Koike-Akino · 2022 [cited by examiner]
US 20220237744A1 · Kwon et al. · 2022 [cited by applicant]
US 20230104262A1 · Lin · 2023 [cited by examiner]
US 20230162023A1 · Koike Akino · 2023 [cited by examiner]
US 20230229892A1 · Biswas · 2023 [cited by examiner]
US 20230289984A1 · Reaungamornrat · 2023 [cited by examiner]
US 20230344962A1 · Tran · 2023 [cited by examiner]
US 20240127504A1 · Melnik · 2024 [cited by examiner]
US 20250200374A1 · Graef · 2025 [cited by examiner]
CN 112215353A · 2021 [cited by applicant]
CN 114549373A · 2022 [cited by applicant]
CN 116912083A · 2023 [cited by applicant]
WO 2022126333 · 2022 [cited by applicant]
Zhou et al, UNet++: A Nested U-Net Architecture for Medical Image Segmentation, Jul. 18, 2018 (Year: 2018). [cited by examiner]
Berrar, D. P., Dubitzky, W., & Granzow, M. (Eds.), (2003), A practical approach to microarray data analysis (pp. 15-19), New York: Kluwer academic publishers, Access on the internet at https://link.springer.com/book/10.… [cited by applicant]
Dastjerdi, et al., “Guided Co-Modulated GAN for 360 {\deg} Field of View Extrapolation”, arXiv preprint arXiv:2204.07286, 2022, 18 pages. [cited by applicant]
He, et al., “Channel Pruning for Accelerating Very Deep Neural Networks”, In Proceedings of the IEEE international conference on computer vision, (2017), (pp. 1389-1397). [cited by applicant]
Karras, et al., “Analyzing and Improving the Image Quality of StyleGAN”, In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition, (2020), (pp. 8110-8119). [cited by applicant]
Malik, et al., “Low-Rank Tucker Decomposition of Large Tensors Using TensorSketch”, Advances in neural information processing systems, 31, (2018), 11 pages. [cited by applicant]
Zhao, et al., “Large Scale Image Completion Via Co-Modulated Generative Adversarial Networks”, arXiv preprint arXiv:2103.10428v1 [cs.CV] Mar. 18, 2021, 25 pages. [cited by applicant]
Combined Search and Examination Report dated Feb. 20, 2024 in corresponding Great Britain Patent Application No. 2313454.7 (7 pages). [cited by applicant]