IP Library Granted Patent US 12,457,300
Granted Patent B2
US 12,457,300 · App. 18/309,410 · Granted Oct 28, 2025

Generating stylized images on mobile devices

Inventors: Wentian Zhao (San Jose, CA); Kun Wan (San Jose, CA); Xin Lu (Mountain View, CA); Jen-Chan Jeff Chien (Saratoga, CA)
Assignee: Adobe Inc.
H04N5/2621G06T5/73G06V10/40G06V10/56G06V10/82H04N5/265H04N23/631H04N23/632
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,457,300
App. No.
18/309,410
Granted
Oct 28, 2025
Kind
B2
Abstract

Methods, systems, and non-transitory computer readable media are disclosed for generating artistic images by applying an artistic-effect to one or more frames of a video stream or digital images. In one or more embodiments, the disclosed system captures a video stream utilizing a camera of a computing device. The disclosed system deploys a distilled artistic-effect neural network on the computing device to generate an artistic version of the captured video stream at a first resolution in real time. The disclosed system can provide the artistic video stream for display via the computing device. Based on an indication of a capture event, the disclosed system utilizes the distilled artistic-effect neural network to generate an artistic image at a higher resolution than the artistic video stream. Furthermore, the disclosed system tunes and utilizes an artistic-effect patch generative adversarial neural network to modify parameters for the distilled artistic-effect neural network.

Claims (42)

1. A system comprising:

a camera;

one or more memory devices comprising an artistic-effect neural network; and

at least one processing device coupled to the one or more memory devices, the at least one processing device to perform operations comprising:

capturing, utilizing the camera, a video stream comprising a plurality of frames;

generating, in real-time utilizing the artistic-effect neural network, a synthesized artistic video stream at a first resolution by applying an artistic-effect to the plurality of frames from the video stream;

providing, for display via a viewfinder as the video stream is being captured, the synthesized artistic video stream at the first resolution; and

generating, based on an indication of a capture event and utilizing the artistic-effect neural network, an artistic image comprising a frame from the synthesized artistic video stream with the artistic-effect at a second resolution that is higher than the first resolution.

2. The system as recited in claim 1 , wherein generating, based on the indication of the capture event and utilizing the artistic-effect neural network, the artistic image comprises generating the artistic image at a 4K resolution.

3. The system as recited in claim 1 , wherein the system comprises a mobile computing device.

4. The system as recited in claim 3 , wherein the artistic-effect comprises an anime style.

5. The system as recited in claim 3 , wherein the artistic-effect neural network comprises convolutional blocks comprising a single upsampling block and efficient separable convolutions.

6. The system as recited in claim 3 , wherein the operations further comprise generating, based on an indication of a selection of a second artistic-effect, in real time utilizing a second artistic-effect neural network, a second synthesized artistic video stream at the first resolution by applying the second artistic-effect to the plurality of frames from the video stream.

7. A non-transitory computer-readable medium storing executable instructions, which when executed by a processing device, cause the processing device to perform operations comprising:

capturing, utilizing a camera, a video stream comprising a plurality of frames;

generating, utilizing an artistic-effect neural network, a synthesized artistic video stream at a first resolution by applying an artistic-effect to the plurality of frames from the video stream;

providing, for display via a viewfinder, the synthesized artistic video stream at the first resolution as the video stream is being captured; and

generating, based on an indication of a capture event and utilizing the artistic-effect neural network, an artistic image comprising a frame from the synthesized artistic video stream with the artistic-effect at a second resolution that is higher than the first resolution.

8. The non-transitory computer-readable medium of claim 7 , wherein generating, in real-time utilizing the artistic-effect neural network, the synthesized artistic video stream at the first resolution allows for a real-time display of the synthesized artistic video stream as the video stream is being captured.

9. The non-transitory computer-readable medium of claim 7 , wherein generating, utilizing the artistic-effect neural network, the artistic image at the second resolution comprises generating the artistic image at a resolution that is twice the first resolution.

10. The non-transitory computer-readable medium of claim 7 , wherein the operations further comprise generating and displaying via a graphical user interface including the viewfinder, a plurality of different artistic effect previews; and

wherein generating, utilizing the artistic-effect neural network, the synthesized artistic video stream at the first resolution by applying the artistic-effect to the plurality of frames from the video stream is in response to a selection of an artistic effect preview corresponding to an artistic effect.

11. The non-transitory computer-readable medium of claim 10 , wherein the operations further comprise:

receiving a selection of a second artistic effect preview;

generating, based on the selection of the second artistic effect preview, utilizing a second artistic-effect neural network, a second synthesized artistic video stream at the first resolution by applying a second artistic-effect to the plurality of frames from the video stream; and

replacing the display of the synthesized artistic video stream in the viewfinder with the second synthesized artistic video stream.

12. The non-transitory computer-readable medium of claim 7 , wherein the operations further comprise displaying the artistic image in place of the synthesized artistic video stream.

13. The non-transitory computer-readable medium of claim 7 , wherein generating, utilizing the artistic-effect neural network, the artistic image comprises upsampling within the artistic-effect neural network.

14. A method comprising:

capturing, utilizing a camera, a video stream comprising a plurality of frames;

generating, in real-time utilizing an artistic-effect neural network, a synthesized artistic video stream at a first resolution by applying an artistic-effect to the plurality of frames from the video stream;

providing, for display via a viewfinder as the video stream is being captured, the synthesized artistic video stream at the first resolution; and

generating, based on an indication of a capture event and utilizing the artistic-effect neural network, an artistic image comprising a frame from the synthesized artistic video stream with the artistic-effect at a second resolution that is higher than the first resolution.

15. The method of claim 14 , wherein generating, based on the indication of the capture event and utilizing the artistic-effect neural network, the artistic image comprises generating the artistic image at a 4K resolution.

16. The method of claim 14 , wherein:

capturing, utilizing the camera, the video stream comprises capturing the video stream utilizing a mobile phone; and

providing, for display via the viewfinder, the synthesized artistic video stream, comprises displaying the synthesized artistic video stream on the mobile phone.

17. The method of claim 14 , wherein generating, the artistic image with the artistic-effect comprises generating an artistic image with a painting style.

18. The method of claim 14 , further comprising generating, based on an indication of a selection of a second artistic-effect, in real time utilizing a second artistic-effect neural network, a second synthesized artistic video stream at the first resolution by applying the second artistic-effect to the plurality of frames from the video stream.

19. The method of claim 14 , further comprising generating and displaying via a graphical user interface including the viewfinder, a plurality of different artistic effect previews; and

wherein generating, in real-time utilizing the artistic-effect neural network, the synthesized artistic video stream at the first resolution by applying the artistic-effect to the plurality of frames from the video stream is in response to a selection of an artistic effect preview corresponding to an artistic effect.

20. The method of claim 14 , wherein generating, utilizing the artistic-effect neural network, the artistic image comprises upsampling within the artistic-effect neural network.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 28, 2023
From: ZHAO, WENTIAN; LU, XIN; WAN, KUN; CHIEN, JEN-CHAN JEFF
To: ADOBE INC.
Reel/Frame 063482/0461 →
Continuity (2)
Division 17073697 · Oct 19, 2020
Related Publication 20230262189A1 · Aug 17, 2023
References Cited (23)
US 10909657B1 · Rossi · 2021 [cited by examiner]
US 20180075581A1 · Shi · 2018 [cited by examiner]
US 20180103213A1 · Holzer · 2018 [cited by examiner]
US 20190130229A1 · Lu et al. · 2019 [cited by applicant]
US 20200258206A1 · Shen et al. · 2020 [cited by applicant]
US 20200380639A1 · Rossi · 2020 [cited by examiner]
JP 2020010331A · 2020 [cited by applicant]
Ahn et al, Texture Enhancement via High-Resolution Style Transfer for Single Image Super Resolution, 2016, IEEE Transactions on Consumer Electronics, arXiv: 1612.00085, pp. 1-10. (Year: 2016). [cited by examiner]
Ye et al, Japanese Animation Style Transfer Using Deep Neural Networks, 2017, International Conference on Information, Communication and Engineering, pp. 1-5. (Year: 2017). [cited by examiner]
Pandey et al, Computationally Efficient Approaches for Image Style Transfer, 2018, 15th ieeE India Council International Conference, pp. 1-7. (Year: 2018). [cited by examiner]
Chen et al, StyleBank: An Explicit Representation for Neural Image Style Transfer, 2017, arXiv:1703.09210v2, pp. 1-10. (Year: 2017). [cited by examiner]
Cao et al, Development of Real-Time Style Transfer for Video Stream, 2019, 3rd International Conference on Circuits, Systems and Simulation, pp. 1-6. (Year: 2019). [cited by examiner]
Texler et al, Interactive Video Stylization Using Few-Shot Patch-Based Training, 2020, arXiv:2004.14489v1, pp. 1-11. (Year: 2020). [cited by examiner]
J. Johnson, A. Alahi, and L. Fei-Fei “Perceptual losses for real-time style transfer and superresolution,” in Computer Vision—ECCV 2016—14th European Conference, Amsterdam, The Netherlands, Oct. 11-14, 2016, Proceedings… [cited by applicant]
J. Zhu, T. Park, P. Isola, and A. A. Efros, “Unpaired image-to-image translation using cycleconsistent adversarial networks,” in IEEE International Conference on Computer Vision, ICCV 2017, Venice, Italy, Oct. 22-29, 20… [cited by applicant]
X. Mao, Q. Li, H. Xie, R. Y. K. Lau, Z. Wang and S. P. Smolley, “Least Squares Generative Adversarial Networks,” 2017 IEEE International Conference on Computer Vision (ICCV), Venice, 2017, pp. 2813-2821, doi: 10.1109/IC… [cited by applicant]
P. Isola, J. Zhu, T. Zhou and A. A. Efros, “Image-to-Image Translation with Conditional Adversarial Networks,” 2017 IEEE Conference on Computer Vision and Pattern Recognition (CVPR), Honolulu, HI, 2017, pp. 5967-5976, d… [cited by applicant]
Li et al, Convolutional Neural Network-Based Block Up-Sampling for Intra Frame Coding, 2018, IEEE, 28:9; 2316-2330. (Year: 2018 ) (Year: 2018). [cited by applicant]
Li et al, Precomputed Real-Time Texture Synthesis with Markovian Generative Adversarial Networks, 2016, ECCV 2016, Part III, LNCS 9907,pp. 702-716. (Year: 2016). [cited by applicant]
Lin et al., Convolutional Neural Network-Based Block Up-Sampling for HEVC, 2019, IEEE, 29:12; 3701-3715. (Year: 2019) (Year: 2019). [cited by applicant]
U.S. Appl. No. 17/073,697, Aug. 17, 2022, Preinterview 1st Office Action. [cited by applicant]
U.S. Appl. No. 17/073,697, Oct. 3, 2022, 1st Action Office Action. [cited by applicant]
U.S. Appl. No. 17/073,697, Feb. 1, 2023, Notice of Allowance. [cited by applicant]