IP Library › Granted Patent US 12,524,983
Granted Patent B2
US 12,524,983 · App. 18/141,142 · Granted Jan 13, 2026

Image processing apparatus and operation method thereof

Inventors: Iljun Ahn (Suwon-si, KR); Jaeyeon Park (Suwon-si, KR); Hanul Shin (Suwon-si, KR); Soomin Kang (Suwon-si, KR); Youngchan Song (Suwon-si, KR); Tammy Lee (Suwon-si, KR)
Assignee: SAMSUNG ELECTRONICS CO., LTD.
G06V10/44G06T3/40G06T7/0002G06V10/761G06V10/764G06T2207/20084G06T2207/30168
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,524,983
App. No.
18/141,142
Granted
Jan 13, 2026
Kind
B2
Abstract

An image processing apparatus for performing an image by using one or more neural networks may include a memory storing one or more instructions and at least one processor configured to execute the one or more instructions to obtain classification information of a first image and first feature information of the first image, generate a first feature image for the first image by performing first image processing on the classification information and the first feature information, obtain second feature information by performing second image processing on the classification information and the first feature information, obtain fourth feature information by performing third image processing on third feature information extracted during the first image processing, generate a second feature image for the first image, based on the second feature information and the fourth feature information, and generate a second image based on the first feature image and the second feature image.

Claims (60)

1 . An image processing apparatus comprising:

a memory storing at least one instruction; and

at least one processor configured to execute the at least one instruction to:

obtain classification information of a first image and first feature information of the first image,

generate a first feature image for the first image by performing first image processing on the classification information and the first feature information,

obtain second feature information by performing second image processing on the classification information and the first feature information,

obtain fourth feature information by performing third image processing on third feature information extracted during the first image processing,

generate a second feature image for the first image, based on the second feature information and the fourth feature information, and

generate a second image based on the first feature image and the second feature image.

2 . The image processing apparatus of claim 1 , wherein the first feature image comprises features of non-high frequency components in the first image, and

wherein the second feature image comprises features of high frequency components in the first image.

3 . The image processing apparatus of claim 1 , wherein a quality of the second image is higher than a quality of the first image.

4 . The image processing apparatus of claim 1 , wherein the at least one processor is further configured to execute the at least one instruction to obtain the classification information and the first feature information by using at least one convolutional neural network.

5 . The image processing apparatus of claim 1 , wherein the first image processing comprises upsampling the first feature information, and

wherein the first image, the first feature image, and the second feature image have a same size.

6 . The image processing apparatus of claim 1 , wherein the second image processing is performed by a multi-layer perceptron (MLP) module comprising at least one fully connected layer,

wherein the at least one processor is further configured to execute the at least one instruction to perform a multiplication operation between an input data fed to a fully connected layer and a weight matrix in the fully connected layer and an addition operation with biases in the fully connected layer.

7 . The image processing apparatus of claim 1 , wherein the at least one processor is further configured to execute the at least one instruction to:

obtain a sub-feature information by performing downscaling and upscaling on the third feature information;

obtain a difference information between the sub-feature information and the third feature information; and

generate the fourth feature information by performing a convolution operation between the difference information and a pre-trained weight.

8 . The image processing apparatus of claim 1 , wherein the at least one processor is further configured to execute the at least one instruction to:

obtain fifth feature information by performing a first operation on the second feature information; and

generate the second feature image by performing a second operation on the fifth feature information and the fourth feature information.

9 . The image processing apparatus of claim 8 , wherein the first operation comprises an adaptive instance normalization (AdaIn) operation, and

wherein the second operation comprises a spatial feature transform (SFT) operation.

10 . A method of operating an image processing apparatus comprising:

obtaining classification information of a first image and first feature information of the first image;

generating a first feature image for the first image by performing first image processing on the classification information and the first feature information;

obtaining second feature information by performing second image processing on the classification information and the first feature information;

obtaining fourth feature information by performing third image processing on third feature information extracted during the first image processing;

generating a second feature image for the first image, based on the second feature information and the fourth feature information; and

generating a second image based on the first feature image and the second feature image.

11 . The method of claim 10 , wherein the first feature image comprises features of non-high frequency components in the first image, and

wherein the second feature image comprises features of high frequency components in the first image.

12 . The method of claim 10 , wherein a quality of the second image is higher than a quality of the first image.

13 . The method of claim 10 , wherein the obtaining the classification information and the first feature information comprises:

obtaining the classification information and the first feature information of the first image by using at least one convolutional neural network.

14 . The method of claim 10 , wherein the first image processing comprises upsampling of the first feature information, and

wherein the first image, the first feature image, and the second feature image have a same size.

15 . The method of claim 10 , wherein the second image processing is performed by a multi-layer perceptron (MLP) module comprising at least one fully connected layer, and

wherein the obtaining the second feature information by performing the second image processing on the classification information and the first feature information comprises:

performing a multiplication operation between input data fed to a fully connected layer and a weight matrix in the fully connected layer and an addition operation with biases in the fully connected layer.

16 . The method of claim 10 , wherein the obtaining the fourth feature information by performing the third image processing on the third feature information extracted during the first image processing comprises:

obtaining sub-feature information by performing downscaling and upscaling on the third feature information;

obtaining difference information between the third feature information and the sub-feature information; and

generating the fourth feature information by performing a convolution operation between the difference information and a pre-trained weight.

17 . The method of claim 10 , wherein the generating the second feature image for the first image based on the second feature information and the fourth feature information comprises:

obtaining fifth feature information by performing a first operation on the second feature information; and

generating the second feature image by performing a second operation on the fifth feature information and the fourth feature information.

18 . The method of claim 17 , wherein the first operation comprises an adaptive instance normalization (AdaIn) operation, and

wherein the second operation comprises a spatial feature transform (SFT) operation.

19 . The method of claim 10 , wherein the third feature information comprises a plurality of pieces of intermediate data that is output by a process of generating the first feature image.

20 . A non-transitory computer-readable recording medium storing computer readable program code or instructions which are executable by a processor to perform a method of image processing, the method comprising:

obtaining classification information of a first image and first feature information of the first image;

generating a first feature image for the first image by performing first image processing on the classification information and the first feature information;

obtaining second feature information by performing second image processing on the classification information and the first feature information;

obtaining fourth feature information by performing third image processing on third feature information extracted during the first image processing;

generating a second feature image for the first image, based on the second feature information and the fourth feature information; and

generating a second image based on the first feature image and the second feature image.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 28, 2023
From: AHN, ILJUN; PARK, JAEYEON; KANG, SOOMIN; SONG, YOUNGCHAN; LEE, TAMMY
To: SAMSUNG ELECTRONICS CO., LTD.
Reel/Frame 063482/0558 →
Priority Claims (2)
KR 10-2022-0056890 · May 9, 2022 · national
KR 10-2022-0127170 · Oct 5, 2022 · national
Continuity (2)
Continuation PCTKR2023004673 · Apr 6, 2023
Related Publication 20230360359A1 · Nov 9, 2023
References Cited (31)
US 7280600B2 · Orhand et al. · 2007 [cited by applicant]
US 9014260B2 · Alshin et al. · 2015 [cited by applicant]
US 9288495B2 · Kim et al. · 2016 [cited by applicant]
US 10607120B2 · Bai et al. · 2020 [cited by applicant]
US 10832450B2 · Tong et al. · 2020 [cited by applicant]
US 10853986B2 · Elgammal · 2020 [cited by applicant]
US 11594056B2 · Wakui · 2023 [cited by applicant]
US 11694306B2 · Baek et al. · 2023 [cited by applicant]
US 11836890B2 · Lee · 2023 [cited by examiner]
US 12315111B2 · Wang · 2025 [cited by examiner]
US 20200364486A1 · Park et al. · 2020 [cited by applicant]
US 20210334578A1 · Lim et al. · 2021 [cited by applicant]
US 20230031143A1 · Baek · 2023 [cited by examiner]
US 20230360169A1 · Park · 2023 [cited by examiner]
US 20230360359A1 · Ahn · 2023 [cited by examiner]
US 20230360382A1 · Kang · 2023 [cited by examiner]
US 20250166345A1 · Beye · 2025 [cited by examiner]
JP 7195220B2 · 2022 [cited by applicant]
KR 101675116B1 · 2016 [cited by applicant]
KR 1020170087734A · 2017 [cited by applicant]
KR 101807170B1 · 2017 [cited by applicant]
KR 101887558B1 · 2018 [cited by applicant]
KR 1020200015095A · 2020 [cited by applicant]
KR 1020200132304A · 2020 [cited by applicant]
KR 102221225B1 · 2021 [cited by applicant]
KR 1020210154684A · 2021 [cited by applicant]
WO 2021188254A1 · 2021 [cited by applicant]
WO 2022057837A1 · 2022 [cited by applicant]
International Search Report and Written Opinion issued Jul. 13, 2023 by the International Searching Authority in counterpart International Patent Application No. PCT/KR2023/004673. (PCT/ISA/220, PCT/ISA/210 and PCT/ISA/… [cited by applicant]
Karras, Tero et al., “A Style-Based Generator Architecture for Generative Adversarial Networks”, CVPR, 2019. (10 pages total). [cited by applicant]
Wang, Xintao et al., “Towards Real-World Blind Face Restoration with Generative Facial Prior”, CVPR, 2021. (11 pages total). [cited by applicant]