IP Library › Granted Patent US 12,530,738
Granted Patent B2
US 12,530,738 · App. 18/147,371 · Granted Jan 20, 2026

Neural network training method, image processing method, and apparatus

Inventors: Dehua Song (Moscow, RU); Yunhe Wang (Beijing, CN); Hanting Chen (Shenzhen, CN); Chunjing Xu (Shenzhen, CN)
Assignee: HUAWEI TECHNOLOGIES CO., LTD.
G06T3/4046G06T3/4015G06T3/4053G06T5/20G06T5/73G06T2207/20081G06T2207/20084
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,530,738
App. No.
18/147,371
Granted
Jan 20, 2026
Kind
B2
Abstract

A neural network training method, includes: obtaining an input feature map of a training image; performing feature extraction processing on the input feature map by using a feature extraction core of a neural network to obtain a first candidate feature map; adding the first candidate feature map and a second candidate feature map to obtain an output feature map, where the second candidate feature map is a feature map obtained after a value corresponding to each element in the input feature map is increased by N times, and N is greater than 0; determining an image processing result of the training image based on the output feature map; and adjusting a parameter of the neural network based on the image processing result.

Claims (62)

1 . A neural network training method comprising:

obtaining an input feature map of a training image;

performing feature extraction processing on the input feature map by using a feature extraction core of a neural network to obtain a first intermediate feature map, wherein the feature extraction processing enables each element in the first intermediate feature map to be an L 1 regular distance between the feature extraction core and data at a corresponding location in the input feature map;

adding the first intermediate feature map and a second intermediate feature map to obtain an output feature map, wherein the second intermediate feature map is a feature map obtained after a value corresponding to each element in the input feature map is increased by N times, and N is an integer that is greater than 0;

determining an image processing result of the training image based on the output feature map; and

adjusting a parameter of the neural network based on the image processing result;

wherein the adding of the first intermediate feature map and the second intermediate feature map to obtain the output feature map comprises:

processing the first intermediate feature map by using an activation function to obtain a processed first intermediate feature map; and

adding the processed first intermediate feature map and the second intermediate feature map to obtain the output feature map, and

wherein the processing of the first intermediate feature map by using the activation function comprises:

enhancing high-frequency texture information of the first intermediate feature map by using a power activation function, wherein the parameter of the neural network comprises a parameter of the power activation function.

2 . The method according to claim 1 , wherein the determining of the image processing result of the training image based on the output feature map comprises:

processing the output feature map by using another activation function to obtain a processed output feature map; and

determining the image processing result of the training image by using the processed output feature map.

3 . The method according to claim 2 , wherein the processing of the output feature map by using the another activation function comprises:

enhancing high-frequency texture information of the output feature map by using another power activation function, wherein the parameter of the neural network comprises a parameter of the another power activation function.

4 . The method according to claim 1 , wherein the power activation function is

( Y )=sign( Y )·| Y| α , wherein

Y is a feature map input into the power activation function, sign(⋅) is a symbolic function, |⋅| is an absolute value operation, α is a parameter of the power activation function, and α>0.

5 . An image processing method comprising:

obtaining an input feature map of a training image;

performing feature extraction processing on the input feature map by using a feature extraction core of a neural network to obtain a first intermediate feature map, wherein the feature extraction processing enables each element in the first intermediate feature map to be an L 1 regular distance between the feature extraction core and data at a corresponding location in the input feature map;

adding the first intermediate feature map and a second intermediate feature map to obtain an output feature map, wherein the second intermediate feature map is a feature map obtained after a value corresponding to each element in the input feature map is increased by N times, and N is an integer that is greater than 0;

determining an image processing result of the training image based on the output feature map;

adjusting a parameter of the neural network based on the image processing result to obtain a trained neural network;

obtaining an input feature map of a to-be-processed image; and

performing image processing on the input feature map of the to-be-processed image by using the trained neural network to obtain an image processing result of the to-be-processed image;

wherein the adding of the first intermediate feature map a and the second intermediate feature map to obtain the output feature map comprises:

processing the first intermediate feature map by using an activation function to obtain a processed first intermediate feature map, and

adding the processed first intermediate feature map and the second intermediate feature map to obtain the output feature map; and

serein the processing of the first intermediate feature map by using the activation function comprises:

enhancing high-frequency texture information of the first intermediate feature map by using a power activation function, wherein the parameter of the neural network comprises a parameter of the power activation function.

6 . The method according to claim 5 , wherein the image processing comprises at least one of image super-resolution processing, image denoising processing, image demosaicing processing, or image deblurring processing.

7 . The method according to claim 5 , wherein the determining of the image processing result of the training image based on the output feature map comprises:

processing the output feature map by using another activation function to obtain a processed output feature map; and

determining the image processing result of the training image by using the processed output feature map.

8 . The method according to claim 7 , wherein the processing of the output feature map by using the another activation function comprises:

enhancing high-frequency texture information of the output feature map by using another power activation function, wherein the parameter of the neural network comprises a parameter of the another power activation function.

9 . The method according to claim 5 , wherein the power activation function is

( Y )=sign( Y )·| Y| α , wherein

Y is a feature map input into the power activation function, sign(⋅) is a symbolic function, |⋅| is an absolute value operation, α is a parameter of the power activation function, and α>0.

10 . A neural network training apparatus comprising:

a processor and a non-transitory computer-readable memory storing instructions which upon execution by the processor cause the neural network training apparatus to at least be configured to:

obtain an input feature map of a training image;

perform feature extraction processing on the input feature map by using a feature extraction core of a neural network to obtain a first intermediate feature map, wherein the feature extraction processing enables each element in the first intermediate feature map to be an L 1 regular distance between the feature extraction core and data at a corresponding location in the input feature map;

perform processing including adding the first intermediate feature map and a second intermediate feature map to obtain an output feature map, wherein the second intermediate feature map is a feature map obtained after a value corresponding to each element in the input feature map is increased by N times, and N is an integer that is greater than 0;

perform image processing including determining an image processing result of the training image based on the output feature map; and

adjust a parameter of the neural network based on the image processing result,

wherein in order to perform the processing, the neural network training apparatus is further configured to:

process the first intermediate feature map by using an activation function to obtain a processed first intermediate feature map; and

add the processed first intermediate feature map and the second intermediate feature map to obtain the output feature map; and

wherein in order to perform the processing, the neutral network training apparatus is further configured to:

enhance high-frequency texture information of the first intermediate feature man by using a power activation function, wherein the parameter of the neural network comprises a para meter of the power activation function.

11 . The apparatus according to claim 10 , wherein in order to perform the image processing, the neural network training apparatus is further configured to:

process the output feature map by using another activation function to obtain a processed output feature map; and

determine the image processing result of the training image by using the processed output feature map.

12 . The apparatus according to claim 11 , wherein in order to perform the image processing, the neural network training apparatus is further configured to:

enhance high-frequency texture information of the output feature map by using another power activation function, wherein the parameter of the neural network comprises a parameter of the another power activation function.

13 . The apparatus according to claim 10 , wherein the power activation function is

( Y )=sign( Y )·| Y| α , wherein

Y is a feature map input into the power activation function, sign(⋅) is a symbolic function, |⋅| is an absolute value operation, α is a parameter of the power activation function, and α>0.

14 . The method according to claim 1 , wherein the parameter of the power activation function is learnable such that the power activation function is a learnable power activation function trained to help the neural network adapt to different tasks and/or scenarios.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 1, 2023
From: SONG, DEHUA; WANG, YUNHE; CHEN, HANTING; XU, CHUNJING
To: HUAWEI TECHNOLOGIES CO., LTD.
Reel/Frame 064443/0709 →
Priority Claims (1)
CN 202010616988.X · Jun 30, 2020 · national
Continuity (2)
Continuation PCTCN2021092581 · May 10, 2021
Related Publication 20230177641A1 · Jun 8, 2023
References Cited (23)
US 20210019627A1 · Zhang · 2021 [cited by examiner]
US 20220004849A1 · Chen · 2022 [cited by examiner]
US 20230214458A1 · Marsden · 2023 [cited by examiner]
CN 109035260A · 2018 [cited by applicant]
CN 109191476A · 2019 [cited by applicant]
CN 110096960A · 2019 [cited by applicant]
CN 110111366A · 2019 [cited by applicant]
CN 110543890A · 2019 [cited by applicant]
CN 111145107A · 2020 [cited by applicant]
CN 111311629A · 2020 [cited by applicant]
CN 111914997A · 2020 [cited by applicant]
CN 111914997B · 2024 [cited by applicant]
KR 20120100171A · 2012 [cited by applicant]
WO 2018214195A1 · 2018 [cited by applicant]
WO 2019020075A1 · 2019 [cited by applicant]
WO 2019184462A1 · 2019 [cited by applicant]
Lin et al., “Dense-Add Net: An Novel Convolutional Neural Network for Remote Sensing Image Inpainting”, 2018, IEEE Xplore, IGARSS 2018, pp. 4985-4988 (Year: 2018). [cited by examiner]
Hanting Chen et al: “AdderNet: Do We Really Need Multiplications in Deep Learning?”, arxiv.org, Cornell University Library, 201 Olin Library Cornell University Ithaca, NY 14853, Dec. 31, 2019 (Dec. 31, 2019), p. 1-p. 5,… [cited by applicant]
Lin Daoyu et al: “Dense-Add Net: An Novel Convolutional Neural Network for Remote Sensing Image Inpainting”, IGARSS 2018-2018 IEEE International Geoscience and Remote Sensing Symposium, IEEE, Jul. 22, 2018 (Jul. 22, 201… [cited by applicant]
Dehua Song et al: “AdderSR: Towards Energy Efficient Image Super-Resolution”, arxiv.org, Sep. 21, 2020 (Sep. 21, 2020), total 9pages, XP081766129. [cited by applicant]
Office Action dated Jul. 8, 2023, issued for Chinese Application No. 202010616988.X (8 pages). [cited by applicant]
Extended European Search Report dated Oct. 16, 2023, issued for European Application No. 21832925.8 (15 pages). [cited by applicant]
International Search Report and Written Opinion of the International Searching Authority for International Application No. PCT/CN2021/092581 dated Aug. 11, 2021 (10 pages). [cited by applicant]