IP Library › Granted Patent US 12,456,172
Granted Patent B2
US 12,456,172 · App. 17/972,961 · Granted Oct 28, 2025

Systems and methods for image denoising using deep convolutional networks

Inventors: Zengli Yang (San Jose, CA); Long Bao (San Diego, CA); Shuangquan Wang (San Diego, CA); Dongwoon Bai (San Diego, CA); Jungwon Lee (San Diego, CA)
Assignee: Samsung Electronics Co., Ltd.
G06T5/70G06N3/045G06N3/084G06T5/30G06T5/50G06T2207/20016G06T2207/20081G06T2207/20084G06T2207/20224
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,456,172
App. No.
17/972,961
Granted
Oct 28, 2025
Kind
B2
Abstract

A method includes: computing noise data by subtracting, by a processing circuit, a noisy image from a corresponding ground truth image; clustering, by the processing circuit, a plurality of noise values of the noise data based on intensity values of the corresponding ground truth image; permuting, by the processing circuit, a plurality of locations of the noise values of the noise data within each cluster; generating, by the processing circuit, a synthetic noise image based on the permuted locations of the noise values; adding, by the processing circuit, the synthetic noise image to the corresponding ground truth image to generate a synthetic noisy image; and augmenting an image dataset for training a neural network to perform image denoising with the synthetic noisy image.

Claims (74)

1. A method comprising:

computing noise data by subtracting, by a processing circuit, a noisy image from a corresponding ground truth image;

clustering, by the processing circuit, a plurality of noise values of the noise data based on intensity values of the corresponding ground truth image;

permuting, by the processing circuit, a plurality of locations of the noise values of the noise data within each cluster;

generating, by the processing circuit, a synthetic noise image based on the permuted locations of the noise values;

adding, by the processing circuit, the synthetic noise image to the corresponding ground truth image to generate a synthetic noisy image; and

augmenting an image dataset for training a neural network to perform image denoising with the synthetic noisy image.

2. The method of claim 1 , further comprising:

permuting, by the processing circuit, the plurality of locations of the noise values of the noise data within each cluster to compute second permuted locations of the noise values;

generating, by the processing circuit, a second synthetic noise image based on the second permuted locations of the noise values;

adding, by the processing circuit, the second synthetic noise image to the corresponding ground truth image to generate a second synthetic noisy image different from the synthetic noisy image; and

augmenting the image dataset with the second synthetic noisy image.

3. The method of claim 1 , wherein the noisy image is one of a plurality of noisy images associated with the corresponding ground truth image; and

wherein the corresponding ground truth image is generated by fusing the plurality of noisy images to generate a synthetic noise-free image.

4. The method of claim 3 , further comprising:

generating a first plurality of synthetic noisy images each based on a pairing of one of the plurality of noisy images and the synthetic noise-free image; and

augmenting the image dataset with the plurality of synthetic noisy images.

5. The method of claim 4 , further comprising:

generating a second plurality of synthetic noisy images based on the pairing of one of the plurality of noisy images and the synthetic noise-free image, each of the second plurality of synthetic noisy images corresponding to a different permutation of the locations of the noise values; and

augmenting the image dataset with the second plurality of synthetic noisy images.

6. The method of claim 4 , further comprising:

generating a synthetic noisy image for each pairing of the plurality of noisy images and the corresponding ground truth images; and

augmenting the image dataset with the synthetic noisy image for each pairing of the plurality of noisy images and the corresponding ground truth images.

7. The method of claim 5 , further comprising generating a plurality of synthetic noisy images for each pairing of the plurality of noisy images and the corresponding ground truth images, each of the plurality of synthetic noisy images corresponding to a different permutation of the locations of the noise values.

8. The method of claim 1 , further comprising training the neural network to perform image denoising, the training being performed using the image dataset augmented by the synthetic noisy image,

wherein the neural network comprises a multi-scale residual dense block (MRDB), the MRDB comprising:

a residual dense block (RDB) comprising a plurality of convolutional modules; and

an atrous spatial pyramid pooling (ASPP) module.

9. A non-transitory computer-readable medium storing instructions that, when executed by a processor, cause the processor to:

compute noise data by subtracting a noisy image from a corresponding ground truth image;

cluster a plurality of noise values of the noise data based on intensity values of the corresponding ground truth image;

permute a plurality of locations of the noise values of the noise data within each cluster;

generate a synthetic noise image based on the permuted locations of the noise values;

add the synthetic noise image to the corresponding ground truth image to generate a synthetic noisy image; and

augment an image dataset for training a neural network to perform image denoising with the synthetic noisy image.

10. The non-transitory computer-readable medium of claim 9 , further storing instructions that, when executed by the processor, cause the processor to:

permute the plurality of locations of the noise values of the noise data within each cluster to compute second permuted locations of the noise values;

generate a second synthetic noise image based on the second permuted locations of the noise values;

add the second synthetic noise image to the corresponding ground truth image to generate a second synthetic noisy image different from the synthetic noisy image; and

augment the image dataset with the second synthetic noisy image.

11. The non-transitory computer-readable medium of claim 9 , wherein the noisy image is one of a plurality of noisy images associated with the corresponding ground truth image; and

wherein the corresponding ground truth image is generated by fusing the plurality of noisy images to generate a synthetic noise-free image.

12. The non-transitory computer-readable medium of claim 11 , further storing instructions that, when executed by the processor, cause the processor to:

generate a first plurality of synthetic noisy images each based on a pairing of one of the plurality of noisy images and the synthetic noise-free image; and

augment the image dataset with the plurality of synthetic noisy images.

13. The non-transitory computer-readable medium of claim 12 , further storing instructions that, when executed by the processor, cause the processor to:

generate a second plurality of synthetic noisy images based on the pairing of one of the plurality of noisy images and the synthetic noise-free image, each of the second plurality of synthetic noisy images corresponding to a different permutation of the locations of the noise values; and

augment the image dataset with the second plurality of synthetic noisy images.

14. The non-transitory computer-readable medium of claim 12 , further storing instructions that, when executed by the processor, cause the processor to:

generate a synthetic noisy image for each pairing of the plurality of noisy images and the corresponding ground truth images; and

augment the image dataset with the synthetic noisy image for each pairing of the plurality of noisy images and the corresponding ground truth images.

15. The non-transitory computer-readable medium of claim 14 , further storing instructions that, when executed by the processor, cause the processor to generate a plurality of synthetic noisy images for each pairing of the plurality of noisy images and the corresponding ground truth images, each of the plurality of synthetic noisy images corresponding to a different permutation of the locations of the noise values.

16. A system comprising:

a processing circuit; and

a memory storing instructions that, when executed by the processing circuit, cause the processing circuit to:

compute noise data by subtracting a noisy image from a corresponding ground truth image;

cluster a plurality of noise values of the noise data based on intensity values of the corresponding ground truth image;

permute a plurality of locations of the noise values of the noise data within each cluster;

generate a synthetic noise image based on the permuted locations of the noise values;

add the synthetic noise image to the corresponding ground truth image to generate a synthetic noisy image; and

augment an image dataset for training a neural network to perform image denoising with the synthetic noisy image.

17. The system of claim 16 , wherein the memory further stores instructions that, when executed by the processing circuit, cause the processing circuit to:

permute the plurality of locations of the noise values of the noise data within each cluster to compute second permuted locations of the noise values;

generate a second synthetic noise image based on the second permuted locations of the noise values;

add the second synthetic noise image to the corresponding ground truth image to generate a second synthetic noisy image different from the synthetic noisy image; and

augment the image dataset with the second synthetic noisy image.

18. The system of claim 16 , wherein the noisy image is one of a plurality of noisy images associated with the corresponding ground truth image; and

wherein the corresponding ground truth image is generated by fusing the plurality of noisy images to generate a synthetic noise-free image.

19. The system of claim 18 , wherein the memory further stores instructions that, when executed by the processing circuit, cause the processing circuit to:

generate a first plurality of synthetic noisy images each based on a pairing of one of the plurality of noisy images and the synthetic noise-free image; and

augment the image dataset with the plurality of synthetic noisy images.

20. The system of claim 19 , wherein the memory further stores instructions that, when executed by the processing circuit, cause the processing circuit to:

generate a second plurality of synthetic noisy images based on the pairing of one of the plurality of noisy images and the synthetic noise-free image, each of the second plurality of synthetic noise images corresponding to a different permutation of the locations of the noise values; and

augment the image dataset with the second plurality of synthetic noisy images.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 8, 2022
From: YANG, ZENGLI; BAO, LONG; WANG, SHUANGQUAN; BAI, DONGWOON; LEE, JUNGWON
To: SAMSUNG ELECTRONICS CO., LTD.
Reel/Frame 062027/0152 →
Continuity (4)
Division 17010670 · Sep 2, 2020
Provisional Application 62987802 · Mar 10, 2020
Provisional Application 62988844 · Mar 12, 2020
Related Publication 20230043310A1 · Feb 9, 2023
References Cited (85)
US 5068914A · Klees · 1991 [cited by examiner]
US 7853099B2 · Shinmei · 2010 [cited by examiner]
US 8593483B2 · Cote et al. · 2013 [cited by applicant]
US 9332239B2 · Cote et al. · 2016 [cited by applicant]
US 9633274B2 · Tuzel et al. · 2017 [cited by applicant]
US 9672601B2 · Fergus et al. · 2017 [cited by applicant]
US 10043243B2 · Matviychuk et al. · 2018 [cited by applicant]
US 10430708B1 · Hu · 2019 [cited by examiner]
US 10438117B1 · Ross et al. · 2019 [cited by applicant]
US 10635927B2 · Chen · 2020 [cited by examiner]
US 10726525B2 · El-Khamy et al. · 2020 [cited by applicant]
US 11282208B2 · Cohen · 2022 [cited by examiner]
US 20070165112A1 · Shinmei · 2007 [cited by examiner]
US 20140153819A1 · Lin · 2014 [cited by examiner]
US 20150178591A1 · Fergus · 2015 [cited by examiner]
US 20170256038A1 · Lee · 2017 [cited by examiner]
US 20180253622A1 · Chen · 2018 [cited by examiner]
US 20190205758A1 · Zhu · 2019 [cited by examiner]
US 20200034948A1 · Park · 2020 [cited by examiner]
US 20200068151A1 · Pourreza Shahri · 2020 [cited by examiner]
US 20200074234A1 · Tong · 2020 [cited by examiner]
US 20200074271A1 · Liang · 2020 [cited by examiner]
US 20200088856A1 · Ida · 2020 [cited by examiner]
US 20200126190A1 · Lebel · 2020 [cited by examiner]
US 20200151546A1 · Liu · 2020 [cited by examiner]
US 20200286208A1 · Halupka · 2020 [cited by examiner]
US 20200311878A1 · Matsuura · 2020 [cited by examiner]
US 20200349675A1 · Park · 2020 [cited by examiner]
US 20210027098A1 · Ge · 2021 [cited by examiner]
US 20210027162A1 · Yang · 2021 [cited by examiner]
US 20210104021A1 · Sohn · 2021 [cited by examiner]
US 20210124990A1 · Lian · 2021 [cited by examiner]
US 20210241429A1 · Pan · 2021 [cited by examiner]
US 20210248718A1 · Yu · 2021 [cited by examiner]
US 20210272240A1 · Litwiller · 2021 [cited by examiner]
US 20220028085A1 · Vasilev · 2022 [cited by examiner]
US 20220301114A1 · Marras · 2022 [cited by examiner]
TW 201944357A · 2019 [cited by applicant]
TW 201947464A · 2019 [cited by applicant]
WO WO2020000171A1 · 2020 [cited by applicant]
Moreno-Barea, Francisco J., et al. “Forward noise adjustment scheme for data augmentation.” 2018 IEEE symposium series on computational intelligence (SSCI). IEEE, 2018. [cited by examiner]
Taiwanese Office Action dated Jun. 20, 2024, issued in corresponding Taiwanese Patent Application No. 110108457 (10 pages). [cited by applicant]
Buades, A., Coll, B., & Morel, J. M. (Jun. 2005). A non-local algorithm for image denoising. In [cited by applicant]
Dabov, K., Foi, A., Katkovnik, V., & Egiazarian, K. (2007). Image denoising by sparse 3-D transform-domain collaborative filtering. [cited by applicant]
Abdelhamed, A., Timofte, R., & Brown, M. S. (2019). Ntire 2019 challenge on real image denoising: Methods and results. In [cited by applicant]
Brooks, T., Mildenhall, B., Xue, T., Chen, J., Sharlet, D., & Barron, J. T. (2019). Unprocessing images for learned raw denoising. In [cited by applicant]
Timofte, R., Agustsson, E., Van Gool, L., Yang, M. H., & Zhang, L. (2017). Ntire 2017 challenge on single image super-resolution: Methods and results. In [cited by applicant]
Martin, D., Fowlkes, C., Tal, D., & Malik, J., Jan. 2001, A Database of Human Segmented Natural Images and its Application to Evaluating Segmentation Algorithms and Measuring Ecological Statistics, ICCV, Computer Scienc… [cited by applicant]
R. Franzen. Kodak lossless true color image suite. Source: http://r0k.us/graphics/kodak, 3 pages. [cited by applicant]
Gu, S., Li, Y., Gool, L. V., & Timofte, R. (2019). Self-guided network for fast image denoising. In [cited by applicant]
K. Zhang, W. Zuo, Y. Chen, D. Meng, and L. Zhang. Beyond a gaussian denoiser: Residual learning of deep cnn for image denoising. [cited by applicant]
Anaya, J., & Barbu, A. (2018). RENOIR-A dataset for real low-light image noise reduction. [cited by applicant]
Nam, S., Hwang, Y., Matsushita, Y., & Joo Kim, S. (2016). A holistic approach to cross-channel image noise modeling and its application to image denoising. In [cited by applicant]
Plotz, T., & Roth, S. (2017). Benchmarking denoising algorithms with real photographs. In [cited by applicant]
Guo, S., Yan, Z., Zhang, K., Zuo, W., & Zhang, L. (2019). Toward convolutional blind denoising of real photographs. In [cited by applicant]
Xu, J., Li, H., Liang, Z., Zhang, D., & Zhang, L. (2018). Real-world noisy image denoising: A new benchmark. arXiv preprint arXiv:1804.02603, 13 pages. [cited by applicant]
Foi, A., Trimeche, M., Katkovnik, V., & Egiazarian, K. (2008). Practical Poissonian-Gaussian noise modeling and fitting for single-image raw-data. [cited by applicant]
Abdelhamed, A., Lin, S., & Brown, M. S. (2018). A high-quality denoising dataset for smartphone cameras. In [cited by applicant]
Zhang, Y., Tian, Y., Kong, Y., Zhong, B., & Fu, Y. (2020). Residual dense network for image restoration. [cited by applicant]
Zini, S., Bianco, S., & Schettini, R. (2019). Deep Residual Autoencoder for quality independent JPEG restoration. arXiv preprint arXiv:1903.06117, 10 pages. [cited by applicant]
Lefkimmiatis, S. (2017). Non-local color image denoising with convolutional neural networks. In [cited by applicant]
Anwar, S., Huynh, C. P., & Porikli, F. (2017). Chaining identity mapping modules for image denoising. arXiv preprint arXiv:1712.02933, 10 pages. [cited by applicant]
Burger, H. C., Schuler, C. J., & Harmeling, S. (Jun. 2012). Image denoising: Can plain neural networks compete with BM3D ?. In [cited by applicant]
Zhang, K., Zuo, W., Gu, S., & Zhang, L. (2017). Learning deep CNN denoiser prior for image restoration. In [cited by applicant]
Anwar, S., & Barnes, N. (2019). Real image denoising with feature attention. In [cited by applicant]
Haris, M., Shakhnarovich, G., & Ukita, N. (2018). Deep back-projection networks for super- resolution. In [cited by applicant]
Park, B., Yu, S., & Jeong, J. (2019). Densely connected hierarchical network for image denoising. In [cited by applicant]
Kim, D. W., Ryun Chung, J., & Jung, S. W. (2019). Grdn: Grouped residual dense network for real image denoising and gan-based real-world noise modeling. In [cited by applicant]
Xia, X., & Kulis, B. (2017). W-net: A deep model for fully unsupervised image segmentation. arXiv preprint arXiv:1711.08506, 13 pages. [cited by applicant]
Liu, J., Wu, C. H., Wang, Y., Xu, Q., Zhou, Y., Huang, H., . . . & Wang, J. (2019). Learning raw image denoising with bayer pattern unification and bayer preserving augmentation. In [cited by applicant]
Abdelhamed, A., Brubaker, M. A., & Brown, M. S. (2019). Noise flow: Noise modeling with conditional normalizing flows. In [cited by applicant]
Abdelhamed, A., Afifi, M., Timofte, R., & Brown, M. S. (2020). Ntire 2020 challenge on real image denoising: Dataset, methods and results. In [cited by applicant]
Kingma, D. P., & Ba, J. (2014). Adam: A method for stochastic optimization. arXiv preprint arXiv:1412.6980, 9 pages. [cited by applicant]
Aharon, M., Elad, M., & Bruckstein, A. (2006). K-SVD: An algorithm for designing overcomplete dictionaries for sparse representation. [cited by applicant]
Zhang, L., Dong, W., Zhang, D., & Shi, G. (2010). Two-stage image denoising by principal component analysis with local pixel grouping. [cited by applicant]
S. Roth and M. J. Black. Fields of experts. [cited by applicant]
Gu, S., Zhang, L., Zuo, W., & Feng, X. (2014). Weighted nuclear norm minimization with application to image denoising. In [cited by applicant]
Talebi, H., & Milanfar, P. (2013). Global image denoising. [cited by applicant]
Y. Chen and T. Pock. Trainable nonlinear reaction diffusion: A flexible framework for fast and effective image restoration. IEEE TPAMI, 39(6):1256-1272, 2017. [cited by applicant]
Zoran, D., & Weiss, Y. (Nov. 2011). From learning models of natural image patches to whole image restoration. In [cited by applicant]
Elad, M., & Aharon, M. (2006). Image denoising via sparse and redundant representations over learned dictionaries. [cited by applicant]
Zamir, S. W., Arora, A., Khan, S., Hayat, M., Khan, F. S., Yang, M. H., & Shao, L. (2020). CycleISP: Real Image Restoration via Improved Data Synthesis. In [cited by applicant]
Chen, Liang-Chieh, et al. “Encoder-decoder with atrous separable convolution for semantic image segmentation.” [cited by applicant]
Zhang, Yulun, et al. “Residual dense network for image super-resolution.” [cited by applicant]
Song, Y., et al., “Dynamic Residual Dense Network for Image Denoising”, Sensors 2019, 19, 3809; doi: 10.3390/s 19173809. [cited by applicant]