IP Library › Granted Patent US 12,444,025
Granted Patent B2
US 12,444,025 · App. 18/452,634 · Granted Oct 14, 2025

Synthetic bracketing for exposure correction

Inventors: Iuri Frosio (Bergamo, IT); Mayoore Selvarasa Jaiswal (Bothell, WA); Jan Kautz (Lexington, MA); Jianyuan Min (Santa Clara, CA)
Assignee: NVIDIA Corporation
G06T5/50H04N23/743G06T2207/10016G06T2207/20081G06T2207/20221
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,444,025
App. No.
18/452,634
Granted
Oct 14, 2025
Kind
B2
Abstract

Systems and methods are disclosed related to synthetic bracketing for exposure correction. A deep learning based method and system produces a set of differently exposed images from a single input image. The images in the set may be combined to produce an output image with improved global and local exposure compared with the input image. An image encoder applies learned parameters to each input image to generate a set of image features including local exposure estimates for each of two or more regions of the input image and a low resolution latent representation of the input image. A decoder receives the local exposure estimates, the latent representation, and target enhancements that are processed to generate synthesized transformations. When applied to the input image, the synthesized transformations produce the set of transformed images. Each transformed image is a version of the input image synthesized to correspond to a respective target enhancement.

Claims (31)

1. A computer-implemented method, comprising:

encoding an input image into a latent representation for the input image and, for each of two or more regions of the input image, a respective local exposure estimate;

computing synthesized transformations using the local exposure estimates, latent representation, and target enhancement levels; and

applying the synthesized transformations to the input image to produce a set of exposure transformed images, wherein each exposure transformed image in the set is associated with a different one of the target enhancement levels.

2. The computer-implemented method of claim 1 , wherein the target enhancement levels are generated from the latent representation and the local exposure estimates.

3. The computer-implemented method of claim 1 , wherein the target enhancement levels are generated from the latent representation and enhancement signals provided by a user.

4. The computer-implemented method of claim 3 , wherein the enhancement signals include at least one of an indication of whether the input image is acquired with denoising on or off, a level of exposure correction to be applied, a brightness level control, a minimum luminance boosting amount, a noise reduction control, a color saturation control, a black point control, a local edge enhancement control, or a chroma control.

5. The computer-implemented method of claim 1 , further comprising combining the set of exposure transformed images to produce an output image with corrected exposure.

6. The computer-implemented method of claim 5 , wherein the input image is a frame in a video sequence, and further comprising repeating the encoding, computing, applying, and combining for additional input images included in the video sequence to produce additional output images with corrected exposures.

7. The computer-implemented method of claim 5 , wherein the combining is performed based on the target enhancement levels.

8. The computer-implemented method of claim 1 , wherein the encoding and computing are performed according to parameters that are learned using a training dataset comprising training input images and associated training target enhancements and ground truth sets of exposure transformed images.

9. The computer-implemented method of claim 8 , further comprising adjusting the parameters based on differences between the ground truth sets of exposure transformed images and sets of exposure transformed images.

10. The computer-implemented method of claim 9 , further comprising adjusting the parameters based on differences between a latent exposure channel included in the synthesized transformations and exposure and saturation of the training input images.

11. The computer-implemented method of claim 5 , wherein at least one of the steps of encoding, computing, or applying is performed on a server or in a data center to generate the output image, and the output image is streamed to a user device.

12. The computer-implemented method of claim 1 , wherein at least one of the steps of encoding, computing, or applying is performed within a cloud computing environment.

13. The computer-implemented method of claim 1 , wherein at least one of the steps of encoding, computing, or applying is performed for training, testing, or certifying a neural network employed in a machine, robot, or autonomous vehicle.

14. The computer-implemented method of claim 1 , wherein at least one of the steps of encoding, computing, or applying is performed on a virtual machine comprising a portion of a graphics processing unit.

15. A system, comprising:

a memory that stores an input image; and

a processor that is connected to the memory, wherein the processor is configured to:

encode the input image into a latent representation for the input image and, for each of two or more regions of the input image, a respective local exposure estimate;

compute synthesized transformations using the local exposure estimates, latent representation, and target enhancement levels; and

apply the synthesized transformations to the input image to produce a set of exposure transformed images, wherein each exposure transformed image in the set is associated with a different one of the target enhancement levels.

16. The system of claim 15 , wherein the target enhancement levels are generated from the latent representation and the local exposure estimates.

17. The system of claim 15 , wherein the target enhancement levels are generated from the latent representation and enhancement signals provided by a user.

18. The system of claim 15 , wherein the processor is further configured to combine the set of exposure transformed images to produce an output image with corrected exposure.

19. A non-transitory computer-readable media storing computer instructions that, when executed by one or more processors, cause the one or more processors to perform the steps of:

encoding an input image into a latent representation for the input image and, for each of two or more regions of the input image, a respective local exposure estimate;

computing synthesized transformations using the local exposure estimates, latent representation, and target enhancement levels; and

applying the synthesized transformations to the input image to produce a set of exposure transformed images, wherein each exposure transformed image in the set is associated with a different one of the target enhancement levels.

20. The non-transitory computer-readable media of claim 19 , further comprising combining the set of exposure transformed images to produce an output image with corrected exposure.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 21, 2023
From: FROSIO, IURI; JAISWAL, MAYOORE SELVARASA; KAUTZ, JAN; MIN, JIANYUAN
To: NVIDIA CORPORATION
Reel/Frame 064648/0364 →
Continuity (1)
Related Publication 20250069191A1 · Feb 27, 2025
References Cited (10)
US 11037278B2 · Zamir · 2021 [cited by examiner]
US 20230171509A1 · Varghese · 2023 [cited by examiner]
US 20240320809A1 · Varadarajan · 2024 [cited by examiner]
HDR image reconstruction from a single exposure using deep CNNs, by Eilertsen et al., ACM Transactions on Graphics, vol. 36, No. 6, Article 178. Publication date: Nov. 2017 (Year: 2017). [cited by examiner]
Afifi, M., et al., “Learning multi-scale photo exposure correction,” In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), pp. 9157-9167, Jun. 2021. [cited by applicant]
Chen, C., et al., “Learning to see in the dark,” In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, pp. 3291-3300, 2018. [cited by applicant]
Mertens, T., et al., “Exposure fusion,” In 15th Pacific Conference on Computer Graphics and Applications (PG 2007), pp. 382-390, IEEE. [cited by applicant]
Qu, L., et al., “TransMEF: A transformer-based multi-exposure image fusion framework using self-supervised multi- task learning,” In Proceedings of the AAAI Conference on Artificial Intelligence, vol. 36, pp. 2126-2134,… [cited by applicant]
Yuan, L., et al., “Automatic exposure correction of consumer photographs,” In Computer Vision-ECCV 2012: 12th European Conference on Computer Vision, Florence, Italy, Oct. 7-13, 2012, Proceedings, Part IV 12, pp. 771-78… [cited by applicant]
Xie, ZF, et al., “Multi-exposure motion estimation based on deep convolutional networks,” Journal of Computer Science and Technology, 33(3):487, 2018. [cited by applicant]