IP Library Granted Patent US 12,445,746
Granted Patent B2
US 12,445,746 · App. 18/464,987 · Granted Oct 14, 2025

Systems, methods, and media for high dynamic range imaging using single-photon and conventional image sensor data

Inventors: Felipe Gutierrez Barragan (Alameda, CA); Yuhao Liu (Madison, WI); Atul Ingle (Madison, WI); Mohit Gupta (Madison, WI); Andreas Velten (Madison, WI)
Assignee: Wisconsin Alumni Research Foundation
H04N25/585H10F39/809H10F39/182
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,445,746
App. No.
18/464,987
Granted
Oct 14, 2025
Kind
B2
Abstract

In accordance with some embodiments, systems, methods, and media for high dynamic range imaging using single-photon and conventional image sensor data are provided. In some embodiments, the system comprises: first detectors configured to detect a level of photons proportional to incident photon flux; second detectors configured to detect arrival of individual photons; a processor programmed to: receive, from the first detectors, first values indicative of photon flux from a scene with a first resolution; receive, from the second detectors, second values indicative of photon flux from the scene with a lower resolution; provide a first encoder of a trained machine learning model first flux values based on the first values, provide the second encoder of the model second flux values; receive, as output, values indicative of photon flux from the scene; and generate a high dynamic range image based on the third plurality of values.

Claims (67)

1. A system for generating high dynamic range digital images, comprising:

a communication connection configured to receive readout data from an image data source, the image data source comprising at least one image sensor;

a processor; and

at least one memory connected to receive and store the readout data received by the communication connection, and having stored thereon a set of software instructions which, when executed by the processor, cause the processor to:

receive, via the communication connection, first readout data for a scene, the first readout data indicative of flux detected by a first detector of the at least one image sensor having a first dynamic range and a first resolution;

receive, via the communication connection, second readout data for the scene, the second readout data indicative of flux at a second detector of the image sensor, the second detector having a second dynamic range that is lower than the first dynamic range and a second resolution that is higher than the first resolution;

provide the first readout data and the second readout data as inputs to a trained machine learning model, wherein the trained machine learning model was trained using a dataset comprising image data for scenes with corresponding estimated and known flux values from sensors having different dynamic ranges;

receive, as output from the trained machine learning model, third readout data based on properties of both the first readout data and the second readout data; and

generate an image corresponding to the scene, wherein the image has a third resolution that is higher than the first resolution and a third dynamic range that is higher than the second dynamic range.

2. The system of claim 1 , wherein the second detector comprises a plurality of second detectors which capture color and brightness information for a scene.

3. The system of claim 2 , wherein the plurality of second detectors comprises a complementary semiconductor metal oxide (CMOS) pixel array.

4. The system of claim 1 , wherein the first detector comprises at least one single photon detector.

5. The system of claim 3 , wherein the first detector comprises at least one single-photon avalanche diode (SPAD) detector, and the flux detected by the first detector is based on detection of individual photons by the SPAD detector during a given time period.

6. The system of claim 1 , wherein the trained machine learning model includes a first skip connection between a layer of a first encoder, configured to process the first readout data indicative of flux detected by the first detector, and a layer of a decoder, and a second skip connection between a layer of a second encoder, configured to process the second readout data indicative of flux detected by the second detector, and the layer of the decoder, wherein the trained machine learning model is configured to concatenate values from the layer of the first encoder and values from the layer of the second encoder.

7. The system of claim 1 , wherein the processor is further programmed to:

estimate a first plurality of flux values associated with the first readout data using a relationship:

Φ

^

CMOS

=

N

^

T

CMOS

q

CMOS

T

,

Where {circumflex over (ϕ)} CMOS is the estimated flux for a portion of the scene, {circumflex over (N)} T CMOS is a value output by the first detector, q CMOS is a sensitivity of the first detector, and T is exposure time; and

estimate a second plurality of flux values associated with the second readout data using a relationship:

Φ

^

SPC

=

N

^

T

SPC

SPC

/

q

SPAD

T

SPC

-

τ

d

n

^

τ

SPC

SPC

,

where {circumflex over (ϕ)} SPC is the estimated flux for the portion of the scene, T SPC is exposure time, {circumflex over (N)} T SPC SPC is a photon count corresponding to a number of photon detections in exposure time T SPC , q SPAD is a sensitivity of the detector, and Ta is a dead time of the detector.

8. A method for generating high dynamic range digital images, the method comprising:

receiving a first plurality of image data for a scene;

determining a first plurality of photon flux indications;

receiving a second plurality of image data for the scene;

determining a second plurality of photon flux indications;

providing, as an input to a first encoder of a trained machine learning model, the first plurality of photon flux indications;

concatenating, using the trained machine learning model, values from a first encoder layer and values from a second encoder layer;

receiving, as an output from the trained machine learning model, a third plurality of photon flux indications for the scene; and

generating a high dynamic range image based on the third plurality of photon flux indications.

9. The method of claim 8 , wherein the output is an HDR output.

10. The method of claim 8 , wherein concatenating values from a first encoder layer and values from a second encoder layer further comprises introducing a plurality of attention gates.

Assignments (2)
CONFIRMATORY LICENSE Recorded Jun 23, 2025
From: WISCONSIN ALUMNI RESEARCH FOUNDATION
To: NATIONAL SCIENCE FOUNDATION
Reel/Frame 071687/0824 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Sep 21, 2023
From: VELTEN, ANDREAS; INGLE, ATUL; LIU, YUHAO; GUTIERREZ BARRAGAN, FELIPE; GUPTA, MOHIT
To: WISCONSIN ALUMNI RESEARCH FOUNDATION
Reel/Frame 064985/0307 →
Continuity (2)
Continuation 17572236 · Jan 10, 2022
Related Publication 20240259706A1 · Aug 1, 2024
References Cited (24)
US 10616512B2 · Ingle · 2020 [cited by applicant]
US 11170549B2 · Gupta · 2021 [cited by applicant]
US 11539895B1 · Seets · 2022 [cited by applicant]
US 20210319606A1 · Gupta · 2021 [cited by applicant]
US 20220012815A1 · Kearney · 2022 [cited by examiner]
US 20240185428A1 · Lilaonitkul · 2024 [cited by examiner]
JP 2023123474A · 2023 [cited by applicant]
Asatsuma, et al., “Sub-pixel architecture of cmos image sensor achieving over 120 db dynamic range with less motion artifact characteristics,” In Proceedings of the 2019 International Image Sensor Workshop, vol. 1, 2019. [cited by applicant]
Burt et al., “A multiresolution spline with application to image mosaics,” ACM transactions on Graphics (1983). [cited by applicant]
Debevec, et al., “Recovering high dynamic range radiance maps from photographs,” In ACM SIGGRAPH 2008 classes, pp. 1-10. 2008. [cited by applicant]
Eilertsen, et al., “Hdr image reconstruction from a single exposure using deep cnns,” ACM transactions on graphics (TOG), 36(6):1-15, 2017. [cited by applicant]
Fossum, et al., “The quanta image sensor: Every photon Counts,” Sensors, (2016). [cited by applicant]
Funt, et al., “The rehabilitation of maxrgb,” in Color and Imaging Conference (2010)). [cited by applicant]
Funt, et al., “The effect of exposure on maxrgb color constancy,” in Human Vision and Electronic Imaging (2010). [cited by applicant]
Gardner, et al., “Learning to predict indoor illumination from a single image,” arXiv preprint arXiv:1704.00090 (2017) and are available at http(colon)//indoor(dot)hdrdb(dot)com. [cited by applicant]
Hasinoff, et al., “Noise-optimal capture for high dynamic range photography,” In 2010 IEEE Computer Society Conference on Computer Vision and Pattern Recognition, pp. 553-560. IEEE, 2010. [cited by applicant]
Han, et al., “Neuromorphic camera guided high dynamic range imaging,” In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp. 1730-1739, 2020. [cited by applicant]
Liu, et al., “Single-image hdr reconstruction by learning to reverse the camera pipeline,” In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp. 1651-1660, 2020. [cited by applicant]
Mann, et al., “On being undigital with digital cameras: Extending dynamic range by combining differently exposed pictures,” In Proceedings of IS&T, pp. 442-448, 1995. [cited by applicant]
Marnerides, et al., “Expandnet: A deep convolutional neural network for high dynamic range expansion from low dynamic range content,” In Computer Graphics Forum, vol. 37, pp. 37-49. Wiley Online Library, 2018. [cited by applicant]
Mase, et al., “A wide dynamic range cmos image sensor with multiple exposure-time signal outputs and 12-bit column-parallel cyclic a/d converters,” IEEE Journal of Solid-State Circuits, 40(12):2787-2795, 2005. [cited by applicant]
Nayar, et al., “High dynamic range imaging: Spatially varying pixel exposures,” In Proceedings IEEE Conference on Computer Vision and Pattern Recognition. CVPR 2000 (Cat. No. PR00662), vol. 1, pp. 472-479. IEEE, 2000. [cited by applicant]
Santos, et al., “Single image hdr reconstruction using a cnn with masked features and perceptual loss,” arXiv preprint arXiv:2005.07335 (2020). [cited by applicant]
Wang, et al., “Esgran: Enhanced super-resolution generative adversarial networks,” arXiv preprint arXiv:1809.00219 (2018). [cited by applicant]