IP Library › Granted Patent US 12,586,221
Granted Patent B2
US 12,586,221 · App. 18/297,396 · Granted Mar 24, 2026

Method and apparatus for estimating depth information of images

Inventors: Soon-Heung Jung (Daejeon, KR); David Crandall (Bloomington, IN); Vibhas Kumar Vats (Bloomington, IN); Shaurya Shubham (Bloomington, IN); Md Alimoor Reza (Bloomington, IN); Chuhua Wang (Bloomington, IN)
Assignees: Electronics and Telecommunications Research Institute; The Trustees of Indiana University
G06T7/50G06V10/761G06T2207/20081
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,586,221
App. No.
18/297,396
Granted
Mar 24, 2026
Kind
B2
Abstract

A method and apparatus for estimating depth information of an image are disclosed. The depth information estimation method of the image includes providing a confidence map for a ground truth depth map; and learning a depth information estimation model for estimating depth information of an image based on the ground truth depth map and the confidence map.

Claims (51)

1 . A method of estimating the depth information of an image, the method comprising:

providing, by an apparatus including a memory and a processor, a confidence map for a ground truth depth map; and

learning, by the apparatus, a depth information estimation model that estimates depth information of an image based on the ground truth depth map and the confidence map,

wherein the providing the confidence map includes:

filling, by the apparatus, an empty depth value of the ground truth depth map using a depth completion technique and providing the confidence map by generating the confidence map for the ground truth depth map filled with the empty depth value,

wherein the generating the confidence map includes:

configuring, by the apparatus, a confidence value to 1 for a depth value generated in a same pixel position as the ground truth depth map; and

generating, by the apparatus, the confidence map for the ground truth depth map filled with the empty depth value, by configuring the confidence value based on a distance from a number of pixels having a confidence value present in a certain radius for a depth value of a remaining pixel position.

2 . The method of claim 1 ,

wherein the providing the confidence map includes:

generating, by the apparatus, the confidence map using a pre-learned confidence estimation model that inputs at least one image corresponding to the ground truth depth map and the ground truth depth map.

3 . The method of claim 1 ,

wherein the providing the confidence map includes:

for a ground truth depth map with a predetermined ratio of the empty depth value among the ground truth depth map, filling, by the apparatus, the empty depth value using the depth completion technique.

4 . The method of claim 1 ,

wherein the learning the depth information estimation model includes:

learning, by the apparatus, the depth information estimation model by reflecting the confidence map in a loss function.

5 . A method for estimating depth information of image, the method comprising:

obtaining, by an apparatus including a memory and a processor, a ground truth depth map;

filling, by the apparatus, an empty depth value for each of the obtained ground truth depth map using a depth completion technique;

generating, by the apparatus, a confidence map for the ground truth depth map filled with the empty depth value; and

learning, by the apparatus, a depth information estimation model for estimating depth information of an image based on the ground truth depth map filled with the empty depth value and the confidence map,

wherein the providing the confidence map includes:

configuring, by the apparatus, a confidence value to 1 for a depth value generated in a same pixel position as the ground truth depth map; and

generating, by the apparatus, the confidence map for the ground truth depth map filled with the empty depth value, by configuring a confidence value based on a distance from a number of pixels having a confidence value present in a certain radius for a depth value of a remaining pixel position.

6 . The method of claim 5 ,

wherein the generating the confidence map includes:

generating the confidence map based on the ground truth depth map filled with the empty depth value and at least one image corresponding to the obtained ground truth depth map.

7 . The method of claim 6 , wherein:

the filling using the depth completion technique includes:

for a ground truth depth map with a predetermined ratio of the empty depth value among the ground truth depth map, filling, by the apparatus, the empty depth value using the depth completion technique.

8 . An apparatus for estimating depth information of an image, the apparatus comprising:

a memory;

a transceiver; and

a processor,

wherein the processor is configured to:

provide a confidence map for a ground truth depth map;

learn a depth information estimation model for estimating depth information of an image based on the ground truth depth map and the confidence map; and

fill an empty depth value of the ground truth depth map using depth completion technology and provide the confidence map by generating the confidence map for the ground truth depth map filled with the empty depth value,

wherein the processor is configured to:

configure a confidence value to 1 for a depth value generated in a same pixel position as the ground truth depth map; and

generate the confidence map for the ground truth depth map filled with the empty depth value, by configuring the confidence value based on a distance from a number of pixels having the confidence value present in a certain radius for a depth value of a remaining pixel position.

9 . The apparatus of claim 8 ,

wherein the processor is configured to:

generate the confidence map using a pre-learned confidence estimation model that inputs at least one image corresponding to the ground truth depth map and the ground truth depth map.

10 . The apparatus of claim 8 ,

wherein the processor is configured to:

for a ground truth depth map with a predetermined ratio of the empty depth value among the ground truth depth map, fill the empty depth value using the depth completion technology.

11 . The apparatus of claim 8 ,

wherein the processor is configured to:

learn the depth information estimation model by reflecting the confidence map in a loss function.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 7, 2023
From: JUNG, SOON-HEUNG; CRANDALL, DAVID; VATS, VIBHAS KUMAR; SHUBHAM, SHAURYA; REZA, MD ALIMOOR; WANG, CHUHUA
To: ELECTRONICS AND TELECOMMUNICATIONS RESEARCH INSTITUTE; THE TRUSTEES OF INDIANA UNIVERSITY
Reel/Frame 063262/0064 →
Priority Claims (1)
KR 10-2022-0043542 · Apr 7, 2022 · national
Continuity (1)
Related Publication 20230326051A1 · Oct 12, 2023
References Cited (15)
US 10091485B2 · Lim · 2018 [cited by applicant]
US 10929994B2 · Yu et al. · 2021 [cited by applicant]
US 11164326B2 · Liu · 2021 [cited by examiner]
US 12100173B2 · Bhutani · 2024 [cited by examiner]
US 20200193630A1 · Mousavian · 2020 [cited by examiner]
US 20200327686A1 · Zatzarinni · 2020 [cited by examiner]
KR 101784620B1 · 2017 [cited by applicant]
KR 102110690B1 · 2020 [cited by applicant]
KR 1020210073435A · 2021 [cited by applicant]
Tian et al., “Semi-supervised Depth Estimation from a Single Image Based on Confidence Learning,” ICASSP 2019—2019 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP), Brighton, UK, 2019, p… [cited by examiner]
Yang et al., “Fast Depth Prediction and Obstacle Avoidance on a Monocular Drone Using Probabilistic Convolutional Neural Network,” in IEEE Transactions on Intelligent Transportation Systems, vol. 22, No. 1, pp. 156-167,… [cited by examiner]
Alex Kendall et al., What Uncertainties Do We Need in Bayesian Deep Learning for Computer Vision?, 31st Conference on Neural Information Processing Systems (NIPS 2017), Long Beach, CA, USA. [cited by applicant]
Hyesong Choi et al., Adaptive confidence thresholding for monocular depth estimation, ICCV 2021, Computer Vision Foundation, pp. 12808-12818. [cited by applicant]
Inwook Shim et al., High-Fidelity Depth Upsampling Using the Self-Learning Framework, Sensors 2019, 81. https://doi.org/10.3390/s19010081, Dec. 27, 2018. [cited by applicant]
Yufan Zhu et al., Robust Depth Completion with Uncertainty-Driven Loss Functions, Computer Vision and Pattern Recognition, arXiv:2112.07895v2 [cs.CV], Dec. 28, 2021. [cited by applicant]