IP Library › Granted Patent US 12,322,127
Granted Patent B2
US 12,322,127 · App. 18/098,904 · Granted Jun 3, 2025

Depth completion method and apparatus using a spatial-temporal

Inventors: Myung Sik Yoo (Seoul, KR); Minh Tri Nguyen (Seoul, KR)
Assignee: FOUNDATION OF SOONGSIL UNIVERSITY-INDUSTRY COOPERATION
G06T7/579G06N3/0442G06N3/0455G06T7/248G06T7/521G06V10/82G06V20/58G06T2207/10024G06T2207/10028G06T2207/20084G06T2207/20221G06T2207/30241
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,322,127
App. No.
18/098,904
Granted
Jun 3, 2025
Kind
B2
Abstract

Provided are a depth completion method and apparatus using spatial-temporal information. The depth completion apparatus according to the present invention comprises a processor; and a memory connected to the processor, wherein the memory stores program instructions executable by the processor for performing operations comprising receiving an RGB image and a sparse image through a camera and LiDAR, generating a dense first depth map by processing color information of the RGB image through a first branch based on an encoder-decoder, generating a dense second depth map by up-sampling the sparse image through a second branch based on an encoder-decoder, generating a third depth map by fusing the first depth map and the second depth map, and generating a final depth map including a trajectory of a moving object included in an RGB image continuously captured during movement by inputting the third depth map to a convolution long term short memory (LSTM).

Claims (21)

1. A depth completion apparatus using spatial-temporal information comprising:

a processor; and

a memory connected to the processor,

wherein the memory stores program instructions executable by the processor for performing operations comprising:

receiving an RGB image and a sparse image through a camera and LiDAR,

generating a dense first depth map by processing color information of the RGB image through a first branch based on an encoder-decoder,

generating a dense second depth map by up-sampling the sparse image through a second branch based on an encoder-decoder,

generating a third depth map by fusing the first depth map and the second depth map, and

generating a final depth map including a trajectory of a moving object included in an RGB image continuously captured during movement by inputting the third depth map to a convolution long-term short memory (LSTM).

2. The depth completion apparatus of claim 1 , wherein a first encoder of the first branch and a second encoder of the second branch include a plurality of layers,

wherein the first and second encoders include a convolutional layer and a plurality of residual blocks having a skip connection.

3. The depth completion apparatus of claim 2 , wherein each layer of the first encoder is connected to each layer of the second encoder to help preserve rich features of the RGB image.

4. The depth completion apparatus of claim 1 , wherein the convolution LSTM calculates a future state of an input pixel according to a past state of a corresponding local neighboring pixel in a previous frame of the RGB image.

5. The depth completion apparatus of claim 4 , wherein the convolutional LSTM calculates a hidden state and cell statistics in the previous frame, and calculates a future state of a cell using a window sliding the RGB image to obtain a trajectory of a moving object.

6. A depth completion method using spatial-temporal information in an apparatus including a processor and a memory comprising:

receiving an RGB image and a sparse image through a camera and LiDAR;

generating a dense first depth map by processing color information of the RGB image through a first branch based on an encoder-decoder;

generating a dense second depth map by up-sampling the sparse image through a second branch based on an encoder-decoder;

generating a third depth map by fusing the first depth map and the second depth map; and

generating a final depth map including a trajectory of a moving object included in an RGB image continuously captured during movement by inputting the third depth map to a convolution long-term short memory (LSTM).

7. A non-transitory computer-readable medium storing a program for performing the method according to claim 6 .

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 19, 2023
From: YOO, MYUNG SIK; NGUYEN, MINH TRI
To: FOUNDATION OF SOONGSIL UNIVERSITY-INDUSTRY COOPERATION
Reel/Frame 062423/0523 →
Priority Claims (1)
KR 10-2022-0008217 · Jan 20, 2022 · national
Continuity (1)
Related Publication 20230230269A1 · Jul 20, 2023
References Cited (11)
US 20200293064A1 · Wu · 2020 [cited by examiner]
US 20220067950A1 · Lv · 2022 [cited by examiner]
US 20220262023A1 · Vignard · 2022 [cited by examiner]
KR 1020210073416A · 2021 [cited by applicant]
WO WO2021013334A1 · 2021 [cited by applicant]
Fu, Chen, et al. “Depth completion via inductive fusion of planar lidar and monocular camera.” 2020 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS). IEEE, 2020. (Year: 2020). [cited by examiner]
Ma, Fangchang, Guilherme Venturelli Cavalheiro, and Sertac Karaman. “Self-supervised sparse-to-dense: Self-supervised depth completion from lidar and monocular camera.” 2019 international conference on robotics and auto… [cited by examiner]
Prakash, Aditya, Kashyap Chitta, and Andreas Geiger. “Multi-Modal Fusion Transformer for End-to-End Autonomous Driving.” 2021 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR). IEEE, 2021. (Year: 202… [cited by examiner]
Shivakumar, Shreyas S., et al. “Dfusenet: Deep fusion of rgb and sparse depth information for image guided dense depth completion.” 2019 IEEE Intelligent Transportation Systems Conference (ITSC). IEEE, 2019. (Year: 2019… [cited by examiner]
Nguyen, Tri, and Myungsik Yoo. “Dense-depth-net: a spatial-temporal approach on depth completion task.” 2021 IEEE Region 10 Symposium (TENSYMP). IEEE, 2021. (Year: 2021). [cited by examiner]
1st Office Action issued on Dec. 5, 2023 for Korean Patent Application No. 10-2022-0008217. [cited by applicant]