IP Library › Granted Patent US 12,361,729
Granted Patent B2
US 12,361,729 · App. 17/545,237 · Granted Jul 15, 2025

Single-stage 3-dimension multi-object detecting apparatus and method for autonomous driving

Inventors: Gon Woo Kim (Daejeon, KR); Loc Duy Hoang (Cheongju-si, KR)
Assignee: CHUNGBUK NATIONAL UNIVERSITY INDUSTRY-ACADEMIC COOPERATION FOUNDATION
G06V20/64G01S17/89G06N3/08G06V10/766G06V10/82
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,361,729
App. No.
17/545,237
Granted
Jul 15, 2025
Kind
B2
Abstract

According to at least one embodiment, the present disclosure provides an apparatus for single-stage three-dimensional (3D) multi-object detection by using a LiDAR sensor to detect 3D multiple objects, comprising: a data input module configured to receive raw point cloud data from the LiDAR sensor; a BEV image generating module configured to generate bird's eye view (BEV) images from the raw point cloud data; a learning module configured to perform a deep learning algorithm-based learning task to extract a fine-grained feature image from the BEV images; and a localization module configured to perform a regression operation and a localization operation to find 3D candidate boxes and classes corresponding to the 3D candidate boxes for detecting 3D objects from the fine-grained feature image.

Claims (19)

1. An apparatus for single-stage three-dimensional (3D) multi-object detection by using a LiDAR sensor to detect 3D multiple objects, comprising:

a data input module configured to receive raw 3D point cloud data from the LiDAR sensor;

a BEV image generating module configured to:

generate four feature map images based on a height, a density, an intensity, and a distance of the raw 3D point cloud data, by encoding the raw 3D point cloud data, and

generate bird's eye view (BEV) images by projecting the raw 3D point cloud data into 2D pseudo-images and discretizing a result of the projecting;

a learning module configured to perform a deep learning algorithm-based learning task to extract a fine-grained feature image from the BEV images; and

a localization module configured to perform a regression operation and a localization operation to find 3D candidate boxes and classes corresponding to the 3D candidate boxes for detecting 3D objects from the fine-grained feature image.

2. The apparatus of claim 1 , wherein the learning module is configured to perform a convolutional neural network-based (CNN-based) learning task.

3. A method performed by an apparatus for single-stage three-dimensional (3D) multi-object detection by using a LiDAR sensor to detect 3D multiple objects, the method comprising:

performing a data input operation by receiving raw 3D point cloud data from the LiDAR sensor;

generating bird's eye view (BEV) images from the raw point cloud data, comprising:

generating four feature map images based on a height, a density, an intensity, and a distance of the raw 3D point cloud data, by encoding the raw 3D point cloud data, and

generating the BEV images by projecting the raw 3D point cloud data into 2D pseudo-images and discretizing a result of the projecting;

performing a deep learning algorithm-based learning task to extract a fine-grained feature image from the BEV images; and

performing a regression operation and a localization operation to find 3D candidate boxes and classes corresponding to the 3D candidate boxes for detecting 3D objects from the fine-grained feature image.

4. The method of claim 3 , wherein the performing of the deep learning algorithm-based learning task comprises:

performing a convolutional neural network-based (CNN-based) learning task.

5. The apparatus of claim 1 , wherein the learning module employs a center regression, offset regression, orientation regression, Z-axis location regression and a size regression.

6. The method of claim 3 , wherein the performing the regression operation comprises performing a center regression, offset regression, orientation regression, Z-axis location regression and a size regression.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 8, 2021
From: KIM, GON WOO; HOANG, LOC DUY
To: CHUNGBUK NATIONAL UNIVERSITY INDUSTRY-ACADEMIC COOPERATION FOUNDATION
Reel/Frame 058336/0205 →
Priority Claims (1)
KR 10-2021-108154 · Aug 17, 2021 · national
Continuity (1)
Related Publication 20230071437A1 · Mar 9, 2023
References Cited (19)
US 20190003836A1 · Zhang et al. · 2019 [cited by applicant]
US 20190188541A1 · Wang · 2019 [cited by examiner]
US 20200081124A1 · Shi · 2020 [cited by applicant]
US 20210048529A1 · Roy Chowdhury · 2021 [cited by examiner]
US 20210082181A1 · Shi et al. · 2021 [cited by applicant]
US 20210157006A1 · Sun et al. · 2021 [cited by applicant]
US 20220153310A1 · Yang · 2022 [cited by examiner]
JP 2017166971A · 2017 [cited by applicant]
JP 2019532433A · 2019 [cited by applicant]
JP 202042009A · 2020 [cited by applicant]
JP 202182296A · 2021 [cited by applicant]
JP 2021532442A · 2021 [cited by applicant]
KR 101655606B1 · 2016 [cited by applicant]
Bin Yang, “Auto4D: Learning to Label 4D Objects from Sequential Point Clouds”, arXiv:2101.06586v2 [cs.CV] Mar. 11, 2021, pp. 1-8. (Year: 2021). [cited by examiner]
Bin Yang et al., “PIXOR: Real-time 3D Object Detection from Point Clouds”, IEEE/CVF Conference on Computer Vision and Pattern Recognition, 2018, pp. 7652-7660 (9 pages total). [cited by applicant]
Yi Zhang et al., “Accurate and Real-time Object Detection based on Bird's Eye View on 3D Point Clouds”, International Conference on 3D Vision (3DV), 2019, pp. 214-221 (8 pages total). [cited by applicant]
Guojun Wang et al., “CenterNet3D: An Anchor free Object Detector for Autonomous Driving”, Arxiv.Org, 2020, pp. 1-9 (9 pages total). [cited by applicant]
Guojun Wang et al., “CenterNet3D: An Anchor Free Object Detector for Point Cloud”, IEEE Transactions on Intelligent Transportation Systems, 2020, pp. 1-13 (13 pages total). [cited by applicant]
Extended European Search Report dated May 24, 2022 in European Application No. 21213697.2. [cited by applicant]