IP Library Granted Patent US 12,223,677
Granted Patent B1
US 12,223,677 · App. 18/354,415 · Granted Feb 11, 2025

Map-anchored object detection

Inventor: Louis Foucard (San Francisco, CA)
Assignee: AURORA OPERATIONS, INC.
G06T7/74B60W60/001B60W2420/00B60W2554/00B60W2556/00G06T2207/20081G06T2207/30204G06T2207/30256
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,223,677
App. No.
18/354,415
Granted
Feb 11, 2025
Kind
B1
Abstract

An example method includes (a) obtaining sensor data descriptive of an environment of an autonomous vehicle; (b) obtaining a plurality of travel way markers from map data descriptive of the environment; (c) determining, using a machine-learned object detection model and based on the sensor data, an association between one or more travel way markers of the plurality of travel way markers and an object in the environment; and (d) generating, using the machine-learned object detection model, an offset with respect to the one or more travel way markers of a spatial region of the environment associated with the object.

Claims (68)

1. A computer-implemented method, comprising:

(a) obtaining sensor data descriptive of an environment of an autonomous vehicle;

(b) obtaining a plurality of travel way markers from map data descriptive of the environment;

(c) determining, using a machine-learned object detection model and based on the sensor data, an association between one or more travel way markers of the plurality of travel way markers and an object in the environment; and

(d) generating, using the machine-learned object detection model, an offset with respect to the one or more travel way markers of a spatial region of the environment associated with the object;

wherein (c) comprises:

inputting the travel way markers and the sensor data to the machine-learned object detection model; and

obtaining object data from the machine-learned object detection model at projected locations of the travel way markers in a reference frame of the sensor data, wherein the object data indicates that the object is likely to be present at a projected location of the one or more travel way markers.

2. The computer-implemented method of claim 1 , wherein the travel way markers include lane markers.

3. The computer-implemented method of claim 2 , wherein the lane markers comprise centerline markers.

4. The computer-implemented method of claim 1 , wherein obtaining the object data comprises subsampling, based on the travel way markers, a detection map generated by the machine-learned object detection model.

5. The computer-implemented method of claim 1 , wherein one or more portions of the machine-learned object detection model are configured to sparsely use portions of an output layer based on locations in the sensor data corresponding to the projected locations.

6. The computer-implemented method of claim 1 , wherein the machine-learned object detection model is trained by:

obtaining ground truth travel way marker labels indicating a ground truth association between the object and one or more of the travel way markers; and

determining, based on comparing the object data and the ground truth travel way marker labels, a sparse loss for the machine-learned object detection model.

7. The computer-implemented method of claim 1 , wherein (d) comprises:

determining an offset of a centroid of a boundary of the spatial region; and

determining one or more dimensions of the boundary.

8. The computer-implemented method of claim 1 , wherein (d) comprises:

determining a first offset of a centroid of a first boundary of the spatial region in two dimensions; and

determining a second offset of a centroid of a second boundary of the spatial region in three dimensions.

9. The computer-implemented method of claim 1 , comprising:

based on determining that a velocity of the object is below a threshold, outputting a characteristic for the object indicating that the object is a static object; and

outputting the characteristic to a motion planning system of the autonomous vehicle.

10. The computer-implemented method of claim 1 , comprising:

based on determining that a velocity of the object is below a threshold and that the object is located adjacent to a travel way in the environment, outputting a characteristic for the object indicating that the object is a static object; and

outputting the characteristic to a motion planning system of the autonomous vehicle.

11. The computer-implemented method of claim 1 , wherein (b) comprises:

sampling discrete travel way markers from continuous travel way map data.

12. The computer-implemented method of claim 1 , wherein the spatial region of the environment is beyond an effective range of a LIDAR sensor of the autonomous vehicle.

13. The computer-implemented method of claim 1 , wherein:

the machine-learned object detection model was trained using training sensor data having a training field of view and training travel way markers having a training resolution;

the sensor data is associated with a runtime field of view; and

the travel way markers are obtained in (c) at a runtime resolution selected based on a comparison of the training field of view and the runtime field of view.

14. The computer-implemented method of claim 1 , comprising:

projecting, using a projection transform, the travel way markers into a reference frame of the sensor data;

determining one or more offsets of the spatial region with respect to the travel way markers;

based on the determined one or more offsets, determining a projection error for the projected travel way markers; and

recalibrating the projection transform based on the determined projection error.

15. The computer-implemented method of claim 1 , comprising:

identifying a lane in which the object is located based on the one or more travel way markers.

16. The computer-implemented method of claim 1 , wherein (d) comprises:

regressing, by the machine-learned object detection model, the offset.

17. An autonomous vehicle control system for controlling an autonomous vehicle, the autonomous vehicle control system comprising:

one or more processors; and

one or more non-transitory computer-readable media storing instructions that are executable by the one or more processors to cause the autonomous vehicle control system to perform operations, the operations comprising:

(a) obtaining sensor data descriptive of an environment of the autonomous vehicle;

(b) obtaining a plurality of travel way markers from map data descriptive of the environment;

(c) determining, using a machine-learned object detection model and based on the sensor data, an association between one or more travel way markers of the plurality of travel way markers and an object in the environment; and

(d) generating, using the machine-learned object detection model, an offset with respect to the one or more travel way markers of a spatial region of the environment associated with the object;

wherein (c) comprises:

inputting the travel way markers and the sensor data to the machine-learned object detection model; and

obtaining object data from the machine-learned object detection model at projected locations of the travel way markers in a reference frame of the sensor data, wherein the object data indicates that the object is likely to be present at a projected location of the one or more travel way markers.

18. The autonomous vehicle control system of claim 17 , wherein obtaining the object data comprises subsampling, based on the travel way markers, a detection map generated by the machine-learned object detection model.

19. The autonomous vehicle control system of claim 17 , wherein (d) comprises:

determining a first offset of a centroid of a first boundary of the spatial region in two dimensions; and

determining a second offset of a centroid of a second boundary of the spatial region in three dimensions.

20. The autonomous vehicle control system of claim 17 , wherein the machine-learned object detection model is trained by:

obtaining ground truth travel way marker labels indicating a ground truth association between the object and one or more of the travel way markers; and

determining, based on comparing the object data and the ground truth travel way marker labels, a sparse loss for the machine-learned object detection model.

21. One or more non-transitory computer-readable media storing instructions that are executable by one or more processors to cause an autonomous vehicle control system to perform operations, the operations comprising:

(a) obtaining sensor data descriptive of an environment of an autonomous vehicle;

(b) obtaining a plurality of travel way markers from map data descriptive of the environment;

(c) determining, using a machine-learned object detection model and based on the sensor data, an association between one or more travel way markers of the plurality travel way markers and an object in the environment; and

(d) generating, using the machine-learned object detection model, an offset with respect to the one or more travel way markers of a spatial region of the environment associated with the object;

wherein (c) comprises:

inputting the travel way markers and the sensor data to the machine-learned object detection model; and

obtaining object data from the machine-learned object detection model at projected locations of the travel way markers in a reference frame of the sensor data, wherein the object data indicates that the object is likely to be present at a projected location of the one or more travel way markers.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jul 18, 2023
From: FOUCARD, LOUIS
To: AURORA OPERATIONS, INC.
Reel/Frame 064302/0863 →
References Cited (95)
US 11209820B2 · Isele · 2021 [cited by examiner]
US 11738772B1 · Beller · 2023 [cited by examiner]
US 20180340790A1 · Kislovskiy · 2018 [cited by examiner]
US 20180341261A1 · Kislovskiy · 2018 [cited by examiner]
US 20180341888A1 · Kislovskiy · 2018 [cited by examiner]
US 20180365533A1 · Sathyanarayana · 2018 [cited by examiner]
US 20190354782A1 · Kee · 2019 [cited by examiner]
US 20200027241A1 · Tong et al. · 2020 [cited by applicant]
US 20200159225A1 · Zeng · 2020 [cited by examiner]
US 20200409358A1 · Gogna · 2020 [cited by examiner]
US 20210150230A1 · Smolyanskiy · 2021 [cited by examiner]
US 20210158696A1 · McNew · 2021 [cited by examiner]
US 20210181758A1 · Das · 2021 [cited by examiner]
US 20210182609A1 · Arar · 2021 [cited by examiner]
US 20210197858A1 · Zhang · 2021 [cited by examiner]
US 20210208273A1 · Yu et al. · 2021 [cited by applicant]
US 20210253128A1 · Nister · 2021 [cited by examiner]
US 20210287459A1 · Cella · 2021 [cited by examiner]
US 20210383533A1 · Zhao · 2021 [cited by examiner]
US 20220001892A1 · Fairley · 2022 [cited by examiner]
US 20220051017A1 · Choi · 2022 [cited by examiner]
US 20220114888A1 · Napanda et al. · 2022 [cited by applicant]
US 20220138568A1 · Smolyanskiy · 2022 [cited by examiner]
US 20220144304A1 · Ehsanibenafati · 2022 [cited by examiner]
US 20220164602A1 · Frtunikj · 2022 [cited by examiner]
US 20220188695A1 · Zhu · 2022 [cited by examiner]
US 20220230021A1 · Muehlenstaedt · 2022 [cited by examiner]
US 20220289180A1 · Casas · 2022 [cited by examiner]
US 20220379913A1 · Rodriguez Hervas · 2022 [cited by examiner]
US 20230020135A1 · Ferguson · 2023 [cited by examiner]
US 20230037767A1 · Yang · 2023 [cited by examiner]
US 20230049567A1 · Popov · 2023 [cited by examiner]
US 20230054759A1 · Robinson · 2023 [cited by examiner]
US 20230099494A1 · Kocamaz · 2023 [cited by examiner]
US 20230112004A1 · Hari · 2023 [cited by examiner]
US 20230192141A1 · Zhang · 2023 [cited by examiner]
US 20230229889A1 · Urtasun · 2023 [cited by examiner]
US 20230258794A1 · Karrman · 2023 [cited by examiner]
US 20230286539A1 · Malloch · 2023 [cited by examiner]
US 20230296758A1 · Akbarzadeh · 2023 [cited by examiner]
US 20230298361A1 · Yang · 2023 [cited by examiner]
US 20230324194A1 · Akbarzadeh · 2023 [cited by examiner]
US 20230351638A1 · Wu · 2023 [cited by examiner]
US 20230351769A1 · Wu · 2023 [cited by examiner]
US 20230386323A1 · Georgiou · 2023 [cited by examiner]
US 20230388481A1 · Yin · 2023 [cited by examiner]
US 20240020968A1 · Haskin · 2024 [cited by examiner]
US 20240043040A1 · Liu · 2024 [cited by examiner]
US 20240126268A1 · Yang · 2024 [cited by examiner]
US 20240127579A1 · Capellier · 2024 [cited by examiner]
US 20240220675A1 · Aghaei · 2024 [cited by examiner]
AU 2024202759A1 · 2024 [cited by examiner]
CN 117274944A · 2023 [cited by examiner]
EP 3623761A1 · 2020 [cited by applicant]
GB 2610663A · 2023 [cited by examiner]
GB 2612857A · 2023 [cited by examiner]
KR 20210030136A · 2021 [cited by applicant]
WO WO2023126914A2 · 2023 [cited by examiner]
WO WO2023235094A1 · 2023 [cited by examiner]
International Search Report and Written Opinion for Application No. PCT/US2023/086418, mailed Apr. 30, 2024, 10 pages. [cited by applicant]
Aurora Blog. “FMCW lidar: The self-driving game-changer”, 2020. 6 pgs. [cited by applicant]
Aurora Blog, “Continuous real-time sensor recalibration: A long-range perception game-changer”, Mar. 15, 2023, https://blog.aurora.tech/engineering/continuous-real-time-sensor-recalibration, retrieved on May 2, 2024, 8 … [cited by applicant]
Bertoni, et al. “MonoLoco: Monocular 3d Pedestrian Localization and Uncertainty Estimation”, https://openaccess.thecvf.com/content_ICCV_2019/papers/Bertoni_MonoLoco_Monocular_3D_Pedestrian_Localization_and_Uncertainty_E… [cited by applicant]
California Department of Transportation. “Move over. It's the law.” News Release. https://dot.ca.gov/news-releases/news-release-2019-080. Oct. 16, 2019. 2 pgs. [cited by applicant]
Casas, et al. “Intentnet: Learning to predict intention from raw sensor data.” Proceedings of Machine Learning Research, Oct. 29-31, 2018. https://proceedings.mlr.press/v87/casas18a/casas18a.pdf. 10 pgs. [cited by applicant]
Chen, et al. “Multi-View 3D Object Detection Network for Autonomous Driving.” https://arxiv.org/pdf/1611.07759, Jun. 22, 2017. 9 pgs. [cited by applicant]
Federal Motor Carrier Safety Administration. “Long stopping distances”, federal motor carrier safety administration. Video. https://www.fmcsa.dot.gov/ourroads/, 2022. [cited by applicant]
Ku et al. “Joint 3D Proposal Generation and Object Detection from View Aggregation” https://arxiv.org/pdf/1712.02294, Jul. 12, 2018. 8 pgs. [cited by applicant]
Lang et al. “PointPillars: Fast Encoders for Object Detection From Point Clouds” https://openaccess.thecvf.com/content_CVPR_2019/papers/Lang_PointPillars_Fast_Encoders_for_Object_Detection_From_Point_Clouds_CVPR_2019_pa… [cited by applicant]
Li et al. “GS3D: An Efficient 3D Object Detection Framework for Autonomous Driving” https://arxiv.org/pdf/1903.10955. Mar. 27, 2019. 9 pgs. [cited by applicant]
Liang et al. “BEVFusion: A Simple and Robust LiDAR-Camera Fusion Framework” https://arxiv.org/pdf/2205.13790. Nov. 11, 2022. 17 pgs. [cited by applicant]
Liang et al. “Deep Continuous Fusion for Multi-Sensor 3D Object Detection” https://arxiv.org/pdf/2012.10992. Dec. 20, 2020. 16 pgs. [cited by applicant]
Liang et al. “Multi-Task Multi-Sensor Fusion for 3D Object Detection” https://arxiv.org/pdf/2012.12397. Dec. 22, 2020. 11 pgs. [cited by applicant]
Liang et al. “PnPNet: End-to-End Perception and Prediction with Tracking in the Loop” https://openaccess.thecvf.com/content_CVPR_2020/papers/Liang_PnPNet_End-to-End_Perception_and_Prediction_With_Tracking_in_the_Loop_CV… [cited by applicant]
Luo et al. “Fast and Furious: Real Time End-to-End 3D Detection, Tracking and Motion Forecasting with a Single Convolutional Net” https://arxiv.org/pdf/2012.12395. Dec. 22, 2020. 9 pgs. [cited by applicant]
Meyer et al. “Lasernet: An Efficient Probabilistic 3D Object Detector for Autonomous Driving” https://arxiv.org/pdf/1903.08701. Mar. 20, 2019. 10 pgs. [cited by applicant]
Meyer et al. “Sensor Fusion for Joint 3D Object Detection and Semantic Segmentation” https://arxiv.org/pdf/1904.11466. Apr. 25, 2019. 8 pgs. [cited by applicant]
Mousavian et al. “3D Bounding Box Estimation Using Deep Learning and Geometry” https://arxiv.org/pdf/1612.00496. Apr. 10, 2017. 10 pgs. [cited by applicant]
Pfeuffer et al. “Robust Semantic Segmentation in Adverse Weather Conditions by means of Sensor Data Fusion” https://arxiv.org/pdf/1905.10117. May 24, 2019. 8 pgs. [cited by applicant]
Pierrottet et al. “Linear FMCW Laser Radar for Precision Range and Vector Velocity Measurements” https://ntrs.nasa.gov/api/citations/20080026181/downloads/20080026181.pdf. 9 pgs. [cited by applicant]
Qi et al. “PointNet: Deep Learning on Point Sets for 3D Classification and Segmentation” https://arxiv.org/pdf/1612.00593. Apr. 10, 2017. 19 pgs. [cited by applicant]
Qi et al. “PointNet++: Deep Hierarchical Feature Learning on Point Sets in a Metric Space” https://arxiv.org/pdf/1706.02413. Jun. 7, 2017. 14 pgs. [cited by applicant]
Qin et al. “MonoGRNet: A Geometric Reasoning Network for Monocular 3D Object Localization” https://arxiv.org/pdf/1811.10247. Mar. 31, 2020. 9 pgs. [cited by applicant]
Qin et al. “Triangulation Learning Network: from Monocular to Stereo 3D Object Detection” https://arxiv.org/pdf/1906.01193. Jun. 4, 2019. 9 pgs. [cited by applicant]
Shi et al. “PointRCNN: 3D Object Proposal Generation and Detection from Point Cloud” https://arxiv.org/pdf/1812.04244. May 16, 2019. 10 pgs. [cited by applicant]
Vora et al. “PointPainting: Sequential Fusion for 3D Object Detection” https://arxiv.org/pdf/1911.10150. May 6, 2020. 11 pgs. [cited by applicant]
Wang et al. “Deep Parametric Continuous Convolutional Neural Networks” https://arxiv.org/pdf/2101.06742. Jan. 17, 2021. 18 pgs. [cited by applicant]
Wang et al. “Pseudo-LiDAR from Visual Depth Estimation: Bridging the Gap in 3D Object Detection for Autonomous Driving” https://arxiv.org/pdf/1812.07179. Feb. 22, 2020. 16 pgs. [cited by applicant]
Xu et al. “Multi-Level Fusion based 3D Object Detection from Monocular Images” https://openaccess.thecvf.com/content_cvpr_2018/papers/Xu_Multi-Level_Fusion_Based_CVPR_2018_paper.pdf. Jun. 2018. 9 pgs. [cited by applicant]
Yang et al. “HDNET:Exploiting HD Maps for 3D Object Detection” https://arxiv.org/pdf/2012.11704. Dec. 21, 2020. 10 pgs. [cited by applicant]
Zhou et al. “Objects as Points” https://arxiv.org/pdf/1904.07850. Apr. 25, 2019. 12 pgs. [cited by applicant]
Zhou et al. “Voxelnet: End-to-end learning for point cloud based 3d object detection” https://arxiv.org/pdf/1711.06396, Nov. 17, 2017. 10 pgs. [cited by applicant]
Bai, et al. “TransFusion: Robust lidar-camera fusion for 3D object detection with transformers.” Computer Vision Foundation. Conference on Computer Vision and Pattern Recognition, 2020. 10 pages. [cited by applicant]
Shi, et al. “PV-RCNN: Point-voxel feature set abstraction for 3D object detection.” arXiv:2012.00463v3. Nov. 7, 2022. 19 pages. [cited by applicant]
Zhu, et al. “ConQueR: query contrast voxel-DETR for 3D object detection.” arXiv:2212.07289v1, Dec. 14, 2022. 11 pages. [cited by applicant]
Cited By (1)
US 12,631,469