IP Library › Granted Patent US 12,450,918
Granted Patent B2
US 12,450,918 · App. 17/823,916 · Granted Oct 21, 2025

Automatic lane marking extraction and classification from lidar scans

Inventors: Dhananjai Sharma (Singapore, SG); Venice Erin Baylon Liong (Singapore, SG); Sergi Adipraja Widjaja (Singapore, SG); Edouard Francois Marc Capellier (Singapore, SG)
Assignee: Motional AD LLC
G06V20/588G01S17/89G06T7/10G06V10/82G06T2207/10028
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,450,918
App. No.
17/823,916
Granted
Oct 21, 2025
Kind
B2
Abstract

Provided are methods, systems, and computer program products for generating an output map indicating a likelihood of individual elements of an image as corresponding to particular road elements, such as lane dividers, road dividers, and road boundaries. An example method may include applying a machine learning architecture to the image, which architecture includes a convolutional neural network and a sub-network capturing global context from feature maps generated by the convolutional neural network.

Claims (35)

1. A computer-implemented method implemented by at least one processor, the computer-implemented method comprising:

receiving, by the at least one processor, an image representing data from a liDAR scan of a vehicle environment;

convoluting, by the at least one processor, the image to generate a first feature map;

converting, by the at least one processor, the first feature map into a first feature vector;

applying, by the at least one processor, one or more additional convolutions to the first feature map to generate a second feature map;

converting, by the at least one processor, the second feature map into a second feature vector;

passing, by the at least one processor, an input representing at least the first feature vector and the second feature vector through a neural network to produce a scene feature vector;

generating, by the at least one processor, an output map based on at least the scene feature vector and the second feature map, wherein the output map includes a plurality of pixels and indicates a likelihood of each pixel corresponding to a road element; and

generating one or more polylines from the output map, each of the one or more polylines identifying a road element in the vehicle environment, wherein generating the one or more polylines comprises:

converting one or more adjacent pixels in the output map with a classification value satisfying a threshold into an undirected graph; and

identifying a longest path within the undirected graph as a polyline of the one or more polylines.

2. The computer-implemented method of claim 1 , wherein the image is a two-dimensional pseudo-image generated from a three-dimensional liDAR point cloud image.

3. The computer-implemented method of claim 2 , wherein the three-dimensional liDAR point cloud image is generated from an aggregate three-dimensional liDAR point cloud image representing multiple liDAR scans of the vehicle environment.

4. The computer-implemented method of claim 1 , wherein generating the output map comprises generating a segmentation map that associates a classification value to each pixel of the plurality of pixels, the classification value indicating the likelihood that the pixel corresponds to the road element.

5. The computer-implemented method of claim 4 , wherein the road element is one of a plurality of road elements, and wherein the output map associates multiple classification values to each pixel, each of the multiple classification values corresponding to a respective road element of the plurality of road elements.

6. The computer-implemented method of claim 5 , wherein the plurality of road elements comprises at least one of a lane divider indicating a division between lanes of traffic in a same direction, a road divider indicating a division between lanes of traffic in different directions, a road boundary element indicating a boundary of a roadway, stop lines, pedestrian markings, bike markings, and chevron markings.

7. The computer-implemented method of claim 1 , wherein applying one or more additional convolutions to the first feature map to generate the second feature map comprises applying a first additional convolution to the first feature map to generate an intermediary feature map and applying a second additional convolution to the intermediary feature map to generate the second feature map.

8. The computer-implemented method of claim 1 , wherein converting the first feature map into the first feature vector comprises reducing a dimensionality of the first feature map using a pooling operation.

9. The computer-implemented method of claim 1 , wherein the input representing at least the first feature vector and the second feature vector is a concatenation of at least the first feature vector and the second feature vectors.

10. The computer-implemented method of claim 1 , wherein generating the output map based on the scene feature vector and the second feature map includes deconvolving the second feature map based on at least the scene feature vector.

11. The computer-implemented method of claim 1 , further comprising generating one or more polylines from the output map, each of the one or more polylines identifying a road element in the vehicle environment.

12. A system comprising:

one or more non-transitory data stores including computer-executable instructions; and

one or more hardware processors configured to execute the computer-executable instructions to:

receive an image representing data from a liDAR scan of a vehicle environment;

convolute the image to generate a first feature map;

convert the first feature map into a first feature vector;

apply one or more additional convolutions to the first feature map to generate a second feature map;

convert the second feature map into a second feature vector;

pass an input representing at least the first feature vector and the second feature vector through a neural network to produce a scene feature vector; and

generate an output map based on at least the scene feature vector and the second feature map, wherein the output map includes a plurality of pixels and indicates a likelihood of each pixel corresponding to a road element; and

generate one or more polylines from the output map, each of the one or more polylines identifying a road element in the vehicle environment, wherein to generate the one or more polylines, the computer-executable instructions cause the system to:

convert one or more adjacent pixels in the output map with a classification value satisfying a threshold into an undirected graph; and

identify a longest path within the undirected graph as a polyline of the one or more polylines.

13. The system of claim 12 , wherein to generate the one or more polylines, the computer-executable instructions, when executed by the one or more hardware processors, further cause the system to skeletonize the output map.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Sep 13, 2022
From: SHARMA, DHANANJAI; LIONG, VENICE ERIN BAYLON; WIDJAJA, SERGI ADIPRAJA; CAPELLIER, EDOUARD FRANCOIS MARC
To: MOTIONAL AD LLC
Reel/Frame 061078/0949 →
Continuity (2)
Provisional Application 63365698 · Jun 1, 2022
Related Publication 20240096109A1 · Mar 21, 2024
References Cited (28)
US 9710714B2 · Chen · 2017 [cited by examiner]
US 10558222B2 · Fridman · 2020 [cited by examiner]
US 12122420B2 · Carillo Peña · 2024 [cited by examiner]
US 20170039436A1 · Chen et al. · 2017 [cited by applicant]
US 20190156206A1 · Graham · 2019 [cited by examiner]
US 20200025935A1 · Liang · 2020 [cited by examiner]
US 20200104690A1 · Bai · 2020 [cited by examiner]
US 20200271450A1 · Gorur Sheshagiri · 2020 [cited by examiner]
US 20210027634A1 · Li · 2021 [cited by applicant]
US 20210046861A1 · Li · 2021 [cited by examiner]
US 20210082181A1 · Shi · 2021 [cited by examiner]
US 20210150230A1 · Smolyanskiy · 2021 [cited by examiner]
US 20220073090A1 · Kakeshita · 2022 [cited by examiner]
US 20220300745A1 · Yang · 2022 [cited by examiner]
US 20230089897A1 · Pan · 2023 [cited by examiner]
US 20230139606A1 · Kim · 2023 [cited by examiner]
US 20240378137A1 · Frickenstein · 2024 [cited by examiner]
US 20240416897A1 · Iwabuchi · 2024 [cited by examiner]
CN 112241963A · 2021 [cited by applicant]
Li, et al, A Deep Learning-Based Hybrid Framework for Object Detection and Recognition in Autonomous Driving, 2020, IEEE (Year: 2020). [cited by examiner]
Lang, PointPillars: Fast Encoders for Object Detection From Point Clouds, 2019 (Year: 2019). [cited by examiner]
Lang, A. et al., “PointPillars: Fast Encoders for Object Detection from Point Clouds”, CVPR 2019, May 2019, arXiv:1812.05784v2, in 9 pages. [cited by applicant]
SAE On-Road Automated Vehicle Standards Committee, “SAE International's Standard J3016: Taxonomy and Definitions for Terms Related to Driving Automation Systems for On-Road Motor Vehicles”, Jun. 2018, in 35 pages. [cited by applicant]
Great Britain Office Action issued for Application No. GB 2218175.4, dated Dec. 15, 2022. [cited by applicant]
Great Britain Office Action issued for Application No. GB 2218175.4, dated Jun. 1, 2023. [cited by applicant]
Great Britain Office Action (Preliminary Examination) issued for Application No. GB 2414748.0, dated Oct. 11, 2024. [cited by applicant]
Great Britain Office Action issued for Application No. GB 2414748.0, dated Nov. 12, 2024. [cited by applicant]
Great Britain Office Action issued for Application No. GB 2414748.0, dated May 14, 2025. [cited by applicant]