IP Library Granted Patent US 10,528,823
Granted Patent B2
US 10,528,823 · App. 15/822,597 · Granted Jan 7, 2020

System and method for large-scale lane marking detection using multimodal sensor data

Inventors: Dazhou Guo (San Diego, CA); Yujie Wei (San Diego, CA); Xue Mei (San Diego, CA); Xiaodi Hou (San Diego, CA)
Assignee: TUSIMPLE
G06K9/00798G05D1/0248G06K9/00718G06K9/6289G06T7/10G08G1/167G05D2201/0213G06K9/209G06T2207/10028G06T2207/30256
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,528,823
App. No.
15/822,597
Granted
Jan 7, 2020
Kind
B2
Abstract

A system and method for large-scale lane marking detection using multimodal sensor data are disclosed. A particular embodiment includes: receiving image data from an image generating device mounted on a vehicle; receiving point cloud data from a distance and intensity measuring device mounted on the vehicle; fusing the image data and the point cloud data to produce a set of lane marking points in three-dimensional (3D) space that correlate to the image data and the point cloud data; and generating a lane marking map from the set of lane marking points.

Claims (35)

1. A system comprising:

a data processor; and

a multimodal lane detection module, executable by the data processor, the multimodal lane detection module being configured to perform a multimodal lane detection operation configured to:

receive image data from an image generating device mounted on a vehicle;

fit piecewise lines for each lane marking object detected in the received image data;

receive point cloud data from a distance and intensity measuring device mounted on the vehicle;

fuse the image data and the point cloud data to produce a set of lane marking points in three-dimensional (3D) space that correlate to the image data and the point cloud data, wherein the multimodal lane detection operation is configured to project 3D point cloud data on to two-dimensional (2D) image data, and to add a 3D point cloud point to the set of lane marking points if a distance between the position of the projected 3D point cloud point in 2D space and the position of at least one of the piecewise lines is within a pre-determined threshold, the pre-determined threshold being linearly decreased based on perspective depth; and

generate a lane marking map from the set of lane marking points.

2. The system of claim 1 being configured to perform a semantic segmentation operation on the received image data to identify and label objects in the image data with object category labels on a per-pixel basis.

3. The system of claim 2 being configured to train a neural network to perform the semantic segmentation operation.

4. The system of claim 1 wherein the image generating device is one or more cameras.

5. The system of claim 1 wherein the distance and intensity measuring device is one or more laser light detection and ranging (LIDAR) devices.

6. The system of claim 1 being configured to receive vehicle metrics from a vehicle subsystem.

7. The system of claim 1 being further configured to output the lane marking map to a vehicle control subsystem of the vehicle.

8. A method comprising:

receiving image data from an image generating device mounted on a vehicle;

fitting piecewise lines for each lane marking object detected in the received image data;

receiving point cloud data from a distance and intensity measuring device mounted on the vehicle;

fusing the image data and the point cloud data to produce a set of lane marking points in three-dimensional (3D) space that correlate to the image data and the point cloud data, wherein the fusing includes projecting 3D point cloud data on to two-dimensional (2D) image data, and adding a 3D point cloud point to the set of lane marking points if a distance between the position of the projected 3D point cloud point in 2D space and the position of at least one of the piecewise lines is within a pre-determined threshold, the pre-determined threshold being linearly decreased based on perspective depth; and

generating a lane marking map from the set of lane marking points.

9. The method of claim 8 including performing a semantic segmentation operation on the received image data to identify and label objects in the image data with object category labels on a per-pixel basis.

10. The method of claim 9 including training a neural network to perform the semantic segmentation operation.

11. The method of claim 8 wherein the image generating device is one or more cameras.

12. The method of claim 8 wherein the distance and intensity measuring device is one or more laser light detection and ranging (LIDAR) devices.

13. The method of claim 8 including receiving vehicle metrics from a vehicle subsystem.

14. The method of claim 8 including outputting the lane marking map to a vehicle control subsystem of the vehicle.

15. A non-transitory machine-useable storage medium embodying instructions which, when executed by a machine, cause the machine to:

receive image data from an image generating device mounted on a vehicle;

fit piecewise lines for each lane marking object detected in the received image data;

receive point cloud data from a distance and intensity measuring device mounted on the vehicle;

fuse the image data and the point cloud data to produce a set of lane marking points in three-dimensional (3D) space that correlate to the image data and the point cloud data, wherein the instructions are configured to project 3D point cloud data on to two-dimensional (2D) image data, and to add a 3D point cloud point to the set of lane marking points if a distance between the position of the projected 3D point cloud point in 2D space and the position of at least one of the piecewise lines is within a pre-determined threshold, the pre-determined threshold being linearly decreased based on perspective depth; and

generate a lane marking map from the set of lane marking points.

16. The non-transitory machine-useable storage medium of claim 15 being configured to perform a semantic segmentation operation on the received image data to identify and label objects in the image data with object category labels on a per-pixel basis.

17. The non-transitory machine-useable storage medium of claim 15 wherein the image generating device is one or more cameras.

18. The non-transitory machine-useable storage medium of claim 15 wherein the distance and intensity measuring device is one or more laser light detection and ranging (LIDAR) devices.

Assignments (3)
CHANGE OF NAME Recorded Dec 3, 2025
From: TUSIMPLE, INC.
To: CREATEAI, INC.
Reel/Frame 073832/0485 →
CHANGE OF NAME Recorded Jan 30, 2020
From: TUSIMPLE
To: TUSIMPLE, INC.
Reel/Frame 051757/0470 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 11, 2018
From: GUO, DAZHOU; WEI, YUJIE; MEI, XUE; HOU, XIAODI
To: TUSIMPLE
Reel/Frame 047468/0143 →
Continuity (1)
Related Publication 20190163989A1 · May 30, 2019
Cited By (1)
US 12,235,651