IP Library › Granted Patent US 12,347,209
Granted Patent B2
US 12,347,209 · App. 17/884,716 · Granted Jul 1, 2025

Training of 3D lane detection models for automotive applications

Inventors: Mina Alibeiginabi (Gothenburg, SE); Erik Brorsson (Gothenburg, SE); Silas Ulander (Lerum, SE); Benjamin Waubert (Gothenburg, SE)
Assignee: ZENSEACT AB
G06V20/588G06V10/7753G06V10/82G06V20/64
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,347,209
App. No.
17/884,716
Granted
Jul 1, 2025
Kind
B2
Abstract

The present invention relates to a method for training artificial neural network configured for 3D lane detection based on unlabelled image data from camera. The method includes generating a first set of 3D lane boundaries in first coordinate system based on first image, generating a second set of 3D lane boundaries in second coordinate system based on second image, transforming at least one of the second set of 3D lane boundaries and first set of 3D lane boundaries based on positional data associated with first image and second image, evaluating the first set of 3D lane boundaries against second set of 3D lane boundaries in common coordinate system in order to find matching lane pairs of first set of 3D lane boundaries and second set of 3D lane boundaries, and updating one or more model parameters of an artificial neural network based on a spatio-temporal consistency loss.

Claims (35)

1. A method for training an artificial neural network configured for 3D lane detection based on unlabelled image data from a vehicle-mounted camera, the method comprising:

generating, by means of the artificial neural network, a first set of 3D lane boundaries in a first coordinate system based a first image captured by the vehicle-mounted camera;

generating, by means of the artificial neural network, a second set of 3D lane boundaries in a second coordinate system based on a second image captured by the vehicle mounted camera, wherein the second image is captured at a later moment in time compared to the first image and wherein the first image and the second image contain at least partly overlapping road portions;

transforming at least one of the second set of 3D lane boundaries and the first set of 3D lane boundaries based on positional data associated with the first image and the second image, such that the first set of 3D lane boundaries and the second set of 3D lane boundaries have a common coordinate system;

evaluating the first set of 3D lane boundaries against the second set of 3D lane boundaries in the common coordinate system in order to find matching lane pairs of the first set of 3D lane boundaries and the second set of 3D lane boundaries; and

updating one or more model parameters of the artificial neural network based on a spatio-temporal consistency loss between the found matching lane pairs of the first set of 3D lane boundaries and the second set of 3D lane boundaries.

2. The method according to claim 1 , wherein the step of updating one or more parameters of the artificial neural network comprises penalizing the artificial neural network based on the spatio-temporal consistency loss between matching lane pairs of the first set of 3D lane boundaries and the second set of 3D lane boundaries.

3. The method according to claim 1 , wherein the spatio-temporal consistency loss is based on a Euclidean distance between matching lane pairs of the first set of 3D lane boundaries and the second set of 3D lane boundaries.

4. The method according to claim 1 , wherein the step of evaluating the first set of 3D lane boundaries against the second set of 3D lane boundaries in the common coordinate system comprises:

determining an average distance between each lane of the first set of 3D lane boundaries and each lane of the second set of 3D lane boundaries in the common coordinate system;

comparing the determined average distances against an average distance threshold; and

determining each lane pair having a determined average distance below the average distance threshold as a matching lane pair.

5. The method according to claim 1 , wherein the first image and the second image are captured at different geographical positions.

6. The method according to claim 1 , wherein the artificial neural network is configured to generate a set of anchor lines where each anchor line is associated with a confidence value indicative of a presence of a lane boundary at the corresponding anchor line.

7. The method according to claim 6 , further comprising:

disregarding any anchor line associated with a confidence value below a confidence value threshold in the transformation step.

8. The method according to claim 6 , wherein the step of updating the one or more model parameters of the artificial neural network is further based on the confidence values of any anchor lines associated with matching lane pairs and any anchor lines associated with non-matching lane pairs, such that any anchor lines associated with matching lane pairs should predict a normalized confidence level of 1,0 and any anchor lines associated with non-matching lane pairs should predict a normalized confidence value of 0.

9. A non-transitory computer-readable storage medium storing one or more instructions configured to be executed by one or more processors of a processing system, the one or more instructions for performing the method according to claim 1 .

10. An apparatus for training an artificial neural network configured for 3D lane detection based on unlabelled image data from a vehicle-mounted camera, the apparatus comprising control circuitry configured to:

generate, by means of the artificial neural network, a first set of 3D lane boundaries in a first coordinate system based a first image captured by the vehicle-mounted camera;

generate, by means of the artificial neural network, a second set of 3D lane boundaries in a second coordinate system based on a second image captured by the vehicle mounted camera, wherein the second image is captured at a later moment in time compared to the first image and wherein the first image and the second image contain at least partly overlapping road portions;

transform at least one of the second set of 3D lane boundaries and the first set of 3D lane boundaries based on positional data associated with the first image and the second image, such that the first set of 3D lane boundaries and the second set of 3D lane boundaries have a common coordinate system;

evaluate the first set of 3D lane boundaries against the second set of 3D lane boundaries in the common coordinate system in order to find matching lane pairs of the first set of 3D lane boundaries and the second set of 3D lane boundaries; and

update one or more model parameters of the artificial neural network based on a spatio-temporal consistency loss between the found matching lane pairs of the first set of 3D lane boundaries and the second set of 3D lane boundaries.

11. A vehicle comprising:

a camera configured to capture images of at least a portion of a surrounding environment of the vehicle;

a localization system configured to monitor a geographical position of the vehicle; and

an apparatus for training an artificial neural network configured for 3D lane detection based on unlabelled image data from the camera, the apparatus comprising control circuitry configured to:

generate, by means of the artificial neural network, a first set of 3D lane boundaries in a first coordinate system based a first image captured by the vehicle-mounted camera;

generate, by means of the artificial neural network, a second set of 3D lane boundaries in a second coordinate system based on a second image captured by the vehicle mounted camera, wherein the second image is captured at a later moment in time compared to the first image and wherein the first image and the second image contain at least partly overlapping road portions;

transform at least one of the second set of 3D lane boundaries and the first set of 3D lane boundaries based on positional data associated with the first image and the second image, such that the first set of 3D lane boundaries and the second set of 3D lane boundaries have a common coordinate system;

evaluate the first set of 3D lane boundaries against the second set of 3D lane boundaries in the common coordinate system in order to find matching lane pairs of the first set of 3D lane boundaries and the second set of 3D lane boundaries; and

update one or more model parameters of the artificial neural network based on a spatio-temporal consistency loss between the found matching lane pairs of the first set of 3D lane boundaries and the second set of 3D lane boundaries.

12. A remote server comprising the apparatus according to claim 10 .

13. A cloud environment comprising one or more remote servers according to claim 12 .

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Sep 6, 2022
From: ALIBEIGINABI, MINA; BRORSSON, ERIK; ULANDER, SILAS; WAUBERT, BENJAMIN
To: ZENSEACT AB
Reel/Frame 060995/0578 →
Priority Claims (1)
EP 21191605 · Aug 17, 2021 · regional
Continuity (1)
Related Publication 20230074419A1 · Mar 9, 2023
References Cited (19)
US 10867190B1 · Vajna · 2020 [cited by examiner]
US 20190171223A1 · Liang et al. · 2019 [cited by applicant]
US 20190286153A1 · Rankawat · 2019 [cited by examiner]
US 20190325595A1 · Stein · 2019 [cited by examiner]
US 20190384304A1 · Towal · 2019 [cited by examiner]
US 20200089232A1 · Gdalyahu · 2020 [cited by examiner]
US 20200218907A1 · Baik · 2020 [cited by examiner]
US 20200218909A1 · Myeong · 2020 [cited by examiner]
US 20200249684A1 · Onofrio et al. · 2020 [cited by applicant]
US 20200310450A1 · Reschka · 2020 [cited by examiner]
US 20220024518A1 · Yang · 2022 [cited by examiner]
WO 2008091565A1 · 2008 [cited by applicant]
Extended European Search Report mailed Feb. 9, 2022 for European Application No. 21191605.1, 7 pages. [cited by applicant]
Ghafoorian, Mohsen et al.; “EL-GAN: Embedding Loss Driven Generative Adversarial Networks for Lane Detection”; Cornell University Library; www.arxiv.org; Jun. 14, 2018; 14 pages. [cited by applicant]
Garnett, Noa et al.; “3D-LaneNet: End-to-End 3D Multiple Lane Detection”; Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV), 2019, pp. 2921-2930 (10 pages). [cited by applicant]
Guo, Yuliang et al.; “Gen-LaneNet: A Generalized and Scalable Approach for 3D Lane Detection”; European Conference on Computer Vision (ECCV); 2020; ages 666-681 (16 pages). [cited by applicant]
Communication pursuant to Article 94(3) EPC mailed Jan. 17, 2025 for European Patent Application No. 21181605.1, 6 pages. [cited by applicant]
Wang, Guangming et al.; “Unsupervised Learning of 3D Scene Flow from Monocular Camera”; 2021 IEEE International Conference on Robotics and Automation (IRCA 2021); Xi'an, China; May 31-Jun. 4, 2021; XP033989496; pp. 4325… [cited by applicant]
Communication pursuant to Article 94(3) EPC mailed Jan. 17, 2025 for European Patent Application No. 21191605.1, 6 pages. [cited by applicant]