IP Library › Granted Patent US 12,243,327
Granted Patent B2
US 12,243,327 · App. 18/054,844 · Granted Mar 4, 2025

Generation and update of HD maps using data from heterogeneous sources

Inventor: Gil Arditi (Palo Alto, CA)
Assignee: Lyft, Inc.
G06V20/58G01C21/32G01C21/3841G01C21/3878G05D1/0274G06F16/29G06N3/08G06N20/00
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,243,327
App. No.
18/054,844
Granted
Mar 4, 2025
Kind
B2
Abstract

In one embodiment, a method includes a computing system collecting first sensor data, of a particular geographic location, from a first type of sensor. The computing system may process the first sensor data to identify one or more first objects at the particular geographic location. The computing system may access an existing high-definition (HD) map associated with the particular geographic location. The existing HD map includes one or more second objects and is generating using second sensor data collected with a second type of sensor. The computing system may determine whether the one or more first objects are included in the existing HD map. In response to determining that the one or more first objects are not included in the existing HD map, the computing system may update the existing HD map to generate an updated HD map that includes the first objects and the second objects.

Claims (64)

1. A method comprising, by a computing system:

collecting, by the computing system of a first vehicle, first sensor data, of a particular geographic location, from a first type of sensor of the first vehicle associated with a fleet of vehicles;

processing, by the computing system of the first vehicle, the first sensor data to identify one or more first objects at the particular geographic location;

accessing, by the computing system of the first vehicle, an existing high-definition (HD) map associated with the particular geographic location;

determining, by the computing system of the first vehicle, whether the one or more first objects are included in the existing HD map; and

in response to determining that the one or more first objects are not included in the existing HD map, sending, by the computing system of the first vehicle, the first sensor data to a server to transform the first sensor data into a latent representation, wherein transforming the first sensor data into the latent representation comprises:

encoding, using a trained machine-learning model, the first sensor data collected from the first type of sensor of the first vehicle into the latent representation in a common data space to minimize discrepancies between the first sensor data collected from the first type of sensor of the first vehicle and second sensor data collected from a second type of sensor of a second vehicle;

receiving, by the computing system of the first vehicle, the latent representation of the first sensor data from the server;

decoding, by the computing system of the first vehicle, the latent representation to generate map data corresponding to the first sensor data; and

associating, by the computing system of the first vehicle, the map data corresponding to the first sensor data with the existing HD map to generate an updated HD map.

2. The method of claim 1 , wherein:

the first sensor data comprises camera data and the first type of sensor is a camera sensor; and

the second sensor data comprises light detection and ranging (LiDAR) data and the second type of sensor is a LiDAR sensor.

3. The method of claim 1 , wherein the first type of sensor and the second type of sensor are same.

4. The method of claim 1 , wherein the first type of sensor and the second type of sensor are different.

5. The method of claim 1 , further comprising:

determining that the one or more first objects are inanimate or stationary objects.

6. The method of claim 5 , wherein determining that the one or more first objects are inanimate or stationary objects comprises:

receiving additional sensor data, of the particular geographic location, from other vehicles associated with the fleet of vehicles at different times within a certain time frame; and

determining that the one or more first objects identified at the particular geographic location are consistent across the first sensor data associated with the first vehicle, the second sensor data associated with the second vehicle, and the additional sensor data associated with the other vehicles of the fleet of vehicles.

7. The method of claim 1 , wherein the trained machine-learning model comprises a convolutional neural network and a deconvolutional neural network, wherein an output of the convolutional neural network is configured to be an input of the deconvolutional neural network.

8. The method of claim 7 , wherein:

the encoding the first sensor data into the latent representation is performed using the convolutional neural network; and

the decoding the latent representation is performed using the deconvolutional neural network.

9. The method of claim 1 , wherein processing the first sensor data to identify the one or more first objects at the particular geographic location comprises:

detecting a plurality of objects based at least in part on the first sensor data;

classifying, using an object classifier, the plurality of objects into different classification types; and

determining, based on the different classification types, stationary objects from the plurality of objects, wherein the one or more first objects are the stationary objects on a road.

10. The method of claim 1 , wherein determining whether the one or more first objects are included in the existing HD map comprises:

generating, for each first object of the one or more first objects, a confidence score representing a likelihood of the first object being included in the existing HD map; and

comparing, for each first object of the one or more first objects, the confidence score with a threshold.

11. The method of claim 10 , wherein confidence score generated for the first object is based on one or more of a similarity comparison of a measured size, dimensions, classification, or location of the first object with known objects in the existing HD map.

12. The method of claim 1 , wherein determining whether the one or more first objects are included in the existing HD map comprises:

determining whether there is a mismatch between the first sensor data of the particular geographic location and the existing HD map associated with particular geographic location based on comparing map data corresponding to the first sensor data with corresponding map data in the existing HD map.

13. A system, comprising:

one or more processors and one or more computer-readable non-transitory storage media in communication with the one or more processors, the one or more computer-readable non-transitory storage media comprising instructions operable when executed by the one or more processors to cause the system to perform operations comprising:

collecting, by a computing system of a first vehicle, first sensor data, of a particular geographic location, from a first type of sensor of the first vehicle associated with a fleet of vehicles;

processing, by the computing system of the first vehicle, the first sensor data to identify one or more first objects at the particular geographic location;

accessing, by the computing system of the first vehicle, an existing high-definition (HD) map associated with the particular geographic location;

determining, by the computing system of the first vehicle, whether the one or more first objects are included in the existing HD map; and

in response to determining that the one or more first objects are not included in the existing HD map, sending, by the computing system of the first vehicle, the first sensor data to a server to transform the first sensor data into a latent representation, wherein transforming the first sensor data into the latent representation comprises:

encoding, using a trained machine-learning model, the first sensor data collected from the first type of sensor of the first vehicle into the latent representation in a common data space to minimize discrepancies between the first sensor data collected from the first type of sensor of the first vehicle and second sensor data collected from a second type of sensor of a second vehicle;

receiving, by the computing system of the first vehicle, the latent representation of the first sensor data from the server;

decoding, by the computing system of the first vehicle, the latent representation to generate map data corresponding to the first sensor data; and

associating, by the computing system of the first vehicle, the map data corresponding to the first sensor data with the existing HD map to generate an updated HD map.

14. The system of claim 13 , wherein:

the first sensor data comprises camera data and the first type of sensor is a camera sensor; and

the second sensor data comprises light detection and ranging (LiDAR) data and the second type of sensor is a LiDAR sensor.

15. The system of claim 13 , wherein the first type of sensor and the second type of sensor are different.

16. The system of claim 13 , wherein the first type of sensor and the second type of sensor are same.

17. One or more computer-readable non-transitory storage media including instructions that are operable when executed to cause one or more processors to perform operations comprising:

collecting, by a computing system of a first vehicle, first sensor data, of a particular geographic location, from a first type of sensor of the first vehicle associated with a fleet of vehicles;

processing, by the computing system of the first vehicle, the first sensor data to identify one or more first objects at the particular geographic location;

accessing, by the computing system of the first vehicle, an existing high-definition (HD) map associated with the particular geographic location;

determining, by the computing system of the first vehicle, whether the one or more first objects are included in the existing HD map; and

in response to determining that the one or more first objects are not included in the existing HD map, sending, by the computing system of the first vehicle, the first sensor data to a server to transform the first sensor data into a latent representation, wherein transforming the first sensor data into the latent representation comprises:

encoding, using a trained machine-learning model, the first sensor data collected from the first type of sensor of the first vehicle into the latent representation in a common data space to minimize discrepancies between the first sensor data collected from the first type of sensor of the first vehicle and second sensor data collected from a second type of sensor of a second vehicle;

decoding, by the computing system of the first vehicle, the latent representation to generate map data corresponding to the first sensor data; and

associating, by the computing system of the first vehicle, the map data corresponding to the first sensor data with the existing HD map to generate an updated HD map.

18. The one or more computer-readable non-transitory storage media of claim 17 , wherein:

the first sensor data comprises camera data and the first type of sensor is a camera sensor; and

the second sensor data comprises light detection and ranging (LiDAR) data and the second type of sensor is a LiDAR sensor.

19. The one or more computer-readable non-transitory storage media of claim 17 , wherein the first type of sensor and the second type of sensor are different.

20. The one or more computer-readable non-transitory storage media of claim 17 , wherein the first type of sensor and the second type of sensor are same.

Continuity (2)
Continuation 15811489 · Nov 13, 2017
Related Publication 20230132889A1 · May 4, 2023
References Cited (9)
US 20080062287A1 · Agrawal · 2008 [cited by examiner]
WO WO2016130719A2 · 2016 [cited by examiner]
Yasrab et al., An Encoder-Decoder Based Convolution Neural Network (CNN) for Future Advanced Driver Assistance System (ADAS), Applied Sciences, Mar. 17, 2017. [cited by examiner]
Du, Car Detection for Autonomous Vehicle Lidar and Vision Fusion Approach through Deep Learning Framework, 2017 IEEE International Conference on Intelligent Robots and System, Sep. 2017 (Year: 2017). [cited by examiner]
Sachan, Keras Tensorflow Tutorial: Practical Guide from Getting Started to Developing Complex Deep Neural Network, CV-Tricks.com, May 2017 (Year: 2017). [cited by examiner]
Audebert, Fusion of Heterogeneous Data in Convolutional Network for Urban Semantic, JURSE 2017 Joint Urban Remote Sensing Event, Mar. 2017 (Year: 2017). [cited by examiner]
Todas, Multimodal Machine Learning: A Survey and Taxonomy, arXiv, 2017 (Year: 2017). [cited by examiner]
Feichtenhofer, Convolutional Two-Stream Network Fusion for Video Action Recognition, arXiv, 2016 (Year: 2016). [cited by examiner]
Badrinarayanan, SegNet a Deep Convolutional Encoder-Decoder Architecture for Image Segmentation, arXiv, 2016 (Year: 2016). [cited by examiner]
Cited By (1)
US 12,662,137