IP Library Granted Patent US 12,469,160
Granted Patent B2
US 12,469,160 · App. 18/440,764 · Granted Nov 11, 2025

Artificial intelligence modeling techniques for vision-based occupancy determination

Inventors: Pengfei Duan (Austin, TX); Nishant Desai (Austin, TX); Philip Lee (Austin, TX); Ashok Elluswamy (Austin, TX)
Assignee: Tesla, Inc.
G06T7/62G06T15/08B60W60/001B60W2420/403G06T2207/20081G06T2207/30252
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,469,160
App. No.
18/440,764
Granted
Nov 11, 2025
Kind
B2
Abstract

Disclosed herein are methods and systems for using artificial intelligence modeling techniques to train and execute an artificial intelligence model to analyze camera feed received from an ego to generate an occupancy data indicating whether different voxels within the ego's surroundings are occupied by an object having mass. A method comprises inputting, using a camera of an ego object, image data of a space around the ego object into an artificial intelligence model; predicting, by executing the artificial intelligence model, an occupancy attribute of a plurality of voxels; and generating a dataset based on the plurality of voxels and their corresponding occupancy attribute.

Claims (41)

1 . A method for generating a three-dimensional occupancy grid of a space around an ego object based on two-dimensional visual data, the method comprising:

inputting, by a processor using a camera of an ego object configured to navigate a space, periodically captured two-dimensional (2D) visual data from the camera of the space around the ego object into an artificial intelligence model to cause the artificial intelligence model to generate an output using image data comprising only the 2D visual data from the camera;

periodically predicting, by the processor executing the artificial intelligence model using only the captured 2D visual data from the camera, an occupancy attribute of 3D occupancy data; and

generating, by the processor using only the captured 2D visual data from the camera, a dataset based on the 3D occupancy data and their corresponding occupancy attribute.

2 . The method of claim 1 , further comprising:

generating, by the processor, an output representing an environment of the ego object and illustrating the 3D occupancy data and corresponding occupancy attributes, wherein the output comprises a graphical indicator of the occupancy attribute for at least a portion of the 3D occupancy data.

3 . The method of claim 2 , wherein the graphical indicator corresponds to a detected object associated with the at least the portion of the plurality of vexels 3D occupancy data.

4 . The method of claim 2 , further comprising:

displaying, by the processor, the output on a screen associated with the ego object.

5 . The method of claim 1 , wherein the dataset is a queryable dataset configured to transmit the occupancy attribute of the 3D occupancy data to an autonomous driving protocol of the ego object.

6 . The method of claim 1 , wherein the artificial intelligence model is trained using a sensor attribute of the 3D occupancy data.

7 . The method of claim 1 , wherein the ego object is an autonomous vehicle executing a driving protocol based on the dataset.

8 . The method of claim 1 , further comprising:

featurizing, by the processor, the 2D visual data prior to executing the artificial intelligence model.

9 . The method of claim 1 , wherein the 2D visual data comprises a plurality of camera feeds from a plurality of cameras of the ego object, the method further comprising:

temporally aligning, by the processor, the plurality of camera feeds.

10 . An ego object comprising:

a camera;

a first processor;

a second processor;

a non-transitory computer-readable medium containing an artificial intelligence model configured to be executed by the first processor, wherein the first processor is configured to:

input, using the camera of the ego object configured to navigate a space, periodically captured two-dimensional (2D) visual data from the camera of the space around the ego object into the artificial intelligence model to cause the artificial intelligence model to generate an output using image data comprising only the 2D visual data from the camera;

periodically predict, executing the artificial intelligence model using only the captured 2D visual data from the camera, an occupancy attribute of a plurality of voxels 3D occupancy data; and

generate, using only the captured 2D visual data from the camera, a dataset based on the 3D occupancy data and their corresponding occupancy attribute,

wherein the second processor is configured to:

autonomously navigate the ego object using the dataset.

11 . The ego object of claim 10 , wherein the first processor is further configured to:

generate an output representing an environment of the ego object and illustrating the 3D occupancy data and their corresponding occupancy attribute, wherein the output comprises a graphical indicator of the occupancy attribute for at least a portion of the 3D occupancy data.

12 . The ego object of claim 11 , wherein the graphical indicator corresponds to a detected object associated with the at least the portion of the 3D occupancy data.

13 . The ego object of claim 11 , wherein the first processor is further configured to:

display the output on a screen associated with the ego object.

14 . The ego object of claim 10 , wherein the artificial intelligence model is trained using a sensor attribute of the 3D occupancy data.

15 . The ego object of claim 10 , wherein the ego object is an autonomous vehicle executing a driving protocol based on the dataset.

16 . A method comprising:

training, by a processor, an artificial intelligence model using a training dataset comprising first two-dimensional (2D) visual data received from a camera of an ego object, the training dataset having a first set of data points where each data point within the set of data points corresponds to a location and an image attribute of 3D occupancy data of space around the ego object,

whereby the artificial intelligence model correlates each data point within the first set of data points with a corresponding data point within a second set of data points using locations for each data point,

whereby, when the artificial intelligence model is trained, the artificial intelligence model is configured to receive a camera feed comprising second 2D visual data from a second ego object configured to navigate a space and periodically predict, using only the second 2D visual data, a third set of data points using only the camera feed, where each data point within the third set of data points corresponds an occupancy attribute indicating whether at least a portion of the 3D occupancy data of space around the second ego object is occupied by any object having mass.

17 . The method of claim 16 , wherein the artificial intelligence model is further configured to generate an output representing an environment of the ego object and illustrating the 3D occupancy data and their corresponding occupancy attribute.

18 . The method of claim 16 , wherein the training dataset further comprises a second set of data points where each data point within the second set of data points corresponds to the location and a sensor attribute of 3D occupancy data of the space around the ego object.

19 . The method of claim 17 , wherein a graphical indicator corresponds to a detected object associated with at least portion of the 3D occupancy data.

20 . The method of claim 17 , wherein the artificial intelligence model uses a three-dimensional multiview reconstruction protocol to generate the output.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Feb 13, 2024
From: DUAN, PENGFEI PHIL; DESAI, NISHANT; LEE, PHILIP; ELLUSWAMY, ASHOK
To: TESLA, INC.
Reel/Frame 066456/0300 →
Continuity (4)
Continuation PCTUS2023032214 · Sep 7, 2023
Provisional Application 63377954 · Sep 30, 2022
Provisional Application 63375199 · Sep 9, 2022
Related Publication 20240185445A1 · Jun 6, 2024
References Cited (57)
US 9390128B1 · Seetala · 2016 [cited by applicant]
US 10216217B1 · Santan et al. · 2019 [cited by applicant]
US 20040105573A1 · Neumann et al. · 2004 [cited by applicant]
US 20060182431A1 · Kobayashi et al. · 2006 [cited by applicant]
US 20080307240A1 · Dahan et al. · 2008 [cited by applicant]
US 20100115047A1 · Briscoe et al. · 2010 [cited by applicant]
US 20110187924A1 · Toraichi et al. · 2011 [cited by applicant]
US 20170004157A1 · Varadarajan et al. · 2017 [cited by applicant]
US 20170162177A1 · Lebeck et al. · 2017 [cited by applicant]
US 20180336481A1 · Guttmann · 2018 [cited by examiner]
US 20190035101A1 · Kwant et al. · 2019 [cited by applicant]
US 20190197778A1 · Sachdeva et al. · 2019 [cited by applicant]
US 20190286478A1 · Sengupta et al. · 2019 [cited by applicant]
US 20190310627A1 · Halder et al. · 2019 [cited by applicant]
US 20190310650A1 · Halder · 2019 [cited by examiner]
US 20190382007A1 · Casas et al. · 2019 [cited by applicant]
US 20200135014A1 · Gonzalez et al. · 2020 [cited by applicant]
US 20200232800A1 · Bai et al. · 2020 [cited by applicant]
US 20200327702A1 · Wang et al. · 2020 [cited by applicant]
US 20200377105A1 · Murashkin et al. · 2020 [cited by applicant]
US 20200380257A1 · He et al. · 2020 [cited by applicant]
US 20200387799A1 · Vivekraja et al. · 2020 [cited by applicant]
US 20200410259A1 · Srinivasan · 2020 [cited by applicant]
US 20210009166A1 · Li et al. · 2021 [cited by applicant]
US 20210049465A1 · Bogdan et al. · 2021 [cited by applicant]
US 20210061294A1 · Doemling et al. · 2021 [cited by applicant]
US 20210081780A1 · Tawari et al. · 2021 [cited by applicant]
US 20210101590A1 · Finelt et al. · 2021 [cited by applicant]
US 20210103776A1 · Jiang · 2021 [cited by examiner]
US 20210213973A1 · Carillo Peña · 2021 [cited by examiner]
US 20210255635A1 · Vora et al. · 2021 [cited by applicant]
US 20210271258A1 · Tran · 2021 [cited by examiner]
US 20210272308A1 · Dinh · 2021 [cited by examiner]
US 20210281867A1 · Golinski et al. · 2021 [cited by applicant]
US 20210357791A1 · Loginov · 2021 [cited by applicant]
US 20210398338A1 · Philion · 2021 [cited by examiner]
US 20220024485A1 · Theverapperuma · 2022 [cited by examiner]
US 20220026920A1 · Ebrahimi Afrouzi · 2022 [cited by examiner]
US 20220044114A1 · Sriram et al. · 2022 [cited by applicant]
US 20220044359A1 · Harrison · 2022 [cited by applicant]
US 20220139618A1 · Kang et al. · 2022 [cited by applicant]
US 20220147808A1 · Ming Chang et al. · 2022 [cited by applicant]
US 20220292699A1 · Zhu · 2022 [cited by examiner]
CN 104200445B · 2014 [cited by applicant]
Foreign Office Action on PCT PCT/US2023/032214 dated Oct. 31, 2023 (2 pages). [cited by applicant]
International Search Report and Written Opinion on PCT App. PCT/US2023/034091 dated Feb. 2, 2024 (14 pages). [cited by applicant]
International Search Report and Written Opinion on PCT App. PCT/US2023/075631 dated Jan. 19, 2024 (14 pages). [cited by applicant]
International Search Report and Written Opinion on PCT matter PCT/US2023/032214 dated Jan. 31, 2024 (20 pages). [cited by applicant]
ISR and WO on PCT patent application No. PCR/US2023/034168 dated Jan. 9, 2024 (27 pages). [cited by applicant]
ISR and WO on PCT patent application No. PCT/US2023/034173 dated Jan. 3, 2024 (14 pages). [cited by applicant]
ISR and WO on PCT patent application No. PCT/US2023/034235 dated Jan. 11, 2024 (7 pages). [cited by applicant]
ISR and WO on PCT patent application No. PCT/US2023/075632 dated Jan. 12, 2024 (15 pages). [cited by applicant]
ISR and WO on PCT patent application No. PCT/US2023/34233 dated Jan. 12, 2024 (14 pages). [cited by applicant]
International Search Report and Written Opinion on PCT app. PCT/US2023/075626 dated Feb. 14, 2024 (9 pages). [cited by applicant]
Jo et al., “Vehicle Trajectory Prediction Using Hierarchical Graph Neural Network for Considering Interaction among Multimodal Maneuvers”, Sensors, Aug. 9, 2021. [cited by applicant]
Mo et al., “ReCoG: A Deep Learning Framework with Heterogeneous Graph for Interaction-Aware Trajectory Prediction”, arXiv, Dec. 19, 2020. [cited by applicant]
International Search Report and Written Opinion for international App. No. PCT/US2023/034019. [cited by applicant]