IP Library Granted Patent US 12,505,682
Granted Patent B2
US 12,505,682 · App. 17/468,799 · Granted Dec 23, 2025

Close following detection using machine learning models

Inventors: Ali Hassan (Lahore, PK); Afsheen Rafaqat Ali (Gujranwala, PK); Hussam Ullah Khan (Lahore, PK); Ijaz Akhter (Lahore, PK)
Assignee: MOTIVE TECHNOLOGIES, INC.
G06V20/588B60R1/24G06F18/2431G06N3/08G06T11/20B60R2001/1253G06T2210/12
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,505,682
App. No.
17/468,799
Granted
Dec 23, 2025
Kind
B2
Abstract

Described are embodiments for training and using a close following classifier. In the example embodiments, a system includes a backbone network configured to receive an image; and at least one prediction head communicatively coupled to the backbone network, the at least one prediction head configured to receive an output from the backbone network, wherein the at least one prediction head includes a classifier configured to classify the image as including a close-following event, the classifier receiving the output of the backbone network and a vehicle speed as inputs.

Claims (32)

1 . A system comprising:

a backbone network configured to receive an image as an input, the backbone network comprising a machine learning model; and

at least one prediction head communicatively coupled to the backbone network, the at least one prediction head configured to

receive a vector output from the backbone network,

generate a combined input vector that concatenates the vector output from the backbone network and a vehicle speed recorded at a time associated with the image,

input the combined input vector into a classifier, and

obtain a classification of the image as representing a close-following event based on processing of the combined input vector by the classifier.

2 . The system of claim 1 , wherein the at least one prediction head includes a plurality of prediction heads and the plurality of prediction heads is trained using a joint loss function aggregating losses of each of the plurality of prediction heads.

3 . The system of claim 2 , wherein the plurality of prediction heads comprises a camera obstruction detection head, a lane detection head, an object detection head, and a distance estimation head.

4 . The system of claim 3 , wherein the classifier is configured to receive inputs from the lane detection head and the distance estimation head.

5 . The system of claim 4 , wherein the distance estimation head is configured to receive an input from the object detection head.

6 . The system of claim 3 , wherein the plurality of prediction heads further comprises an intermediate neural network, the intermediate neural network configured to process the output of the backbone network and transmit the output to the classifier.

7 . The system of claim 3 , further comprising an intermediate neural network configured to process the output of the backbone network and transmit the output to each of the plurality of prediction heads.

8 . The system of claim 2 , wherein the plurality of prediction heads comprises a camera obstruction detection head, a lane detection head, and an object bounding box, lane number and distance estimation head.

9 . The system of claim 8 , wherein the classifier is configured to receive inputs from the lane detection head and the object bounding box, lane number and distance estimation head.

10 . The system of claim 8 , wherein the object bounding box, lane number and distance estimation head is configured to receive an input from the lane detection head.

11 . A system comprising:

a backbone network configured to receive an image as an input, the backbone network comprising a machine learning model; and

a classifier configured to output a classification of the image as including a close-following event or not, the classifier receiving, as input, a combined input vector that comprises a concatenation of a vector output from the backbone network and a vehicle speed recorded at a time associated with the image, the classifier trained using a neural network, the neural network comprising at least one prediction head including the classifier, the at least one prediction head communicatively coupled to the backbone network based on processing of the combined input vector by the classifier.

12 . The system of claim 11 , further comprising an intermediate neural network communicatively coupled to the backbone network and the classifier, the intermediate neural network configured to process the output of the backbone network prior to transmitting the output to the classifier.

13 . The system of claim 11 , wherein the backbone network is configured to receive a video frame from a camera.

14 . The system of claim 13 , wherein the camera comprises a camera situated in a dash-mounted or windshield-mounted device.

15 . The system of claim 14 , wherein the backbone network and the classifier are executed on the dash-mounted or windshield-mounted device.

16 . A non-transitory computer-readable storage medium for tangibly storing computer program instructions capable of being executed by a computer processor, the computer program instructions comprising steps of:

receiving an image;

processing the image using a backbone network as an input, the backbone network comprising a machine learning model, an output of the backbone network comprising a set of features;

generating combined input vectors that each concatenate the set of features from the backbone network and a respective set of vehicle speeds recorded at a time associated with the set of features, wherein the at least one prediction head generates a prediction vector, and wherein the at least one prediction head includes a classifier configured to classify the image as including a close-following event based on processing of the combined input vectors; and

adjusting parameters of the backbone network and the at least one prediction head using a loss function.

17 . The non-transitory computer-readable storage medium of claim 16 , wherein the at least one prediction head further comprises an intermediate neural network, the intermediate neural network configured to process the output of the backbone network and transmit the output to the classifier.

18 . The non-transitory computer-readable storage medium of claim 16 , further comprising processing, using an intermediate neural network, the output of the backbone network, and transmitting the output to the at least one prediction head.

19 . The non-transitory computer-readable storage medium of claim 16 , wherein the at least one prediction head comprises a plurality of prediction heads including a camera obstruction detection head, a lane detection head, and an object bounding box, lane number and distance estimation head.

20 . The non-transitory computer-readable storage medium of claim 19 , wherein the loss function comprises a joint loss function aggregating individual loss functions of the plurality of prediction heads, each of the individual loss functions associated with a corresponding prediction head in the plurality of prediction heads.

Assignments (2)
CHANGE OF NAME Recorded Apr 12, 2022
From: KEEP TRUCKIN, INC.
To: MOTIVE TECHNOLOGIES, INC.
Reel/Frame 059965/0872 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Sep 8, 2021
From: HASSAN, ALI; ALI, AFSHEEN RAFAQAT; KHAN, HUSSAM ULLAH; AKHTER, IJAZ
To: KEEP TRUCKIN, INC.
Reel/Frame 057408/0835 →
Continuity (1)
Related Publication 20230077207A1 · Mar 9, 2023
References Cited (23)
US 10984290B1 · Goel et al. · 2021 [cited by applicant]
US 11069082B1 · Ebrahimi Afrouzi · 2021 [cited by examiner]
US 11637998B1 · Pieper · 2023 [cited by examiner]
US 20190266418A1 · Xu · 2019 [cited by examiner]
US 20190384304A1 · Towal · 2019 [cited by examiner]
US 20200209864A1 · Chen · 2020 [cited by examiner]
US 20210031772A1 · Gomora · 2021 [cited by applicant]
US 20210042575A1 · Firner · 2021 [cited by examiner]
US 20210056306A1 · Hu · 2021 [cited by examiner]
US 20210086364A1 · Handa · 2021 [cited by examiner]
US 20210142160A1 · Mohseni · 2021 [cited by examiner]
US 20210191395A1 · Gao et al. · 2021 [cited by applicant]
US 20210271259A1 · Karpathy · 2021 [cited by examiner]
US 20210350150A1 · An · 2021 [cited by examiner]
US 20210383533A1 · Zhao · 2021 [cited by examiner]
CN 107909016A · 2018 [cited by examiner]
CN 110345924A · 2018 [cited by examiner]
WO 2021046578A1 · 2021 [cited by applicant]
WO 2021092702A1 · 2021 [cited by applicant]
International Search Report and Written Opinion to corresponding International Search Report PCT/US2022/076007 mailed Nov. 22, 2022 (13 pages). [cited by applicant]
Extended EP Search Report issued in corresponding EP Appln. No. 22868250.6, mailed Apr. 7, 2025, 9 pages. [cited by applicant]
Ishihara et al., “Multi-task Learning with Attention for End-to-end Autonomous Driving,” Arxic. Org., Cornell University Library, (2021). [cited by applicant]
Wenyuan et al., “DSDNet” Deep Structured Self-driving Network, pp. 156-172 (2020). [cited by applicant]