IP Library Granted Patent US 12,258,016
Granted Patent B2
US 12,258,016 · App. 18/584,410 · Granted Mar 25, 2025

Methods and devices for triggering vehicular actions based on passenger actions

Inventors: Min-An Chao (Munich, DE); Neslihan Kose Cihangir (Munich, DE); Rafael Rosales (Unterhaching, DE)
Assignee: Mobileye Vision Technologies Ltd.
B60W30/16B60W40/08B60W60/001G06N3/04G06T7/50G06V10/764G06V10/82G06V20/40G06V20/58G06V20/593G06V20/597B60W2420/403G06T2207/10016G06T2207/30252
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,258,016
App. No.
18/584,410
Granted
Mar 25, 2025
Kind
B2
Abstract

Autonomous driving system methods and devices which trigger vehicular actions based on the monitoring of one or more occupants of a vehicle are presented. The methods, and corresponding devices, may include identifying a plurality of features in a plurality of subsets of image data detailing the one or more occupants; tracking changes over time of the plurality of features over the plurality of subsets of image data; determining a state, from a plurality of states, of the one or more occupants based on the tracked changes; and triggering the vehicular action based on the determined state.

Claims (42)

1. A system for triggering a vehicular action based on monitoring of one or more occupants of a vehicle, the system comprising:

at least one processor comprising circuitry and a memory, wherein the memory includes instructions that when executed by the circuitry cause the at least one processor to:

receive, from an image acquisition device, a plurality of subsets of image data detailing the one or more occupants;

identify a plurality of features of the one or more occupants in the plurality of subsets of image data;

identify, based on the plurality of subsets of image data, at least one change in the plurality of features over time, wherein identifying the at least one change includes inputting at least some of the plurality of subsets of image data into a spatio-temporal model;

determine a state of the one or more occupants from a plurality of states based on the at least one change, wherein the spatio-temporal model includes an output layer providing a score for each of the plurality of states, and wherein the state that is associated with the one or more occupants is based on the score;

determine, based on the determined state of the one or more occupants and an assumed minimum braking rate of the vehicle, a theoretical safe distance between the vehicle and at least one object; and

trigger the vehicular action based on the determined theoretical safe distance.

2. The system of claim 1 , wherein each of the plurality of states is associated with a corresponding time value and wherein the theoretical safe distance is determined based on a time value associated with the determined state of the one or more occupants.

3. The system of claim 2 , wherein the time value represents an estimated reaction time of the driver.

4. The system of claim 2 , wherein the time value is determined using a look-up table correlating the plurality of states to the corresponding time values.

5. The system of claim 1 , wherein at least one of the plurality of states corresponds to a classification of an action being performed by the one or more occupants while driving.

6. The system of claim 1 , wherein determining the state of the one or more occupants includes selecting a state from the plurality of states having a highest score.

7. The system of claim 6 , wherein triggering the vehicular action based on the determined theoretical safe distance includes comparing the theoretical safe distance to the real physical distance.

8. The system of claim 1 , wherein the memory incudes instructions that when executed by the circuitry further cause the at least one processor to determine a real physical distance to the at least one object.

9. The system of claim 8 , wherein triggering the vehicular action includes causing a braking response of the vehicle until the physical distance exceeds the theoretical safe distance.

10. The system of claim 1 , wherein the theoretical safe distance between the vehicle and at least one object is further determined based on an assumed maximum acceleration value of the vehicle.

11. The system of claim 1 , wherein triggering the vehicular action includes causing a notification to be presented to the at least one occupant.

12. The system of claim 11 , wherein the notification is presented via a user interface of the vehicle.

13. The system of claim 1 , wherein the vehicular action includes a braking response of the vehicle.

14. The system of claim 1 , wherein identifying the plurality of features includes applying a Feature-Level Temporal Filtering based model.

15. The system of claim 1 , wherein identifying the plurality of features includes inputting the plurality of subsets of image data into a neural network.

16. A method for triggering a vehicular action based on monitoring of one or more occupants of a vehicle, the method comprising:

receiving, from an image acquisition device, a plurality of subsets of image data detailing the one or more occupants;

identifying a plurality of features of the one or more occupants in the plurality of subsets of image data;

identifying, based on the plurality of subsets of image data, at least one change in the plurality of features over time, wherein identifying the at least one change includes inputting at least some of the plurality of subsets of image data into a spatio-temporal model;

determining a state of the one or more occupants from a plurality of states based on the at least one change, wherein the spatio-temporal model includes an output layer providing a score for each of the plurality of states, and wherein the state that is associated with the one or more occupants is based on the score;

determining, based on the determined state of the one or more occupants and an assumed minimum braking rate of the vehicle, a theoretical safe distance between the vehicle and at least one object; and

triggering the vehicular action based on the determined theoretical safe distance.

17. The method of claim 16 , wherein each of the plurality of states is associated with a corresponding time value and wherein the theoretical safe distance is determined based on a time value associated with the determined state of the one or more occupants.

18. The method of claim 17 , wherein the time value represents an estimated reaction time of the driver.

19. The method of claim 16 , wherein the method further includes determining a real physical distance to the at least one object and wherein triggering the vehicular action based on the determined theoretical safe distance includes comparing the theoretical safe distance to the real physical distance.

20. A non-transitory computer readable medium containing instructions that when executed by at least one processor, cause the at least one processor to perform a method for triggering a vehicular action based on monitoring of one or more occupants of a vehicle, the method comprising:

receiving, from an image acquisition device, a plurality of subsets of image data detailing the one or more occupants;

identifying a plurality of features of the one or more occupants in the plurality of subsets of image data;

identifying, based on the plurality of subsets of image data, at least one change in the plurality of features over time, wherein identifying the at least one change includes inputting at least some of the plurality of subsets of image data into a spatio-temporal model;

determining a state of the one or more occupants from a plurality of states based on the at least one change, wherein the spatio-temporal model includes an output layer providing a score for each of the plurality of states, and wherein the state that is associated with the one or more occupants is based on the score;

determining, based on the determined state of the one or more occupants and an assumed minimum braking rate of the vehicle, a theoretical safe distance between the vehicle and at least one object; and

triggering the vehicular action based on the determined theoretical safe distance.

21. The non-transitory computer readable medium of claim 20 , wherein at least one of the plurality of states corresponds to a classification of an action being performed by the one or more occupants while driving.

22. The non-transitory computer readable medium of claim 20 , wherein the theoretical safe distance between the vehicle and at least one object is further determined based on an assumed maximum acceleration value of the vehicle.

23. The non-transitory computer readable medium of claim 20 , wherein triggering the vehicular action includes causing a notification to be presented to the at least one occupant.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 17, 2025
From: INTEL CORPORATION
To: MOBILEYE VISION TECHNOLOGIES LTD.
Reel/Frame 069910/0069 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 17, 2025
From: CHAO, MIN-AN; KOSE CIHANGIR, NESLIHAN; ROSALES, RAFAEL
To: INTEL CORPORATION
Reel/Frame 069940/0001 →
Continuity (3)
Continuation 18077258 · Dec 8, 2022
Continuation 16830312 · Mar 26, 2020
Related Publication 20240190431A1 · Jun 13, 2024
References Cited (29)
US 8384531B2 · Szczerba et al. · 2013 [cited by applicant]
US 9111255B2 · Nurmi · 2015 [cited by applicant]
US 9235751B2 · Sukegawa et al. · 2016 [cited by applicant]
US 9405982B2 · Zhang · 2016 [cited by examiner]
US 10909721B2 · Noble et al. · 2021 [cited by applicant]
US 11062125B2 · Ishii · 2021 [cited by applicant]
US 11093788B2 · Possos et al. · 2021 [cited by applicant]
US 11161470B2 · Saito et al. · 2021 [cited by applicant]
US 11288789B1 · Chen et al. · 2022 [cited by applicant]
US 11295151B2 · Naidu et al. · 2022 [cited by applicant]
US 11295458B2 · Dasgupta et al. · 2022 [cited by applicant]
US 20120075450A1 · Ding et al. · 2012 [cited by applicant]
US 20180126901A1 · Levkova et al. · 2018 [cited by applicant]
US 20200216027A1 · Deng et al. · 2020 [cited by applicant]
US 20200216080A1 · Soltanian · 2020 [cited by examiner]
US 20210056306A1 · Hu · 2021 [cited by examiner]
DE 102021125234A1 · 2022 [cited by applicant]
JP 2004504219A · 2004 [cited by applicant]
Wang et al., “Temporal segment networks: Towards good practices for deep action recognition”; European Conference on Computer Vision (ECCV); Aug. 2, 2016; pp. 1-16. [cited by applicant]
Wang et al., “Temporal segment networks for action recognition in videos”; IEEE Transactions on Pattern Analysis and Machine Intelligence, Nov. 2019; pp. 2740-2755; vol. 41, No. 11. [cited by applicant]
Carreira et al., “Quo vadis, action recognition? A new model and the kinetics dataset”; IEEE Conference on Computer Vision and Pattern Recognition (CVPR); Honolulu, USA; Jul. 2017; pp. 4724-4733. [cited by applicant]
Donahue et al., “Long-term recurrent convolutional networks for visual recognition and description”; IEEE Conference on Computer Vision and Pattern Recognition (CVPR); 2015; pp. 2625-2634. [cited by applicant]
Yue-Hei Ng et al., “Beyond short snippets: Deep networks for video classification”; IEEE Conference on Computer Vision nd Pattern Recognition (CVPR); 2015; pp. 4694-4702. [cited by applicant]
Kopuklu et al., “Motion fused frames: Data level fusion strategy for hand gesture recognition”; IEEE Conference on Computer Vision and Pattern Recognition (CVPR) Workshop; 2018; pp. 2216-2224. [cited by applicant]
Russakovsky et al., “Imagenet large scale visual recognition challenge”; International journal of computer vision; Jan. 30, 2015; 43 pages. [cited by applicant]
Liwicki et al., “Efficient Online Subspace Learning with an Indefinite Kernel for Visual Tracking and Recognition”, IEEE, Oct. 2012, pp. 1624-1636, IEEE Transactions on Neural Networks and Learning Systems, vol. 23, Iss… [cited by applicant]
Communication issued from the European Patent Office in Application No. 20205118.1, dated Feb. 3, 2025 (12 pages). [cited by applicant]
Van Engelen et al., “A survey on semi-supervised learning,” Machine Learning, Springer US, vol. 109, No. 2, Nov. 15, 2019, pp. 373-440. [cited by applicant]
Krizhevsky et al., “ImageNet classification with deep convolutional neural networks,” Advances in Neural Information Processing Systems 25: 26th Annual Conference on Neural Information Processing Systems 2012, Dec. 6, 2… [cited by applicant]