IP Library Granted Patent US 12,190,625
Granted Patent B2
US 12,190,625 · App. 18/479,797 · Granted Jan 7, 2025

Modular predictions for complex human behaviors

Inventors: Dominic Noy (London, GB); Matthew Cameron Angus (London, GB); James Over Everard (London, GB); Wassim El Youssoufi (London, GB); Raunaq Bose (London, GB); Leslie Cees Nooteboom (London, GB); Maya Audrey Lara Pindeus (London, GB)
Assignee: Humanising Autonomy Limited
G06V40/103B60W60/0027G06F18/295G06N3/08G06N7/01G06V10/764G06V10/82G06V20/56G06V20/58B60W2554/4046
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,190,625
App. No.
18/479,797
Granted
Jan 7, 2025
Kind
B2
Abstract

A device performs operations including determining a probability that a vulnerable road user (VRU) will continue on a current path (e.g., in connection with controlling an autonomous vehicle). The device receives an image depicting a vulnerable road user (VRU). The device inputs at least a portion of the image into a model, and receives, as output from the model, a plurality of probabilities describing the VRU, each of the probabilities corresponding to a probability that the VRU is in a given state. The device determines, based on at least some of the plurality of probabilities, a probability that the VRU will exhibit a behavior, and outputs the probability that the VRU will exhibit the behavior to a control system.

Claims (49)

1. A method comprising:

receiving a plurality of sequential images comprising a first image depicting a human captured at a first time and a second image depicting the human captured at a second time later than the first time;

inputting at least a portion of the first image into a hybrid model comprising a deep learning model and a probabilistic graphical model, the deep learning model comprising a multi-task model having different branches, each different branch trained to determine a different feature;

receiving, as output from the hybrid model, a plurality of probabilities corresponding to a probability for a given variable feature corresponding to different states of the human;

inputting the plurality of probabilities and the second image into the hybrid model and receiving, as output from the hybrid model, a confidence value that the human will exhibit a behavior; and

outputting the confidence value that the human will exhibit the behavior to a control system.

2. The method of claim 1 , wherein the behavior is continuing on a current path, and wherein the method further comprises:

determining, based on at least two or more of the plurality of probabilities, a probability that the human is distracted, wherein determining the probability that the human will continue on the current path is based on the probability that the human is distracted.

3. The method of claim 2 , wherein determining, based on the at least two or more of the plurality of probabilities, the probability that the human is distracted comprises:

determining that a given probability of the plurality of probabilities has a conditional independency, where the given probability exceeds a threshold; and

responsive to determining that the given probability exceeds the threshold, determining the probability that the human is distracted by using the given probability to an exclusion of at least one other probability of the plurality of probabilities.

4. The method of claim 3 , further comprising:

determining the at least one other probability of the plurality of probabilities that is excluded based on a data structure that defines the conditional independency.

5. The method of claim 1 , wherein the probabilistic graphical model is at least one of a Markov network and a Bayesian network.

6. The method of claim 1 , wherein the confidence value that the human will exhibit the behavior is determined based on a posterior predictive distribution of possible outcomes, and wherein the confidence value is determined based on a spread of the posterior predictive distribution.

7. The method of claim 1 , wherein the control system determines whether to issue a control signal based on the confidence value score.

8. The method of claim 1 , wherein the hybrid model is applicable in a plurality of domains without a need for domain-specific training data.

9. A non-transitory computer-readable medium comprising memory with instructions encoded thereon, the instructions causing one or more processors to perform operations when executed, the instructions comprising instructions to:

receiving a plurality of sequential images comprising a first image depicting a human captured at a first time and a second image depicting the human captured at a second time later than the first time;

inputting at least a portion of the first image into a hybrid model comprising a deep learning model and a probabilistic graphical model, the deep learning model comprising a multi-task model having different branches, each different branch trained to determine a different feature;

receiving, as output from the hybrid model, a plurality of probabilities corresponding to a probability for a given variable feature corresponding to different states of the human;

inputting the plurality of probabilities and the second image into the hybrid model and receiving, as output from the hybrid model, a confidence value that the human will exhibit a behavior; and

outputting the confidence value that the human will exhibit the behavior to a control system.

10. The non-transitory computer-readable medium of claim 9 , wherein the behavior is continuing on a current path, and wherein the instructions further comprise instructions to:

determine, based on at least two or more of the plurality of probabilities, a probability that the human is distracted, wherein determining the probability that the human will continue on the current path is based on the probability that the human is distracted.

11. The non-transitory computer-readable medium of claim 10 , wherein the instructions to determine, based on the at least two or more of the plurality of probabilities, the probability that the human is distracted comprise instructions to:

determine that a given probability of the plurality of probabilities has a conditional independency, where the given probability exceeds a threshold; and

responsive to determining that the given probability exceeds the threshold, determine the probability that the human is distracted by using the given probability to an exclusion of at least one other probability of the plurality of probabilities.

12. The non-transitory computer-readable medium of claim 11 , the instructions further comprising instructions to:

determine the at least one other probability of the plurality of probabilities that is excluded based on a data structure that defines the conditional independency.

13. The non-transitory computer-readable medium of claim 9 , wherein the probabilistic graphical model is at least one of a Markov network and a Bayesian network.

14. The non-transitory computer-readable medium of claim 9 , wherein the confidence value that the human will exhibit the behavior is determined based on a posterior predictive distribution of possible outcomes, and wherein the confidence value is determined based on a spread of the posterior predictive distribution.

15. The non-transitory computer-readable medium of claim 9 , wherein the control system determines whether to issue a control signal based on the confidence value.

16. The non-transitory computer-readable medium of claim 9 , wherein the hybrid model is applicable in a plurality of domains without a need for domain-specific training data.

17. A system comprising:

memory with instructions encoded thereon; and

one or more processors that, when executing the instructions, are caused to perform operations comprising:

receiving a plurality of sequential images comprising a first image depicting a human captured at a first time and a second image depicting the human captured at a second time later than the first time;

inputting at least a portion of the first image into a hybrid model comprising a deep learning model and a probabilistic graphical model, the deep learning model comprising a multi-task model having different branches, each different branch trained to determine a different feature;

receiving, as output from the hybrid model, a plurality of probabilities corresponding to a probability for a given variable feature corresponding to different states of the human;

inputting the plurality of probabilities and the second image into the hybrid model and receiving, as output from the hybrid model, a confidence value that the human will exhibit a behavior; and

outputting the confidence value that the human will exhibit the behavior to a control system.

18. The system of claim 17 , wherein the behavior is continuing on a current path, and wherein the operations further comprise:

determining, based on at least two or more of the plurality of probabilities, a probability that the human is distracted, wherein determining the probability that the human will continue on the current path is based on the probability that the human is distracted.

19. The system of claim 18 , wherein determining, based on the at least two or more of the plurality of probabilities, the probability that the human is distracted comprises:

determining that a given probability of the plurality of probabilities has a conditional independency, where the given probability exceeds a threshold; and

responsive to determining that the given probability exceeds the threshold, determining the probability that the human is distracted by using the given probability to an exclusion of at least one other probability of the plurality of probabilities.

20. The system of claim 19 , further comprising:

determining the at least one other probability of the plurality of probabilities that is excluded based on a data structure that defines the conditional independency.

Assignments (2)
SECURITY INTEREST Recorded May 27, 2026
From: PORTABLE MULTIMEDIA LIMITED
To: IGF BUSINESS CREDIT LIMITED
Reel/Frame 075653/0536 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 3, 2023
From: NOY, DOMINIC; ANGUS, MATTHEW CAMERON; EVERARD, JAMES OVER; EL YOUSSOUFI, WASSIM; BOSE, RAUNAQ; NOOTEBOOM, LESLIE CEES; PINDEUS, MAYA AUDREY LARA
To: HUMANISING AUTONOMY LIMITED
Reel/Frame 065104/0231 →
Continuity (3)
Continuation 17011854 · Sep 3, 2020
Provisional Application 62896487 · Sep 5, 2019
Related Publication 20240029467A1 · Jan 25, 2024
References Cited (26)
US 11436537B2 · Jaenisch et al. · 2022 [cited by applicant]
US 20080312833A1 · Greene et al. · 2008 [cited by applicant]
US 20110246156A1 · Zecha et al. · 2011 [cited by applicant]
US 20170057497A1 · Laur et al. · 2017 [cited by applicant]
US 20180233048A1 · Andersson et al. · 2018 [cited by applicant]
US 20180326982A1 · Paris et al. · 2018 [cited by applicant]
US 20190012574A1 · Anthony et al. · 2019 [cited by applicant]
US 20190065939A1 · Bourgoin et al. · 2019 [cited by applicant]
US 20190176820A1 · Pindeus et al. · 2019 [cited by applicant]
US 20200247434A1 · Kim et al. · 2020 [cited by applicant]
US 20200380273A1 · Saez et al. · 2020 [cited by applicant]
US 20220153296A1 · Yu et al. · 2022 [cited by applicant]
US 20220227367A1 · Kario · 2022 [cited by examiner]
CN 102096803A · 2011 [cited by examiner]
WO WO2006070865A1 · 2006 [cited by applicant]
WO WO2016121053A1 · 2016 [cited by applicant]
B. Ivanovic, M. Pavone, “The Trajectron: Probablistic Multi-Agent Trajectory Modeling with Dynamic Spatiotemporal Graphs”, 2018 (Year: 2018). [cited by examiner]
Arnab, A., “Conditional Random Fields Meet Deep Neural Networks for Semantic Segmentation” 2018 (Year: 2018). [cited by applicant]
De Nicolao et al., “Onboard Sensor-Based Collision Risk Assessment to Improve Pedestrians' Safety,” [cited by applicant]
International Search Report and Written Opinion, Patent Cooperation Treaty Application No. PCT/IB2020/000713, Jan. 21, 2021, sixteen pages. [cited by applicant]
Kooij et al., “Context-Based Path Prediction for Targets with Switching Dynamics,” [cited by applicant]
Rasouli, A., et al., “Autonomous Vehicles That Interact with Pedestrians: A Survey of Theory and Practice,” Mar. 2019 (Year: 2019). [cited by applicant]
Ridel, D., et al., “A Literature Review on the Prediction of Pedestrian Behavior in Urban Scenarios,” 2018 (Year: 2018). [cited by applicant]
United States Office Action, U.S. Appl. No. 17/011,854, filed Mar. 22, 2023, 11 pages. [cited by applicant]
United States Office Action, U.S. Appl. No. 17/011,854, filed Sep. 12, 2022, 19 pages. [cited by applicant]
Japan Patent Office, Office Action, JP Patent Application No. 2022-514728, Aug. 27, 2024, eight pages. [cited by applicant]
Cited By (1)
US 12,461,532