IP Library Granted Patent US 12,249,184
Granted Patent B2
US 12,249,184 · App. 18/235,025 · Granted Mar 11, 2025

Methods and systems to predict activity in a sequence of images

Inventors: Alexandru Malaescu (Bucharest, RO); Dan Filip (Bucharest, RO); Mihai Ciuc (Bucharest, RO); Liviu-Cristian Dutu (Bucharest, RO); Madalin Dumitru-Guzu (Bucharest, RO)
Assignee: Tobii Technologies Limited
G06V40/20G06F18/2113G06F18/214G06F18/22G06T7/97G06V10/751G06V20/13G06V20/30
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,249,184
App. No.
18/235,025
Granted
Mar 11, 2025
Kind
B2
Abstract

A method to determine activity in a sequence of successively acquired images of a scene, comprises: acquiring the sequence of images; for each image in the sequence of images, forming a feature block of features extracted from the image and determining image specific information including a weighting for the image; normalizing the determined weightings to form a normalized weighting for each image in the sequence of images; for each image in the sequence of images, combining the associated normalized weighting and associated feature block to form a weighted feature block; passing a combination of the weighted feature blocks through a predictive module to determine an activity in the sequence of images; and outputting a result comprising the determined activity in the sequence of images.

Claims (40)

1. A driver monitoring system to determine activity in a sequence of successively acquired images of a scene within a vehicle, comprising:

memory; and

one or more processors configured to perform operations comprising:

acquiring the sequence of images;

forming, for each image in the sequence of images, a feature block of features extracted from the image;

determining, for each image in the sequence of images, image specific information, wherein the image specific information includes a weighting indicating image importance for the image and one or more likelihoods of one or more activities, wherein the weighting is determined based on retrieving a previously determined stored feature block and image specific information of the previously determined stored feature block;

storing the formed feature blocks and the determined image specific information;

passing a plurality of weighted feature blocks through a predictive model to determine an activity in the sequence of images;

comparing the determined activities with a most likely image activity in the image specific information;

validating the determined activity based on the comparison; and

controlling the vehicle according to the validated activity.

2. The system of claim 1 , wherein determining image specific information comprises determining respective likelihoods of a plurality of predetermined activities occurring in the image, and the weighting for the image relates to the highest of the determined likelihoods of the plurality of predetermined activities.

3. The system of claim 2 , wherein the operations further comprise:

comparing the determination of the activity in the sequence of images against the most likely image activity in the image specific information of at least one image in the sequence of images; and

validating the determination of the activity if no difference is found in the comparison.

4. The system of claim 3 , wherein the at least one image in the sequence of images is all images in the sequence of images.

5. The system of claim 3 , wherein operations further comprise triggering a further action if the comparison reveals at least one difference.

6. The system of claim 5 , wherein triggering a further action comprises at least one of issuing a warning; and adjusting the determination of the activity in the sequence of images to a warning or a default value.

7. The system of claim 5 , wherein triggering a further action comprises:

counting how many compared images reveal a difference in the comparison to find a number of disagreeing images; and

responsive to the number of disagreeing images being greater than half a number of images in the sequence, outputting a new determination of the activity in the sequence of images.

8. The system of claim 7 , wherein outputting a new determination of the activity in the sequence of images comprises:

when the disagreeing images all have the same most likely image activity in the image specific information, outputting the most likely image activity of the disagreeing images as the activity in the sequence of images.

9. The system of claim 2 , wherein the image specific information is determined for at least one frame.

10. The system of claim 1 , wherein forming a feature block of features extracted from the image and determining image specific information including a weighting for the image comprises:

passing each image through a feature encoding convolutional neural network to form a feature block; and

passing each feature block through an image-based module comprising at least one fully connected layer.

11. The system of claim 1 , wherein passing the plurality of weighted feature blocks through a predictive model to determine an activity in a sequence of images comprises:

passing a concatenation of the weighted feature blocks through a time-based model comprising a convolutional neural network.

12. The system of claim 1 , wherein the operations further comprise normalizing the determined weightings to form a normalized weighting for each image in the sequence of images and passing the determined weightings through a SoftMax module.

13. A method for determining an activity in a sequence of successively acquired images of a scene within a vehicle, comprising:

acquiring the sequence of images;

forming, for each image in the sequence of images, a feature block of features extracted from the image;

determining, for each image in the sequence of images, image specific information, wherein the image specific information includes a weighting indicating image importance for the image and one or more likelihoods of one or more activities, wherein the weighting is determined based on retrieving a previously determined stored feature block and image specific information of the previously determined stored feature block;

storing the formed feature blocks and the determined image specific information;

passing a plurality of weighted feature blocks through a predictive model to determine an activity in the sequence of images;

comparing the determined activities with a most likely image activity in the image specific information;

validating the determined activity based on the comparison; and

controlling the vehicle according to the validated activity.

14. The method of claim 13 , wherein determining image specific information comprises determining respective likelihoods of a plurality of predetermined activities occurring in the image, and the weighting for the image relates to the highest of the determined likelihoods of the plurality of predetermined activities.

Assignments (2)
CHANGE OF NAME Recorded May 19, 2025
From: FOTONATION LIMITED
To: TOBII TECHNOLOGIES LIMITED
Reel/Frame 071292/0964 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 18, 2023
From: MALAESCU, ALEXANDRU; FILIP, DAN; CIUC, MIHAI; DUTU, LIVIU-CRISTIAN; DUMITRU-GUZU, MADALIN
To: FOTONATION LIMITED
Reel/Frame 064631/0004 →
Continuity (2)
Continuation 16929051 · Jul 14, 2020
Related Publication 20230419727A1 · Dec 28, 2023
References Cited (55)
US 10152642B2 · Chang et al. · 2018 [cited by applicant]
US 10528047B1 · Trujillo · 2020 [cited by examiner]
US 10607463B2 · Pan et al. · 2020 [cited by applicant]
US 10684681B2 · Lemley et al. · 2020 [cited by applicant]
US 10853675B2 · Wang et al. · 2020 [cited by applicant]
US 10946873B2 · Deng et al. · 2021 [cited by applicant]
US 11044404B1 · Persiantsev et al. · 2021 [cited by applicant]
US 11222198B2 · Kwatra et al. · 2022 [cited by applicant]
US 20030026340A1 · Divakaran et al. · 2003 [cited by applicant]
US 20030095602A1 · Divakaran et al. · 2003 [cited by applicant]
US 20030220725A1 · Harter et al. · 2003 [cited by applicant]
US 20050220191A1 · Choi et al. · 2005 [cited by applicant]
US 20060088191A1 · Zhang et al. · 2006 [cited by applicant]
US 20100053480A1 · Jaworski et al. · 2010 [cited by applicant]
US 20140079297A1 · Tadayon et al. · 2014 [cited by applicant]
US 20140093174A1 · Zhang et al. · 2014 [cited by applicant]
US 20150186714A1 · Ren et al. · 2015 [cited by applicant]
US 20160055381A1 · Adsumilli et al. · 2016 [cited by applicant]
US 20160101784A1 · Olson et al. · 2016 [cited by applicant]
US 20170185846A1 · Hwangbo et al. · 2017 [cited by applicant]
US 20170308756A1 · Sigal et al. · 2017 [cited by applicant]
US 20170351922A1 · Campbell · 2017 [cited by applicant]
US 20180144636A1 · Becker · 2018 [cited by examiner]
US 20180239975A1 · Tamrakar et al. · 2018 [cited by applicant]
US 20190065873A1 · Wang et al. · 2019 [cited by applicant]
US 20190138855A1 · Sohn et al. · 2019 [cited by applicant]
US 20190163978A1 · Yang et al. · 2019 [cited by applicant]
US 20200218959A1 · Srinivasa · 2020 [cited by applicant]
US 20200226751A1 · Jin et al. · 2020 [cited by applicant]
US 20200302185A1 · Hussein et al. · 2020 [cited by applicant]
US 20210004575A1 · Pescaru et al. · 2021 [cited by applicant]
US 20210081689A1 · Weyers · 2021 [cited by examiner]
US 20210117659A1 · Foroozan et al. · 2021 [cited by applicant]
US 20210158027A1 · Kwatra et al. · 2021 [cited by applicant]
US 20210350121A1 · Sridhar · 2021 [cited by applicant]
US 20220019776A1 · Malaescu · 2022 [cited by examiner]
US 20220076039A1 · Li · 2022 [cited by examiner]
US 20240020992A1 · Barth · 2024 [cited by examiner]
EP 3440833B1 · 2019 [cited by applicant]
WO 2019145516A1 · 2019 [cited by applicant]
WO 2019180033A1 · 2019 [cited by applicant]
J. C. Chen et al., “Driver Behavior Analysis via Two-Stream Deep Convolutional Neural Network” Applied Sciences 2020, 10(6), p. 1908. [cited by applicant]
M. Granat, “How to Use Convolutional Neural Networks for Time Series Classification” published on https://towardsdatascience.com on Oct. 4, 2019. [cited by applicant]
L. Meng, et al., “Interpretable Spatio-Temporal Attention for Video Action Recognition” Jul. 3, 2019 International Conference on Computer Vision Workshop p. 1513, 10 pages. [cited by applicant]
K. Li et al.,.: “Tell Me Where to Look: Guided Attention Inference Network” Feb. 27, 2018, arXiv:1802.10171v1, 10 pages. [cited by applicant]
C. Posch, et al., “Retinomorphic event-based vision sensors: bioinspired cameras with spiking output”, Proceedings of the IEEE, 102(10), 1470-1484, (2014). [cited by applicant]
C. Toga, et al.: U.S. Appl. No. 63/017,165, filed Apr. 29, 2020 titled:“Image Processing System”, Only Specification, Drawings, Abstract, Claims Considered. [cited by applicant]
C. Ryan, et al., U.S. Appl. No. 16/904,122, filed Jun. 17, 2020 titled: “Object Detection for Event Cameras”, Only Specification, Drawings, Abstract, Claims Considered. [cited by applicant]
Intemational Search Authority, European Patent Office: International Search Report and Written Opinion of Intemational Application No. PCT/EP2021/066305 filed Jun. 16, 2021. ISR-WO mailed Sep. 29, 2021, 14 pages. [cited by applicant]
Fernando Basura et al: Weakly Supervised Gaussian Networks for Action Detection 2020 IEEE Winter Conference on Applications of Computer Vision {WACV), IEEE, Mar. 1, 2020 {Mar. 1, 2020), pp. 526-535,XP033770913, DOI: 10.… [cited by applicant]
Nguyen Phuc et al: Weakly Supervise Action Localization by Sparse Temporal Pooling Network, Eye In-Painting With Exemplar Generative Adverserial Networks [Online]Apr. 3, 2018 {Apr. 3, 2018), pp. 5752-6761, XP055842612, … [cited by applicant]
Pilhyeon et al: “Background Suppression Network for Weakly-supervised Temporal Action Localization” arxiv.org. Cornel University Library, 201 Olin Library Cornel University Ithaca, NY, 14853, Nov. 22, 2019 {Nov. 22, 201… [cited by applicant]
“SoftMax Function”, Wikipedia, Version Dated Jul. 7, 2020 (Year: 2020). [cited by applicant]
Xu et al., “Recurrent Convolutional Neural Network for Video Classification”, IEEE, 2016 (Year: 2016). [cited by applicant]
Cheron et al. “P-CN N: Pose-Based CNN Features for Action Recognition” (Year: 2015). [cited by applicant]