IP Library Granted Patent US 11,232,292
Granted Patent B2
US 11,232,292 · App. 16/799,749 · Granted Jan 25, 2022

Activity recognition systems and methods

Inventors: Kamil Wnuk (Playa Del Rey, CA); Nicholas J. Witchey (Laguna Hills, CA)
Assignee: Nant Holdings IP, LLC
G06K9/00342B25J9/1697G06F16/2228G06F16/23G06F16/24578G06F16/285G06F16/7837G06F16/9024G06K9/00664G06K9/469G06K9/4671G06K9/6215G06T11/206
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,232,292
App. No.
16/799,749
Granted
Jan 25, 2022
Kind
B2
Abstract

An activity recognition system is disclosed. A plurality of temporal features is generated from a digital representation of an observed activity using a feature detection algorithm. An observed activity graph comprising one or more clusters of temporal features generated from the digital representation is established, wherein each one of the one or more clusters of temporal features defines a node of the observed activity graph. At least one contextually relevant scoring technique is selected from similarity scoring techniques for known activity graphs, the at least one contextually relevant scoring technique being associated with activity ingestion metadata that satisfies device context criteria defined based on device contextual attributes of the digital representation, and a similarity activity score is calculated for the observed activity graph as a function of the at least one contextually relevant scoring technique, the similarity activity score being relative to at least one known activity graph.

Claims (51)

1. An activity recognition method using at least one processor and at least one memory, the method comprising:

receiving a digital representation comprising at least video data of an observable activity of at least one object;

generating from the digital representation, using at least one feature detection algorithm, a plurality of features related to the observable activity;

establishing an observed activity data object based on the plurality of features;

determining a similarity for the observed activity data object relative to at least one known activity data object based on a context relevant to the plurality of features;

accessing an activity recognition results set based on the similarity; and

triggering an action based on the activity recognition results set,

wherein a known activity associated with the at least one known activity data object includes at least one of a body movement, an interaction, a hand movement, and a facial movement or expression.

2. The method of claim 1 , wherein the context comprises location data.

3. The method of claim 1 , wherein the at least one known activity data object comprises at least a part of a template for interactions.

4. The method of claim 1 , further comprising:

converting aspects of the digital representation to an observed activity graph; and

comparing the observed activity graph to known activity graphs.

5. The method of claim 1 , wherein the digital representation further comprises at least one of image data, still image data, audio data, accelerometer data, tactile data, kinesthetic data, temperature data, kinematic data, 3D registration data, and radio signal or wireless data.

6. The method of claim 5 , wherein the image data comprises at least one of ultrasound, infrared, and visible spectrum data.

7. The method of claim 1 , wherein the digital representation comprises a video of a procedure.

8. The method of claim 7 , wherein contextual relevance relates to one or more of when the procedure is performed, information about the procedure, a provider associated with the procedure, and a location of the procedure.

9. The method of claim 1 , wherein the at least one feature detection algorithm includes one of the following: a scale-invariant feature transform (SIFT), Fast Retina Keypoint (FREAK), Histograms of Oriented Gradient (HOG), Speeded Up Robust Features (SURF), DAISY, Binary Robust Invariant Scalable Keypoints (BRISK), FAST, Binary Robust Independent Elementary Features (BRIEF), Harris Corners, Edges, Gradient Location and Orientation Histogram (GLOH), Energy of image Gradient (EOG), and Transform Invariant Low-rank Textures (TILT) feature detection algorithm.

10. The method of claim 1 , wherein the digital representation comprises video data captured over a time period or within a time frame.

11. The method of claim 10 , wherein at least some of the plurality of features describe a temporal or spatial relationship among comparable events in time.

12. The method of claim 1 , further comprising determining contextual relevance based on ingestion metadata.

13. The method of claim 12 , further comprising selecting the ingestion metadata used to determine contextual relevance based on one or more domain-specific attributes.

14. The method of claim 12 , wherein the ingestion metadata conforms to a defined attribute namespace.

15. The method of claim 1 , further comprising:

recognizing one or more objects in the digital representation using at least some of the plurality of features; and

retrieving object information related to the one or more recognized objects.

16. The method of claim 15 , further comprising using the object information to determine contextual relevance.

17. The method of claim 1 , wherein the similarity includes at least one of a Euclidean distance, linear kernel, polynomial kernel, Chi-squared kernel, Cauchy kernel, histogram intersection kernel, Hellinger's kernel, Jensen-Shannon kernel, hyperbolic tangent (sigmoid) kernel, rational quadratic kernel, multiquadratic kernel, inverse multiquadratic kernel, circular kernel, spherical kernel, wave kernel, power kernel, log kernel, spline kernel, Bessel kernel, generalized T-Student kernel, Bayesian kernel, wavelet kernel, radial basis function (RBF), exponential kernel, Laplacian kernel, ANOVA kernel and B-spline kernel function.

18. The method of claim 1 , further comprising selecting the similarity according to a data modality.

19. The method of claim 1 , wherein the similarity reflects a relative confidence of data from each of a plurality of sensing modalities.

20. The method of claim 1 , wherein the activity recognition results set comprises at least one of an activity identifier, a search result, a classification, a recommendation, an anomaly, a warning, a segmentation, a command, a ranking, context relevant information, content information, and an action prediction.

21. The method of claim 20 , wherein the action prediction is based on variations of known activities.

22. The method of claim 1 , wherein initiating the action comprises executing a command.

23. The method of claim 1 , wherein initiating the action comprises generating an alert.

24. The method of claim 1 , wherein the video data relates to a volumetric space.

25. An activity recognition device having a processor, wherein, upon execution of software instructions stored on a non-transitory computer readable medium, the processor:

captures a digital representation comprising at least video data of an observable activity of at least one object;

generates from the digital representation, using at least one feature detection algorithm, a plurality of features related to the observable activity;

establishes an observed activity data object based on the plurality of features;

determines a similarity for the observed activity data object relative to at least one known activity data object based on a context relevant to the plurality of features;

accesses an activity recognition results set based on the similarity; and

triggers an action based on the activity recognition results set,

wherein a known activity associated with the at least one known activity data object includes at least one of a body movement, an interaction, a hand movement, and a facial movement or expression.

26. A non-transitory computer readable medium comprising instructions executable by a computer processor to perform activity recognition by performing operations comprising:

capturing a digital representation comprising at least video data of an observable activity of at least one object;

generating from the digital representation, using at least one feature detection algorithm, a plurality of features related to the observable activity;

establishing an observed activity data object based on the plurality of features;

determining a similarity for the observed activity data object relative to at least one known activity data object based on a context relevant to the plurality of features;

accessing an activity recognition results set based on the similarity; and

triggering an action based on the activity recognition results set,

wherein a known activity associated with the at least one known activity data object includes at least one of a body movement, an interaction, a hand movement, and a facial movement or expression.

Assignments (4)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Feb 25, 2020
From: NANTWORKS, LLC
To: NANT HOLDINGS IP, LLC
Reel/Frame 051924/0747 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Feb 25, 2020
From: NANT VISION, INC.
To: NANT HOLDINGS IP, LLC
Reel/Frame 051924/0784 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Feb 25, 2020
From: WNUK, KAMIL
To: NANT VISION, INC.
Reel/Frame 051924/0815 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Feb 25, 2020
From: WITCHEY, NICHOLAS J.
To: NANTWORKS, LLC
Reel/Frame 051924/0860 →
Continuity (6)
Continuation 16284972 · Feb 25, 2019
Division 15875681 · Jan 19, 2018
Continuation 15374300 · Dec 9, 2016
Continuation 14741830 · Jun 17, 2015
Provisional Application 62013508 · Jun 17, 2014
Related Publication 20200193151A1 · Jun 18, 2020
Cited By (1)
US 12,243,354