IP Library Granted Patent US 12,513,215
Granted Patent B2
US 12,513,215 · App. 18/167,482 · Granted Dec 30, 2025

Intelligent reasoning framework for user intent extraction

Inventor: Newton Howard (Washington, DC)
Assignee: Genesis Intelligence, LLC
H04L67/306G06F16/9535G06N5/025G06N20/00G08B19/00H04L12/2829H04L67/535
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,513,215
App. No.
18/167,482
Filed
Feb 10, 2023
Granted
Dec 30, 2025
Kind
B2
Examiner
KIM, CHONG G
Art Unit
2443
USPC
707/734
Abstract

Embodiments of the present systems and methods may provide an intelligent systems framework for analysis of user-generated content from various capture points to determine user intent. For example, a method may be implemented in a computer system comprising a processor, memory accessible by the processor, and computer program instructions stored in the memory and executable by the processor, the method may comprise receiving, at the computer system, data relating to a plurality of aspects of at least one person, including data from at least one of physical or physiological sensors and communicatively connected devices, extracting, at the computer system, from the received data, features relevant to events relating to at least one person, extracting, at the computer system, at least one intent of at least one event relating to at least one person, and performing, at the computer system, an action based on the extracted at least one intent.

Claims (39)

1 . A method implemented in a computer system comprising a processor, memory accessible by the processor, and computer program instructions stored in the memory and executable by the processor, the method comprising:

receiving, at the computer system, data relating to a plurality of aspects of at least one person, including data from at least one of physical or physiological sensors and communicatively connected devices;

extracting, at the computer system, from the received data, features relevant to events relating to at least one person using a models processor adapted to select and process models for feature extraction;

extracting, at the computer system, at least one intent of at least one event relating to at least one person using an intent extractor adapted to use ensemble models to process and integrate persona data, ontology data from an ontology database comprising user domain data, intent domain data, context domain data, and world data, and extracted features, wherein the ensemble models are selected to find an optimal stack of predictors, and wherein extracting the at least one intent comprises performing intention awareness processing to form multidimensional intention awareness vectors and determining a component of at least one multidimensional intention awareness vector as a barycenter of the weighted points of the vector, including consistent tracking and extrapolation of objects in an environment and circumstantial semantics, forming multidimensional intention vectors and determining a component of at least one multidimensional intention vector as a barycenter of the weighted points of the vector, and generating an ensemble datastream representing the extracted at least one intent by fusing the intention awareness vectors and the intention vectors; and

performing, at the computer system, an action based on the extracted at least one intent.

2 . The method of claim 1 , wherein the data relating to a plurality of aspects of at least one person comprises live data retrieved from a plurality of capture points that at least one person is exposed to and interacts with;

the communicatively connected device comprises at least one of a microphone, room camera, fridge camera, smart mobile, smart refrigerator, smart watch, smart fire alarm, smart door lock, smart bicycle, medical sensor, fitness tracker, smart security system, voice controller, dash button, doorbell cam, mobile robot, smart light switch, and air quality monitor;

the physical and physiological sensor comprises at least one of an audio sensor, video sensor, electro-encephalogram sensor, electro-cardiogram sensor, heart rate sensor, breathing rate sensor, blood pressure sensor, body temperature sensor, head movement sensor, body posture sensor, and blood oxygenation level sensor, humidity sensor, biometric sensor; and

the data further comprises at least one of browsing history, bookmarks, browsing behavior, time spent on particular web pages, location history, calendar with past and upcoming events, social media activity, posts, social network graph, text, audio, video, social media content, chat, Short Message Service (SMS), email, medical visits, current diseases, medical treatments, real-time movements, breathing, cardiac frequency, and sleep patterns.

3 . The method of claim 2 , wherein the features relevant to events relating to at least one person are extracted using artificial intelligence and machine learning models trained using data relating to a plurality of aspects of at least one person, wherein data relating to actions that are highly-specific predictors for particular intents are tagged.

4 . The method of claim 3 , wherein at least one intent of at least one event relating to at least one person is extracted using ontologies data relating to subject areas that shows the properties and the relations between the subject area; and

ontologies methodology is used to create a supporting framework for intention query by defining concepts, sub-concepts, relationships, and aggregations of multiple concepts.

5 . The method of claim 4 , wherein the physical action comprises at least one of generating an alarm, providing information or suggestions, and providing narrativization of events.

6 . A system comprising a processor, memory accessible by the processor, and computer program instructions stored in the memory and executable by the processor to perform:

receiving data relating to a plurality of aspects of at least one person, including data from at least one of physical or physiological sensors and communicatively connected devices;

extracting from the received data, features relevant to events relating to at least one person using a models processor adapted to select and process models for feature extraction, the processed models including at least ensemble models;

extracting, at the computer system, at least one intent of at least one event relating to at least one person using an intent extractor adapted to use ensemble models to process and integrate persona data, ontology data from an ontology database comprising user domain data, intent domain data, context domain data, and world data, and extracted features, wherein the ensemble models are selected to find an optimal stack of predictors, and wherein extracting the at least one intent comprises performing intention awareness processing to form multidimensional intention awareness vectors and determining a component of at least one multidimensional intention awareness vector as a barycenter of the weighted points of the vector, including consistent tracking and extrapolation of objects in an environment and circumstantial semantics, forming multidimensional intention vectors and determining a component of at least one multidimensional intention vector as a barycenter of the weighted points of the vector, and generating an ensemble datastream representing the extracted at least one intent by fusing the intention awareness vectors and the intention vectors; and

performing an action based on the extracted at least one intent.

7 . The system of claim 6 , wherein the data relating to a plurality of aspects of at least one person comprises live data retrieved from a plurality of capture points that at least one person is exposed to and interacts with;

the communicatively connected device comprises at least one of a microphone, room camera, fridge camera, smart mobile, smart refrigerator, smart watch, smart fire alarm, smart door lock, smart bicycle, medical sensor, fitness tracker, smart security system, voice controller, dash button, doorbell cam, mobile robot, smart light switch, and air quality monitor;

the physical and physiological sensor comprises at least one of an audio sensor, video sensor, electro-encephalogram sensor, electro-cardiogram sensor, heart rate sensor, breathing rate sensor, blood pressure sensor, body temperature sensor, head movement sensor, body posture sensor, and blood oxygenation level sensor, humidity sensor, biometric sensor; and

the data further comprises at least one of browsing history, bookmarks, browsing behavior, time spent on particular web pages, location history, calendar with past and upcoming events, social media activity, posts, social network graph, text, audio, video, social media content, chat, Short Message Service (SMS), email, medical visits, current diseases, medical treatments, real-time movements, breathing, cardiac frequency, and sleep patterns.

8 . The system of claim 7 , wherein the features relevant to events relating to at least one person are extracted using artificial intelligence and machine learning models trained using data relating to a plurality of aspects of at least one person, wherein data relating to actions that are highly-specific predictors for particular intents are tagged.

9 . The system of claim 8 , wherein at least one intent of at least one event relating to at least one person is extracted using ontologies data relating to subject areas that shows the properties and the relations between the subject area; and

ontologies methodology is used to create a supporting framework for intention query by defining concepts, sub-concepts, relationships, and aggregations of multiple concepts.

10 . The method of claim 9 , wherein the physical action comprises at least one of generating an alarm, providing information or suggestions, and providing narrativization of events.

11 . A computer program product comprising a non-transitory computer readable storage having program instructions embodied therewith, the program instructions executable by a computer, to cause the computer to perform a method comprising:

receiving, at the computer system, data relating to a plurality of aspects of at least one person, including data from at least one of physical or physiological sensors and communicatively connected devices;

extracting, at the computer system, from the received data, features relevant to events relating to at least one person using a models processor adapted to select and process models for feature extraction, the processed models including at least ensemble models;

extracting, at the computer system, at least one intent of at least one event relating to at least one person using an intent extractor adapted to use ensemble models to process and integrate persona data, ontology data from an ontology database comprising user domain data, intent domain data, context domain data, and world data, and extracted features, wherein the ensemble models are selected to find an optimal stack of predictors, wherein extracting the at least one intent comprises performing intention awareness processing to form multidimensional intention awareness vectors and determining a component of at least one multidimensional intention awareness vector as a barycenter of the weighted points of the vector, including consistent tracking and extrapolation of objects in an environment and circumstantial semantics, forming multidimensional intention vectors and determining a component of at least one multidimensional intention vector as a barycenter of the weighted points of the vector, and generating an ensemble datastream representing the extracted at least one intent by fusing the intention awareness vectors and the intention vectors; and

performing, at the computer system, an action based on the extracted at least one intent.

12 . The computer program product of claim 11 , wherein the data relating to a plurality of aspects of at least one person comprises live data retrieved from a plurality of capture points that at least one person is exposed to and interacts with;

the communicatively connected device comprises at least one of a microphone, room camera, fridge camera, smart mobile, smart refrigerator, smart watch, smart fire alarm, smart door lock, smart bicycle, medical sensor, fitness tracker, smart security system, voice controller, dash button, doorbell cam, mobile robot, smart light switch, and air quality monitor;

the physical and physiological sensor comprises at least one of an audio sensor, video sensor, electro-encephalogram sensor, electro-cardiogram sensor, heart rate sensor, breathing rate sensor, blood pressure sensor, body temperature sensor, head movement sensor, body posture sensor, and blood oxygenation level sensor, humidity sensor, biometric sensor; and

the data further comprises at least one of browsing history, bookmarks, browsing behavior, time spent on particular web pages, location history, calendar with past and upcoming events, social media activity, posts, social network graph, text, audio, video, social media content, chat, Short Message Service (SMS), email, medical visits, current diseases, medical treatments, real-time movements, breathing, cardiac frequency, and sleep patterns.

13 . The computer program product of claim 12 , wherein the features relevant to events relating to at least one person are extracted using artificial intelligence and machine learning models trained using data relating to a plurality of aspects of at least one person, wherein data relating to actions that are highly-specific predictors for particular intents are tagged.

14 . The computer program product of claim 13 , wherein at least one intent of at least one event relating to at least one person is extracted using ontologies data relating to subject areas that shows the properties and the relations between the subject area; and

ontologies methodology is used to create a supporting framework for intention query by defining concepts, sub-concepts, relationships, and aggregations of multiple concepts.

15 . The computer program product of claim 14 , wherein the physical action comprises at least one of generating an alarm, providing information or suggestions, and providing narrativization of events.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 28, 2025
From: HOWARD, NEWTON
To: GENESIS INTELLIGENCE, LLC
Reel/Frame 073374/0001 →
Continuity (3)
Continuation 16520673 · Jul 24, 2019
Provisional Application 62702815 · Jul 24, 2018
Related Publication 20230188612A1 · Jun 15, 2023
References Cited (45)
US 6745168B1 · Enomoto · 2004 [cited by applicant]
US 8380902B2 · Howard · 2013 [cited by examiner]
US 8606735B2 · Cho · 2013 [cited by applicant]
US 9124694B2 · Monegan · 2015 [cited by applicant]
US 9652797B2 · Vijayaraghavan · 2017 [cited by applicant]
US 10482521B2 · Vijayaraghavan · 2019 [cited by examiner]
US 10846601B1 · Howard · 2020 [cited by examiner]
US 10990645B1 · Shi · 2021 [cited by examiner]
US 20080155147A1 · Howard · 2008 [cited by examiner]
US 20090037832A1 · Falchuk · 2009 [cited by applicant]
US 20100090835A1 · Liu · 2010 [cited by applicant]
US 20120016678A1 · Gruber · 2012 [cited by examiner]
US 20140079195A1 · Srivastava · 2014 [cited by applicant]
US 20150100943A1 · Gabel et al. · 2015 [cited by applicant]
US 20150256675A1 · Sri · 2015 [cited by examiner]
US 20150269150A1 · Carper · 2015 [cited by applicant]
US 20150356405A1 · Sanchez · 2015 [cited by applicant]
US 20160042359A1 · Singh · 2016 [cited by examiner]
US 20170199866A1 · Gunaratna · 2017 [cited by applicant]
US 20180012163A1 · Smith · 2018 [cited by examiner]
US 20180039745A1 · Chevalier et al. · 2018 [cited by applicant]
US 20180114527A1 · Zilotti · 2018 [cited by applicant]
US 20180211175A1 · Mendels · 2018 [cited by applicant]
US 20190155577A1 · Prabha · 2019 [cited by applicant]
US 20190215290A1 · Kozloski · 2019 [cited by applicant]
US 20190251626A1 · Jezewski · 2019 [cited by applicant]
US 20190332647A1 · Rincon Opden Bosch · 2019 [cited by applicant]
US 20210256345A1 · Mars · 2021 [cited by examiner]
US 20250165995A1 · Hamedi · 2025 [cited by examiner]
Cambria et al., Sentic blending: Scalable multimodal fusion for the continuous interpretation of semantics and sentics, 2013 IEEE Symposium on Computational Intelligence for Human-like Intelligence (CIHLI), Singapore, 2… [cited by examiner]
Howard et al., Intention awareness: improving upon situation awareness in human-centric environments. Hum. Cent. Comput. Inf. Sci. 3, 9 (2013) (Year: 2013). [cited by examiner]
Howard et al., Application of intention awareness and sentic computing for sensemaking in joint-cognitive systems, 2013 IEEE Symposium on Intelligent Agents (IA), Singapore, 2013, pp. 1-4 (Year: 2013). [cited by examiner]
Cambria et al., Semantic Multidimensional Scaling for Open-Domain Sentiment Analysis, IEEE Intelligent Systems, vol. 29, No. 02, pp. 44-51, 2014 (Year: 2014). [cited by examiner]
Ethan Fast, Binbin Chen, Michael S. Bernstein, “Empath: Understanding Topic Signals in Large-Scale Text”, Chi: ACM Conference on Human Factors in Computing Systems May 2016, arXiv:1602.06979, https://doi.org/10.48550/ar… [cited by applicant]
Karol Kurach, Sylvain Gelly, Michal Jastrzebski, Philip Haeusser, Olivier Teytaud, Damien Vincent, Olivier Bousquet, “Better Text Understanding Through Image-To-Text Transfer”, May 2017, arXiv:1705.08386, https://doi.or… [cited by applicant]
Hema Swetha Koppula, Rudhir Gupta and Ashutosh Saxena, “Learning human activities and object affordances from RGB-D videos”, The International Journal of Robotics Research, May 2013, arXiv:1210.1207, https://doi.org/10.… [cited by applicant]
Koppula, H.S. & Saxena, A .. (2013). “Learning spatio-temporal structure from RGB-D videos for human activity detection and anticipation.” 30th International Conference on Machine Learning, ICML Jan. 2013. 1829-1837. [cited by applicant]
Ng, Joe & Hausknecht, Matthew & Vijayanarasimhan, Sudheendra & Vinyals, Oriol & Monga, Rajat & Toderici, George. (Jun. 2015). “Beyond short snippets: Deep networks for video classification”. 4694-4702. 10.1109/ CVPR.201… [cited by applicant]
Kuehne, H et al. “Hmdb: A Large Video Database for Human Motion Recognition.” IEEE, 2011. 2556-2563. Web. Apr. 11, 2012. @ 2012 Institute of Electrical and Electronics Engineers. [cited by applicant]
Tran, Du & Bourdev, Lubomir & Fergus, Rob & Torresani, Lorenzo & Paluri, Manohar. (Dec. 2015). Learning Spatiotemporal Features with 3D Convolutional Networks. 4489-4497. 10.1109/ICCV.2015.510. [cited by applicant]
Xiaolong Wang, Ali Farhadi and Abhinav Gupta, “Actions ˜ Transformations”, Proc. of IEEE Conference on Computer Vision and Pattern Recognition (CVPR), 2016. [cited by applicant]
N. Howard, “Application of intention awareness and sen tic computing for sensemaking in joint-cognitive systems,” 2013 IEEE Symposium on Intelligent Agents (IA), 2013, pp. 1-4 (Year: 2013). [cited by applicant]
Howard, N., Cambria, E. Intention awareness: improving upon situation awareness in human-centric environments. Hum. Cent. Comput. Inf. Sci. 3, 9 (2013) (Year: 2013). [cited by applicant]
Written Opinion of the International Searching Authority dated Oct. 9, 2019, received in International application No. PCT/US19/43168 (3 pages). [cited by applicant]
Notification of Transmittal of the International Search Report dated Oct. 9, 2019, received in International application No. PCT/US19/43168 (4 pages). [cited by applicant]