IP Library › Granted Patent US 11,250,602
Granted Patent B2
US 11,250,602 · App. 17/137,651 · Granted Feb 15, 2022

Generating concept images of human poses using machine learning models

Inventors: Samarth Bharadwaj (Bangalore, IN); Saneem Chemmengath (Bangalore, IN); Suranjana Samanta (Bangalore, IN); Karthik Sankaranarayanan (Bangalore, IN)
Assignee: International Business Machines Corporation
G06T11/20G06K9/00335G06K9/6256G06K9/6263G06N20/00G06T7/70G06T2207/20081G06T2207/30196
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,250,602
App. No.
17/137,651
Granted
Feb 15, 2022
Kind
B2
Abstract

Methods, systems, and computer program products for generating concept images of human poses using machine learning models are provided herein. A computer-implemented method includes identifying events from input data by applying a machine learning recognition model to at least a portion of the input data, wherein the identifying comprises (i) detecting multiple entities from the input data and (ii) determining behavioral relationships among at least a portion of the multiple entities; generating, using a machine learning interpretability model and at least a portion of the identified events, images illustrating human poses related to at least a portion of the identified events; outputting at least a portion of the generated images to a user; and updating the machine learning recognition model based at least in part on (i) at least a portion of the generated images and (ii) input from the user.

Claims (30)

1. A computer-implemented method comprising:

identifying one or more events from input data by applying a machine learning recognition model to at least a portion of the input data, wherein said identifying comprises (i) detecting multiple entities from the input data (ii) determining one or more behavioral relationships among at least a portion of the multiple entities, and (iii) generating, via the machine learning recognition model, a vector representation of one or more human poses related to at least a portion of the identified events;

generating, using a machine learning interpretability model and at least a portion of the one or more identified events, one or more images illustrating one or more human poses related to the at least a portion of the one or more identified events, wherein said generating the one or more images comprises converting the vector representation to one or more human pose representations over one or more human form images;

outputting at least a portion of the one or more generated images to at least one user; and

updating the machine learning recognition model based at least in part on (i) at least a portion of the one or more generated images and (ii) input from the at least one user;

wherein the method is carried out by at least one computing device.

2. The computer-implemented method of claim 1 , comprising:

training the machine learning recognition model using data pertaining to human pose structures and image texture information.

3. The computer-implemented method of claim 1 , wherein the input data comprise image data.

4. The computer-implemented method of claim 1 , wherein the input data comprise text data.

5. The computer-implemented method of claim 1 , wherein the input data comprise multimodal data.

6. The computer-implemented method of claim 1 , wherein the one or more images comprise one or more stick figure images.

7. A computer program product comprising a computer readable storage medium having program instructions embodied therewith, the program instructions executable by a computing device to cause the computing device to:

identify one or more events from input data by applying a machine learning recognition model to at least a portion of the input data, wherein said identifying comprises (i) detecting multiple entities from the input data (ii) determining one or more behavioral relationships among at least a portion of the multiple entities, and (iii) generating, via the machine learning recognition model, a vector representation of one or more human poses related to at least a portion of the identified events;

generate, using a machine learning interpretability model and at least a portion of the one or more identified events, one or more images illustrating one or more human poses related to the at least a portion of the one or more identified events, wherein said generating the one or more images comprises converting the vector representation to one or more human pose representations over one or more human form images;

output at least a portion of the one or more generated images to at least one user; and

update the machine learning recognition model based at least in part on (i) at least a portion of the one or more generated images and (ii) input from the at least one user.

8. The computer program product of claim 7 , wherein the input data comprise image data.

9. The computer program product of claim 7 , wherein the input data comprise text data.

10. The computer program product of claim 7 , wherein the input data comprise multimodal data.

11. A system comprising:

a memory; and

at least one processor operably coupled to the memory and configured for:

identifying one or more events from input data by applying a machine learning recognition model to at least a portion of the input data, wherein said identifying comprises (i) detecting multiple entities from the input data (ii) determining one or more behavioral relationships among at least a portion of the multiple entities, and (iii) generating, via the machine learning recognition model, a vector representation of one or more human poses related to at least a portion of the identified events;

generating, using a machine learning interpretability model and at least a portion of the one or more identified events, one or more images illustrating one or more human poses related to the at least a portion of the one or more identified events, wherein said generating the one or more images comprises converting the vector representation to one or more human pose representations over one or more human form images;

outputting at least a portion of the one or more generated images to at least one user; and

updating the machine learning recognition model based at least in part on (i) at least a portion of the one or more generated images and (ii) input from the at least one user.

12. The system of claim 11 , wherein the input data comprise image data.

13. The system of claim 11 , wherein the input data comprise text data.

14. The system of claim 11 , wherein the input data comprise multimodal data.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 30, 2020
From: BHARADWAJ, SAMARTH; CHEMMENGATH, SANEEM; SAMANTA, SURANJANA; SANKARANARAYANAN, KARTHIK
To: INTERNATIONAL BUSINESS MACHINES CORPORATION
Reel/Frame 054775/0297 →
Continuity (2)
Continuation 16548388 · Aug 22, 2019
Related Publication 20210118206A1 · Apr 22, 2021