IP Library Granted Patent US 12,136,158
Granted Patent B2
US 12,136,158 · App. 18/060,449 · Granted Nov 5, 2024

Body pose estimation

Inventors: Avihay Assouline (Tel Aviv, IL); Itamar Berger (Hod Hasharon, IL); Yuncheng Li (Los Angeles, CA)
Assignee: SNAP INC.
G06T13/40G06V40/107
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,136,158
App. No.
18/060,449
Granted
Nov 5, 2024
Kind
B2
Abstract

Aspects of the present disclosure involve a system comprising a computer-readable storage medium storing a program and a method for detecting a pose of a user. The program and method include receiving a monocular image that includes a depiction of a body of a user; detecting a plurality of skeletal joints of the body depicted in the monocular image; and determining a pose represented by the body depicted in the monocular image based on the detected plurality of skeletal joints of the body. A pose of an avatar is modified to match the pose represented by the body depicted in the monocular image by adjusting a set of skeletal joints of a rig of an avatar based on the detected plurality of skeletal joints of the body; and the avatar having the modified pose that matches the pose represented by the body depicted in the monocular image is generated for display.

Claims (58)

1. A method comprising:

storing, as a screenshot, a frame of a video that simultaneously depicts a first user and a first avatar;

concurrently with displaying the screenshot, displaying a separate region comprising a blank space indicating that receipt of a corresponding screenshot from a second user is pending; and

receiving input comprising a selection of the blank space.

2. The method of claim 1 , further comprising:

receiving the video depicting the first user;

determining that a predetermined set of skeletal joints of the first user is visible in the video;

adjusting one or more portions of the first avatar to represent a pose of the predetermined set of skeletal joints;

causing display of the video that simultaneously depicts the first user and the first avatar representing the pose of the predetermined set of skeletal joints; and

during display of the video that simultaneously depicts the first user and the first avatar, detecting a condition for capturing a screenshot.

3. The method of claim 2 , further comprising receiving input that selects the first avatar from a plurality of avatars.

4. The method of claim 2 , wherein the condition comprises passage of a threshold period of time.

5. The method of claim 2 , wherein the condition comprises receiving a spoken command from the first user to capture the screenshot.

6. The method of claim 2 , wherein the condition comprises maintaining a particular pose for a threshold period of time.

7. The method of claim 2 , further comprising:

generating, for display, a plurality of identical avatars having a first set of different poses; and

detecting that the pose corresponds to a specified pose.

8. The method of claim 7 , further comprising:

causing a list of recipients to be presented in response to receiving the input;

receiving input that selects a designated recipient from the list of recipients as a second user; and

transmitting the screenshot to the second user in response to receiving the input that selects the designated recipient.

9. The method of claim 8 , further comprising:

causing receiving of a video of the second user;

causing detecting a condition for capturing a second screenshot; and

causing capturing of a second screenshot of the second user and a second avatar in response to detecting the condition.

10. The method of claim 9 , further comprising:

causing automatically sending of the second screenshot to the first user; and

causing the second screenshot to be presented in place of the blank space to the first user.

11. The method of claim 1 , further comprising:

generating, for display, a plurality of avatars having a first set of different poses.

12. The method of claim 11 further comprising, in response to detecting that a pose of the first user corresponds to a specified pose:

modifying the first set of different poses of the plurality of avatars to have identical poses that match the pose of the first user; and

generating for display the plurality of avatars having the identical poses.

13. The method of claim 12 further comprising animating modification of the first set of different poses of the plurality of avatars in the generated display.

14. A system comprising:

at least one processor configured to perform operations comprising:

storing, as a screenshot, a frame of a video that simultaneously depicts a first user and a first avatar;

concurrently with displaying the screenshot, displaying a separate region comprising a blank space indicating that receipt of a corresponding screenshot from a second user is pending; and

receiving input comprising a selection of the blank space.

15. The system of claim 14 , wherein the operations further comprise:

receiving the video depicting the first user;

determining that a predetermined set of skeletal joints of the first user is visible in the video;

adjusting one or more portions of the first avatar to represent a pose of the predetermined set of skeletal joints;

causing display of the video that simultaneously depicts the first user and the first avatar representing the pose of the predetermined set of skeletal joints; and

during display of the video that simultaneously depicts the first user and the first avatar, detecting a condition for capturing a screenshot.

16. The system of claim 15 , wherein the condition comprises passage of a threshold period of time.

17. The system of claim 15 , wherein the condition comprises receiving a spoken command from the first user to capture the screenshot.

18. The system of claim 15 , wherein the condition comprises maintaining a particular pose for a threshold period of time.

19. A non-transitory machine-readable storage medium that includes instructions that, when executed by one or more processors of a machine, cause the machine to perform operations comprising:

storing, as a screenshot, a frame of a video that simultaneously depicts a first user and a first avatar;

concurrently with displaying the screenshot, displaying a separate region comprising a blank space indicating that receipt of a corresponding screenshot from a second user is pending; and

receiving input comprising a selection of the blank space.

20. The non-transitory machine-readable storage medium of claim 19 , wherein the operations further comprise:

receiving the video depicting the first user;

determining that a predetermined set of skeletal joints of the first user is visible in the video;

adjusting one or more portions of the first avatar to represent a pose of the predetermined set of skeletal joints;

causing display of the video that simultaneously depicts the first user and the first avatar representing the pose of the predetermined set of skeletal joints; and

during display of the video that simultaneously depicts the first user and the first avatar, detecting a condition for capturing a screenshot.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 7, 2022
From: ASSOULINE, AVIHAY; BERGER, ITAMAR; LI, YUNCHENG
To: SNAP INC.
Reel/Frame 062007/0936 →
Continuity (3)
Continuation 17212555 · Mar 25, 2021
Continuation 16269312 · Feb 6, 2019
Related Publication 20230090086A1 · Mar 23, 2023