IP Library › Granted Patent US 7,961,910
Granted Patent B2
US 7,961,910 · App. 12/621,013 · Granted Jun 14, 2011

Systems and methods for tracking a model

Assignee: Microsoft Corporation
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 7,961,910
App. No.
12/621,013
Granted
Jun 14, 2011
Kind
B2
Abstract

An image such as a depth image of a scene may be received, observed, or captured by a device. A grid of voxels may then be generated based on the depth image such that the depth image may be downsampled. A model may be adjusted based on a location or position of one or more extremities estimated or determined for a human target in the grid of voxels. The model may also be adjusted based on a default location or position of the model in a default pose such as a T-pose, a DaVinci pose, and/or a natural pose.

Claims (31)

1. A method for tracking a user in a scene, the method comprising:

receiving a depth image;

generating a grid of voxels based on the depth image;

determining whether a location or position has been estimated for an extremity of a human target included in the grid of voxels; and

adjusting a body part of a model associated with the extremity to the location or position when, based on the determination, the location or position has been estimated for the extremity.

2. The method of claim 1 , wherein the model comprises a skeletal model having joints and bones.

3. The method of claim 2 , wherein adjusting the body part of the model associated with the extremity to the location or position comprises adjusting a joint of the skeletal model associated with the extremity to the location or position.

4. The method of claim 1 , further comprising determining whether the location or position has been estimated for the extremity of the human target included the grid of voxels further comprises determining whether the location or position is valid, and wherein the body part of the model associated with the extremity is adjusted to the location or position when, based on the determinations, the location or position has been estimated for the extremity and the location or position estimated for the extremity is valid.

5. The method of claim 1 , further comprising adjusting the model based on a geometric constraint.

6. The method of claim 1 , further comprising relaxing the body part of the model associated with the extremity to a default location or position associated with a default pose when, based on the determination, the location or position has not been determined for the estimated extremity.

7. The method of claim 6 , wherein the default pose comprises at least one of the following: a T-Pose, a Di Vinci pose, and a natural pose.

8. The method of claim 6 , further comprising adjusting the body part of the model to a location or position of a voxel of the human target closest to the default location or position associated with the default pose.

9. The method of claim 1 , further comprising determining whether a pose associated with the model and the one or more adjusted body parts is valid.

10. The method of claim 1 , further comprising refining a position of the body part of the model based on one or more pixels associated with the extremity of the human target in the received depth image.

11. A computer-readable storage medium having stored thereon computer executable instructions for tracking a user in a scene, the computer executable instructions comprising instructions for:

receiving a depth image;

generating a grid of voxels based on the depth image;

determining whether a location or position has been estimated for an extremity of a human target included in the grid of voxels;

relaxing a body part of a model associated with the extremity to a default location or position associated with a default pose when, based on the determination, the location or position has not been determined for the estimated extremity; and

magnetizing the body part of the model to a location or position of a voxel of the human target closest to the default location or position associated with the default pose.

12. The computer-readable storage medium of claim 11 , wherein the default pose comprises at least one of the following: a T-Pose, a Di Vinci pose, and a natural pose.

13. The computer-readable storage medium of claim 12 , further comprising instructions for refining a position of the body part of the model based on one or more pixels associated with the extremity of the human target in the received depth image.

14. A computer-readable storage medium of claim 11 , further comprising instructions for adjusting the model based on a geometric constraint.

15. The computer-readable storage medium of claim 14 , wherein the geometric constraint comprises at least one of the following: an angle of the body part for a typical human, a length of the body part for the typical human, a width of the body part for the typical human, an angle determined for the extremity, a length determined for the extremity, and a width determined for the extremity.

16. The computer-readable storage medium of claim 11 , further comprising instructions for adjusting a body part of a model associated with the extremity to the location or position estimated for the extremity when, based on the determination, the location or position has been estimated for the extremity.

17. A system for tracking a user in a scene, the system comprising:

a capture device, wherein the capture device comprises a camera component that receives a depth image of the scene; and

a computing device in operative communication with the capture device, wherein the computing device comprises a processor that generates a grid of voxels based on the depth image; determines whether a location or position has been estimated for an extremity of a human target included in the grid of voxels; adjusts a body part of a model associated with the extremity to the location or position when, based on the determination, the location or position has been estimated for the extremity; and adjusts the body part of the model to a closest voxel associated with the human target when, based on the determination, the location or position has not been estimated for the extremity.

18. The system of claim 17 , wherein the processor further adjusts the body part of the model to a closest voxel associated with the human target when, based on the determination, the location or position has not been estimated for the extremity by relaxing the body part of the model associated with the extremity to a default location or position associated with a default pose; and adjusting the body part of the model to a location or position of a voxel of the human target closest to the default location or position associated with the default pose.

19. The system of claim 17 , wherein the processor further refines a position of the body part of the model based on one or more pixels associated with the extremity of the human target in the received depth image.

20. The system of claim 17 , wherein the processor further adjusts the model based on a geometric constraint.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 9, 2014
From: MICROSOFT CORPORATION
To: MICROSOFT TECHNOLOGY LICENSING, LLC
Reel/Frame 034564/0001 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jul 13, 2010
From: LEE, JOHNNY CHUNG; LEYVAND, TOMMER; STACHNIAK, SIMON PIOTR; PEEPER, CRAIG
To: MICROSOFT CORPORATION
Reel/Frame 024671/0144 →
Continuity (2)
Continuation In Part 12575388 · Oct 7, 2009
Related Publication 20110081045A1 · Apr 7, 2011