IP Library › Granted Patent US 12,205,350
Granted Patent B2
US 12,205,350 · App. 17/800,582 · Granted Jan 21, 2025

Apparatus for separating feature points for each object, method for separating feature points for each object and computer program

Inventors: Masaaki Matsumura (Musashino, JP); Hajime Noto (Musashino, JP); Yoshinori Kusachi (Musashino, JP)
Assignee: NIPPON TELEGRAPH AND TELEPHONE CORPORATION
G06V10/7715
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,205,350
App. No.
17/800,582
Granted
Jan 21, 2025
Kind
B2
Abstract

An object-specific keypoint separation apparatus includes: an inference execution unit configured to receive a captured image capturing an object as an input and use a pre-trained model that has been trained in order to output a plurality of first maps and a plurality of second maps generated from the input captured image to output the plurality of first maps and the plurality of second maps, the plurality of first maps storing a vector describing a connection relationship of a keypoint of the object only around the keypoint, and the plurality of second maps representing a heat map configured to have a peak at coordinates at which the keypoint of the object appears; a map correction unit configured to correct the plurality of second maps using the plurality of first maps and the plurality of second maps; an upsampling unit configured to upsample the plurality of first maps; and an object-specific keypoint separation unit configured to separate keypoints for each object based on the plurality of upsampled first maps and the plurality of corrected second maps.

Claims (17)

1. An object-specific keypoint separation apparatus comprising:

a processor; and

a storage medium having computer program instructions stored thereon, when executed by the processor, perform to:

receive a captured image capturing an object as an input and use a pre-trained model that has been trained in order to output a plurality of first maps and a plurality of second maps generated from the input captured image to output the plurality of first maps and the plurality of second maps, the plurality of first maps storing a vector describing a connection relationship of a keypoint of the object only around the keypoint, and the plurality of second maps representing a heat map configured to have a peak at coordinates at which the keypoint of the object appears;

correct the plurality of second maps using the plurality of first maps and the plurality of second maps output from the inference execution unit;

upsample the plurality of first maps; and

separate keypoints for each object based on the plurality of upsampled first maps and the plurality of corrected second maps.

2. The object-specific keypoint separation apparatus according to claim 1 ,

wherein the plurality of first maps and the plurality of second maps output have a low resolution, and wherein the computer program instructions further perform to upsamples the plurality of second maps to cause the plurality of second maps to have an equal resolution.

3. The object-specific keypoint separation apparatus according to claim 2 , wherein the computer program instructions further uses coordinates of a keypoint indicated by a vector on the first map to generate a vector density map representing a degree of vector density, and normalizes the vector density map by dividing a value on the vector density map by a maximum value on the generated vector density map.

4. The object-specific keypoint separation apparatus according to claim 3 , wherein the computer program instructions further perform to multiplies the normalized vector density map and each of the plurality of upsampled second maps by a predetermined percentage value and adds up multiplication results to generate the plurality of corrected second maps.

5. An object-specific keypoint separation method comprising:

receiving a captured image capturing an object as an input and using a pre-trained model that has been trained in order to output a plurality of first maps and a plurality of second maps generated from the input captured image to output the plurality of first maps and the plurality of second maps, the plurality of first maps storing a vector describing a connection relationship of a keypoint of the object only around the keypoint, and the plurality of second maps representing a heat map configured to have a peak at coordinates at which the keypoint of the object appears;

correcting the plurality of second maps using the plurality of output first maps and the plurality of output second maps;

upsampling the plurality of output first maps; and

separating the keypoints for each object based on the plurality of upsampled first maps and the plurality of corrected second maps.

6. A non-transitory computer-readable medium having computer-executable instructions that, upon execution of the instructions by a processor of a computer, cause the computer to function as the object-specific keypoint separation apparatus according to claim 1 .

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 18, 2022
From: MATSUMURA, MASAAKI; NOTO, HAJIME; KUSACHI, YOSHINORI
To: NIPPON TELEGRAPH AND TELEPHONE CORPORATION
Reel/Frame 060840/0131 →
Continuity (1)
Related Publication 20230101653A1 · Mar 30, 2023
References Cited (7)
US 12094159B1 · Akbas · 2024 [cited by examiner]
US 20210049356A1 · Gu · 2021 [cited by examiner]
US 20240252841A1 · Labarbe · 2024 [cited by examiner]
Cao et al., “OpenPose: Realtime Multi-Person 2D Pose Estimation using Part Affinity Fields,” arXiv:1812.08008 (Year: 2018). [cited by examiner]
Papandreou et al., “PersonLab: Person Pose Estimation and Instance Segmentation with a Bottom-Up, Part-Based, Geometric Embedding Model,” arXiv:1803.08225v1 [cs.CV] Mar. 22, 2018 (Year: 2018). [cited by examiner]
Zhe Cao et al., OpenPose: Realtime Multi-Person 2D Pose Estimation using Part Affinity Fields, arXiv, Dec. 18, 2018. [cited by applicant]
George Papandreou et al., PersonLab: Person pose estimation and instance segmentation with a bottom-up, part-based, geometric embedding model, arXiv, Mar. 22, 2018. [cited by applicant]