IP Library Granted Patent US 10,026,017
Granted Patent B2
US 10,026,017 · App. 15/289,830 · Granted Jul 17, 2018

Scene labeling of RGB-D data with interactive option

Inventor: Tao Luo (Beijing, CN)
Assignee: THOMSON LICENSING
G06K9/6263G06K9/624G06K9/628G06K9/6262G06K9/6293
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,026,017
App. No.
15/289,830
Granted
Jul 17, 2018
Kind
B2
Abstract

A method for generating labels for an image of low quality is described. The method includes mapping image data and depth information of said image to a 3D point cloud; segmenting the 3-D point cloud and the image into super voxels and image patches, fusing features obtained from the super voxels and image patches, by using a fusion model, applying classifiers to fused features, wherein the fusion model and the classifiers are generated from a dataset including image data of selected quality and quantity, corresponding point cloud and image labels, and generating scene labels based on applied classifiers.

Claims (24)

1. A method for generating labels for an image, comprising:

mapping image data and depth information of said image to a three-dimensional (3D) point cloud;

segmenting the 3D point cloud and the image into super voxels and image patches, respectively;

fusing features obtained from the super voxels and image patches by using a fusion model;

applying classifiers to the fused features, wherein the fusion model and the classifiers are generated from a stored dataset including image data having a selected quality and quantity along with corresponding point clouds and image labels; and

generating scene labels based on the applied classifiers.

2. The method according to claim 1 wherein the image data and depth information comprises red, green and blue color data and depth information.

3. The method according to claim 1 wherein the features are obtained by using one of an unsupervised or a supervised feature learning framework on the image and the 3D point cloud separately.

4. The method according to claim 1 wherein the classifiers are obtained using an exemplar Support Vector Machine (SVM) method.

5. The method according to claim 1 further including interactively correcting the generated scene labels based on user input.

6. The method according to claim 1 wherein the classifiers are updated based on user input.

7. An Apparatus for scene labeling an image, comprising:

at least one processor configured to

(a) map image data and depth information of said image to a three-dimensional-(3D) point cloud;

(b) segment the 3D point cloud and the image into super voxels and image patches, respectively;

(c) fuse features that are obtained from the super voxels and image patches by using a fusion model;

(d) apply classifiers to the fused features; and

(e) generate scene labels based on the applied pre-trained classifiers, wherein the fusion model and the classifiers are generated from a stored dataset including image data having a selected quality and quantity along with corresponding point clouds and image labels.

8. The apparatus according to claim 7 wherein the image data and the depth information comprises red, green and blue color data and depth information.

9. The apparatus according to claim 7 wherein the fusion model comprises a bimodal Deep Boltzman Machine model.

10. The apparatus according to claim 7 wherein the features are obtained using one of an unsupervised or a supervised feature learning framework on the image and the 3D point cloud separately.

11. The apparatus according to claim 7 wherein the classifiers are obtained using an exemplar Support Vector Machine (SVM) apparatus.

12. The apparatus according to claim 7 wherein the processor interactively corrects the generated scene labels based on user input.

13. The apparatus according to claim 7 wherein the processor updates the classifiers based on user input.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 20, 2020
From: THOMSON LICENSING S.A.S.
To: MAGNOLIA LICENSING LLC
Reel/Frame 053570/0237 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 3, 2017
From: LUO, TAO
To: THOMSON LICENSING
Reel/Frame 041825/0472 →
Priority Claims (1)
EP 15306661 · Oct 16, 2015 · regional
Continuity (1)
Related Publication 20170109611A1 · Apr 20, 2017
Cited By (1)
US 12,272,107