IP Library Granted Patent US 8,761,510
Granted Patent B2
US 8,761,510 · App. 13/676,494 · Granted Jun 24, 2014

Object-centric spatial pooling for image classification

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 8,761,510
App. No.
13/676,494
Granted
Jun 24, 2014
Kind
B2
Abstract

A method is provided for classifying an image. The method includes inferring location information of an object of interest in an input representation of the image. The method further includes determining foreground object features and background object features from the input representation of the image. The method additionally includes pooling the foreground object features separately from the background object features using the location information to form a new representation of the image. The new representation is different than the input representation of the image. The method also includes classifying the image based on the new representation of the image.

Claims (22)

1. A method for classifying an image, comprising:

inferring location information of an object of interest in an input representation of the image;

determining foreground object features and background object features from the input representation of the image;

pooling the foreground object features separately from the background object features using the location information to form a new representation of the image, the new representation being different than the input representation of the image;

classifying the image based on the new representation of the image;

performing an outer iterative loop, the outer iterative loop comprising:

initializing a background region from the input representation of the image; and

incrementally shrinking at least a smallest bounding box from among a plurality of varying sized bounding boxes applied to the background region to incrementally increase a size of the background region with respect to at least the smallest bounding box;

performing an inner iterative loop that cooperates with the outer iterative loop, the inner iterative loop comprising:

inferring the location information of the object of interest from a current version of the smallest bounding box; and

from among a set of positive bounding boxes from positive images that are known to include the object of interest therein and from among a set of negative bounding boxes from negative images that are known to omit the object of interest therein, training a support vector machine classifier to discriminate the positive bounding boxes from the negative bounding boxes to refine the location information of the object of interest,

wherein said classifying step classifies the image based on discrimination results applied to a binary label of the image.

2. The method of claim 1 , further comprising separating the foreground object features from the background object features using only image-level class labels.

3. The method of claim 2 , wherein the image-level class levels are inferred using one or more bounding boxes.

4. The method of claim 2 , further comprising training an object classification model using the foreground object features, the background object features, and the image-level class labels.

5. The method of claim 1 , wherein said pooling and classifying steps combine the location information with classifying information in a joint model for both localization and classification.

6. The method of claim 1 , wherein at least the smallest bounding box is incrementally shrunk until a predetermined condition is reached.

7. The method of claim 6 , wherein the predetermined condition comprises at least the smallest bounding box entirely encompassing the object of interest including salient and non-salient portions thereof.

8. The method of claim 1 , wherein at least the smallest bounding box is positioned around a most prominent instance of an object class within the background region, the most prominent instant of the object class being equated to the object of interest.

9. The method of claim 1 , wherein the support vector machine classifier uses only features in the set of positive bounding boxes to the exclusion of features from an entirety of the image that are outside the positive bounding boxes.

10. The method of claim 1 , further comprising retraining an object classification model initially used to separate the foreground object features from the background object features based on the new representation of the image.

11. The method of claim 10 , further comprising updating a hypothesis of the location information of the object using the retrained object classification model, and repeating said determining and pooling steps to form another new representation of the Image.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 13, 2015
From: NEC LABORATORIES AMERICA, INC.
To: NEC CORPORATION
Reel/Frame 034765/0565 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 14, 2012
From: RUSSAKOVSKY, OLGA; LIN, YUANQING; YU, KAI; LI, FEI-FEI
To: NEC LABORATORIES AMERICA, INC.
Reel/Frame 029298/0488 →