IP Library Granted Patent US 9,183,226
Granted Patent B2
US 9,183,226 · App. 14/336,692 · Granted Nov 10, 2015

Image classification

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 9,183,226
App. No.
14/336,692
Granted
Nov 10, 2015
Kind
B2
Abstract

An image classification system trains an image classification model to classify images relative to text appearing with the images. Training images are iteratively selected and classified by the image classification model according to feature vectors of the training images. An independent model is trained for unique n-grams of text. The image classification system obtains text appearing with an image and parses the text into candidate labels for the image. The image classification system determines whether an image classification model has been trained for the candidate labels. When an image classification model corresponding to a candidate label has been trained, the image classification subsystem classifies the image relative to the candidate label. The image is labeled based on candidate labels for which the image is classified as a positive image.

Claims (52)

1. A computer-implemented method, comprising:

obtaining a user query;

identifying, by one or more computers, images that have been assigned, by one or more image classification models for one or more corresponding n-grams that each match the user query, the n-gram as a text label of the image, wherein the image classification models assign n-grams as labels to images based on the images being associated with text matching the corresponding n-gram of the image classification model and a feature vector of the image;

applying, by one or more computers, a boost value to a relevance score of each of at least one of the images to obtain an adjusted relevance score, wherein the boost value applied is based, at least in part, on a strength of the match between the user query and the text label of the image;

ranking, by one or more computers, the images based, at least in part, on the adjusted relevance score; and

outputting, by one or more computers, at least a portion of the images for presentation on a search results page responsive to the user query.

2. The method of claim 1 , further comprising assigning a magnitude to the boost value based on a strength of the match between the user query and the text label, wherein the magnitude of the boost value is proportional to the strength of the match.

3. The method of claim 2 , wherein assigning a magnitude to the boost value comprises assigning a highest magnitude to the boost value when the user query exactly matches the text label.

4. The method of claim 1 , further comprising training the classification model.

5. The method of claim 4 , wherein training the classification model comprises training the classification model based on feature vectors of positive training images having relevance measures, for the n-grams, that satisfy a relevance threshold and negative training images having relevance measures, for the n-grams, that do not satisfy the relevance threshold.

6. The method of claim 5 , wherein training the image classification model comprises:

for each of one or more training images from the positive training images or the negative training images:

computing a classification score for the training image based on an initial image classification model;

classifying the training image based on the classification score;

determining that the training image is incorrectly classified; and

computing a minimum kernel approximation based on the feature vector for the training image to adjust the initial image classification model.

7. The method of claim 5 , further comprising assigning an image a label based on a feature vector for the image and the image classification model for an n-gram matching the label.

8. A system comprising:

a data store storing images; and

one or more computers that interact with the data store and execute instructions that cause the one more computers to perform operations comprising:

obtaining a user query;

identifying images that have been assigned, by one or more image classification models for one or more corresponding n-grams that each match the user query, the n-gram as a text label of the image, wherein the image classification models assign n-grams as labels to images based on the images being associated with text matching the corresponding n-gram of the image classification model and a feature vector of the image;

applying a boost value to a relevance score of each of at least one of the images to obtain an adjusted relevance score, wherein the boost value applied is based, at least in part, on a strength of the match between the user query and the text label of the image;

ranking the images based, at least in part, on the adjusted relevance score; and

outputting at least a portion of the images for presentation on a search results page responsive to the user query.

9. The system of claim 8 , wherein the instructions cause the one more computers to perform operations further comprising assigning a magnitude to the boost value based on a strength of the match between the user query and the text label, wherein the magnitude of the boost value is proportional to the strength of the match.

10. The system of claim 9 , wherein assigning a magnitude to the boost value comprises assigning a highest magnitude to the boost value when the user query exactly matches the text label.

11. The system of claim 8 , wherein the instructions cause the one more computers to perform operations further comprising training the classification model.

12. The method of claim 11 , wherein training the classification model comprises training the classification model based on feature vectors of positive training images having relevance measures, for the n-grams, that satisfy a relevance threshold and negative training images having relevance measures, for the n-grams, that do not satisfy the relevance threshold.

13. The system of claim 12 , wherein training the image classification model comprises:

for each of one or more training images from the positive training images or the negative training images:

computing a classification score for the training image based on an initial image classification model;

classifying the training image based on the classification score;

determining that the training image is incorrectly classified; and

computing a minimum kernel approximation based on the feature vector for the training image to adjust the initial image classification model.

14. The system of claim 12 , wherein the instructions cause the one more computers to perform operations further comprising assigning an image a label based on a feature vector for the image and the image classification model for an n-gram matching the label.

15. A non-transitory computer readable medium encoded with a computer program comprising instructions that when executed cause one or more computers to perform operations comprising:

obtaining a user query;

identifying images that have been assigned, by one or more image classification models for one or more corresponding n-grams that each match the user query, the n-gram as a text label of the image, wherein the image classification models assign n-grams as labels to images based on the images being associated with text matching the corresponding n-gram of the image classification model and a feature vector of the image;

applying a boost value to a relevance score of each of at least one of the images to obtain an adjusted relevance score, wherein the boost value applied is based, at least in part, on a strength of the match between the user query and the text label of the image;

ranking the images based, at least in part, on the adjusted relevance score; and

outputting at least a portion of the images for presentation on a search results page responsive to the user query.

16. The computer readable medium of claim 15 , wherein the instructions cause the one more computers to perform operations further comprising assigning a magnitude to the boost value based on a strength of the match between the user query and the text label, wherein the magnitude of the boost value is proportional to the strength of the match.

17. The computer readable medium of claim 16 , wherein assigning a magnitude to the boost value comprises assigning a highest magnitude to the boost value when the user query exactly matches the text label.

18. The computer readable medium of claim 15 , wherein the instructions cause the one more computers to perform operations further comprising training the classification model.

19. The computer readable medium of claim 18 , wherein training the classification model comprises training the classification model based on feature vectors of positive training images having relevance measures, for the n-grams, that satisfy a relevance threshold and negative training images having relevance measures, for the n-grams, that do not satisfy the relevance threshold.

20. The computer readable medium of claim 19 , wherein training the image classification model comprises:

for each of one or more training images from the positive training images or the negative training images:

computing a classification score for the training image based on an initial image classification model;

classifying the training image based on the classification score;

determining that the training image is incorrectly classified; and

computing a minimum kernel approximation based on the feature vector for the training image to adjust the initial image classification model.

Assignments (2)
CHANGE OF NAME Recorded Oct 2, 2017
From: GOOGLE INC.
To: GOOGLE LLC
Reel/Frame 044334/0466 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 1, 2014
From: YEE, YANGLI HECTOR; BENGIO, SAMY; ROSENBERG, CHARLES J.; MURPHY-CHUTORIAN, ERIK
To: GOOGLE INC.
Reel/Frame 033441/0029 →