IP Library Granted Patent US 9,489,577
Granted Patent B2
US 9,489,577 · App. 12/804,518 · Granted Nov 8, 2016

Visual similarity for video content

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 9,489,577
App. No.
12/804,518
Granted
Nov 8, 2016
Kind
B2
Abstract

Methods and apparatus, including computer program products, for visual similarity. A method includes receiving a stream of video content, generating interpretations of the received video content using speech/natural language processing (NLP), associating the interpretations of the received video content with images extracted from video content based on timeline, and using the interpretations to obtain interpretations of other images or other video content.

Claims (31)

1. A method comprising:

receiving a stream of video content;

generating speech to text for the received video content;

generating passage level annotations from the generated text using natural language processing (NLP);

associating the passage level annotations with a timeline; and

associating imagery with the text to generate thumbnails at periodic time intervals resulting in a database of annotations to imagery and imagery to annotations.

2. The method of claim 1 further comprising using database of annotations to imagery and imagery to annotations to obtain interpretations of other digital images to improve the interpretation derived from the speech.

3. The method of claim 1 wherein the annotations comprise names, concepts, and/or logical substructure.

4. The method of claim 1 wherein associating imagery with the annotated text to generate thumbnails at periodic time intervals comprises applying a machine learning technique.

5. The method of claim 4 wherein the machine learning technique includes one of a neural networks approach or a classification and regression tree (CART) approach to classify an image into an annotation or to classify an annotation into an image.

6. The method of claim 4 wherein the machine learning technique is a principal component analysis (PCA) of facial imagery.

7. The method of claim 4 wherein associating imagery with the annotated text to generate thumbnails at periodic time intervals further comprises clustering the images.

8. The method of claim 1 further comprising generating a master catalog of images for selected annotations.

9. The method of claim 8 further comprising comparing videos to determine whether the videos contain speech.

10. The method of claim 9 wherein comparing videos comprises pixel comparison.

11. An apparatus comprising:

a local computing system linked to a network of interconnected computer systems, the local computing system comprising a processor, a memory and a storage device;

the memory comprising an operating system and a visual similarity process, the visual similarity process comprising:

receiving a stream of video content;

generating speech to text for the received video content;

generating passage level annotations from the generated text using natural language processing (NLP);

associating the passage level annotations with the text time aligned to result in text, annotations and a time stamp; and

associating imagery with the annotated text to generate thumbnails at periodic time intervals resulting in a database of annotations to imagery and imagery to annotations.

12. The apparatus of claim 11 wherein the annotations comprise names, concepts, and/or logical substructure.

13. The apparatus of claim 11 wherein associating imagery with the annotated text to generate thumbnails at periodic time intervals comprises applying a machine learning technique.

14. The apparatus of claim 13 wherein the machine learning technique includes one of a neural networks approach or a classification and regression tree (CART) approach to classify an image into an annotation or to classify an annotation into an image.

15. The apparatus of claim 13 wherein the machine learning technique is a principal component analysis (PCA) of facial imagery.

16. The apparatus of claim 11 wherein associating imagery with the annotated text to generate thumbnails at periodic time intervals further comprises clustering the images.

17. The apparatus of claim 11 wherein the visual similarity process further comprises generating a master catalog of images for selected annotations.

18. The apparatus of claim 17 wherein the visual similarity process further comprises comparing videos to determine whether the videos contain speech.

19. The apparatus of claim 18 wherein comparing videos comprises pixel comparison.

Assignments (6)
RELEASE OF SECURITY INTEREST Recorded Jan 2, 2025
From: SIXTH STREET SPECIALTY LENDING INC.
To: PIANO SOFTWARE B.V.
Reel/Frame 069722/0531 →
SECURITY INTEREST Recorded Oct 3, 2022
From: PIANO SOFTWARE B.V.
To: SIXTH STREET SPECIALTY LENDING, INC.
Reel/Frame 061290/0590 →
PATENT RELEASE AND REASSIGNMENT Recorded Feb 26, 2021
From: WESTERN ALLIANCE BANK
To: CXENSE, INC.
Reel/Frame 055424/0862 →
SECURITY INTEREST Recorded Oct 3, 2019
From: CXENSE, INC.
To: WESTERN ALLIANCE BANK, AN ARIZONA CORPORATION
Reel/Frame 050612/0905 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 30, 2015
From: RAMP HOLDINGS INC.
To: CXENSE ASA
Reel/Frame 037018/0816 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Sep 13, 2010
From: WILDE, THOMAS
To: RAMP HOLDINGS, INC.
Reel/Frame 024981/0785 →