IP Library Granted Patent US 8,763,038
Granted Patent B2
US 8,763,038 · App. 12/321,856 · Granted Jun 24, 2014

Capture of stylized TV table data via OCR

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 8,763,038
App. No.
12/321,856
Granted
Jun 24, 2014
Kind
B2
Abstract

In certain implementations consistent with the present invention, a method of detecting text in a television video display table involves saving a frame of video to a memory device; determining that the frame of video contains a table having cells containing text; storing a working copy of the frame of video to a memory; isolating text in the table by: removing any table boundaries from the image; removing any cell boundaries from the image; determining if the image has three dimensional or shadow attributes and removing any three dimensional or shadow attributes identified; thereby producing text isolated against a contrasting color background; and processing the isolated text using an optical character recognition (OCR) engine to extract the text as data. This abstract is not to be considered limiting, since other embodiments may deviate from the features described in this abstract.

Claims (48)

1. A method of detecting text in a television video display table, comprising:

saving a frame of video representing an image to a memory device;

determining that the frame of video contains a table having cells containing text;

storing a working copy of the frame of video to a memory;

isolating text in the table by:

removing any table boundaries from the image;

removing any cell boundaries from the image;

determining if the image has three dimensional or shadow attributes in the table boundaries or cell boundaries and removing any three dimensional or shadow attributes identified, wherein determining if the image has three dimensional or shadow attributes in the table boundaries or cell boundaries is carried out by finding line patterns adjacent and outside table or cell boundaries that track a table or cell boundary;

where determining if a cell has three dimensional or shadow attributes is carried out by subtracting a variable rectangular band of pixels from other cell values to see what size band maximizes cancellation in order to distinguish the cell from the text area;

thereby producing text isolated against a contrasting color background; and

processing the isolated text using an optical character recognition (OCR) engine to extract the text as data.

2. The method according to claim 1 , further comprising converting the text isolated against a contrasting color background to black text on a white background.

3. The method according to claim 1 , wherein determining that the frame of video contains a table having cells containing text is carried out by tracking remote control commands transmitted from a remote control to identify commands that result in display of a table having cells containing text.

4. The method according to claim 3 , wherein the commands comprise commands that cause display of a program guide, list of content, PVR list, dialogue box, closed captioning, or setup table.

5. The method according to claim 1 , wherein determining that the frame of video contains a table having cells containing text is carried out by detecting rectangular shapes of size adequate to contain legible text.

6. The method according to claim 1 , wherein determining that the frame of video contains a table having cells containing text is carried out by matching a display to a template.

7. The method according to claim 1 , wherein removing the boundaries of the cells and the tables is carried out by matching a display to a template in order to locate the boundaries.

8. The method according to claim 1 , further comprising storing the extracted text as data to a metadata database.

9. The method according to claim 1 , where the frame of video containing the table has cells containing text of a foreground color against a background color.

10. A method of detecting text in a television video display table, comprising:

saving a frame of video representing an image to a memory device;

determining that the frame of video contains a table having cells containing text by detecting rectangular shapes of size adequate to contain legible text;

storing a working copy of the frame of video to a memory;

isolating text in the table by:

removing any table boundaries from the image;

removing any cell boundaries from the image by matching a display to a template in order to locate the boundaries;

determining if the image has three dimensional or shadow attributes in the table boundaries or cell boundaries and removing any three dimensional or shadow attributes identified, wherein determining if the image has three dimensional or shadow attributes in the table boundaries or cell boundaries is carried out by finding line patterns adjacent and outside table or cell boundaries that track a table or cell boundary;

where determining if a cell has three dimensional or shadow attributes is carried out by subtracting a variable rectangular band of pixels from other cell values to see what size band maximizes cancellation in order to distinguish the cell from the text area;

thereby producing text isolated against a contrasting color background; and

processing the isolated text using an optical character recognition (OCR) engine to extract the text as data to a metadata database.

11. The method according to claim 10 , further comprising converting the text isolated against a contrasting color background to black text on a white background.

12. The method according to claim 10 , wherein determining that the frame of video contains a table having cells containing text is carried out by tracking remote control commands issued to identify commands that result in display of a table having cells containing text.

13. The method according to claim 12 , wherein the commands comprise commands that cause display of a program guide, list of content, PVR list, dialogue box, closed captioning, or setup table.

14. The method according to claim 10 , wherein determining that the frame of video contains a table having cells containing text further comprises matching a display to a template.

15. The method according to claim 10 , where the frame of video containing the table has cells containing text of a foreground color against a background color.

16. A method of detecting text in a television video display table, comprising:

saving a frame of video representing an image to a memory device;

determining that the frame of video contains a table having cells containing text by detecting rectangular shapes of size adequate to contain legible text and then by matching a display to a template;

storing a working copy of the frame of video to a memory;

isolating text in the table by:

removing any table boundaries from the image;

removing any cell boundaries from the image by matching a display to a template in order to locate the boundaries;

determining if the image has three dimensional or shadow attributes in the table boundaries or cell boundaries and removing any three dimensional or shadow attributes identified, wherein determining if the image has three dimensional or shadow attributes in the table boundaries or cell boundaries is carried out by finding line patterns adjacent and outside table or cell boundaries that track a table or cell boundary;

where determining if a cell has three dimensional or shadow attributes is further carried out by subtracting a variable rectangular band of pixels from other cell values to see what size band maximizes cancellation in order to distinguish the cell from the text area;

thereby producing text isolated against a contrasting color background;

converting the text isolated against a contrasting color background to black text on a white background; and

processing the isolated text using an optical character recognition (OCR) engine to extract the text as data to a metadata database.

17. The method according to claim 16 , wherein determining that the frame of video contains a table having cells containing text is carried out by tracking remote control commands issued to identify commands that result in display of a table having cells containing text, wherein the commands comprise commands that cause display of a program guide, list of content, PVR list, dialogue box, closed captioning, or setup table.

Assignments (4)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 15, 2019
From: SONY CORPORATION
To: SATURN LICENSING LLC
Reel/Frame 048974/0222 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 11, 2015
From: SONY ELECTRONICS INC.
To: SONY CORPORATION
Reel/Frame 036330/0420 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 24, 2014
From: CANDELORE, BRANT L.
To: SONY CORPORATION; SONY ELECTRONICS INC.
Reel/Frame 032745/0106 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Feb 23, 2009
From: CANDELORE, BRANT L.
To: SONY CORPORATION; SONY ELECTRONICS INC.
Reel/Frame 022295/0614 →