IP Library Granted Patent US 10,679,091
Granted Patent B2
US 10,679,091 · App. 16/525,354 · Granted Jun 9, 2020

Image box filtering for optical character recognition

Inventors: Arnaud G. Flament (Sunnyvale, CA); Guillaume B. Koch (San Jose, CA)
Assignee: Open Text Corporation
G06K9/46G06K9/18G06K2009/4666
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,679,091
App. No.
16/525,354
Granted
Jun 9, 2020
Kind
B2
Abstract

A method for box filtering includes obtaining, by a computing device, a form image, and identifying, by the computing device, a region of the form image that includes boxes. Vertical lines in the region of the form image are detected. The boxes in the region are detected according to the plurality of vertical lines, and image content is extracted from the boxes.

Claims (63)

1. A computer program product comprising a non-transitory computer readable medium storing a set of computer-readable instructions, the set of computer-readable instructions comprising instructions executable by a processor to:

obtain a form image;

identify a region of the obtained form image, the region comprising a plurality of boxes included in the obtained form image;

load a set of box removal parameters, the set of box removal parameters comprising a sliding window parameter specifying a size of a sliding window;

detect a plurality of vertical lines in the region of the obtained form image from a plurality of pixels in the region, the plurality of pixels comprising pixels representing the plurality of vertical lines, wherein each vertical line in the plurality of vertical lines is detected from a respective aggregation of pixels in the sliding window that comply with a color requirement;

detect the plurality of boxes in the region based on the detected plurality of vertical lines;

extract image content from the plurality of boxes; and

generate a preprocessed form with the plurality of boxes removed, the preprocessed form including the image content extracted from the plurality of boxes.

2. The computer program product of claim 1 , wherein the set of computer-readable instructions further comprise instructions executable to:

perform optical character recognition on the image content in the preprocessed form to generate a text recognized form;

extract character content from the text recognized form to generate extracted character content; and

store the extracted character content.

3. The computer program product of claim 1 , wherein the set of computer-readable instructions further comprise instructions executable to:

project the plurality of pixels in the region onto a horizontal axis to create a horizontal axis projection; and

identifying a first plurality of peaks in the horizontal axis projection, wherein the plurality of vertical lines are detected using the first plurality of peaks.

4. The computer program product of claim 3 , wherein the set of computer-readable instructions further comprise instructions executable to:

project the plurality of pixels on a vertical axis to create a vertical axis projection; and

identify a second plurality of peaks in the vertical axis projection to detect a top and a bottom of each of the plurality of boxes.

5. The computer program product of claim 4 , wherein the set of computer-readable instructions further comprise instructions executable to remove a sub-region of the region between adjacent peaks in the first plurality of peaks based on lacking the top and the bottom.

6. The computer program product of claim 3 , wherein the set of box removal parameters comprises a box type parameter.

7. The computer program product of claim 6 , wherein the set of computer-readable instructions further comprise instructions executable to:

project the plurality of pixels on a vertical axis to create a vertical axis projection; and

based on a determination that the box type parameter indicates a comb box, identify a second plurality of peaks in the vertical axis projection to detect a bottom of each of the plurality of boxes.

8. The computer program product of claim 3 , wherein the set of computer-readable instructions further comprise instructions executable to:

detect a standard width of the plurality of boxes based on the first plurality of peaks; and

remove at least one peak from the first plurality of peaks that defines a boundary of a box that is less than the standard width by a threshold amount.

9. The computer program product of claim 3 , wherein the set of computer-readable instructions further comprise instructions executable to:

detect a standard width of the plurality of boxes based on the first plurality of peaks; and

remove at least one peak from the first plurality of peaks that defines a boundary of a box that is greater than the standard width but not a multiple of the standard width.

10. The computer program product of claim 3 , wherein extracting the image content from the plurality of boxes comprises:

copying an image in each of the plurality of boxes within a defined margin.

11. A method for box filtering comprising:

obtaining, by a computing device, a form image, the obtained from image comprising a plurality of boxes and image content in the plurality of boxes;

identifying, by the computing device, a region of the obtained form image comprising the plurality of boxes included in the obtained form image;

loading a set of box removal parameters, the set of box removal parameters comprising a sliding window parameter specifying a size of a sliding window;

detecting a plurality of vertical lines in the region of the obtained form image from a plurality of pixels in the region, the plurality of pixels comprising pixels representing the plurality of vertical lines, wherein each vertical line in the plurality of vertical lines is detected from a respective aggregation of pixels in the sliding window that comply with a color requirement;

detecting, by the computing device, the plurality of boxes in the region based on the detected plurality of vertical lines;

extracting, by the computing device, the image content from the plurality of boxes; and

generating a preprocessed form with the plurality of boxes removed, the preprocessed form including the image content extracted from the plurality of boxes.

12. The method of claim 11 , further comprising:

performing optical character recognition on the image content in the preprocessed form to generate a text recognized form;

extracting character content from the text recognized form to generate extracted character content; and

storing the extracted character content.

13. The method of claim 11 , wherein detecting the plurality of vertical lines comprises:

projecting, by the computing device, the plurality of pixels in the region onto a horizontal axis to create a horizontal axis projection; and

identifying, by the computing device, a first plurality of peaks in the horizontal axis projection, wherein the plurality of vertical lines are detected using the first plurality of peaks.

14. The method of claim 13 , wherein detecting the plurality of boxes in the region comprises:

projecting the plurality of pixels on a vertical axis to create a vertical axis projection; and

identifying a second plurality of peaks in the vertical axis projection to detect a top and a bottom of each of the plurality of boxes.

15. The method of claim 14 , wherein detecting the plurality of boxes in the region comprises:

removing a sub-region of the region between adjacent peaks in the first plurality of peaks based on lacking the top and the bottom.

16. The method of claim 13 , wherein the set of box removal parameters comprises a box type parameter.

17. The method of claim 16 , wherein detecting the plurality of boxes in the region comprises:

projecting the plurality of pixels on a vertical axis to create a vertical axis projection; and

based on a determination that the box type parameter indicates a comb box, identifying a second plurality of peaks in the vertical axis projection to detect a bottom of each of the plurality of boxes.

18. The method of claim 13 , wherein detecting the plurality of boxes comprises:

detecting a standard width of the plurality of boxes based on the first plurality of peaks; and

removing at least one peak from the first plurality of peaks that defines a boundary of a box that is less than the standard width by a threshold amount.

19. The method of claim 13 , wherein detecting the plurality of boxes comprises:

detecting a standard width of the plurality of boxes based on the first plurality of peaks; and

removing at least one peak from the first plurality of peaks that defines a boundary of a box that is greater than the standard width but not a multiple of the standard width.

20. The method of claim 11 , wherein extracting the image content from the plurality of boxes comprises:

copying an image in each of the plurality of boxes within a defined margin.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 5, 2020
From: EMC CORPORATION
To: OPEN TEXT CORPORATION
Reel/Frame 052026/0295 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 5, 2020
From: FLAMENT, ARNAUD G.; KOCH, GUILLAUME B.
To: EMC CORPORATION
Reel/Frame 052096/0978 →
Continuity (3)
Continuation 14788170 · Jun 30, 2015
Provisional Application 62158775 · May 8, 2015
Related Publication 20190354792A1 · Nov 21, 2019
Cited By (1)
US 12,412,413