IP Library › Granted Patent US 10,445,615
Granted Patent B2
US 10,445,615 · App. 15/646,512 · Granted Oct 15, 2019

Method and device for extracting images from portable document format (PDF) documents

Inventors: Balaji Jagan (Dindigul, IN); Naveen Kumar Nanjappa (Bengaluru, IN)
Assignee: Wipro Limited
G06K9/469G06K9/00456G06K9/00469G06K9/2009G06K9/4604G06K2209/27
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,445,615
App. No.
15/646,512
Granted
Oct 15, 2019
Kind
B2
Abstract

A method and device for extracting images from PDF documents are disclosed. The method includes performing a text recognition process on a PDF document that includes one or more images. The text recognition process replaces the one or more images with a plurality of contiguous newlines. The method further includes storing a location of each of the one or more images within the PDF document based on occurrence of the plurality of contiguous newlines within the PDF document. The method includes converting each page of the PDF document to an image format in order to generate an image document corresponding to the PDF document. The method further includes extracting each of the one or more images from the image document based on the location stored for each of the one or more images within the PDF document.

Claims (31)

1. A method for extracting images from Portable Document Format (PDF) documents, the method comprising:

performing, by an image extraction device, a text recognition process on a PDF document comprising one or more images, wherein the text recognition process replaces the one or more images with a plurality of contiguous newlines;

storing, by the image extraction device, a location of each of the one or more images within the PDF document based on occurrence of the plurality of contiguous newlines within the PDF document;

converting, by the image extraction device, each page of the PDF document to an image format in order to generate an image document corresponding to the PDF document;

incrementally scanning, by the image extraction device, a page of the image document comprising the image wherein the scanning comprises tracing contour of the image based on the coordinates of corners of the image within the page in at least one of a square and a rectangle pattern; and

extracting, by the image extraction device, each of the one or more images from the image document based on the location stored for each of the one or more images within the PDF document.

2. The method of claim 1 , wherein at least one of the one or more images is a vector graphic image.

3. The method of claim 1 , wherein storing a location of an image from the one or more images within the PDF document comprises associating a location metadata with the PDF document, wherein the location metadata comprises information related to the location of the image.

4. The method of claim 1 , wherein a location of an image from the one or more images comprises a page number of a page including the image and coordinates of corners of the image within the page.

5. The method of claim 1 further comprising storing each of the one or more images in a predefined format in response to extracting each of the one or more images from the image document.

6. The method of claim 5 , wherein an extracted image is tagged with an associated location metadata indicating location of the extracted image within the PDF document.

7. The method of claim 1 , wherein the text recognition process is performed using an Open source Computer Vision (OpenCV) tool.

8. An image extraction device for extracting images from Portable Document Format (PDF) documents, the image extraction device comprising:

at least one processor; a memory communicatively coupled to the processor, wherein the memory stores processor instructions, which, on execution, causes the processor to:

perform a text recognition process on a PDF document comprising one or more images, wherein the text recognition process replaces the one or more images with a plurality of contiguous newlines;

store a location of each of the one or more images within the PDF document based on occurrence of the plurality of contiguous newlines within the PDF document;

convert each page of the PDF document to an image format in order to generate an image document corresponding to the PDF document;

incrementally scan a page of the image document comprising the image wherein the scanning comprises tracing the contour of the image based on the coordinates of corners of the image within the page in at least one of a square and a rectangle pattern; and

extract each of the one or more images from the image document based on the location stored for each of the one or more images within the PDF document.

9. The image extraction device of claim 8 , wherein at least one of the one or more images is a vector graphic image.

10. The image extraction device of claim 8 , wherein to store a location of an image from the one or more images within the PDF document, the processor instructions further cause the processor to associate a location metadata with the PDF document, wherein the location metadata comprises information related to the location of the image.

11. The image extraction device of claim 8 , wherein a location of an image from the one or more images comprises a page number of a page including the image and coordinates of corners of the image within the page.

12. The image extraction device of claim 8 , wherein the processor instructions further cause the processor to store each of the one or more images in a predefined format in response to extracting each of the one or more images from the image document.

13. The image extraction device of claim 12 , wherein an extracted image is tagged with an associated location metadata indicating location of the extracted image within the PDF document.

14. The image extraction device of claim 8 , wherein the text recognition process is performed using an Open source Computer Vision (OpenCV) tool.

15. A non-transitory computer-readable storage medium having stored thereon, a set of computer-executable instructions causing a computer comprising one or more processors to perform steps comprising:

performing a text recognition process on a PDF document comprising one or more images, wherein the text recognition process replaces the one or more images with a plurality of contiguous newlines;

storing a location of each of the one or more images within the PDF document based on occurrence of the plurality of contiguous newlines within the PDF document;

converting each page of the PDF document to an image format in order to generate an image document corresponding to the PDF document;

incrementally scanning, by the image extraction device, a page of the image document comprising the image wherein the scanning comprises tracing contour of the image based on the coordinates of corners of the image within the page in at least one of a square and a rectangle pattern; and

extracting each of the one or more images from the image document based on the location stored for each of the one or more images within the PDF document.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jul 11, 2017
From: JAGAN, BALAJI; KUMAR NANJAPPA, NAVEEN
To: WIPRO LIMITED
Reel/Frame 043152/0183 →
Priority Claims (1)
IN 201741018278 · May 24, 2017 · national
Continuity (1)
Related Publication 20180341830A1 · Nov 29, 2018