IP Library Granted Patent US 11,412,102
Granted Patent B2
US 11,412,102 · App. 17/037,684 · Granted Aug 9, 2022

Information processing apparatus, and non-transitory computer readable medium for splitting documents

Inventors: Takuma Yamamoto (Kanagawa, JP); Aya Kuwano (Kanagawa, JP); Mitsuru Sato (Kanagawa, JP); Toru Takahashi (Kanagawa, JP)
Assignee: FUJIFILM Business Innovation Corp.
H04N1/00641G06V30/40H04N1/0044H04N1/00811H04N1/00824G06V30/10
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,412,102
App. No.
17/037,684
Granted
Aug 9, 2022
Kind
B2
Abstract

An information processing apparatus includes a processor configured to: acquire a read image and item information, the read image being an image obtained by reading a paper medium including plural documents, the item information being information of plural items specified by a user from among plural items contained in the documents; extract plural character strings from the read image, each character string being associated with the corresponding one of the items included in the item information; in response to extracting the character strings associated with the item information from the read image, set a split position, the split position being a position at which to split out a portion of the read image as a set of documents, the portion being a portion of the read image from a page where the extracting has begun to a page containing the last extracted character string; and output the read image split in accordance with the split position.

Claims (46)

1. An information processing apparatus comprising

a processor configured to

acquire a read image and item information, the read image being an image obtained by reading a paper medium including a plurality of documents, the item information being information of a plurality of items specified by a user from among a plurality of items contained in the plurality of documents,

extract a plurality of character strings from the read image, each character string being associated with a corresponding one of the plurality of specified items contained in the plurality of documents and included in the item information,

in response to extracting the plurality of character strings associated with the item information from the read image, set a split position, the split position being a position at which to split out a portion of the read image as a set of documents, the portion being a portion of the read image from a first page where the extracting has begun to a last page containing a last extracted character string in the plurality of documents, and

output the read image split in accordance with the split position.

2. The information processing apparatus according to claim 1 ,

wherein the processor is further configured to acquire a preset value of at least one of a maximum number of pages of the set of documents and a maximum number of copies of the set of documents, and to, in response to the preset value being exceeded, notify that the preset value has been exceeded.

3. The information processing apparatus according to claim 2 ,

wherein the processor is further configured to, in response to the preset value being exceeded, further acquire a user-specified split position, and to, in accordance with the acquired user-specified split position, split and output a set of documents for which the preset value has been exceeded.

4. The information processing apparatus according to claim 2 ,

wherein the processor is further configured to stop a process in response to the preset value being exceeded, the process being a process of extracting the plurality of character strings associated with the item information from the read image.

5. The information processing apparatus according to claim 1 ,

wherein the processor is further configured to display the read image and the split position.

6. The information processing apparatus according to claim 2 ,

wherein the processor is further configured to display the read image and the split position.

7. The information processing apparatus according to claim 3 ,

wherein the processor is further configured to display the read image and the split position.

8. The information processing apparatus according to claim 4 ,

wherein the processor is further configured to display the read image and the split position.

9. The information processing apparatus according to claim 5 ,

wherein the processor is further configured to acquire a user-specified split position, and to, in accordance with the acquired user-specified split position, perform at least one of correction, addition, and deletion of the set split position.

10. The information processing apparatus according to claim 6 ,

wherein the processor is further configured to acquire a user-specified split position, and to, in accordance with the acquired user-specified split position, perform at least one of correction, addition, and deletion of the set split position.

11. The information processing apparatus according to claim 1 ,

wherein the processor is further configured to acquire an extraction region from which each character string is to be extracted, and to extract, from the extraction region of the read image, each character string associated with the item information.

12. The information processing apparatus according to claim 11 ,

wherein the processor is further configured to, if a single page includes a plurality of the extraction regions, set a priority for each extraction region.

13. The information processing apparatus according to claim 11 ,

wherein the processor is further configured to store a region from which each character string has been extracted, and to set the stored region as the extraction region.

14. The information processing apparatus according to claim 1 ,

wherein the processor is further configured to, in response to extracting a plurality of the character strings associated with the item information from a single page contained in the read image, select one of the extracted plurality of character strings.

15. The information processing apparatus according to claim 14 ,

wherein the processor is further configured to set a priority for each item included in the item information, and to select a character string based on the priority.

16. The information processing apparatus according to claim 1 ,

wherein the processor is further configured to set a plurality of necessary items from among the plurality of items included in the item information, and to, in response to extracting all of a plurality of character strings associated with the plurality of necessary items, set a split position, the split position being a position at which a portion of the read image up to a page containing a last extracted character string is to be split out as a set of documents.

17. The information processing apparatus according to claim 16 ,

wherein the processor is further configured to set a plurality of necessary items and a plurality of selective items from among the plurality of items included in the item information, and to, in response to extracting all of a plurality of character strings associated with the plurality of necessary items and extracting at least one of a plurality of character strings associated with the plurality of selective items, set a split position, the split position being a position at which a portion of the read image up to a page containing a last extracted character string is to be split out as a set of documents.

18. The information processing apparatus according to claim 1 ,

wherein the processor is further configured to, in response to extracting the plurality of character strings from the read image, resume extraction from a page immediately following a page containing a last extracted character string.

19. The information processing apparatus according to claim 1 , wherein each of the character strings has a predetermined positional relationship with the corresponding one of the plurality of specified items contained in the plurality of documents.

20. A non-transitory computer readable medium storing a program causing a computer to execute a process for processing information, the process comprising:

acquiring a read image and item information, the read image being an image obtained by reading a paper medium including a plurality of documents, the item information being information of a plurality of items specified by a user from among a plurality of items included in the plurality of documents;

extracting a plurality of character strings from the read image, each character string being associated with a corresponding one of the plurality of specified items contained in the plurality of documents and included in the item information;

in response to extracting the plurality of character strings associated with the item information from the read image, setting a split position, the split position being a position at which to split out a portion of the read image as a set of documents, the portion being a portion of the read image from a first page where the extracting has begun to a last page containing a last extracted character string in the plurality of documents; and

outputting the read image split accordance with the split position.

Assignments (2)
CHANGE OF NAME Recorded May 14, 2021
From: FUJI XEROX CO., LTD.
To: FUJIFILM BUSINESS INNOVATION CORP.
Reel/Frame 056237/0144 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Sep 30, 2020
From: YAMAMOTO, TAKUMA; KUWANO, AYA; SATO, MITSURU; TAKAHASHI, TORU
To: FUJI XEROX CO., LTD.
Reel/Frame 053925/0796 →
Priority Claims (1)
JP JP2020-055029 · Mar 25, 2020 · national
Continuity (1)
Related Publication 20210306494A1 · Sep 30, 2021