IP Library Granted Patent US 11,336,779
Granted Patent B2
US 11,336,779 · App. 17/037,746 · Granted May 17, 2022

Information processing apparatus, and non-transitory computer readable medium

Inventors: Takuma Yamamoto (Kanagawa, JP); Aya Kuwano (Kanagawa, JP); Mitsuru Sato (Kanagawa, JP); Toru Takahashi (Kanagawa, JP)
Assignee: FUJIFILM Business Innovation Corp.
H04N1/00331G06V30/153H04N1/00644H04N2201/0094
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,336,779
App. No.
17/037,746
Granted
May 17, 2022
Kind
B2
Abstract

An information processing apparatus includes a processor configured to: acquire a read image and item information, the read image being an image obtained by reading a paper medium including plural documents, the item information being information of an item specified by a user from among plural items contained in the documents; extract a character string from the read image, the character string being associated with the item information; if a character string contained in a page of the read image and extracted from the page differs from a character string extracted from the previous page immediately preceding the page, set a split position, the split position being a position at which to split out a portion of the read image as a set of documents, the portion being a portion of the read image from a page where extraction has begun to the previous page; and output the read image split in accordance with the split position.

Claims (51)

1. An information processing apparatus comprising

a processor configured to

acquire a read image and item information, the read image being an image obtained by reading a paper medium including a plurality of documents, the item information being information of an item specified by a user from among a plurality of items contained in the plurality of documents,

extract a character string from the read image, the character string being associated with the item information,

if a character string contained in a page of the read image and extracted from the page differs from a character string extracted from a previous page immediately preceding the page, set a split position, the split position being a position at which to split out a portion of the read image as a set of documents, the portion being a portion of the read image from a page where extraction has begun to the previous page, and

output the read image split in accordance with the split position.

2. The information processing apparatus according to claim 1 ,

wherein the processor is further configured to acquire a preset value of at least one of a maximum number of pages of the set of documents and a maximum number of copies of the set of documents, and to, in response to the preset value being exceeded, notify that the preset value has been exceeded.

3. The information processing apparatus according to claim 2 ,

wherein the processor is further configured to, in response to the preset value being exceeded, further acquire a user-specified split position, and to, in accordance with the acquired user-specified split position, split and output a set of documents for which the preset value has been exceeded.

4. The information processing apparatus according to claim 3 ,

wherein the processor is further configured to display the read image and the split position.

5. The information processing apparatus according to claim 4 ,

wherein the processor is further configured to acquire a user-specified split position, and to, in accordance with the acquired user-specified split position, perform at least one of correction, addition, and deletion of the set split position.

6. The information processing apparatus according to claim 2 ,

wherein the processor is further configured to stop a process in response to the preset value being exceeded, the process being a process of extracting the character string associated with the item information from the read image.

7. The information processing apparatus according to claim 6 ,

wherein the processor is further configured to display the read image and the split position.

8. The information processing apparatus according to claim 2 ,

wherein the processor is further configured to display the read image and the split position.

9. The information processing apparatus according to claim 8 ,

wherein the processor is further configured to acquire a user-specified split position, and to, in accordance with the acquired user-specified split position, perform at least one of correction, addition, and deletion of the set split position.

10. The information processing apparatus according to claim 1 ,

wherein the processor is further configured to display the read image and the split position.

11. The information processing apparatus according to claim 10 ,

wherein the processor is further configured to acquire a user-specified split position, and to, in accordance with the acquired user-specified split position, perform at least one of correction, addition, and deletion of the set split position.

12. The information processing apparatus according to claim 1 ,

wherein the processor is further configured to acquire an extraction region from which the character string is to be extracted, and to extract, from the extraction region of the read image, the character string associated with the item information.

13. The information processing apparatus according to claim 12 ,

wherein the extraction region is one of a plurality of extraction regions, and

wherein the processor is further configured to, if a single page includes the plurality of extraction regions, set a priority for each extraction region.

14. The information processing apparatus according to claim 12 ,

wherein the processor is further configured to accept a user's specification of an extraction region.

15. The information processing apparatus according to claim 12 ,

wherein the processor is further configured to store a region from which the character string has been extracted, and to set the stored region as the extraction region.

16. The information processing apparatus according to claim 1 ,

wherein the processor is further configured to

store the character string extracted from the read image, and if a character string extracted from a page of the read image differs from a character string extracted from a previous page immediately preceding the page,

compare the extracted character string with the stored character string, and

add the page containing the extracted character string to a split-out portion of the read image that contains the stored character string.

17. The information processing apparatus according to claim 16 ,

wherein the processor is further configured to, if the extracted character string corresponds to the stored character string, add the page of the read image containing the extracted character string to the split-out portion of the read image that contains the stored character string.

18. The information processing apparatus according to claim 16 ,

wherein the processor is further configured to, if the extracted character string corresponds to the stored character string, notify that the plurality of documents are not in proper order.

19. The information processing apparatus according to claim 1 ,

wherein the processor is further configured to, in response to extracting the character string from the read image, resume extraction from a page immediately following a page containing a last extracted character string.

20. A non-transitory computer readable medium storing a program causing a computer to execute a process for processing information, the process comprising:

acquiring a read image and item information, the read image being an image obtained by reading a paper medium including a plurality of documents, the item information being information of an item specified by a user from among a plurality of items contained in the plurality of documents;

extracting a character string from the read image, the character string being associated with the item information;

if a character string contained in a page of the read image and extracted from the page differs from a character string extracted from a previous page immediately preceding the page, setting a split position, the split position being a position at which to split out a portion of the read image as a set of documents, the portion being a portion of the read image from a page where extraction has begun to the previous page; and

outputting the read image split in accordance with the split position.

Assignments (2)
CHANGE OF NAME Recorded May 14, 2021
From: FUJI XEROX CO., LTD.
To: FUJIFILM BUSINESS INNOVATION CORP.
Reel/Frame 056237/0119 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 5, 2020
From: YAMAMOTO, TAKUMA; KUWANO, AYA; SATO, MITSURU; TAKAHASHI, TORU
To: FUJI XEROX CO., LTD.
Reel/Frame 053968/0449 →
Priority Claims (1)
JP JP2020-030965 · Feb 26, 2020 · national
Continuity (1)
Related Publication 20210266416A1 · Aug 26, 2021