IP Library › Granted Patent US 12,340,607
Granted Patent B2
US 12,340,607 · App. 17/958,262 · Granted Jun 24, 2025

Method and apparatus for form identification and registration

Inventor: Junchao Wei (San Mateo, CA)
Assignee: KONICA MINOLTA BUSINESS SOLUTIONS U.S.A., INC.
G06V30/19147G06V30/1916G06V30/412
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,340,607
App. No.
17/958,262
Granted
Jun 24, 2025
Kind
B2
Abstract

Aspects of the present invention relate to a machine learning system that is trained to identify forms, performing a method that includes receiving a form as an input image; identifying a field in the input image; identifying boundaries of the field; identifying locations of characters in the field; creating a two-dimensional space containing special characters; replacing the special characters with the characters in the field; identifying one or more keywords in the field based on identification of words and/or location of words; and responsive to an indication that the identifying one or more keywords yielded an incorrect result, updating the machine learning system. In another aspect, the machine learning system is used to identify forms, and can identify whether a form requires registration and, if registration is required, performing the registration.

Claims (32)

1. A computer-implemented method of training a machine learning system to identify forms in document and form analysis, the method comprising:

a) receiving a form as an input image;

b) identifying a field in the input image;

c) identifying boundaries of the field;

d) identifying locations of characters in the field;

e) creating a two-dimensional space as a blown up representation of the field, the two-dimensional space containing special characters, the special characters being different from the characters in the field, the special characters being inserted in pixel locations so as to preserve the locations of the characters in the field;

f) replacing the special characters in the two-dimensional space with the characters in the field, the characters being positioned in locations in the two-dimensional space corresponding to the locations of the characters in the field, each character being related to a pixel location;

g) identifying one or more keywords in the field based on identification of words and/or location of words, wherein a) to g) enable form matching for the document and form analysis; and

h) responsive to an indication that the identifying one or more keywords yielded an incorrect result, updating the machine learning system.

2. The method of claim 1 , wherein updating the machine learning system comprises updating weights of nodes in the machine learning system.

3. The method of claim 1 , wherein the input image comprises a synthetic form or an image of a scanned form.

4. The method of claim 1 , wherein the identifying one or more keywords comprises reading the characters in the input image.

5. The method of claim 1 , wherein the identifying one or more keywords comprises using relative locations of characters in the 2D space to determine the one or more keywords.

6. The method of claim 1 , further comprising, responsive to a determination that there are more forms to be processed, receiving the next form and repeating the method.

7. A machine learning system to identify forms in document and form analysis, the machine learning system comprising at least one processor and a non-transitory memory that is programmed for the machine learning system to perform a method of document and form analysis comprising:

a) receiving a form as an input image;

b) identifying a field in the input image;

c) identifying boundaries of the field;

d) identifying locations of characters in the field;

e) creating a two-dimensional space as a blown up representation of the field, the two-dimensional space containing special characters, the special characters being different from the characters in the field, the special characters being inserted in pixel locations so as to preserve the locations of the characters in the field;

f) replacing the special characters in the two-dimensional space with the characters in the field, the characters being positioned in locations in the two-dimensional space corresponding to the locations of the characters in the field, each character being related to a pixel location;

g) identifying one or more keywords in the field based on identification of words and/or location of words, wherein a) to g) enable form matching for the document and form analysis; and

h) responsive to an indication that the identifying one or more keywords yielded an incorrect result, updating the machine learning system.

8. The system of claim 7 , wherein updating the machine learning system comprises updating weights of nodes in the machine learning system.

9. The system of claim 7 , wherein the input image comprises a synthetic form or an image of a scanned form.

10. The system of claim 7 , wherein the identifying one or more keywords comprises reading the characters in the input image.

11. The system of claim 7 , wherein the identifying one or more keywords comprises using relative locations of characters in the 2D space to determine the one or more keywords.

12. The system of claim 7 , further comprising:

responsive to the identifying the one or more keywords, determining a type of the form.

13. The system of claim 12 , further comprising, responsive to the responsive to the determining the type of form, determining whether the form requires registration.

14. The system of claim 7 , further comprising, responsive to a determination that there are more forms to be processed, receiving the next form and repeating the method.

15. The system of claim 13 , wherein, responsive to the determination that the form requires registration, performing registration on the form.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 3, 2022
From: WEI, JUNCHAO
To: KONICA MINOLTA BUSINESS SOLUTIONS U.S.A., INC.
Reel/Frame 061292/0587 →
Continuity (1)
Related Publication 20240112483A1 · Apr 4, 2024
References Cited (24)
US 4392197A · Couper · 1983 [cited by examiner]
US 6002797A · Mori · 1999 [cited by examiner]
US 6751779B1 · Kurosawa · 2004 [cited by examiner]
US 20030085910A1 · Noble · 2003 [cited by examiner]
US 20050201620A1 · Kanamoto · 2005 [cited by examiner]
US 20050278378A1 · Frank · 2005 [cited by examiner]
US 20060271887A1 · Bier · 2006 [cited by examiner]
US 20080188280A1 · Marks · 2008 [cited by examiner]
US 20100275113A1 · Bastos dos Santos · 2010 [cited by examiner]
US 20120136646A1 · Kraenzel · 2012 [cited by examiner]
US 20140207479A1 · Noland · 2014 [cited by examiner]
US 20170308557A1 · Cassidy · 2017 [cited by examiner]
US 20180113858A1 · Peng · 2018 [cited by examiner]
US 20200285558A1 · Kalia · 2020 [cited by examiner]
US 20220318492A1 · Gohari · 2022 [cited by examiner]
US 20230114965A1 · Ferreira · 2023 [cited by examiner]
Ousirimaneechai, Nattapong, and Sukree Sinthupinyo. “Extraction of trend keywords and stop words from thai facebook pages using character n-grams.” International Journal of Machine Learning and Computing 8.6 (2018): 589… [cited by examiner]
Jiang, Jing, and ChengXiang Zhai. “An empirical study of tokenization strategies for biomedical information retrieval.” Information Retrieval 10 (2007): 341-363. (Year: 2007). [cited by examiner]
Shreda, Qais A., and Abualsoud A. Hanani. “Identifying non-functional requirements from unconstrained documents using natural language processing and machine learning approaches.” IEEE Access (2021). (Year: 2021). [cited by examiner]
Sharif O, Hoque M M, Kayes AS, Nowrozy R, Sarker IH. Detecting suspicious texts using machine learning techniques. Applied Sciences. Sep. 18, 2020;10(18):6527. (Year: 2020). [cited by examiner]
Kumar, Rajesh, et al. “A Supervised Method to Find the Relevance of Extracted Keywords Using Deep Learning Approaches.” Emerging Technologies in Data Mining and Information Security: Proceedings of IEMIS 2018, vol. 3. S… [cited by examiner]
Kadhim, Ammar Ismael, Yu-N. Cheah, and Nurul Hashimah Ahamed. “Text document preprocessing and dimension reduction techniques for text document clustering.” 2014 4th international conference on artificial intelligence w… [cited by examiner]
Marmolejos, Licelot, et al. “On the use of textual feature extraction techniques to support the automated detection of refactoring documentation.” Innovations in Systems and Software Engineering (2022): 1-17. (Year: 202… [cited by examiner]
Polato, Mirko, et al. “Efficient Multilingual Deep Learning Model for Keyword Categorization.” 2021 IEEE Symposium Series on Computational Intelligence (SSCI). IEEE, 2021. (Year: 2021). [cited by examiner]