IP Library Granted Patent US 12,373,083
Granted Patent B2
US 12,373,083 · App. 17/169,825 · Granted Jul 29, 2025

Machine learning based extraction of partition objects from electronic documents

Inventors: Dan G. Tecuci (Austin, TX); Ravi Kiran Reddy Palla (Sunnyvale, CA); Hamid Reza Motahari Nezhad (Los Altos, CA); Vincent Poon (Millbrae, CA); Nigel Paul Duffy (San Francisco, CA); Joseph Nipko (Austin, TX)
G06F3/0482G06F18/2148G06F18/2178G06N3/02G06N20/20G06V10/22G06V10/771G06V10/82G06V10/987G06V30/19167G06V30/19173G06V30/40G06V30/43
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,373,083
App. No.
17/169,825
Granted
Jul 29, 2025
Kind
B2
Abstract

An object-extraction method includes generating multiple partition objects based on an electronic document, and receiving a first user selection of a data element via a user interface of a compute device. In response to the first user selection, and using a machine learning model, a first subset of partition objects from the multiple partition objects is detected and displayed via the user interface. A user interaction, via the user interface, with one of the partition objects is detected, and in response, a weight of the machine learning model is modified, to produce a modified machine learning model. A second user selection of the data element is received via the user interface, and in response and using the modified machine learning model, a second subset of partition objects from the multiple partition objects is detected and displayed via the user interface, the second subset different from the first subset.

Claims (53)

1. A computer-implemented method, comprising:

generating a dataset including a plurality of value pairs, each value pair from the plurality of value pairs including an error-free value and an error-containing value, at least one error-free value included in at least one value pair from the plurality of value pairs modified to generate at least one error-containing value included in the at least one value pair from the plurality of value pairs;

training a machine learning model with the dataset to produce a trained machine learning model, training the machine learning model including

applying a semantic model included in the machine learning model and not a similarity model included in the machine learning model,

applying, after applying the semantic model and not the similarity model, the similarity model and the semantic model, and

in response to applying the semantic model and the similarity model, applying the similarity model and a multi-layer perceptron model included in the machine learning model without applying the semantic model;

detecting, via the trained machine learning model, an error in an electronically-stored file;

in response to detecting the error, converting the electronically-stored file, via the trained machine learning model, into a modified electronically-stored file that does not include the error;

identifying a set of objects associated with the modified electronically-stored file; and

in response to a user interaction with a representation of an object from the set of objects, updating the trained machine learning model based on the user interaction, to have more weight for at least one of retrieval or ranking, and to have less weight for semantic searching.

2. The computer-implemented method of claim 1 , wherein the error is an optical character recognition (OCR) induced error.

3. The computer-implemented method of claim 1 , wherein the machine learning model includes at least one of a sequence-to-sequence model or an attention algorithm.

4. The computer-implemented method of claim 1 , wherein the converting the electronically-stored file into the modified electronically-stored file is based on a value pair (1) from the plurality of value pairs and (2) associated with the detected error.

5. The computer-implemented method of claim 4 , wherein the converting the electronically-stored file into the modified electronically-stored file includes replacing an erroneous data segment of the electronically-stored file with the error-free value of the value pair from the plurality of value pairs, the erroneous data segment associated with the detected error.

6. The computer-implemented method of claim 1 , wherein the machine learning model operates within an artificial neural network.

7. The computer-implemented method of claim 1 , wherein the detecting the error in the electronically-stored file includes:

receiving, from a named-entity recognition (NER) model, a signal representing a non-extracted data string that is associated with the electronically-stored file; and

inspecting the non-extracted data string to identify the error.

8. The computer-implemented method of claim 1 , wherein each of the electronically-stored file and the modified electronically-stored file represents an electronic document.

9. An apparatus, comprising:

a processor; and

a memory storing instructions, executable by the processor, to:

generate a dataset including a plurality of value pairs, each value pair from the plurality of value pairs including an error-free value and an error-containing value, at least one error-free value included in at least one value pair from the plurality of value pairs modified to generate at least one error-containing value included in the at least one value pair from the plurality of value pairs;

train a machine learning model with the dataset to produce a trained machine learning model, training the machine learning model including

applying a semantic model included in the machine learning model and not a similarity model included in the machine learning model,

applying, after applying the semantic model and not the similarity model, the similarity model and the semantic model, and

in response to applying the semantic model and the similarity model, applying the similarity model and a multi-layer perceptron model included in the machine learning model without applying the semantic model;

detect, via the trained machine learning model, an error in an electronically-stored file;

in response to detecting the error, convert the electronically-stored file into a modified electronically-stored file that does not include the error;

identify a set of objects associated with the modified electronically-stored file; and

in response to a user interaction with a representation of an object from the set of objects, update the trained machine learning model based on the user interaction, to have more weight for at least one of retrieval or ranking, and to have less weight for semantic searching.

10. The apparatus of claim 9 , wherein the error is an optical character recognition (OCR) induced error.

11. The apparatus of claim 9 , wherein the machine learning model includes at least one of a sequence-to-sequence model or an attention algorithm.

12. The apparatus of claim 9 , wherein the instructions to convert the electronically-stored file into the modified electronically-stored file that does not include the error include instructions to convert the electronically-stored file into the modified electronically-stored file based on a value pair ( 1 ) from the plurality of value pairs and ( 2 ) associated with the detected error.

13. The apparatus of claim 9 , wherein the instructions to convert the electronically-stored file into the modified electronically-stored file that does not include the error include instructions to replace a data segment of the electronically-stored file with the error-free value of the value pair from the plurality of value pairs, the data segment associated with the detected error.

14. The apparatus of claim 9 , wherein the machine learning model operates within an artificial neural network.

15. The apparatus of claim 9 , wherein the instructions to detect the error in the electronically-stored file include instructions to:

receive, from a named-entity recognition (NER) model, a signal representing a non-extracted data string that is associated with the electronically-stored file; and

inspect the non-extracted data string to identify the error.

16. The apparatus of claim 9 , wherein each of the electronically-stored file and the modified electronically-stored file represents an electronic document.

17. A non-transitory, processor-readable medium storing instructions to cause a processor to:

generate a plurality of value pairs, each value pair from the plurality of value pairs including an error-free value and an error-containing value, at least one error-free value included in at least one value pair from the plurality of value pairs modified to generate at least one error-containing value included in the at least one value pair from the plurality of value pairs;

train a machine learning model with the plurality of value pairs to produce a trained machine learning model, training the machine learning model including

applying a semantic model included in the machine learning model and not a similarity model included in the machine learning model,

applying, after applying the semantic model and not the similarity model, the similarity model and the semantic model, and

in response to applying the semantic model and the similarity model, applying the similarity model and a multi-layer perceptron model included in the machine learning model without applying the semantic model;

detect, via the trained machine learning model, an error in an electronically-stored file;

in response to detecting the error, convert the electronically-stored file, via the trained machine learning model, into a modified electronically-stored file that does not include the error;

identify a set of objects associated with the modified electronically-stored file; and

in response to a user interaction with a representation of an object from the set of objects, update the trained machine learning model based on the user interaction, to have more weight for at least one of retrieval or ranking, and to have less weight for semantic searching.

18. The non-transitory, processor-readable medium of claim 17 , wherein the error is an optical character recognition (OCR) induced error.

19. The non-transitory, processor-readable medium of claim 17 , wherein the machine learning model includes at least one of a sequence-to-sequence model or an attention algorithm.

20. The non-transitory, processor-readable medium of claim 17 , wherein the instructions to convert the electronically-stored file into the modified electronically-stored file that does not include the error include instructions to replace a data segment of the electronically-stored file with the error-free value of the value pair from the plurality of value pairs, the data segment associated with the detected error.

Continuity (3)
Division 16790945 · Feb 14, 2020
Division 16382707 · Apr 12, 2019
Related Publication 20210166074A1 · Jun 3, 2021
References Cited (177)
US 5048107A · Tachikawa · 1991 [cited by applicant]
US 5848186A · Wang et al. · 1998 [cited by applicant]
US 5892843A · Zhou et al. · 1999 [cited by applicant]
US 6006240A · Handley · 1999 [cited by applicant]
US 6735748B1 · Teig et al. · 2004 [cited by applicant]
US 6757870B1 · Stinger · 2004 [cited by applicant]
US 7283683B1 · Nakamura et al. · 2007 [cited by applicant]
US 7548847B2 · Acero et al. · 2009 [cited by applicant]
US 8165974B2 · Privault et al. · 2012 [cited by applicant]
US 8731300B2 · Rodriquez et al. · 2014 [cited by applicant]
US 9058536B1 · Yuan et al. · 2015 [cited by applicant]
US 9172842B2 · Booth et al. · 2015 [cited by applicant]
US 9235812B2 · Scholtes · 2016 [cited by applicant]
US 9269053B2 · Naslund et al. · 2016 [cited by applicant]
US 9342892B2 · Booth et al. · 2016 [cited by applicant]
US 9348815B1 · Estes et al. · 2016 [cited by applicant]
US 9703766B1 · Kyre et al. · 2017 [cited by applicant]
US 9875736B2 · Kim et al. · 2018 [cited by applicant]
US 10002129B1 · D'Souza · 2018 [cited by applicant]
US 10062039B1 · Lockett · 2018 [cited by applicant]
US 10241992B1 · Middendorf et al. · 2019 [cited by applicant]
US 10614345B1 · Tecuci et al. · 2020 [cited by applicant]
US 10810709B1 · Tiyyagura · 2020 [cited by applicant]
US 10956786B2 · Tecuci et al. · 2021 [cited by applicant]
US 11106906B2 · Bassu et al. · 2021 [cited by applicant]
US 11113518B2 · Chua et al. · 2021 [cited by applicant]
US 11625934B2 · Tiyyagura et al. · 2023 [cited by applicant]
US 11715313B2 · Chua et al. · 2023 [cited by applicant]
US 11837005B2 · Tiyyagura et al. · 2023 [cited by applicant]
US 11915465B2 · Gangeh et al. · 2024 [cited by applicant]
US 20030097384A1 · Hu et al. · 2003 [cited by applicant]
US 20060288268A1 · Srinivasan et al. · 2006 [cited by applicant]
US 20070041642A1 · Romanoff et al. · 2007 [cited by applicant]
US 20070050411A1 · Hull et al. · 2007 [cited by applicant]
US 20070106494A1 · Detlef · 2007 [cited by examiner]
US 20100174975A1 · Mansfield et al. · 2010 [cited by applicant]
US 20110249905A1 · Singh et al. · 2011 [cited by applicant]
US 20120072859A1 · Wang et al. · 2012 [cited by applicant]
US 20130191715A1 · Raskovic et al. · 2013 [cited by applicant]
US 20140223284A1 · Rankin, Jr. et al. · 2014 [cited by applicant]
US 20150058374A1 · Golubev et al. · 2015 [cited by applicant]
US 20150093021A1 · Xu et al. · 2015 [cited by applicant]
US 20150356461A1 · Vinyals et al. · 2015 [cited by applicant]
US 20160078364A1 · Chiu et al. · 2016 [cited by applicant]
US 20160104077A1 · Jackson, Jr. et al. · 2016 [cited by applicant]
US 20160162456A1 · Munro et al. · 2016 [cited by applicant]
US 20160350280A1 · Lavallee et al. · 2016 [cited by applicant]
US 20160364608A1 · Sengupta et al. · 2016 [cited by applicant]
US 20170083829A1 · Kang et al. · 2017 [cited by applicant]
US 20170177180A1 · Bachmann et al. · 2017 [cited by applicant]
US 20170235848A1 · Van Dusen et al. · 2017 [cited by applicant]
US 20170300472A1 · Parikh et al. · 2017 [cited by applicant]
US 20170300565A1 · Calapodescu et al. · 2017 [cited by applicant]
US 20180060303A1 · Sarikaya et al. · 2018 [cited by applicant]
US 20180068232A1 · Hari Haran et al. · 2018 [cited by applicant]
US 20180129634A1 · Sivaji et al. · 2018 [cited by applicant]
US 20180157723A1 · Chougule et al. · 2018 [cited by applicant]
US 20180181797A1 · Han et al. · 2018 [cited by applicant]
US 20180203674A1 · Dayanandan · 2018 [cited by applicant]
US 20180204360A1 · Bekas et al. · 2018 [cited by applicant]
US 20180260957A1 · Yang et al. · 2018 [cited by applicant]
US 20180336404A1 · Hosabettu et al. · 2018 [cited by applicant]
US 20180341702A1 · Sawruk · 2018 [cited by examiner]
US 20190049540A1 · Odry et al. · 2019 [cited by applicant]
US 20190050381A1 · Agrawal et al. · 2019 [cited by applicant]
US 20190108448A1 · O'Malia et al. · 2019 [cited by applicant]
US 20190147320A1 · Mattyus et al. · 2019 [cited by applicant]
US 20190171704A1 · Buisson et al. · 2019 [cited by applicant]
US 20190171908A1 · Salavon · 2019 [cited by applicant]
US 20190228495A1 · Tremblay et al. · 2019 [cited by applicant]
US 20190251401A1 · Shechtman et al. · 2019 [cited by applicant]
US 20190266394A1 · Yu et al. · 2019 [cited by applicant]
US 20190303663A1 · Krishnapura et al. · 2019 [cited by applicant]
US 20190340240A1 · Duta · 2019 [cited by applicant]
US 20190370323A1 · Davidson · 2019 [cited by examiner]
US 20200005033A1 · Bellert · 2020 [cited by applicant]
US 20200073878A1 · Mukhopadhyay et al. · 2020 [cited by applicant]
US 20200089946A1 · Mallick et al. · 2020 [cited by applicant]
US 20200151444A1 · Price et al. · 2020 [cited by applicant]
US 20200151559A1 · Karras et al. · 2020 [cited by applicant]
US 20200175267A1 · Schäfer et al. · 2020 [cited by applicant]
US 20200175304A1 · Vig · 2020 [cited by examiner]
US 20200250139A1 · Muffat et al. · 2020 [cited by applicant]
US 20200250513A1 · Krishnamoorthy · 2020 [cited by applicant]
US 20200265224A1 · Gurav et al. · 2020 [cited by applicant]
US 20200327373A1 · Tecuci et al. · 2020 [cited by applicant]
US 20200410231A1 · Chua et al. · 2020 [cited by applicant]
US 20210056429A1 · Gangeh et al. · 2021 [cited by applicant]
US 20210064861A1 · Semenov · 2021 [cited by applicant]
US 20210064908A1 · Semenov · 2021 [cited by applicant]
US 20210073325A1 · Angst et al. · 2021 [cited by applicant]
US 20210073326A1 · Aggarwal et al. · 2021 [cited by applicant]
US 20210117668A1 · Zhong et al. · 2021 [cited by applicant]
US 20210150338A1 · Semenov · 2021 [cited by applicant]
US 20210150757A1 · Mustikovela et al. · 2021 [cited by applicant]
US 20210165938A1 · Bailey et al. · 2021 [cited by applicant]
US 20210166016A1 · Sun et al. · 2021 [cited by applicant]
US 20210233656A1 · Tran et al. · 2021 [cited by applicant]
US 20210240976A1 · Tiyyagura et al. · 2021 [cited by applicant]
US 20210365678A1 · Chua et al. · 2021 [cited by applicant]
US 20220027740A1 · Dong et al. · 2022 [cited by applicant]
US 20220058839A1 · Chang et al. · 2022 [cited by applicant]
US 20220148242A1 · Russell et al. · 2022 [cited by applicant]
US 20230237828A1 · Tiyyagura et al. · 2023 [cited by applicant]
US 20240161449A1 · Gangeh et al. · 2024 [cited by applicant]
AU 2018237196A1 · 2019 [cited by applicant]
EP 2154631A2 · 2010 [cited by applicant]
EP 3742346A2 · 2020 [cited by applicant]
WO WO2017163230A1 · 2017 [cited by applicant]
WO WO2018175686A1 · 2018 [cited by applicant]
WO WO2020210791A1 · 2020 [cited by applicant]
WO WO2020264155A1 · 2020 [cited by applicant]
WO WO2021034841A1 · 2021 [cited by applicant]
WO WO2021099006A1 · 2021 [cited by applicant]
WO WO2021156322A1 · 2021 [cited by applicant]
Office Action for U.S. Appl. No. 16/382,707, mailed Sep. 4, 2019, 11 pages. [cited by applicant]
International Search Report and Written Opinion for International Application No. PCT/US2020/027916, mailed Jul. 21, 2020, 16 pages. [cited by applicant]
Office Action for U.S. Appl. No. 16/790,945, mailed Jul. 29, 2020, 13 pages. [cited by applicant]
International Search Report and Written Opinion for International Application No. PCT/US2020/046820, mailed Nov. 11, 2020, 13 pages. [cited by applicant]
Office Action for U.S. Appl. No. 16/456,832, mailed Nov. 6, 2020, 18 pages. [cited by applicant]
International Search Report and Written Opinion for International Application No. PCT/US2020/039611, mailed Oct. 13, 2020, 12 pages. [cited by applicant]
Babatunde, F. F. et al., “Automatic Table Recognition and Extraction from Heterogeneous Documents,” Journal of Computer and Communications, vol. 3, pp. 100-110 (Dec. 2015). [cited by applicant]
International Search Report and Written Opinion for International Application No. PCT/EP2020/076042, mailed Jan. 12, 2021, 11 pages. [cited by applicant]
Dong, C. et al., “Image Super-Resolution Using Deep Convolutional Networks,” arXiv:1501.00092v3 [cs.CV], Jul. 31, 2015, Retrieved from the Internet: <URL: https://arxiv.org/pdf/1501.00092.pdf>, 14 pages. [cited by applicant]
Dong, R. et al., “Multi-input attention for unsupervised OCR correction,” Proceedings of the 56th Annual Meetings of the Association for Computational Linguistics (Long Papers), Melbourne, Australia, Jul. 15-20, 2018, p… [cited by applicant]
Kharb, L. et al., “Embedding Intelligence through Cognitive Services,” International Journal for Research in Applied Science & Engineering Technology (IJRASET), ISSN: 2321-9653; IC Value: 45.98; SJ Impact Factor:6.887, … [cited by applicant]
Paladines, J. et al., “An Intelligent Tutoring System for Procedural Training with Natural Language Interaction,” Conference Paper, DOI: 10.5220/0007712203070314, Jan. 2019, 9 pages. [cited by applicant]
Howard, J. et al., “Universal Language Model Fine-tuning for Text Classification,” arXiv:1801.06146v5 [cs.CL], May 23, 2018, Retrieved from the Internet: <URL: https://arxiv.org/pdf/1801.06146.pdf>, 12 pages. [cited by applicant]
Mac, A. J. et al., “Locating tables in scanned documents for reconstructing and republishing,” arXiv:1412.7689 [cs.CV], Dec. 2014, The 7th International Conference on Information and Automation for Sustainability (ICIAf… [cited by applicant]
Kasar, T. et al., “Learning to Detect Tables in Scanned Document Images Using Line Information,” ICDAR '13: Proceedings of the 2013 12th International Conference on Document Analysis and Recognition, Aug. 2013, Washingt… [cited by applicant]
Isola, P. et al., “Image-to-Image Translation with Conditional Adversarial Networks,” arXiv:1611.07004v3 [cs.CV] Nov. 26, 2018, Retrieved from the Internet: <URL: https://arxiv.org/pdf/1611.07004.pdf>, 17 pages. [cited by applicant]
Klampfl, S. et al., “A Comparison of Two Unsupervised Table Recognition Methods from Digital Scientific Articles,” D-Lib Magazine, vol. 20, No. 11/12, Nov./Dec. 2014, DOI:10.1045/november14-klampfl, 15 pages. [cited by applicant]
Fan, M. et al., “Detecting Table Region in PDF Documents Using Distant Supervision,” arXiv:1506.08891v6 [cs.CV], Sep. 22, 2015, Retrieved from the Internet: <URL: https://arxiv.org/pdf/1506.08891v6.pdf>, 7 pages. [cited by applicant]
Oliveira, H. et al., “Assessing shallow sentence scoring techniques and combinations for single and multi-document summarization,” Expert Systems With Applications, vol. 65 (Dec. 2016) pp. 68-86. [cited by applicant]
Pinto, D. et al., “Table extraction using conditional random fields,” Proceedings of the 26th Annual International ACM SIGIR Conference on Research and Development in Informaion Retrieval (SIGIR '03), ACM, New York, NY,… [cited by applicant]
Hanifah, L. et al., “Table Extraction from Web Pages Using Conditional Random Fields to Extract Toponym Related Data,” Journal of Physics: Conference Series, vol. 801, Issue 1, Article ID 012064, Jan. 2017, 8 pages. [cited by applicant]
Staar, P. W. J. et al., “Corpus conversion service: A machine learning platform to ingest documents at scale,” Applied Data Science Track Paper, KDD 2018, Aug. 19-23, 2018, London, United Kingdom, pp. 774-782. [cited by applicant]
Qasim, S. R. et al., “Rethinking Table Parsing using Graph Neural Networks,” arXiv: 1905.1339lvl [cs.CV]; 2019 International Conference on Document Analysis and Recognition (ICDAR), Sydney, Australia, 2019, pp. 142-147. [cited by applicant]
Handley, J. C., “Table analysis for multi-line cell identification,” Proceedings of SPIE, vol. 4307, Jan. 2001, pp. 34-43. [cited by applicant]
Vincent, P. et al., “Extracting and composing robust features with denoising autoencoders,” in Proceedings of the 25th International Conference on Machine Learning, Helsinki, Finland, 2008, 8 pages. [cited by applicant]
Mao, X-J. et al., “Image Restoration Using Very Deep Convolutional Encoder-Decoder Networks with Symmetric Skip Connections,” 30th Conference on Neural Information Processing Systems (NIPS 2016), Barcelona, Spain, Retri… [cited by applicant]
Mao, X-J. et al., “Image Restoration Using Convolutional Auto-encoders with Symmetric Skip Connections,” arXiv:1606.08921v3 [cs.CV], Aug. 30, 2016, Retrieved from the Internet: <URL: https://arxiv.org/abs/1606.08921>, 1… [cited by applicant]
Lehtinen, J. et al., “Noise2Noise: Learning Image Restoration without Clean Data,” Proceedings of the 35th International Conference on Machine Learning, Stockholm, Sweden, PMLR 80, Jul. 10-15, 2018, Retrieved from the I… [cited by applicant]
Schreiber, S. et al., “DeepDeSRT: Deep Learning for Detection and Structure Recognition of Tables in Document Images,” 2017 14th IAPR International Conference on Document Analysis and Recognition (ICDAR) Nov. 9-15, 2017… [cited by applicant]
Kavasidis, I. et al., “A Saliency-based Convolutional Neural Network for Table and Chart Detection in Digitized Documents,” arXiv.1804.06236v1 [cs.CV], Apr. 17, 2018, Retrieved from the Internet: <URL: https://arxiv.org… [cited by applicant]
Xiang, R., Research Statement, Aug. 2018, 6 pages. [cited by applicant]
Harit, G. et al., “Table Detection in Document Images using Header and Trailer Patterns,” ICVGIP '12, Dec. 16-19, 2012, Mumbai, India, 8 pages. [cited by applicant]
Le Vine, N. et al., “Extracting tables from documents using conditional generative adversarial networks and genetic algorithms,” IJCNN 2019 International Joint Conference on Neural Networks, Budapest, Hungary, Jul. 14-1… [cited by applicant]
Xiao, Y. et al., “Text region extraction in a document image based on the Delaunay tessellation,” Pattern Recognition, vol. 36, No. 3, Mar. 2003, pp. 799-809. [cited by applicant]
Kise, K. et al., “Segmentation of p. Images Using the Area Voronoi Diagram,” Computer Vision and Image Understanding, vol. 70, No. 3, Jun. 1998, pp. 370-382. [cited by applicant]
Eskenazi, S. et al., A comprehensive survey of mostly textual document segmentation algorithms since 2008, Pattern Recognition, vol. 64, Apr. 2017, pp. 1-14. [cited by applicant]
Pellicer, J. P., “Neural networks for document image and text processing,” PHD Thesis, Universitat Politecnica de Valencia, Sep. 2017, 327 pages. [cited by applicant]
Wiraatmaja, C. et al., “The Application of Deep Convolutional Denoising Autoencoder for Optical Character Recognition Preprocessing,” 2017 International Conference on Soft Computing, Intelligent System and Information T… [cited by applicant]
Gangeh, M. J. et al., “Document enhancement system using auto-encoders,” 33rd Conference on Neural Information Processing Systems (NeurIPS 2019), Vancouver, Canada, Nov. 2019, Retrieved from the Internet: <URL:https://o… [cited by applicant]
Hakim, S. M. A. et al., “Handwritten bangla numeral and basic character recognition using deep convolutional neural network,” 2019 International Conference on Electrical, Computer and Communication Engineering (ECCE), I… [cited by applicant]
Non-Final Office Action for U.S. Appl. No. 16/546,938 mailed on Sep. 12, 2022, 9 pages. [cited by applicant]
Paliwal, Shubham Singh, D. Vishwanath, Rohit Rahul, Monika Sharma, and Lovekesh Vig. “Tablenet: Deep learning model for end-to-end table detection and tabular data extraction from scanned document images.” International… [cited by applicant]
Rashid, Sheikh Faisal, Abdullah Akmal, Muhammad Adnan, Ali Adnan Aslam, and Andreas Dengel. “Table recognition in heterogeneous documents using machine learning.” In 2017 14th IAPR International conference on document a… [cited by applicant]
Final Office Action for U.S. Appl. No. 16/546,938 dated Jun. 8, 2023, 12 pages. [cited by applicant]
Non-Final Office Action for U.S. Appl. No. 18/172,461 dated Jun. 20, 2023, 9 pages. [cited by applicant]
Deivalakshmi S., et al., “Detection of table structure and content extraction from scanned documents.” In 2014 International Conference on Communication and Signal Processing, pp. 270-274. IEEE, 2014. (Year: 2014). [cited by applicant]
Notice of Allowance for U.S. Appl. No. 16/781,195, dated Feb. 1, 2023, 10 pages. [cited by applicant]
Sun N., et al., “Faster R-CNN Based Table Detection Combining Corner Locating,” 2019 International Conference on Document Analysis and Recognition (ICDAR), 2019, pp. 1314-1319. [cited by applicant]
International Search Report and Written Opinion for International Application No. PCT/EP2021/052579, mailed May 10, 2021, 15 pages. [cited by applicant]
Li, Y. et al., “A GAN-based Feature Generator for Table Detection,” 2019 International Conference on Document Analysis and Recognition (ICDAR), Conference Paper, IEEE (2019), 6 pages. [cited by applicant]
Office Action for U.S. Appl. No. 16/546,938, mailed May 27, 2022, 11 pages. [cited by applicant]
Office Action for U.S. Appl. No. 16/781,195, mailed Nov. 26, 2021, 26 pages. [cited by applicant]
Office Action for U.S. Appl. No. 16/781,195, mailed Jun. 8, 2022, 29 pages. [cited by applicant]
Ohta, M. et al., “A cell-detection-based table-structure recognition method,” Proceedings of the ACM Symposium on Document Engineering 2019, pp. 1-4. [cited by applicant]
Oro, E. et al., “PDF-TREX: An approach for recognizing and extracting tables from PDF documents,” 2009 10th International Conference on Document Analysis and Recognition, pp. 906-910, IEEE, 2009. [cited by applicant]
Gogar, T. et al., “Deep Neural Networks for Web Page Information Extraction,” Springer International Publishing, In: Iliadis, L., Maglogiannis, I. (eds) Artificial Intelligence Applications and Innovations. AIAI 2016. I… [cited by applicant]
Holecek, M. et al., “Table understanding in structured documents,” 2019 International Conference on Document Analysis and Recognition Workshops (ICDARW), Sep. 2019, pp. 158-164. [cited by applicant]
Non-Final Office Action for U.S. Appl. No. 18/419,946 by Gangeh et al., mailed Dec. 4, 2024; 13 pages. [cited by applicant]
Notice of Allowance for U.S. Appl. No. 16/546,938 by Gangeh et al., mailed Oct. 25, 2023; 13 pages. [cited by applicant]
Notice of Allowance for U.S. Appl. No. 16/546,938 by Gangeh et al., mailed Sep. 7, 2023; 11 pages. [cited by applicant]
Notice of Allowance for U.S. Appl. No. 17/395,201 by Chua et al., mailed Mar. 8, 2023; 8 pages. [cited by applicant]
Zhu, J-Y. et al., “Unpaired Image-to-Image Translation using Cycle-Consistent Adversarial Networks,” 2017 IEEE International Conference on Computer Vision (ICCV), Oct. 2017, pp. 2242-2251. [cited by applicant]