IP Library › Granted Patent US 12,731,424
Granted Patent B2
US 12,731,424 · App. 18/347,983 · Granted Sep 8, 2026

Out of distribution element detection for information extraction

Inventors: Srikant Panda (Bangalore, IN); Amit Agarwal (Bangalore, IN); Gouttham Nambirajan (Bangalore, IN); Kulbhushan Pachauri (Bangalore, IN)
Assignee: Oracle International Corporation
G06V30/19147G06F40/169G06F40/247G06V30/413
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,731,424
App. No.
18/347,983
Filed
Jul 6, 2023
Granted
Sep 8, 2026
Kind
B2
Art Unit
3629
USPC
706/25
Abstract

Techniques for extracting information from unstructured documents that enable an ML model to be trained such that the model can accurately distinguish in-distribution (“in-D”) elements and out-of-distribution (“OO-D”) elements within an unstructured document. Novel training techniques are used that train an ML model using a combination of a regular training dataset and an enhanced augmented training dataset. The regular training dataset is used to train an ML model to identify in-D elements, i.e., to classify an element extracted from a document as belonging to one of the in-D classes contained in the regular training dataset. The augmented training dataset, which is generated based upon the regular training dataset may contain one or more augmented elements which are used to train the model to identify OO-D elements, i.e., to classify an augmented element extracted from a document as belonging to an OO-D class instead of to an in-D class.

Claims (68)

1 . A computer-implemented method comprising:

accessing a first training dataset provided for training a machine learning (ML) model, the first training dataset comprising a first plurality of documents and annotation information for each document in the first plurality of documents, wherein, for each document in the first plurality of documents, the annotation information for the document comprises information indicative of one or more elements in the document, and for each element in the one or more elements, information indicative of an in-distribution (“in-D”) class, from one or more in-D classes, to which the element belongs;

generating a second training dataset based upon the first training dataset, the second training dataset comprising a second plurality of documents and annotation information for each document in the second plurality of documents, wherein, each document in the second plurality of documents includes one or more elements that belong to an out-of-distribution (“OO-D”) class; and

training the ML model using both the first training dataset and the second training dataset to generate a trained machine learning model, wherein, for an element extracted from a particular document, the trained machine learning model is trained to classify the extracted element as belonging to an in-D class or to the OO-D class.

2 . The method of claim 1 , further comprising:

classifying, using the trained ML model, a first element extracted from the particular document as belonging to an in-D class from the one or more in-D classes; and

classifying, using the trained ML model, a second element extracted from the particular document as belonging to the OO-D class.

3 . The method of claim 1 , wherein generating the second training dataset based upon the first training dataset comprises:

identifying a first document in the first plurality of documents;

generating a second document for the second plurality of documents from the first document, wherein the second document comprises a first element; and

generating annotation information for the second document, the annotation information for the second document indicating that the first element belongs to the OO-D class.

4 . The method of claim 3 , wherein generating the second document from the first document comprises:

making a copy of the first document, wherein the second document is the copy of the first document.

5 . The method of claim 3 , wherein the first element is included in the first document.

6 . The method of claim 3 , wherein:

generating the second document from the first document comprises receiving information identifying a region within the second document;

generating the annotation information for the second document comprises including information in the annotation information for the second document indicative that any element located within the region in the second document belongs to the OO-D class; and

the first element is located in the region within the first document and also located in a region in the first document corresponding to the region in the second document.

7 . The method of claim 3 , wherein generating the second document from the first document comprises:

generating the first element;

identifying, based upon the first document, a location within the second document for placing the first element; and

placing the first element in the identified location within the second document.

8 . The method of claim 7 , wherein generating the first element comprises identifying a particular word from a set of words included in the first plurality of documents, wherein the first element is the particular word.

9 . The method of claim 7 , wherein generating the first element comprises identifying a particular word from a set of words included in the second document, wherein the first element is the particular word.

10 . The method of claim 7 , wherein generating the first element comprises:

identifying a particular word from a set of words included in the second 2 document; and

generating a similar word based upon the particular word, wherein the first element is the similar word.

11 . The method of claim 10 , wherein generating the similar word comprises:

generating an embedded representation of the particular word using a word 2 embedding model;

identifying a candidate word using a language model;

determining a similarity measure between the particular word and the candidate word; and

designating the candidate word as the similar word.

12 . The method of claim 7 , wherein generating the first element comprises:

identifying a particular word from a set of words included in the second 2 document; and

determining a synonym of the particular word, wherein the first element is the synonym.

13 . The method of claim 7 , wherein generating the first element comprises:

randomly selecting a word from a corpus of documents, wherein the first element 2 is the randomly selected word.

14 . The method of claim 7 , wherein generating the first element comprises:

identifying a particular word from a set of words included in the second document; and

identifying one or more characteristics associated with the particular word based upon metadata associated with the second document; and

generating a new word based on the one or more identified characteristics, wherein the first element is the new word.

15 . The method of claim 7 , wherein generating the first element comprises modifying a property of the first element, wherein the property includes at least one of: font, color, style, or size.

16 . The method of claim 7 , wherein generating the first element comprises generating an image, wherein the first element is the image.

17 . The method of claim 16 , wherein the image is one of a barcode, QR code, rubber stamp, handwritten text, or a watermark.

18 . A system comprising:

a set of processors;

a memory storing a trained machine learning (“ML”) model, wherein the trained ML model is trained to classify elements in a document as belonging to one of a set of one or more in-distribution (“in-D”) class or to an out-of-distribution (“OO-D”) class, wherein training the trained ML model comprises:

accessing a first training dataset, the first training dataset comprising a first plurality of documents and annotation information for each document in the first plurality of documents, wherein, for each document in the first plurality of documents, the annotation information for the document comprises information indicative of one or more elements in the document, and for each element in the one or more elements, information indicative of an in-D class to which the element belongs;

generating a second training dataset based upon the first training dataset, the second training dataset comprising a second plurality of documents and annotation information for each document in the second plurality of documents, wherein, each document in the second plurality of documents includes one or more elements that belong to the OO-D class; and

training the ML model using both the first training dataset and the second training dataset;

wherein one or more processors from the set of processors are configured to perform processing comprising:

classifying, using the trained ML model, a first element extracted from a document as belonging to an in-D class from the set of one or more in-D classes; and

classifying, using the trained ML model, a second extracted element from the document as belonging to the OO-D class.

19 . The system of claim 18 , wherein generating the second training dataset based upon the first training dataset comprises:

identifying a first document in the first plurality of documents;

generating a second document for the second plurality of documents from the first document, wherein the second document comprises a first element, comprising:

generating the first element;

identifying, based upon the first document, a location within the second document for placing the first element; and

placing the first element in the identified location within the second document; and

generating annotation information for the second document, the annotation information for the second document indicating that the first element belongs to the OO-D class.

20 . A non-transitory computer-readable medium storing computer-executable instructions that, when executed by one or more computer devices, cause the computing devices to perform processing comprising:

accessing a trained machine learning (“ML”) model, wherein the trained ML model is trained to classify elements in a document as belonging to one of a set of one or more in-distribution (“in-D”) class or to an out-of-distribution (“OO-D”) class, wherein training the trained ML model comprises:

accessing a first training dataset, the first training dataset comprising a first plurality of documents and annotation information for each document in the first plurality of documents, wherein, for each document in the first plurality of documents, the annotation information for the document comprises information indicative of one or more elements in the document, and for each element in the one or more elements, information indicative of an in-D class to which the element belongs;

generating a second training dataset based upon the first training dataset, the second training dataset comprising a second plurality of documents and annotation information for each document in the second plurality of documents, wherein, each document in the second plurality of documents includes one or more elements that belong to the OO-D class; and

training the ML model using both the first training dataset and the second training dataset;

wherein one or more processors from the set of processors are configured to perform processing comprising:

classifying, using the trained ML model, a first element extracted from a document as belonging to an in-D class from the set of one or more in-D classes; and

classifying, using the trained ML model, a second extracted element from the document as belonging to the OO-D class.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jul 7, 2023
From: PANDA, SRIKANT; AGARWAL, AMIT; NAMBIRAJAN, GOUTTHAM; PACHAURI, KULBHUSHAN
To: ORACLE INTERNATIONAL CORPORATION
Reel/Frame 064178/0306 →
Continuity (1)
Related Publication 20250014374A1 · Jan 9, 2025
References Cited (112)
US 5912989A · Watanabe · 1999 [cited by applicant]
US 6061652A · Tsuboka et al. · 2000 [cited by applicant]
US 9424668B1 · Petrou et al. · 2016 [cited by applicant]
US 11003959B1 · Levner et al. · 2021 [cited by applicant]
US 11087081B1 · Srivastava et al. · 2021 [cited by applicant]
US 11341367B1 · Barbosa et al. · 2022 [cited by applicant]
US 11430467B1 · Vasudevan et al. · 2022 [cited by applicant]
US 11763092B2 · Duong et al. · 2023 [cited by applicant]
US 11989964B2 · Agarwal et al. · 2024 [cited by applicant]
US 12106595B2 · Agarwal et al. · 2024 [cited by applicant]
US 12182498B1 · Sunkara et al. · 2024 [cited by applicant]
US 20020016798A1 · Sakai et al. · 2002 [cited by applicant]
US 20120062574A1 · Dhoolia et al. · 2012 [cited by applicant]
US 20160103833A1 · Sanders et al. · 2016 [cited by applicant]
US 20160364608A1 · Sengupta et al. · 2016 [cited by applicant]
US 20170351965A1 · Kurniadi et al. · 2017 [cited by applicant]
US 20180114142A1 · Mueller · 2018 [cited by examiner]
US 20190147853A1 · Gunasekara et al. · 2019 [cited by applicant]
US 20200005118A1 · Chen et al. · 2020 [cited by applicant]
US 20200104650A1 · Huang · 2020 [cited by applicant]
US 20200125954A1 · Truong et al. · 2020 [cited by applicant]
US 20200179808A1 · Lee et al. · 2020 [cited by applicant]
US 20200285702A1 · Padhi et al. · 2020 [cited by applicant]
US 20200320053A1 · He et al. · 2020 [cited by applicant]
US 20200380623A1 · Ranjan et al. · 2020 [cited by applicant]
US 20200410231A1 · Chua et al. · 2020 [cited by applicant]
US 20210089587A1 · Gupta et al. · 2021 [cited by applicant]
US 20210133645A1 · Tazi et al. · 2021 [cited by applicant]
US 20210158093A1 · Kaynig-Fittkau et al. · 2021 [cited by applicant]
US 20210248323A1 · Maheshwari et al. · 2021 [cited by applicant]
US 20220092267A1 · Hou et al. · 2022 [cited by applicant]
US 20220156300A1 · Paruchuri · 2022 [cited by examiner]
US 20220171938A1 · Jalaluddin et al. · 2022 [cited by applicant]
US 20220405682A1 · Yoon et al. · 2022 [cited by applicant]
US 20230040084A1 · Cherukara et al. · 2023 [cited by applicant]
US 20230146501A1 · Agarwal et al. · 2023 [cited by applicant]
US 20230153335A1 · McNeill · 2023 [cited by applicant]
US 20230252234A1 · Hoang · 2023 [cited by examiner]
US 20230326224A1 · Agarwal et al. · 2023 [cited by applicant]
US 20230394235A1 · Rahman et al. · 2023 [cited by applicant]
US 20240169272A1 · MacWilliams · 2024 [cited by examiner]
US 20240338813A1 · Kamen · 2024 [cited by examiner]
US 20250157242A1 · Ramaswamy · 2025 [cited by examiner]
US 20250378048A1 · Bhat · 2025 [cited by examiner]
CN 107977345A · 2018 [cited by applicant]
CN 113936340A · 2022 [cited by applicant]
CN 114491010A · 2022 [cited by applicant]
WO 2022078922A1 · 2022 [cited by applicant]
W. Cho, J. Park and J. Choo, “Training Auxiliary Prototypical Classifiers for Explainable Anomaly Detection in Medical Image Segmentation,” 2023 IEEE/CVF Winter Conference on Applications of Computer Vision (WACV), Waik… [cited by examiner]
Mixture Outlier Exposure: Towards Out-of-Distribution Detection in Fine-grained Environments Jingyang Zhang†, Nathan Inkawhich, Randolph Linderman†, Yiran Chen†, Hai Li† †Duke University, Air Force Research Laboratory. … [cited by examiner]
Towards In-distribution Compatibility in Out-of-distribution Detection; Boxi Wu, Jie Jiang, Haidong Ren, Zifan Du, Wenxiao Wang, Zhifeng Li, Deng Cai, Xiaofei He, Binbin Lin, Wei Liu. arXiv:2208.13433 [cs.CV]. [v1] Mon,… [cited by examiner]
U.S. Appl. No. 18/240,480, “Notice of Allowance”, Dec. 30, 2025, 15 pages. [cited by applicant]
U.S. Appl. No. 18/379,091, Non-Final Office Action, Mailed on Jun. 6, 2024, 19 pages. [cited by applicant]
U.S. Appl. No. 18/379,091, Notice of Allowance, Mailed on Jul. 29, 2024, 7 pages. [cited by applicant]
International Application No. PCT/US2024/016876, International Search Report and Written Opinion, Mailed on Jun. 5, 2024, 16 pages. [cited by applicant]
Tang et al., “MatchVIE: Exploiting Match Relevancy Between Entities for Visual Information Extraction”, Available online at: https://arxiv.org/pdf/2106.12940, Jun. 24, 2021, 7 pages. [cited by applicant]
Wei et al., “Robust Layout-Aware IE for Visually Rich Documents with Pre-Trained Language Models”, Cornell University Library, Available online at: https://arxiv.org/pdf/2005.11017, May 22, 2020, 10 pages. [cited by applicant]
Xu et al., “LayoutLMv2: Multi-Modal Pre-Training for Visually-Rich Document Understanding”, Available Online at: https://arxiv.org/pdf/2012.14740v4, Jan. 10, 2022, 13 pages. [cited by applicant]
U.S. Appl. No. 17/217,909, Non-Final Office Action mailed on Apr. 18, 2023, 15 pages. [cited by applicant]
U.S. Appl. No. 17/217,909, Notice of Allowance mailed on Jun. 8, 2023, 27 pages. [cited by applicant]
U.S. Appl. No. 17/714,806, Notice of Allowance mailed on Jul. 26, 2023, 7 pages. [cited by applicant]
Kim et al., “Joint Learning of Domain Classification and Out-of-Domain Detection with Dynamic Class Weighting for Satisficing False Acceptance Rates”, Available online at: https://arxiv.org/pdf/1807.00072.pdf, Jun. 29, … [cited by applicant]
Lane et al., “Out-of-Domain Utterance Detection Using Classification Confidences of Multiple Topics”, Institute of Electrical and Electronics Engineers, Transactions on Audio, Speech, and Language Processing, vol. 15, N… [cited by applicant]
International Application No. PCT/US2021/024917, “International Preliminary Report on Patentability”, Oct. 13, 2022, 9 pages. [cited by applicant]
International Application No. PCT/US2021/024917, “International Search Report and Written Opinion”, Jul. 12, 2021, 13 pages. [cited by applicant]
U.S. Appl. No. 17/524,157 , Notice of Allowance, Mailed on Feb. 28, 2024, 24 pages. [cited by applicant]
U.S. Appl. No. 18/240,480, Non-Final Office Action, Mailed on Jun. 5, 2025, 12 pages. [cited by applicant]
Augmentation Pipeline for Rendering Synthetic Paper Printing, Faxing, Scanning and Copy Machine Processes, Available online at: https://github.com/sparkfish/augraphy, Accessed from Internet Apr. 4, 2022, 13 pages. [cited by applicant]
Bert, Available Online at: https://huggingface.co/docs/transformers/model_doc/bert, Accessed from Internet on Mar. 2, 2022, 114 pages. [cited by applicant]
DataGen—CeDar, Centre for Applied Data Analytics Research, Available online at: https://old.ceadar.ie/wp-content/uploads/CeADAR_Flyer_DataGen_v2.pdf, 1 page. [cited by applicant]
Datagen Synthetic Image Datasets for Computer Vision, Available online at: https://datagen.tech/, Accessed from Internet Jun. 1, 2022, 6 pages. [cited by applicant]
Deterministic Algorithm, Wikipedia, Available Online at: https://en.wikipedia.org/wiki/Deterministic_algorithm, Accessed from Internet on Mar. 2, 2022, 4 pages. [cited by applicant]
DistilBERT, Available Online at: https://huggingface.co/docs/transformers/model_doc/distilbert, Accessed from Internet on Mar. 2, 2022, 62 pages. [cited by applicant]
LayoutLMFT, Available online at https://github.com/microsoft/unilm/tree/master/layoutlmft, Accessed from Internet on Aug. 19, 2021, 2 pages. [cited by applicant]
Public Leader for SROIE, Available online at https://rrc.cvc.uab.es/?ch=13&com=evaluation&task=3, Accessed from Internet on: Aug. 19, 2021, 5 pages. [cited by applicant]
Scipy.Optimize.Linear_Sum_Assignment, Available Online at: https://docs.scipy.org/doc/scipy-0.18.1/reference/generated/scipy.optimize.linear_sum_assignment.html, Sep. 19, 2016, 2 pages. [cited by applicant]
Sklearn.Decomposition.PCA, Available Online at: https://scikit-learn.org/stable/modules/generated/sklearn.decomposition.PCA.html, Accessed from Internet on Mar. 2, 2022, 6 pages. [cited by applicant]
Spaczz: Fuzzy Matching and More for Spacy, Available Online at: https://github.com/gandersen101/spaczz, Accessed from Internet on Mar. 2, 2022, 20 pages. [cited by applicant]
Text Distance, Available Online at: https://github.com/life4/textdistance, Accessed from Internet on Mar. 2, 2022, 9 pages. [cited by applicant]
Welcome to Albumentations Documentation, Available online at: https://albumentations.ai/docs/, Accessed from Internet Apr. 4, 2022, 3 pages. [cited by applicant]
WordNet: A Lexical Database for English, Princeton University, Available Online at: https://wordnet.princeton.edu/, Accessed from Internet on Mar. 2, 2022, 4 pages. [cited by applicant]
U.S. Appl. No. 17/714,806, Notice of Allowance mailed on Jun. 22, 2023, 14 pages. [cited by applicant]
Abdelzad et al., Detecting Out-of-Distribution Inputs in Deep Neural Networks Using an Early-Layer Output, Available online at: https://arxiv.org/pdf/1910.10307.pdf, Oct. 23, 2019, 15 pages. [cited by applicant]
Ba, Meta-data Driven Key-Value Pairs Extraction with Azure Form Recognizer, Available online at: https://techcommunity.microsoft.com/t5/ai-cognitive-services-blog/meta-data-driven-key-value-pairs-extraction-with-azure-f… [cited by applicant]
Biswas et al., DocSynth: A Layout Guided Approach for Controllable Document Image Synthesis, Available online at: https://arxiv.org/pdf/2107.02638.pdf, Sep. 2021, 15 pages. [cited by applicant]
Brems, A One-Stop Shop for Principal Component Analysis, Towards Data Science, Available Online at: https://towardsdatascience.com/a-one-stop-shop-for-principal-component-analysis-5582fb7e0a9c, Apr. 18, 2017, 14 pages. [cited by applicant]
Chogovadze et al., Controllable Data Augmentation Through Deep Relighting, Available online at: https://arxiv.org/pdf/2110.13996.pdf, Oct. 26, 2021, 15 pages. [cited by applicant]
Delalandre et al., Generation of Synthetic Documents for Performance Evaluation of Symbol Recognition & Spotting Systems, International Journal on Document Analysis and Recognition, vol. 13, No. 3, Sep. 2010, pp. 187-20… [cited by applicant]
Devlin et al., BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding, Available Online at: https://arxiv.org/pdf/1810.04805.pdf, May 24, 2019, 16 pages. [cited by applicant]
Gautam, Form Data Augmentation, Available online at: https://github.com/gautam-aayush/form-data-augmentation, Mar. 19, 2021, 11 pages. [cited by applicant]
Ghosh, Invoice Information Extraction Using OCR and Deep Learning, Available online at: https://medium.com/analytics-vidhya/invoice-information-extraction-using-ocr-and-deep-learning-b79464f54d69, Jan. 14, 2021, 31 page… [cited by applicant]
Huang et al., Out-of-Distribution Detection for LiDAR-based 3D Object Detection, Available online at: https://arxiv.org/pdf/2209.14435v1.pdf, Sep. 28, 2022, 7 pages. [cited by applicant]
Jaadi, A Step-by-Step Explanation of Principal Component Analysis (PCA), Builtin.com, Available Online at: https://builtin.com/data-science/step-step-explanation-principal-component-analysis, Apr. 1, 2021, 8 pages. [cited by applicant]
Journet et al., DocCreator: A New Software for Creating Synthetic Ground-Truthed Document Images, Journal of Imaging, vol. 3, Available Online at: https://hal.archives-ouvertes.fr/hal-01668915/file/jimaging.pdf, Dec. 20… [cited by applicant]
Lin et al., An Efficient Data Augmentation Network for Out-of-Distribution Image Detection, Available online at: https://ieeexplore.ieee.org/stamp/stamp.jsp?tp=&arnumber=9363111, Feb. 24, 2021, pp. 35313-35323. [cited by applicant]
Liu et al., Self-Supervised Learning: Generative or Contrastive, Available online at https://arxiv.org/pdf/2006.08218.pdf, Mar. 20, 2021, pp. 1-24. [cited by applicant]
Luan et al., Out-Of-Distribution Detection for Deep Neural Networks with Isolation Forest and Local Outlier Factor, Available online at: https://www.researchgate.net/publication/354189192_Out-Of-Distribution_Detection_f… [cited by applicant]
Ma, NLP Augmentation, Available online at : https://github.com/makcedward/nlpaug, 2019, 4 pages. [cited by applicant]
Ma, nlpaug: Data augmentation for NLP, Available online at: https://github.com/makcedward/nlpaug, Accessed from Internet Apr. 4, 2022, 21 pages. [cited by applicant]
Moore et al., Hungarian Maximum Matching Algorithm, Brilliant Math & Science Wiki, Available Online at: https://brilliant.org/wiki/hungarian-matching/, Accessed from Internet on Mar. 2, 2022, 7 pages. [cited by applicant]
Moore et al., Matching (Graph Theory), Brilliant Math & Science Wiki, Available Online at: https://brilliant.org/wiki/matching/, Accessed from Internet on Mar. 2, 2022, 6 pages. [cited by applicant]
Moore et al., Matching Algorithms (Graph Theory), Brilliant Math & Science Wiki, Available Online at: https://brilliant.org/wiki/matching-algorithms/, Accessed from Internet on Mar. 2, 2022, 5 pages. [cited by applicant]
Rawat et al., PnPOOD: Out-Of-Distribution Detection for Text Classification via Plug and Play Data Augmentation, Available online at: https://arxiv.org/abs/2111.00506, Oct. 31, 2021, 9 pages. [cited by applicant]
Ronneberger et al., U-Net: Convolutional Networks for Biomedical Image Segmentation, International Conference on Medical Image Computing and Computer-Assisted Intervention, Available Online at URL: https://arxiv.org/pdf… [cited by applicant]
Sebastianelli et al., Automatic Dataset Builder for Machine Learning Applications to Satellite Imagery, Software X, vol. 15, Available online at: https://www.sciencedirect.com/science/article/pii/S2352711021000728, Jul.… [cited by applicant]
Sun et al., Spatial Dual-Modality Graph Reasoning for Key Information Extraction, Journal of Latex Class Files, vol. 14, No. 8, Available online at https://arxiv.org/pdf/2103.14470.pdf, Aug. 2015, pp. 1-9. [cited by applicant]
Van Laer, Recognition of Named Entities on Invoices for IxorDocs, Available online at: https://medium.com/ixorthink/recognition-of-named-entities-on-invoices-for-ixordocs-9bef38d24429, Aug. 2, 2018, 14 pages. [cited by applicant]
Veyseh et al., Improving Keyphrase Extraction with Data Augmentation and Information Filtering, Available online at: https://www.researchgate.net/publication/363501653_Improving_Keyphrase_Extraction_with_Data_Augmentati… [cited by applicant]
White, By 2024, 60% of the Data Used for the Development of AI and Analytics Projects Will Be Synthetically Generated, Available online at: https://blogs.gartner.com/andrew_white/2021/07/24/by-2024-60-of-the-data-used-f… [cited by applicant]
Xu et al., LayoutLM: Pre-Training of Text and Layout for Document Image Understanding, Available Online at: https://arxiv.org/pdf/1912.13318.pdf, Aug. 23-27, 2020, 9 pages. [cited by applicant]
You et al., Graph Contrastive Learning with Augmentations, 34th Conference on Neural Information Processing Systems, Available online at https://papers.nips.cc/paper/2020/file/3fe230348e9a12c13120749e3f9fa4cd-Paper.pdf,… [cited by applicant]
Yu et al., PICK: Processing Key Information Extraction from Documents using Improved Graph Learning-Convolutional Networks, arXiv:2004.07464, Available Online at: https://arxiv.org/pdf/2004.07464.pdf, Jul. 18, 2020, 8 p… [cited by applicant]