IP Library › Granted Patent US 12,271,395
Granted Patent B2
US 12,271,395 · App. 18/153,602 · Granted Apr 8, 2025

Focusing unstructured data and generating focused data determinations from an unstructured data set

Inventors: Colum Foley (County Dublin, IE); Paul Ferguson (Dublin, IE)
Assignee: Optum Services (Ireland) Limited
G06F16/258G06N20/00G06Q40/08
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,271,395
App. No.
18/153,602
Granted
Apr 8, 2025
Kind
B2
Abstract

Each year, almost 10% of claims are denied by payers (i.e., health insurance plans). With the cost to recover these denials and underpayments, predicting payer response (likelihood of payment) from claims data with a high degree of accuracy and precision is anticipated to improve healthcare staffs' performance productivity and drive better patient financial experience and satisfaction in the revenue cycle (Barkholz, 2017). However, constructing advanced predictive analytics models has been considered challenging in the last twenty years. That said, we propose a (low-level) context-dependent compact representation of patients' historical claim records by effectively learning complicated dependencies in the (high-level) claim inputs. Built on this new latent representation, we demonstrate that a deep learning-based framework, Deep Claim, can accurately predict various responses from multiple payers using 2,905,026 de-identified claims data from two US health systems. Deep Claim's improvements over carefully chosen baselines in predicting claim denials are most pronounced as 22.21% relative recall gain (at 95% precision) on Health System A, which implies Deep Claim can find 22.21% more denials than the best baseline system.

Claims (55)

1. A computer-implemented method comprising:

integrating, by one or more processors, a first model into a second model by connecting one or more layers of the first model to one or more layers of the second model, wherein the first model has been trained based at least in part on a first data set and the first data set is associated with a first model domain;

training, by the one or more processors, a first portion of the second model based at least in part on the first data set;

freezing, by the one or more processors, the first model integrated into the second model; and

training, by the one or more processors, a second portion of the second model based at least in part on a second data set, wherein the second data set is associated with a second model domain that is different than the first model domain.

2. The computer-implemented method of claim 1 , further comprising:

unfreezing the first model integrated into the second model; and

fine-tuning the second model based at least in part on the second data set.

3. The computer-implemented method of claim 1 , further comprising:

storing the second model.

4. The computer-implemented method of claim 1 , further comprising:

receiving input data associated with an unstructured data set; and

generating, using the second model, output data based at least in part on the input data.

5. The computer-implemented method of claim 1 , further comprising:

increasing a learning rate of the second model upon freezing the first model.

6. The computer-implemented method of claim 1 , further comprising:

unfreezing the first model integrated into the second model; and

decreasing the learning rate of the second model.

7. The computer-implemented method of claim 1 , wherein the first data set comprises waste and error data and wherein the second data set comprises fraud data.

8. A system comprising memory and one or more processors communicatively coupled to the memory, the one or more processors configured to: integrate a first model into a second model by connecting one or more layers of the first model to one or more layers of the second model, wherein the first model is initially trained based at least in part on a first data set that is associated with a first model domain;

train a first portion of the second model based at least in part on the first data set;

freeze the first model integrated into the second model; and

train a second portion of the second model based at least in part on a second data set that is associated with a second model domain different than the first model domain.

9. The computing system of claim 8 , wherein the instructions further configure the apparatus to:

unfreeze the first model integrated into the second model; and

fine-tune the second model based at least in part on the second data set.

10. The computing system of claim 8 , wherein the instructions further configure the apparatus to:

store the second model.

11. The computing system of claim 8 , wherein the instructions further configure the apparatus to:

receive input data associated with an unstructured data set; and

generate, using the second model, output data based at least in part on the input data.

12. The computing system of claim 8 , wherein the instructions further configure the apparatus to:

increase a learning rate of the second model upon freezing the first model.

13. The computing system of claim 8 , wherein the instructions further configure the apparatus to:

unfreeze the first model integrated into the second model; and

decrease the learning rate of the second model.

14. The computing system of claim 8 , wherein the first data set comprises a waste & error data and wherein the second data set comprises fraud data.

15. One or more non-transitory computer-readable storage media including instructions that, when executed by one or more processors, cause the one or more processors to:

integrate a first model into a second model by connecting one or more layers of the first model to one or more layers of the second model, wherein the first model is initially trained based at least in part on a first data set that is associated with a first model domain;

train a first portion of the second model based at least in part on the first data set;

freeze the first model integrated into the second model; and

train a second portion of the second model based at least in part on a second data set that is associated with a second model domain different than the first model domain.

16. The one or more non-transitory computer-readable storage media of claim 15 , wherein the one or more processors are further caused to:

unfreeze the first model integrated into the second model; and

fine-tune the second model based at least in part on the second data set.

17. The one or more non-transitory computer-readable storage media of claim 15 , wherein the one or more processors are further caused to:

store the second model.

18. The one or more non-transitory computer-readable storage media of claim 15 , wherein the one or more processors are further caused to:

receive input data associated with an unstructured data set; and

generate, using the second model, output data based at least in part on the input data.

19. The one or more non-transitory computer-readable storage media of claim 15 , wherein the one or more processors are further caused to:

increase a learning rate of the second model upon freezing the first model.

20. The one or more non-transitory computer-readable storage media of claim 19 , wherein the one or more processors are further caused to:

unfreeze the first model integrated into the second model; and

decrease the learning rate of the second model.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 12, 2023
From: FERGUSON, PAUL; FOLEY, COLUM
To: OPTUM SERVICES (IRELAND) LIMITED
Reel/Frame 062359/0925 →
Continuity (3)
Continuation 18153439 · Jan 12, 2023
Provisional Application 63374252 · Sep 1, 2022
Related Publication 20240078245A1 · Mar 7, 2024
References Cited (46)
US 10102340B2 · Tanner, Jr. et al. · 2018 [cited by applicant]
US 10319468B2 · Ginsburg · 2019 [cited by applicant]
US 10552576B2 · Campbell · 2020 [cited by applicant]
US 11049594B2 · Chintamaneni et al. · 2021 [cited by applicant]
US 11361381B1 · Lehmuth et al. · 2022 [cited by applicant]
US 20110258054A1 · Pandey et al. · 2011 [cited by applicant]
US 20130006655A1 · Van Arkel et al. · 2013 [cited by applicant]
US 20130054259A1 · Wojtusiak et al. · 2013 [cited by applicant]
US 20140081652A1 · Klindworth · 2014 [cited by applicant]
US 20160239617A1 · Farooq et al. · 2016 [cited by applicant]
US 20160253461A1 · Sohr et al. · 2016 [cited by applicant]
US 20170017760A1 · Freese et al. · 2017 [cited by applicant]
US 20170322930A1 · Drew · 2017 [cited by applicant]
US 20190034589A1 · Chen et al. · 2019 [cited by applicant]
US 20200013124A1 · Obee · 2020 [cited by examiner]
US 20200104731A1 · Oliner et al. · 2020 [cited by applicant]
US 20200311601A1 · Robinson et al. · 2020 [cited by applicant]
US 20200381090A1 · Apostolova et al. · 2020 [cited by applicant]
US 20210056113A1 · Mac et al. · 2021 [cited by applicant]
US 20210109915A1 · Godden et al. · 2021 [cited by applicant]
US 20210304749A1 · Singh et al. · 2021 [cited by applicant]
US 20210313022A1 · Chaballout · 2021 [cited by applicant]
US 20220012611A1 · Moradi · 2022 [cited by examiner]
US 20220309592A1 · Zahora et al. · 2022 [cited by applicant]
US 20220392048A1 · Henry · 2022 [cited by examiner]
US 20230048097A1 · Clausen · 2023 [cited by examiner]
US 20230101817A1 · Sinha et al. · 2023 [cited by applicant]
US 20230131694A1 · Saber · 2023 [cited by examiner]
US 20230195443A1 · Eberlein · 2023 [cited by examiner]
US 20230224493A1 · Foley et al. · 2023 [cited by applicant]
US 20230385705A1 · Takehara · 2023 [cited by examiner]
US 20240078609A1 · Foley et al. · 2024 [cited by applicant]
US 20240078610A1 · Foley · 2024 [cited by examiner]
CN 108334935A · 2018 [cited by applicant]
CN 109636061A · 2019 [cited by applicant]
CN 116070630A · 2023 [cited by applicant]
WO 2022057057A1 · 2022 [cited by applicant]
Jain, Ankit. “Claim Analysis and Fraud Detection Using Business Intelligence|Blog,” Nalashaa, Feb. 2, 2017, (3 pages), [Retrieved from the Internet Oct. 7, 2022] <URL: https://www.nalashaa.com/claim-analysis-fraud-detec… [cited by applicant]
Kim, Byung-Hak et al. “Deep Claim: Payer Response Prediction From Claims Data With Deep Learning,” arXiv preprint arXiv:2007.06229v1 [cs.LG], Jul. 13, 2020, (9 pages). [cited by applicant]
Sowah, Robert A. et al. “Decision Support System (DSS) for Fraud Detection in Health Insurance Claims Using Genetic Support Vector Machines (GSVMs),” Hindawi Journal of Engineering, vol. 2019, Article ID 1432597, pp. 1-… [cited by applicant]
Sun, Xu et al. “Feature-Frequency-Adaptive On-Line Training For Fast and Accurate Natural Language Processing,” Computational Linguistics, vol. 40, No. 3, Sep. 1, 2014, pp. 563-586, DOI: 10.1162/COLI_a_00193. [cited by applicant]
Thesmar, David et al. “Combining The Power Of Artificial Intelligence With The Richness Of Healthcare Claims Data: Opportunities and Challenges,” PharmacoEconomics, vol. 37, pp. 745-752, Mar. 8, 2019. [cited by applicant]
Non-Final Rejection Mailed on Feb. 28, 2024 for U.S. Appl. No. 18/153,624, 5 page(s). [cited by applicant]
Non-Final Rejection Mailed on Mar. 26, 2024 for U.S. Appl. No. 17/805,340, 19 page(s). [cited by applicant]
Notice of Allowance and Fees Due (PTOL-85) for U.S. Appl. No. 17/805,340, filed Jul. 26, 2024, 8 pages, United States Patent and Trademark Office, US. [cited by applicant]
Notice of Allowance and Fees Due (PTOL-85) for U.S. Appl. No. 18/153,624 Mailed on Sep. 6, 2024, 8 pages, United States Patent and Trademark Office, US. [cited by applicant]