IP Library › Granted Patent US 12,561,524
Granted Patent B2
US 12,561,524 · App. 18/455,396 · Granted Feb 24, 2026

Training machine learning models to automatically detect and correct contextual and logical errors

Inventors: Jun Su (Beijing, CN); Su Liu (Austin, TX); Yang Liang (Beijing, CN); Richard Anthony LaFrance (Grosse Ile, MI)
Assignee: INTERNATIONAL BUSINESS MACHINES CORPORATION
G06F40/279G06F40/106G06F40/205G06F40/232G06F40/253
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,561,524
App. No.
18/455,396
Granted
Feb 24, 2026
Kind
B2
Abstract

Provided are techniques for training Machine Learning (ML) models to automatically detect and correct contextual and logical errors. A plurality of machine learning models are trained. In response to receiving content, the content is parsed into different elements. One or more knowledge graphs are built based on the different elements. One or more machine learning models are selected from the plurality of machine learning models based on the one or more knowledge graphs and custom criteria in a user profile. The selected one or more machine learning models are used to identify at least one of a contextual error and a logical error and a correction for the at least one of the contextual error and the logical error. The correction for the at least one of the contextual error and the logical error is applied to generate corrected content. The corrected content is rendered.

Claims (64)

1 . A computer-implemented method, comprising operations for:

training a plurality of machine learning models; and

in response to receiving content,

building one or more knowledge graphs based on the content;

selecting a first subset of machine learning models from the plurality of machine learning models based on the one or more knowledge graphs and custom criteria in a user profile;

using the selected one or more machine learning models to identify at least one of a contextual error and a logical error and a correction for the at least one of the contextual error and the logical error;

applying the correction for the at least one of the contextual error and the logical error to generate corrected content;

rendering the corrected content;

adjusting the custom criteria based on user feedback on the corrected content:

in response to receiving new content,

adjusting the one or more knowledge graphs based on the new content; and

selecting a second subset of machine learning models from the plurality of machine learning models based on the one or more adjusted knowledge graphs and the adjusted custom criteria.

2 . The computer-implemented method of claim 1 , wherein the content is parsed into different elements based on time, location, event, subject, verb, and object.

3 . The computer-implemented method of claim 1 , further comprising operations for:

updating the one or more knowledge graphs based on receiving additional content.

4 . The computer-implemented method of claim 1 , further comprising operations for:

receiving user input to configure the custom criteria, wherein the custom criteria comprises identification of monitored applications, identification of types of errors to detect, and how to process the types of errors.

5 . The computer-implemented method of claim 1 , further comprising operations for:

generating a data structure, wherein the data structure comprises columns for: a user identifier, an application identifier, a content identifier, a content type, a content category, a topic, an input buffer, a paragraph identifier, a sentence identifier, an error type, a related sentence identifier, suggested content, and proof reading comments.

6 . The computer-implemented method of claim 1 , wherein each machine learning model of the plurality of machine learning models is retrained using user feedback on the correction.

7 . The computer-implemented method of claim 1 , wherein training data for the plurality of machine learning models comprises sample pairs of a tagged error type and a corresponding correction.

8 . A computer program product, the computer program product comprising a computer readable storage medium having program instructions embodied therewith, the program instructions executable by a processor to cause the processor to perform operations for:

training a plurality of machine learning models; and

in response to receiving content,

building one or more knowledge graphs based on the content;

selecting a first subset of machine learning models from the plurality of machine learning models based on the one or more knowledge graphs and custom criteria in a user profile;

using the selected one or more machine learning models to identify at least one of a contextual error and a logical error and a correction for the at least one of the contextual error and the logical error;

applying the correction for the at least one of the contextual error and the logical error to generate corrected content;

rendering the corrected content;

adjusting the custom criteria based on user feedback on the corrected content;

in response to receiving new content,

adjusting the one or more knowledge graphs based on the new content; and

selecting a second subset of machine learning models from the plurality of machine learning models based on the one or more adjusted knowledge graphs and the adjusted custom criteria.

9 . The computer program product of claim 8 , wherein the content is parsed into different elements based on time, location, event, subject, verb, and object.

10 . The computer program product of claim 8 , wherein the program instructions are executable by the processor to cause the processor to perform operations further comprising:

updating the one or more knowledge graphs based on receiving additional content.

11 . The computer program product of claim 8 , wherein the program instructions are executable by the processor to cause the processor to perform operations further comprising:

receiving user input to configure the custom criteria, wherein the custom criteria comprises identification of monitored applications, identification of types of errors to detect, and how to process the types of errors.

12 . The computer program product of claim 8 , wherein the program instructions are executable by the processor to cause the processor to perform operations further comprising:

generating a data structure, wherein the data structure comprises columns for: a user identifier, an application identifier, a content identifier, a content type, a content category, a topic, an input buffer, a paragraph identifier, a sentence identifier, an error type, a related sentence identifier, suggested content, and proof reading comments.

13 . The computer program product of claim 8 , wherein each machine learning model of the plurality of machine learning models is retrained using user feedback on the correction.

14 . The computer program product of claim 8 , wherein training data for the plurality of machine learning models comprises sample pairs of a tagged error type and a corresponding correction.

15 . A computer system, comprising:

one or more processors, one or more computer-readable memories and one or more computer-readable, tangible storage devices; and

program instructions, stored on at least one of the one or more computer-readable, tangible storage devices for execution by at least one of the one or more processors via at least one of the one or more computer-readable memories, to perform operations comprising:

training a plurality of machine learning models; and

in response to receiving content,

building one or more knowledge graphs based on the content;

selecting a first subset of machine learning models from the plurality of machine learning models based on the one or more knowledge graphs and custom criteria in a user profile;

using the selected one or more machine learning models to identify at least one of a contextual error and a logical error and a correction for the at least one of the contextual error and the logical error;

applying the correction for the at least one of the contextual error and the logical error to generate corrected content;

rendering the corrected content;

adjusting the custom criteria based on user feedback on the corrected content;

in response to receiving new content,

adjusting the one or more knowledge graphs based on the new content; and

selecting a second subset of machine learning models from the plurality of machine learning models based on the one or more adjusted knowledge graphs and the adjusted custom criteria.

16 . The computer system of claim 15 , wherein the content is parsed into different elements based on time, location, event, subject, verb, and object.

17 . The computer system of claim 15 , wherein the operations further comprise:

updating the one or more knowledge graphs based on receiving additional content.

18 . The computer system of claim 15 , wherein the operations further comprise:

receiving user input to configure the custom criteria, wherein the custom criteria comprises identification of monitored applications, identification of types of errors to detect, and how to process the types of errors.

19 . The computer system of claim 15 , wherein the operations further comprise:

generating a data structure, wherein the data structure comprises columns for: a user identifier, an application identifier, a content identifier, a content type, a content category, a topic, an input buffer, a paragraph identifier, a sentence identifier, an error type, a related sentence identifier, suggested content, and proof reading comments.

20 . The computer system of claim 15 , wherein each machine learning model of the plurality of machine learning models is retrained using user feedback on the correction.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 24, 2023
From: SU, JUN; LIU, SU; LIANG, YANG; LAFRANCE, RICHARD ANTHONY
To: INTERNATIONAL BUSINESS MACHINES CORPORATION
Reel/Frame 064698/0067 →
Continuity (1)
Related Publication 20250068843A1 · Feb 27, 2025
References Cited (32)
US 10387565B2 · Hoover et al. · 2019 [cited by applicant]
US 20040002994A1 · Brill · 2004 [cited by examiner]
US 20190361720A1 · Balachandran · 2019 [cited by examiner]
US 20210374329A1 · Deutsch · 2021 [cited by applicant]
US 20220335205A1 · Li · 2022 [cited by examiner]
US 20220374605A1 · Sethi · 2022 [cited by examiner]
US 20230154593A1 · Gao · 2023 [cited by examiner]
US 20240273288A1 · Gashteovski · 2024 [cited by examiner]
CN 109036464B · 2018 [cited by applicant]
KR 20150092879A · 2015 [cited by applicant]
WO 2020100018A1 · 2020 [cited by applicant]
K. Thadani, e al., “A Framework for Identifying Textual Redundancy”, Creative Commons, Proceedings of the International Conference on Computational Linguistics, 2008, 8 pp. (available at https://aclanthology.org/C08-111… [cited by applicant]
A. Islam, et al., “Correcting Different Types of Errors in Texts”, ResearchGate, May 2011, 13 pp. (available at https://www.researchgate.net/publication/225115578). [cited by applicant]
M. Mikowski, “Developing an open-source, rule-based proofreading tool”, ResearchGate, Jan. 2010, 25 pp. (available at https://www.researchgate.net/publication/220282022). [cited by applicant]
H. Stehouwer, et al., “Language models for contextual error detection and correction”, Association for Computational Linguistics, Proceedings of the EACL 2009 Workshop on Computational Linguistic Aspects of Grammatical … [cited by applicant]
J. Zhu, et al., “Machine Learning-Based Grammar Error Detection Method in English Composition”, Hindawi, Scientific Programming vol. 2021, Article ID 4213791, 2021, 10 pp. (available at https://www.hindawi.com/journals/… [cited by applicant]
F. Bannay et al., “Using a SMT solver for risk analysis: detecting logical mistakes in texts”, IEEE, 2014 IEEE 26th International Conference on Tools with Artificial Intelligence, 2014, 8 pp. (available at https://ieeex… [cited by applicant]
Mell, P. et al., “The NIST Definition of Cloud Computing (Draft)”, Sep. 2011, Computer Security Division Information Technology Laboratory National Institute of Standards and Technology, Total 7 pp. [cited by applicant]
Mell, P. et al., “Effectively and Securely Using the Cloud Computing Paradigm”, [online], Oct. 7, 2009, retrieved from the Internet at <URL: http://csrc.nist.gov/groups/SNS/cloud-computing/cloud-computing-v26.ppt>, Tota… [cited by applicant]
“Preps and Pro Pros”, Patent Bots, [online][retrieved Jul. 27, 2023] https://www.patentbots.com/about-patent-proofreading, 5 pp. [cited by applicant]
“Personalized AI, Everywhere You Write”, Grammarly, [online][retrieved Jul. 27, 2023] https://www.grammarly.com/, 7 pp. [cited by applicant]
“There's a better way to write”, ProWritingAid, [online][retrieved Jul. 27, 2023] https://prowritingaid.com/, 10 pp. [cited by applicant]
“Hemingway App makes your writing bold and clear”, Hemingway Editor, [online][retrieved Jul. 27, 2023] https://hemingwayapp.com/, 1 p. [cited by applicant]
Abedini et al. “Correction Tower: A General Embedding Method of The Error Recognition for The Knowledge Graph Correction”, International Journal Of Pattern Recognition And Artificial Intelligence, Jan. 8, 2020, 38 pages… [cited by applicant]
Chen et al. “Converge to the Truth: Factual Error Correction via Iterative Constrained Editing”, arXiv:2211.12130, Feb. 23, 2023, 10 pages. [cited by applicant]
Dong et al. “Integrating Human-In-The-Loop into Swarm Learning for Decentralized Fake News Detection”, Proceedings of The 2022 International Conference On Intelligent Data Science Technologies And Applications (IDSTA'22… [cited by applicant]
Dong et al: “Faithful to the document or to the World? Mitigating hallucinations via entity-linked knowledge in abstractive summarization”, arXiv:2204.13761, Apr. 28, 2022, 12 pages. [cited by applicant]
Fung et al. “InfoSurgeon: Cross-Media Fine-grained Information Consistency Checking for Fake News Detection”, Proceedings of the 59th Annual Meeting of The Association for Computational Linguistics and the 11th Internat… [cited by applicant]
International Searching Authority, “Notification of Transmittal of the International Search Report and the Written Opinion of the International Searching Authority, or Declaration,” Patent Cooperation Treaty, Oct. 24, 2… [cited by applicant]
Lin et al. “A Joint Neural Model for Information Extraction with Global Features”, Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics, Jul. 2020, 7999-8009. [cited by applicant]
Sourati et al. “Case-Based Reasoning with Language Models for Classification of Logical Fallacies”, arXiv:2301.11879 May 17, 2023, 13 pages. [cited by applicant]
Zhu et al. “Enhancing Factual Consistency of Abstractive Summarization”, arXiv:2003.08612, Mar. 15, 2021, 16 pages. [cited by applicant]