IP Library › Granted Patent US 12,579,375
Granted Patent B2
US 12,579,375 · App. 18/378,249 · Granted Mar 17, 2026

Implementing active learning in natural language generation tasks

Inventors: Liat Ein-Dor (Tel Aviv, IL); Yotam Perlitz (Givatayim, IL); Michal Shmueli-Scheuer (Tel Aviv, IL); Dafna Sheinwald (Haifa, IL); Ariel Gera (Givatayim, IL)
Assignee: International Business Machines Corporation
G06F40/40G06N20/00
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,579,375
App. No.
18/378,249
Granted
Mar 17, 2026
Kind
B2
Abstract

Methods, systems, and computer program products for implementing active learning in NLG tasks are provided herein. A computer-implemented method includes generating multiple natural language annotations associated with multiple items of unlabeled data by processing the unlabeled data using at least one artificial intelligence model; determining at least one quality score attributed to at least a portion of the multiple generated natural language annotations based at least in part on at least one quality metric; selecting at least one of the multiple natural language annotations and at least one corresponding item of the multiple items of unlabeled data based at least in part on the at least one determined quality score; and performing one or more automated actions based at least in part on the at least one selected natural language annotation.

Claims (38)

1 . A system comprising:

a memory configured to store program instructions; and

a processor operatively coupled to the memory to execute the program instructions to:

generate multiple natural language annotations associated with multiple items of unlabeled data by processing the unlabeled data using at least one artificial intelligence model;

determine at least one quality score attributed to at least a portion of the multiple generated natural language annotations based at least in part on at least one quality metric;

select at least one of the multiple natural language annotations and at least one corresponding item of the multiple items of unlabeled data based at least in part on the at least one determined quality score, wherein selecting at least one of the multiple natural language annotations and at least one corresponding item of the multiple items of unlabeled data comprises selecting the at least one of the multiple natural language annotations and the at least one corresponding item of the multiple items of unlabeled data based at least in part on at least one difference between the at least one determined quality score and one or more predicted quality scores pertaining to one or more manual annotations associated with the multiple items of unlabeled data; and

perform one or more automated actions based at least in part on the at least one selected natural language annotation, wherein performing one or more automated actions comprises automatically training, using the least one selected natural language annotation and the at least one corresponding item of unlabeled data, the at least one artificial intelligence model to be used in connection with one or more natural language generation (NLG) tasks.

2 . The system of claim 1 , wherein performing one or more automated actions comprises automatically training one or more additional artificial intelligence models to be used in connection with one or more NLG tasks using the least one selected natural language annotation and the at least one corresponding item of unlabeled data.

3 . The system of claim 1 , wherein processing the unlabeled data comprises processing the unlabeled data using at least one artificial intelligence-based natural language generation model.

4 . The system of claim 3 , wherein the at least one artificial intelligence-based natural language generation model comprises at least one transformer-based generative model.

5 . The system of claim 4 , wherein the at least one transformer-based generative model comprises at least one of one or more encoder-decoder models and one or more denoising autoencoders.

6 . The system of claim 1 , wherein generating the multiple natural language annotations comprises processing the unlabeled data using at least one artificial intelligence model in accordance with one or more predetermined characteristics of a given NLG task.

7 . The system of claim 1 , wherein determining at least one quality score comprises determining the at least one quality score attributed to the at least a portion of the multiple generated natural language annotations based at least in part on one or more reference-less evaluation metrics.

8 . The system of claim 1 , wherein selecting at least one of the multiple natural language annotations comprises selecting the at least one natural language annotation which corresponds with improvement associated with the at least one artificial intelligence model, as measured by the at least one quality metric.

9 . The system of claim 1 , wherein the processor is further operatively coupled to the memory to execute the program instructions to:

predict the one or more quality scores pertaining to the one or more manual annotations associated with the multiple items of unlabeled data.

10 . A computer program product comprising a computer readable storage medium having program instructions embodied therewith, the program instructions executable by a computing device to cause the computing device to:

generate multiple natural language annotations associated with multiple items of unlabeled data by processing the unlabeled data using at least one artificial intelligence model;

determine at least one quality score attributed to at least a portion of the multiple generated natural language annotations based at least in part on at least one quality metric;

select at least one of the multiple natural language annotations and at least one corresponding item of the multiple items of unlabeled data based at least in part on the at least one determined quality score, wherein selecting at least one of the multiple natural language annotations and at least one corresponding item of the multiple items of unlabeled data comprises selecting the at least one of the multiple natural language annotations and the at least one corresponding item of the multiple items of unlabeled data based at least in part on at least one difference between the at least one determined quality score and one or more predicted quality scores pertaining to one or more manual annotations associated with the multiple items of unlabeled data; and

perform one or more automated actions based at least in part on the at least one selected natural language annotation, wherein performing one or more automated actions comprises automatically training, using the least one selected natural language annotation and the at least one corresponding item of unlabeled data, the at least one artificial intelligence model to be used in connection with one or more natural language generation (NLG) tasks.

11 . The computer program product of claim 10 , wherein processing the unlabeled data comprises processing the unlabeled data using at least one artificial intelligence-based natural language generation model.

12 . The computer program product of claim 10 , wherein the program instructions executable by a computing device further cause the computing device to:

predict the one or more quality scores pertaining to the one or more manual annotations associated with the multiple items of unlabeled data.

13 . A computer-implemented method comprising:

generating multiple natural language annotations associated with multiple items of unlabeled data by processing the unlabeled data using at least one artificial intelligence model;

determining at least one quality score attributed to at least a portion of the multiple generated natural language annotations based at least in part on at least one quality metric;

selecting at least one of the multiple natural language annotations and at least one corresponding item of the multiple items of unlabeled data based at least in part on the at least one determined quality score, wherein selecting at least one of the multiple natural language annotations and at least one corresponding item of the multiple items of unlabeled data comprises selecting the at least one of the multiple natural language annotations and the at least one corresponding item of the multiple items of unlabeled data based at least in part on at least one difference between the at least one determined quality score and one or more predicted quality scores pertaining to one or more manual annotations associated with the multiple items of unlabeled data; and

performing one or more automated actions based at least in part on the at least one selected natural language annotation, wherein performing one or more automated actions comprises automatically training, using the least one selected natural language annotation and the at least one corresponding item of unlabeled data, the at least one artificial intelligence model to be used in connection with one or more natural language generation (NLG) tasks;

wherein the method is carried out by at least one computing device.

14 . The computer-implemented method of claim 13 , further comprising:

predicting the one or more quality scores pertaining to the one or more manual annotations associated with the multiple items of unlabeled data.

15 . The computer-implemented method of claim 13 , wherein software implementing the method is provided as a service in a cloud environment.

16 . The computer-implemented method of claim 13 , wherein processing the unlabeled data comprises processing the unlabeled data using at least one artificial intelligence-based natural language generation model.

17 . The computer-implemented method of claim 13 , wherein generating the multiple natural language annotations comprises processing the unlabeled data using at least one artificial intelligence model in accordance with one or more predetermined characteristics of a given NLG task.

18 . The computer-implemented method of claim 13 , wherein determining at least one quality score comprises determining the at least one quality score attributed to the at least a portion of the multiple generated natural language annotations based at least in part on one or more reference-less evaluation metrics.

19 . The computer-implemented method of claim 13 , wherein performing one or more automated actions comprises automatically training one or more additional artificial intelligence models to be used in connection with one or NLG tasks using the least one selected natural language annotation and the at least one corresponding item of unlabeled data.

20 . The computer-implemented method of claim 13 , wherein selecting at least one of the multiple natural language annotations comprises selecting the at least one natural language annotation which corresponds with improvement associated with the at least one artificial intelligence model, as measured by the at least one quality metric.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 10, 2023
From: EIN-DOR, LIAT; PERLITZ, YOTAM; SHMUELI-SCHEUER, MICHAL; SHEINWALD, DAFNA; GERA, ARIEL
To: INTERNATIONAL BUSINESS MACHINES CORPORATION
Reel/Frame 065168/0024 →
Continuity (1)
Related Publication 20250117592A1 · Apr 10, 2025
References Cited (17)
US 10977518B1 · Sharma · 2021 [cited by examiner]
US 11048979B1 · Zhdanov · 2021 [cited by examiner]
US 20040111253A1 · Luo et al. · 2004 [cited by applicant]
US 20180240031A1 · Huszar · 2018 [cited by examiner]
US 20230177115A1 · Tommasi · 2023 [cited by examiner]
CN 108805258A · 2018 [cited by examiner]
Dong et al., “A Survey of Natural Language Generation”, Dec. 2022 (Year: 2022). [cited by examiner]
Mohankumar et al., Active Evaluation: Efficient NLG Evaluation with Few Pairwise Comparisons, Annual Meeting of the Association for Computational Linguistics, Published Date: Mar. 11, 2022. [cited by applicant]
Perlitz et al., Active Learning for Natural Language Generation, Published Date: Jul. 7, 2021. [cited by applicant]
Dušek et al., Training a Natural Language Generator From Unaligned Data, Proceedings of the 53rd Annual Meeting of the Association for Computational Linguistics, Published Date: Jul. 31, 2015. [cited by applicant]
Mairesse et al., Phrase-based Statistical Language Generation using Graphical Models and Active Learning, Proceedings of the 48th Annual Meeting of the Association for Computational Linguistics, Published Date: Jul. 16,… [cited by applicant]
Chang et al., On Training Instance Selection for Few-Shot Neural Text Generation, Proceedings of the 59th Annual Meeting of the Association for Computational Linguistics, Published Date: Jul. 7, 2021. [cited by applicant]
Gidiotis et al., Should We Trust This Summary? Bayesian Abstractive Summarization to the Rescue. Findings of the Association for Computational Linguistics: ACL 2022. [cited by applicant]
Tsvigun et al., Active Learning for Abstractive Text Summarization, Findings of the Association for Computational Linguistics: EMNLP 2022. [cited by applicant]
Gao et al., APRIL: Interactively Learning to Summarise by Combining Active Preference Learning and Reinforcement Learning. Proceedings of the 2018 Conference on Empirical Methods in Natural Language Processing. [cited by applicant]
Zhao et al., Active Learning Approaches to Enhancing Neural Machine Translation, Findings of the Association for Computational Linguistics: EMNLP 2020. [cited by applicant]
Karaoguz, E., Adaptive Learning Strategies for Neural Paraphrase Generation, Universitat Hamburg, 2018. [cited by applicant]