IP Library › Granted Patent US 12,530,388
Granted Patent B2
US 12,530,388 · App. 18/175,525 · Granted Jan 20, 2026

Automated enrichment of entity descriptions in unstructured text

Inventors: Marcos Martínez Galindo (Dublin, IE); Leopold Fuchs (Stuttgart, DE); Gabriele Picco (Dublin, IE); Thanh Lam Hoang (Maynooth, IE); Vanessa Lopez Garcia (Dublin, IE); Marco Luca Sbodio (Castaheany, IE)
Assignee: International Business Machines Corporation
G06F16/35G06F16/313G06F40/30
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,530,388
App. No.
18/175,525
Filed
Feb 27, 2023
Granted
Jan 20, 2026
Kind
B2
Art Unit
2154
USPC
707/737
Abstract

Automatically enriching the descriptions of an entity mentioned in a sentence corpus includes generating multiple enriched descriptions corresponding to a label of the entity. Each of the multiple enriched descriptions is ranked. The ranking is generated by a machine learning model that is configured to determine a likelihood that an enriched description correctly describes the entity. The sentence corpus are annotated by coupling each mention of the entity with one or more of the enriched descriptions. The one or more enriched descriptions are selected based on the ranking. As annotated, the sentence corpus can be output.

Claims (60)

1 . A method, comprising:

generating, by a processor of a computer, a plurality of enriched descriptions corresponding to an entity, wherein the entity is mentioned in a sentence corpus input to the computer;

computing, by the processor, using a machine learning model, and based on determining a probability that each mention of the entity in the sentence corpus belongs to one of a plurality of classes, an entropy associated with the plurality of enriched descriptions,

wherein the machine learning model is configured to learn without additional training data, and

wherein the machine learning model identifies and extracts the entity from the sentence corpus based on inferences;

ranking, by the processor, using a semantic similarity model, and based on computing the entropy using the machine learning model, each of the plurality of enriched descriptions, wherein a ranking of an enriched description, of the plurality of enriched descriptions, corresponds to likelihood that a machine learning classifier correctly classifies the entity mentioned in the sentence corpus using information from the enriched description;

selecting, based on the ranking, a subset of the plurality of enriched descriptions that reduces the entropy; and

outputting the subset of the plurality of enriched descriptions selected.

2 . The method of claim 1 ,

wherein the likelihood is determined based on a semantic similarity of embedding space encodings of each of the enriched descriptions and the entity.

3 . The method of claim 1 ,

wherein the method further comprises:

revising the ranking in response to user-provided feedback,

wherein the user-provided feedback includes at least one of a user-generated description or a re-ranking of the plurality of enriched descriptions.

4 . The method of claim 1 ,

wherein the generating includes enriching one or more initial descriptions input to the computer by a user.

5 . The method of claim 4 ,

wherein the enriching one or more initial descriptions includes extending an initial description using a language model that predicts a next word based on one or more previous words.

6 . The method of claim 4 ,

wherein the enriching one or more initial descriptions includes generating a new description using a language model that paraphrases an initial description, wherein language model is trained as a sequence-to-sequence generator.

7 . The method of claim 1 ,

wherein the generating includes automatically generating the enriched descriptions in response to an input of the entity without at least one initial description.

8 . The method of claim 7 ,

wherein the automatically generating is performed using a template and language model configured to generate an enriched description based on the template.

9 . The method of claim 7 ,

wherein the automatically generating is performed using a template and language model configured to use sequence-to-sequence text generation conditioned on an entity label.

10 . A system, comprising:

one or more processors configured to initiate operations including:

generating a plurality of enriched descriptions corresponding to an entity mentioned in a sentence corpus input to a computer;

computing, using a zero-shot machine learning model, and based on determining a probability that each mention of the entity in the sentence corpus belongs to one of a plurality of classes, an entropy associated with the plurality of enriched descriptions;

ranking, using a semantic similarity model and based on computing the entropy using the zero-shot machine learning model, each of the plurality of enriched descriptions, wherein a ranking of an enriched description, of the plurality of enriched descriptions, corresponds to likelihood that a machine learning classifier correctly classifies the entity mentioned in the sentence corpus using information from the enriched description;

selecting, based on the ranking, a subset of the plurality of enriched descriptions that reduces the entropy; and

outputting the subset of the plurality of enriched descriptions selected.

11 . The system of claim 10 ,

wherein the likelihood is determined based on a semantic similarity of embedding space encodings of each of the enriched descriptions and the entity.

12 . The system of claim 10 ,

wherein the one or more processors are configured to initiate operations further including:

revising the ranking in response to user-provided feedback,

wherein the user-provided feedback includes at least one of a user-generated description or a re-ranking of the plurality of enriched descriptions.

13 . The system of claim 10 ,

wherein the generating includes enriching one or more initial descriptions input to the computer by a user.

14 . The system of claim 13 ,

wherein the enriching one or more initial descriptions includes extending an initial description using a language model that predicts a next word based on one or more previous words.

15 . The system of claim 13 ,

wherein the enriching one or more initial descriptions includes generating a new description using a language model that paraphrases an initial description, wherein language model is trained as a sequence-to-sequence generator.

16 . The system of claim 10 ,

wherein the generating includes automatically generating the enriched descriptions in response to an input of the entity without at least one initial description.

17 . The system of claim 16 ,

wherein the automatically generating is performed using a template and language model configured to generate an enriched description based on the template.

18 . A computer program product, the computer program product comprising:

one or more computer-readable storage media and program instructions collectively stored on the one or more computer-readable storage media, the program instructions executable by a processor to cause the processor to initiate operations including:

generating a plurality of enriched descriptions corresponding to an entity mentioned in a sentence corpus input to a computer;

computing, using a machine learning model, and based on determining a probability that each mention of the entity in the sentence corpus belongs to one of a plurality of classes, an entropy associated with the plurality of enriched descriptions;

ranking, using a semantic similarity model based on computing the entropy using the machine learning model, each enriched description of the plurality of enriched descriptions, wherein a ranking of an enriched description, of the plurality of enriched descriptions, corresponds to likelihood that a machine learning classifier correctly classifies the entity mentioned in the sentence corpus using information from the enriched description;

selecting, based on the ranking, a subset of the plurality of the enriched descriptions that reduces the entropy; and

outputting the subset of the plurality of enriched descriptions selected.

19 . The computer program product of claim 18 ,

wherein the machine learning model identifies and extracts the entity from the sentence corpus based on inferences.

20 . The computer program product of claim 18 ,

wherein the likelihood is determined based on a semantic similarity of embedding space encodings of each of the enriched descriptions and the entity.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Feb 27, 2023
From: MARTÍNEZ GALINDO, MARCOS; FUCHS, LEOPOLD; PICCO, GABRIELE; HOANG, THANH LAM; LOPEZ GARCIA, VANESSA; SBODIO, MARCO LUCA
To: INTERNATIONAL BUSINESS MACHINES CORPORATION
Reel/Frame 062819/0064 →
Continuity (1)
Related Publication 20240289371A1 · Aug 29, 2024
References Cited (35)
US 11043214B1 · Hedayatnia · 2021 [cited by examiner]
US 11516158B1 · Luzhnica · 2022 [cited by examiner]
US 11829474B1 · Lu · 2023 [cited by examiner]
US 20020042793A1 · Choi · 2002 [cited by examiner]
US 20100257193A1 · Krupka et al. · 2010 [cited by applicant]
US 20180176661A1 · Varndell · 2018 [cited by examiner]
US 20190122111A1 · Min et al. · 2019 [cited by applicant]
US 20210157990A1 · Lima · 2021 [cited by examiner]
US 20210174003A1 · Meng · 2021 [cited by examiner]
US 20210365502A1 · Hutchins · 2021 [cited by examiner]
US 20220129629A1 · Niu · 2022 [cited by examiner]
US 20220198254A1 · Dalli · 2022 [cited by examiner]
US 20220253871A1 · Miller · 2022 [cited by examiner]
US 20220270721A1 · Schrempf · 2022 [cited by examiner]
US 20230022845A1 · Meng · 2023 [cited by examiner]
US 20230080674A1 · Attali · 2023 [cited by examiner]
US 20230089285A1 · Fan · 2023 [cited by examiner]
US 20230298692A1 · Fant · 2023 [cited by examiner]
CN 110555083B · 2021 [cited by applicant]
CN 114298042A · 2022 [cited by applicant]
“IBM Deep Search,” [online] Copyright © 2022 IBM, Jan. 31, 2023, retrieved from the Internet: < https://ds4sd.github.io/>, 6 pg. [cited by applicant]
Aly, R. et al., “Leveraging Type Descriptions for Zero-shot Named Entity Recognition and Classification,” Proceedings of the 59th Annual Meeting of the Association for Computational Linguistics and the 11th Internationa… [cited by applicant]
Datta, S. et al., “A Relative Information Gain-based Query Performance Prediction Framework with Generated Query Variants,” ACM Transactions on Information Systems, Dec. 21, 2022, vol. 41, No. 2, pp. 1-31. [cited by applicant]
Gadetsky, A. et al., “Conditional generators of words definitions,” arXiv preprint arXiv:1806.10090, Jun. 26, 2018, 6 pg. [cited by applicant]
Noraset, T. et al., “Definition modeling: Learning to define word embeddings in natural language,” InProceedings of the AAAI Conference on Artificial Intelligence, Feb. 12, 2017, vol. 31, No. 1, 8 pg. [cited by applicant]
Cheng, L. et al., Ent-desc: Entity description generation by exploring knowledge graph, arXiv preprint arXiv:2004.14813, Oct. 26, 2020, 11 pg. [cited by applicant]
“Zshot: Zero and Few shost named entitiy recognition plugin for Spacy,” [online] © 2023 GitHub, Inc., GitHub.IBM.Com [retrieved Dec. 15, 2022], retrieved from the Internet: <https://github.ibm.com/Dublin-Research-Lab/zs… [cited by applicant]
Zhang, H. et al. “Improving Interpretability of Word Embeddings by Generating Definition and Usage,” Expert Systems with Applications, Jul. 18, 2020, arXiv: 1912.05898v2, 12 pg. [cited by applicant]
Logeswaran, L. et al., “Zero-shot entity linking by reading entity descriptions,” arXiv preprint arXiv:1906.07348, Jun. 18, 2019, 12 pg. [cited by applicant]
Hu, R.L. et al., “Zero-shot image classification guided by natural language descriptions of classes: A meta-learning approach,” Advances in Neural Information Processing Systems vol. 268, 2018, 4 pg. [cited by applicant]
Yu, H. et al., “Zero-shot learning via simultaneous generating and learning,” Advances in Neural Information Processing Systems, arXiv preprint, arXiv: 1910.09446v1, Oct. 21, 2019, 11 pg. [cited by applicant]
Mell, P. et al., The NIST Definition of Cloud Computing, National Institute of Standards and Technology, U.S. Dept. of Commerce, Special Publication 800-145, Sep. 2011, 7 pg. [cited by applicant]
Aly, R. et al., “Leveraging type descriptions for zero-shot named entity recognition and classification,” In Proceedings of the 59th Annual Meeting of the Association for Computational Linguistics and the 11th Internati… [cited by applicant]
Gao, T. et al., “FewRel 2.0: Towards more challenging few-shot relation classification,” arXiv Preprint, arXiv:1910.07124, Oct. 16, 2019, 6 pg. [cited by applicant]
Han, X. et al., “Fewrel: A large-scale supervised few-shot relation classification dataset with state-of-the-art evaluation,” arXiv Preprint, arXiv:1810.10147, Oct. 24, 2018, 7 pg. [cited by applicant]