IP Library › Granted Patent US 12,566,921
Granted Patent B2
US 12,566,921 · App. 18/087,629 · Granted Mar 3, 2026

Gazetteer integration for neural named entity recognition

Inventors: Tuyen Quang Pham (Melbourne, AU); Cong Duy Vu Hoang (Melbourne, AU); Mark Edward Johnson (Sydney, AU); Thanh Long Duong (Melbourne, AU)
Assignee: Oracle International Corporation
G06F40/295G06F40/205G06F40/284
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,566,921
App. No.
18/087,629
Granted
Mar 3, 2026
Kind
B2
Abstract

Techniques are provided for named entity recognition using a gazetteer incorporated with a neural network. An utterance is received from a user. The utterance is input into a neural network comprising model parameters learned for named entity recognition. The neural network generates a first representation of one or more named entities based on the utterance. A gazetteer is searched based on the input utterance to generate a second representation of one or more named entities identified in the utterance. The first named entity representation is combined with the second named entity representation to generate a combined named entity representation. The combined named entity representation is output for facilitating a response to the user.

Claims (71)

1 . A computer-implemented method comprising:

receiving an utterance from a user;

inputting, by a deep learning named entity recognition (NER) subsystem, the utterance into a neural network comprising a Bidirectional Long Short-Term Memory Conditional Random Field (BiLSTM-CRF) architecture and model parameters learned for named entity recognition;

generating, by the neural network of the NER subsystem, a first representation of one or more named entities based on the utterance;

searching, by an elastic search subsystem, a gazetteer, comprising executing an elastic entity search on the gazetteer based on the utterance to generate an initial search result;

multi-hot encoding, by the elastic search subsystem, the initial search result to generate a second representation of the one or more named entities identified in the utterance, wherein multi-hot encoding the initial search result includes generating the second representation that includes two or more nonzero terms based on a relative location of tokens in the initial search result;

combining, by a conditional random field (CRF) layer of the BiLSTM-CRF, the first representation with the second representation to generate a combined named entity representation;

outputting, by the CRF layer, the combined named entity representation; and

based on the combined named entity representation, generating a response and providing the response to the user as speech output via a speaker and/or text output via a display.

2 . The method of claim 1 , wherein combining the first representation with the second representation comprises:

computing a weighted linear interpolation of the first representation and the second representation.

3 . The method of claim 1 , wherein combining the first representation with the second representation comprises:

concatenating the first representation and the second representation.

4 . The method of claim 1 , wherein combining the first representation with the second representation comprises:

computing a sequence of attention vectors based on the first representation and the second representation.

5 . The method of claim 1 , wherein the method is performed by a machine learning model comprising the neural network, the method further comprising training the machine learning model by:

accessing a training utterance;

inputting the training utterance into the neural network;

generating, by the neural network, a first training representation based on the training utterance;

searching the gazetteer based on the training utterance to generate a second training representation; and

training the machine learning model based on the first training representation and the second training representation.

6 . The method of claim 5 , wherein training the machine learning model further comprises:

identifying a dropout parameter; and

applying the dropout parameter to the first training representation.

7 . A system comprising:

one or more processors; and

a memory coupled to the one or more processors, the memory comprising a plurality of instructions executable by the one or more processors to cause the one or more processors to perform operations comprising:

receiving an utterance from a user;

inputting, by a deep learning named entity recognition (NER) subsystem, the utterance into a neural network comprising a Bidirectional Long Short-Term Memory Conditional Random Field (BiLSTM-CRF) architecture and model parameters learned for named entity recognition;

generating, by the neural network, a first representation of one or more named entities based on the utterance;

searching, by an elastic search subsystem, a gazetteer, comprising executing an elastic entity search on the gazetteer based on the utterance to generate an initial search result;

multi-hot encoding, by the elastic search subsystem, the initial search result to generate a second representation of the one or more named entities identified in the utterance, wherein multi-hot encoding the initial search result includes generating the second representation that includes two or more nonzero terms based on a relative location of tokens in the initial search result;

combining, by a conditional random field (CRF) layer of the BiLSTM-CRF, the first representation with the second representation to generate a combined named entity representation;

outputting, by the CRF layer, the combined named entity representation; and

based on the combined named entity representation, generating a response and providing the response to the user as speech output via a speaker and/or text output via a display.

8 . The system of claim 7 , wherein the operation of combining the first representation with the second representation comprises:

computing a weighted linear interpolation of the first representation and the second representation.

9 . The system of claim 7 , wherein the operation of combining the first representation with the second representation comprises:

concatenating the first representation and the second representation.

10 . The system of claim 7 , wherein the operation of combining the first representation with the second representation comprises:

computing a sequence of attention vectors based on the first representation and the second representation.

11 . The system of claim 7 , wherein the operations are performable by a machine learning model comprising the neural network, and wherein the operations further comprise training the machine learning model by:

accessing a training utterance;

inputting the training utterance into the neural network;

generating, by the neural network, a first training representation based on the training utterance;

searching the gazetteer based on the training utterance to generate a second training representation; and

training the machine learning model based on the first training representation and the second training representation.

12 . The system of claim 11 , wherein the operation of training the machine learning model further comprises:

identifying a dropout parameter; and

applying the dropout parameter to the first training representation.

13 . A non-transitory computer-readable memory comprising a plurality of instructions executable by one or more processors to cause the one or more processors to perform operations comprising:

receiving an utterance from a user;

inputting, by a deep learning named entity recognition (NER) subsystem, the utterance into a neural network comprising a Bidirectional Long Short-Term Memory Conditional Random Field (BiLSTM-CRF) architecture and model parameters learned for named entity recognition;

generating, by the neural network of the NER subsystem, a first representation of one or more named entities based on the utterance;

searching, by an elastic search subsystem, a gazetteer, comprising executing an elastic entity search on the gazetteer based on the utterance to generate an initial search result;

multi-hot encoding, by the elastic search subsystem, the initial search result to generate a second representation of the one or more named entities identified in the utterance, wherein multi-hot encoding the initial search result includes generating the second representation that includes two or more nonzero terms based on a relative location of tokens in the initial search result;

combining, by a conditional random field (CRF) layer of the BiLSTM-CRF, the first representation with the second representation to generate a combined named entity representation;

outputting, by the CRF layer, the combined named entity representation; and

based on the combined named entity representation, generating a response and providing the response to the user as speech output via a speaker and/or text output via a display.

14 . The non-transitory computer-readable memory of claim 13 , wherein the operation of combining the first representation with the second representation comprises:

computing a weighted linear interpolation of the first representation and the second representation.

15 . The non-transitory computer-readable memory of claim 13 , wherein the operation of combining the first representation with the second representation comprises:

concatenating the first representation and the second representation.

16 . The non-transitory computer-readable memory of claim 13 , wherein the operation of combining the first representation with the second representation comprises:

computing a sequence of attention vectors based on the first representation and the second representation.

17 . The non-transitory computer-readable memory of claim 13 , wherein the operations are performable by a machine learning model comprising the neural network, and wherein the operations further comprise training the machine learning model by:

accessing a training utterance;

inputting the training utterance into the neural network;

generating, by the neural network, a first training representation based on the training utterance;

searching the gazetteer based on the training utterance to generate a second training representation; and

training the machine learning model based on the first training representation and the second training representation.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 27, 2022
From: PHAM, TUYEN QUANG; HOANG, CONG DUY VU; JOHNSON, MARK EDWARD; DUONG, THANH LONG
To: ORACLE INTERNATIONAL CORPORATION
Reel/Frame 062213/0083 →
Continuity (2)
Provisional Application 63293440 · Dec 23, 2021
Related Publication 20230205999A1 · Jun 29, 2023
References Cited (26)
US 11604925B1 · Lee · 2023 [cited by examiner]
US 20190213470A1 · Schmidt · 2019 [cited by examiner]
US 20190377790A1 · Redmond et al. · 2019 [cited by applicant]
US 20220012588A1 · Rhee · 2022 [cited by examiner]
US 20230046367A1 · Doyen et al. · 2023 [cited by applicant]
US 20230154455A1 · Vu · 2023 [cited by examiner]
Fetahu, Besnik, et al. “Gazetteer enhanced named entity recognition for code-mixed web queries.” Proceedings of the 44th International ACM SIGIR Conference on Research and Development in Information Retrieval. 2021 (Yea… [cited by examiner]
Lin, Hongyu, et al. “Gazetteer-enhanced attentive neural networks for named entity recognition.” Proceedings of the 2019 Conference on Empirical Methods in Natural Language Processing and the 9th International Joint Con… [cited by examiner]
Gazetteer Enhanced NER (Year: 2021). [cited by examiner]
Gazetteer Enhanced ANN (Year: 2019). [cited by examiner]
Emelyanov et al., “Multilingual Named Entity Recognition Using Pretrained Embeddings, Attention Mechanism and NCRF”, Proceedings of the 7th Workshop on Balto-Slavic Natural Language Processing, Aug. 2, 2019, pp. 94-99. [cited by applicant]
Liu et al., “Towards Improving Neural Named Entity Recognition with Gazetteers”, Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics, Jul. 28-Aug. 2, 2019, pp. 5301-5307. [cited by applicant]
Magnolini et al., “How to Use Gazetteers for Entity Recognition with Neural Models”, Proceedings of the 5th Workshop on Semantic Deep Learning (SemDeep-5), Available Online at: https://aclanthology.org/W19-5807.pdf, Aug… [cited by applicant]
Peshterliev et al., “Self-Attention Gazetteer Embeddings for Named-Entity Recognition”, Available Online at: https://arxiv.org/pdf/2004.04060.pdf, Apr. 18, 2020, 6 pages. [cited by applicant]
Ratinov et al., “Design Challenges and Misconceptions in Named Entity Recognition”, Proceedings of the Thirteenth Conference on Computational Natural Language Learning (CoNLL), Jun. 2009, pp. 147-155. [cited by applicant]
Song et al., “Gazetteer Generation for Neural Named Entity Recognition”, The Thirty-Third International Flairs Conference (Flairs-33), May 2020, pp. 298-301. [cited by applicant]
Song et al., “Improving Neural Named Entity Recognition with Gazetteers”, Available Online at: https://arxiv.org/pdf/2003.03072.pdf, Mar. 2020, 8 pages. [cited by applicant]
Srivastava et al., “Dropout: A Simple Way to Prevent Neural Networks from Overfitting”, Journal of Machine Learning Research, vol. 15, No. 1, Jan. 2014, pp. 1929-1958. [cited by applicant]
Vaswani et al., “Attention Is All You Need”, 31st Conference on Neural Information Processing Systems (NIPS 2017), Jun. 2017, 11 pages. [cited by applicant]
Yang et al., “Drop-Out Conditional Random Fields for Twitter with Huge Mined Gazetteer”, Proceedings of NAACL-HLT, Jun. 12-17, 2016, pp. 282-288. [cited by applicant]
Zhang et al., “Token Drop Mechanism for Neural Machine Translation”, Proceedings of the 28th International Conference on Computational Linguistics, Dec. 8-13, 2020, pp. 4298-4303. [cited by applicant]
U.S. Appl. No. 18/087,647, Non-Final Office Action, mailed on May 9, 2025, 34 pages. [cited by applicant]
Maldonado-Cruz , “Tuning Machine Learning Dropout for Subsurface Uncertainty Model Accuracy”, Journal of Petroleum Science and Engineering, vol. 205, Oct. 2021, 9 pages. [cited by applicant]
Meng et al., “Distantly-Supervised Named Entity Recognition With Noise-Robust Learning and Language Model Augmented Self-Training”, Proceedings of the 2021 Conference on Empirical Methods in Natural Language Processing,… [cited by applicant]
U.S. Appl. No. 18/087,647, “Notice of Allowance”, mailed Nov. 10, 2025, 24 pages. [cited by applicant]
Xue et al., “Convolutional Recurrent Neural Networks with a Self-Attention Mechanism for Personnel Performance Prediction”, Entropy, vol. 21, No., 12, Dec. 16, 2019, 16 pages. [cited by applicant]