IP Library › Granted Patent US 12,361,218
Granted Patent B2
US 12,361,218 · App. 18/286,900 · Granted Jul 15, 2025

Nested named entity recognition

Inventors: Suzanne M Kirch (Waltham, MA); Rajiv Baronia (San Ramon, CA); Vineeth Thanikonda Munirathnam (Bangalore, IN); Jack Porter (Valley Springs, CA)
G06F40/295G06F40/284G06F40/40
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,361,218
App. No.
18/286,900
Granted
Jul 15, 2025
Kind
B2
Abstract

Named Entity Recognition is the identification and classification of named entities within a document. Nested NEs occur when an NE is contained within another NE. The disclosed invention leverages the CapsNet architecture for improved nested NE identification and classification. This includes deriving the features of an input text. The derived features are used to identify and classify any named entities in the text. The system is further configured to identify named entities in the text and perform clustering to group named entities. The disclosed CapsNet considers the context of the whole text to activate higher capsule layers in order to identify the named entities and classify them. The teachings of this invention are applicable to other NER models to improve nested NE identification and classification.

Claims (61)

1. A computer-implemented method for nested named entity recognition, comprising:

receiving, into a stack of neural capsule embedding networks comprised of M number of flat neural capsule embedding networks identified 1 through M, an embedding vector as input, wherein:

a) the embedding vector contains embedding tokens representing words in a natural language text;

b) each neural capsule embedding networks is configured to identify named entities of its assigned word span length, 1 through M;

analyzing, by each neural capsule embedding network simultaneously, the features of each word in context of the embedding vector considering tokens to the left and right of the word;

through dynamic routing of capsules, by each neural capsule embedding network simultaneously, converging to a final capsule layer mapping to each word in the input vector;

generating, by each neural capsule embedding network simultaneously, an output vector, wherein each output vector value:

a) identifies if a word span, of the neural capsule embedding network's assigned word span length, in the input is a named entity or not a named entity;

b) if the word span is a named entity, identifies what cluster the named entity belongs to.

2. The method of claim 1 further comprising:

before receiving, into a stack of neural capsule embedding networks, an embedding vector as input:

a) receiving as input a natural language text;

b) converting words in the natural language text into embeddings and inserting embedding tokens into an embedding vector.

3. The method of claim 2 , wherein converting words in the natural language text into embeddings includes populating with the value of zero any embeddings in the vector that do not correspond to a word.

4. The method of claim 1 , wherein generating, by each neural capsule embedding network, an output vector includes mathematical scaling of output vector values.

5. The method of claim 1 , further comprising:

after receiving, into a stack of neural capsule embedding networks, an embedding vector as input, deriving, by the neural capsule embedding network, features of each word in the context of the natural language text.

6. The method of claim 1 further comprising:

before receiving, into a stack of neural capsule embedding network, an embedding vector as input:

a) receiving as input a natural language text;

b) preprocessing the natural language text to identify features of the natural language text;

c) converting words in the natural language text into embeddings and inserting embedding tokens into an embedding vector.

7. The method of claim 1 , wherein clusters are a predefined set of named entity classes.

8. The method of claim 1 , wherein clusters are determined by the neural capsule embedding network.

9. The method of claim 1 further comprising:

after generating, by each neural capsule embedding network, an output vector, performing, by a neural network layer, mathematical scaling on the output vector values.

10. The method of claim 1 further comprising:

after generating, by each neural capsule embedding network, an output vector, combining all output vectors into a matrix.

11. The method of claim 10 , wherein the output vectors are combined into a matrix by:

for each non-zero integer value in the vector, inserting a 1 in each word span matrix cell, where the column is the position of each word of the word span and the row is the cluster number;

if at least one word span matrix contains a 1 at a cell location, inserting a 1 into a combined output matrix at that cell location.

12. The method of claim 10 , wherein the output vectors are combined into a matrix by inserting each vector into a row in the matrix, where the column is the position of each word of the word span and the row is the word span number.

13. A computer-implemented method for nested named entity recognition, comprising:

receiving, into a stack of named entity recognition models comprised of M number of flat named entity recognition models identified 1 through M, a natural language text as input, wherein each named entity recognition model is configured to identify named entities of its assigned word span length, 1 through M;

running all named entity recognition models simultaneously to generate an output for each named entity recognition model;

converting the output of each named entity recognition model to an output vector, wherein each output vector value:

a) identifies if a word span, of the named entity recognition model's assigned word span length, in the input is a named entity or not a named entity;

b) if the word span is a named entity, identifies what cluster the named entity belongs to.

14. The method of claim 13 further comprising:

after receiving, into a stack of named entity recognition models, an natural language text as input, converting words in the natural language text into an input that the named entity recognition model is configured to receive.

15. The method of claim 13 further comprising:

after converting the output of each named entity recognition model to an output vector, combining all output vectors into a matrix.

16. The method of claim 15 , wherein the output vectors are combined into a matrix by:

for each non-zero integer value in the vector, inserting a 1 in each word span matrix cell, where the column is the position of each word of the word span and the row is the cluster number;

if at least one word span matrix contains a 1 at a cell location, inserting a 1 into a combined output matrix at that cell location.

17. The method of claim 15 , wherein the output vectors are combined into a matrix by inserting each vector into a row in the matrix, where the column is the position of each word of the word span and the row is the word span number.

18. A system for nested named entity recognition, comprising at least one processor, the at least one processor configured to cause the system to at least perform:

receiving, into a stack of neural capsule embedding networks comprised of M number of flat neural capsule embedding networks identified 1 through M, an embedding vector as input, wherein:

a) the embedding vector contains embedding tokens representing words in a natural language text;

b) each neural capsule embedding networks is configured to identify named entities of its assigned word span length, 1 through M;

analyzing, by each neural capsule embedding network simultaneously, the features of each word in context of the embedding vector considering tokens to the left and right of the word;

through dynamic routing of capsules, by each neural capsule embedding network simultaneously, converging to a final capsule layer mapping to each word in the input vector;

generating, by each neural capsule embedding network simultaneously, an output vector, wherein each output vector value:

a) identifies if a word span, of the neural capsule embedding network's assigned word span length, in the input is a named entity or not a named entity;

b) if the word span is a named entity, identifies what cluster the named entity belongs to.

19. The system of claim 18 further comprising:

before receiving, into a stack of neural capsule embedding networks, an embedding vector as input:

a) receiving as input a natural language text;

b) converting words in the natural language text into embeddings and inserting embedding tokens into an embedding vector.

20. The system of claim 18 further comprising:

after generating, by each neural capsule embedding network, an output vector, combining all output vectors into a matrix.

Continuity (2)
Provisional Application 63176217 · Apr 16, 2021
Related Publication 20240193368A1 · Jun 13, 2024
References Cited (14)
US 20190354582A1 · Schäfer · 2019 [cited by examiner]
US 20200334410A1 · Yerebakan · 2020 [cited by examiner]
US 20210209356A1 · Wang · 2021 [cited by examiner]
US 20210406706A1 · Hasan · 2021 [cited by examiner]
US 20220222069A1 · Ravindranath · 2022 [cited by examiner]
WO WO2019205564A1 · 2019 [cited by examiner]
WO WO2020193966A1 · 2020 [cited by examiner]
WO WO2020261234A1 · 2020 [cited by examiner]
WO WO2021003036A1 · 2021 [cited by examiner]
WO WO2022221603A1 · 2022 [cited by examiner]
Deng, Jianfeng & Cheng, Lianglun & Wang, Zhuowei. Self-attention-based BiGRU and capsule network for named entity recognition. (Year: 2020). [cited by examiner]
Wei Zhao et al. Investigating Capsule Networks with Dynamic Routing for Text Classification. In Proceedings of the 2018 Conference on Empirical Methods in Natural Language Processing, pp. 3110-3119, Brussels, Belgium. A… [cited by examiner]
Deng, Jianfeng & Cheng, Lianglun & Wang, Zhuowei. Self-attention-based BilGRU and capsule network for named entity recognition. (Year: 2020). [cited by examiner]
Amit Kumar Jaiswal, Prayag Tiwari, Sahil Garg, M. Shamim Hossain, Entity-aware capsule network for multi-class classification of big data: A deep learning approach. Future Generation Computer Systems, vol. 117, pp. 1-11… [cited by examiner]
Cited By (1)
US 12,718,018