IP Library Granted Patent US 12,153,878
Granted Patent B2
US 12,153,878 · App. 17/718,856 · Granted Nov 26, 2024

Intent detection via multi-hop unified syntactic graph

Inventors: Xuchao Zhang (Elkridge, MD); Yanchi Liu (Monmouth Junction, NJ); Haifeng Chen (West Windsor, NJ)
Assignee: NEC Corporation
G06F40/211G06F16/288G06F40/284G06F40/30
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,153,878
App. No.
17/718,856
Granted
Nov 26, 2024
Kind
B2
Abstract

A method for detecting business intent from a business intent corpus by employing an Intent Detection via Multi-hop Unified Syntactic Graph (IDMG) is presented. The method includes parsing each text sample representing a business need description to extract syntactic information including at least tokens and words, tokenizing the words of the syntactic information to generate sub-words for each of the words by employing a multi-lingual pre-trained language model, aligning the generated sub-words to the tokens of the syntactic information to match ground-truth intent actions and objects to the tokenized sub-words, generating a unified syntactic graph, encoding, via a multi-hop unified syntactic graph encoder, the unified syntactic graph to generate an output, and predicting an intent action and object from the output.

Claims (42)

1. A method for detecting business intent from a business intent corpus by employing an Intent Detection via Multi-hop Unified Syntactic Graph (IDMG), the method comprising:

parsing each text sample representing a business need description to extract syntactic information including at least tokens and words;

tokenizing the words of the syntactic information to generate sub-words for each of the words by employing a multi-lingual pre-trained language model;

aligning the generated sub-words to the tokens of the syntactic information to match ground-truth intent actions and objects to the tokenized sub-words;

generating a unified syntactic graph;

encoding, via a multi-hop unified syntactic graph encoder, the unified syntactic graph to generate an output; and

predicting an intent action and object from the output.

2. The method of claim 1 , wherein the syntactic information further includes at least part-of-speech for each of the words, dependency relation, and abstract meaning relation.

3. The method of claim 1 , wherein generation of the unified syntactic graph includes token node construction, node relation construction, and inter-sentence relation construction.

4. The method of claim 1 , wherein the multi-hop unified syntactic graph encoder includes a syntactic structure attention component and a multi-head semantic attention component.

5. The method of claim 4 , wherein the syntactic structure attention component receives edge relations and node attributes as input from the unified syntactic graph and the multi-head semantic attention component receives sequential embeddings as input.

6. The method of claim 5 , wherein, if a first node and a second node are directly connected, an edge embedding is initialized as a type of dependency relation and fine-tuned during a training process.

7. The method of claim 5 , wherein if a first node and a second node are not directly connected, an edge embedding is a sum of all edge segments.

8. A non-transitory computer-readable storage medium comprising a computer-readable program for detecting business intent from a business intent corpus by employing an Intent Detection via Multi-hop Unified Syntactic Graph (IDMG), wherein the computer-readable program when executed on a computer causes the computer to perform the steps of:

parsing each text sample representing a business need description to extract syntactic information including at least tokens and words;

tokenizing the words of the syntactic information to generate sub-words for each of the words by employing a multi-lingual pre-trained language model;

aligning the generated sub-words to the tokens of the syntactic information to match ground-truth intent actions and objects to the tokenized sub-words;

generating a unified syntactic graph;

encoding, via a multi-hop unified syntactic graph encoder, the unified syntactic graph to generate an output; and

predicting an intent action and object from the output.

9. The non-transitory computer-readable storage medium of claim 8 , wherein the syntactic information further includes at least part-of-speech for each of the words, dependency relation, and abstract meaning relation.

10. The non-transitory computer-readable storage medium of claim 8 , wherein generation of the unified syntactic graph includes token node construction, node relation construction, and inter-sentence relation construction.

11. The non-transitory computer-readable storage medium of claim 8 , wherein the multi-hop unified syntactic graph encoder includes a syntactic structure attention component and a multi-head semantic attention component.

12. The non-transitory computer-readable storage medium of claim 11 , wherein the syntactic structure attention component receives edge relations and node attributes as input from the unified syntactic graph and the multi-head semantic attention component receives sequential embeddings as input.

13. The non-transitory computer-readable storage medium of claim 12 , wherein, if a first node and a second node are directly connected, an edge embedding is initialized as a type of dependency relation and fine-tuned during a training process.

14. The non-transitory computer-readable storage medium of claim 12 , wherein if a first node and a second node are not directly connected, an edge embedding is a sum of all edge segments.

15. A system for detecting business intent from a business intent corpus by employing an Intent Detection via Multi-hop Unified Syntactic Graph (IDMG), the system comprising:

a memory; and

one or more processors in communication with the memory configured to:

parse each text sample representing a business need description to extract syntactic information including at least tokens and words;

tokenize the words of the syntactic information to generate sub-words for each of the words by employing a multi-lingual pre-trained language model;

align the generated sub-words to the tokens of the syntactic information to match ground-truth intent actions and objects to the tokenized sub-words;

generate a unified syntactic graph;

encode, via a multi-hop unified syntactic graph encoder, the unified syntactic graph to generate an output; and

predict an intent action and object from the output.

16. The system of claim 15 , wherein the syntactic information further includes at least part-of-speech for each of the words, dependency relation, and abstract meaning relation.

17. The system of claim 15 , wherein generation of the unified syntactic graph includes token node construction, node relation construction, and inter-sentence relation construction.

18. The system of claim 15 , wherein the multi-hop unified syntactic graph encoder includes a syntactic structure attention component and a multi-head semantic attention component.

19. The system of claim 18 , wherein the syntactic structure attention component receives edge relations and node attributes as input from the unified syntactic graph and the multi-head semantic attention component receives sequential embeddings as input.

20. The system of claim 19 ,

wherein, if a first node and a second node are directly connected, an edge embedding is initialized as a type of dependency relation and fine-tuned during a training process; and

wherein if the first node and the second node are not directly connected, the edge embedding is a sum of all edge segments.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 23, 2024
From: NEC LABORATORIES AMERICA, INC.
To: NEC CORPORATION
Reel/Frame 068981/0866 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 12, 2022
From: ZHANG, XUCHAO; LIU, YANCHI; CHEN, HAIFENG
To: NEC LABORATORIES AMERICA, INC.
Reel/Frame 059575/0821 →
Continuity (2)
Provisional Application 63174716 · Apr 14, 2021
Related Publication 20220343068A1 · Oct 27, 2022