IP Library › Granted Patent US 12,423,385
Granted Patent B2
US 12,423,385 · App. 18/487,425 · Granted Sep 23, 2025

Automatic classification of messages based on keywords

Inventors: Ojuro Yokoyama (Kawasaki, JP); Hiroyuki Sumi (Hadano, JP); Noritoshi Yoshiyama (Yokohama, JP); Anatassios Markas (Chapel Hill, NC)
Assignee: Lenovo (Singapore) Pte. Ltd.
G06F18/24147G06F18/24765
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,423,385
App. No.
18/487,425
Granted
Sep 23, 2025
Kind
B2
Abstract

Electronic communications and a keyword are provided to a machine learning algorithm. Similarity measures are received from the machine learning algorithm. The similarity measures indicate a similarity between the communications and the keyword. The communications are clustered as a function of the similarity measures. False positive communications are removed from a first cluster as a function of a sum of distances between the false positive communication and communications in the first cluster that include the keyword and a sum of distances between the false positive communication and communications in the first cluster that do not include the keyword. False negatives are added to the first cluster as a function of a sum of distances between the false negative and communications in the first cluster that include the keyword and the sum of distances between the false negative and communications in the cluster that do not include the keyword.

Claims (39)

1. A computerized process comprising:

providing electronic communications from a processor to a machine learning algorithm;

providing a keyword to the machine learning algorithm;

receiving similarity measures from the machine learning algorithm, the similarity measures indicating a similarity between the electronic communications and the keyword;

clustering the electronic communications as a function of the similarity measures;

removing a false positive electronic communication from a first cluster as a function of a sum of distances between the false positive electronic communication and electronic communications in the first cluster that include the keyword and a sum of distances between the false positive communication and electronic communications in the first cluster that do not include the keyword; and

adding a false negative to the first cluster as a function of a sum of distances between the false negative and communications in the first cluster that include the keyword and the sum of distances between the false negative and communications in the cluster that do not include the keyword.

2. The computerized process of claim 1 , wherein the machine learning algorithm is an untrained machine learning algorithm.

3. The computerized process of claim 1 , wherein the machine learning algorithm understands ambiguous language.

4. The computerized process of claim 1 , wherein the machine learning algorithm comprises a bidirectional encoder representations from transformers (BERT) model.

5. The computerized process of claim 1 , comprising storing a portion of the electronic communications provided to the machine learning algorithm in a first vector and storing the keyword provided to the machine learning algorithm in a second vector.

6. The computerized process of claim 1 , wherein the electronic communications comprise one or more of an email, a chat and a text.

7. A non-transitory machine-readable medium comprising instructions that when executed by a processor execute a process comprising:

providing electronic communications from the processor to a machine learning algorithm;

providing a keyword to the machine learning algorithm;

receiving similarity measures from the machine learning algorithm, the similarity measures indicating a similarity between the electronic communications and the keyword;

clustering the electronic communications as a function of the similarity measures;

removing a false positive electronic communication from a first cluster as a function of a sum of distances between the false positive electronic communication and electronic communications in the first cluster that include the keyword and a sum of distances between the false positive communication and electronic communications in the first cluster that do not include the keyword; and

adding a false negative to the first cluster as a function of a sum of distances between the false negative and communications in the first cluster that include the keyword and the sum of distances between the false negative and communications in the cluster that do not include the keyword.

8. The non-transitory machine-readable medium of claim 7 , wherein the machine learning algorithm is an untrained machine learning algorithm.

9. The non-transitory machine-readable medium of claim 7 , wherein the machine learning algorithm understands ambiguous language.

10. The non-transitory machine-readable medium of claim 7 , wherein the machine learning algorithm comprises a bidirectional encoder representations from transformers (BERT) model.

11. The non-transitory machine-readable medium of claim 7 , comprising instructions for storing a portion of the electronic communications provided to the machine learning algorithm in a first vector and storing the keyword provided to the machine learning algorithm in a second vector.

12. The non-transitory machine-readable medium of claim 7 , wherein the electronic communications comprise one or more of an email, a chat and a text.

13. A system comprising:

a computer processor; and

a memory coupled to the computer processor;

wherein the computer processor and memory are operable for:

providing electronic communications from the computer processor to a machine learning algorithm;

providing a keyword to the machine learning algorithm;

receiving similarity measures from the machine learning algorithm, the similarity measures indicating a similarity between the electronic communications and the keyword;

clustering the electronic communications as a function of the similarity measures;

removing a false positive electronic communication from a first cluster as a function of a sum of distances between the false positive electronic communication and electronic communications in the first cluster that include the keyword and a sum of distances between the false positive communication and electronic communications in the first cluster that do not include the keyword; and

adding a false negative to the first cluster as a function of a sum of distances between the false negative and communications in the first cluster that include the keyword and the sum of distances between the false negative and communications in the cluster that do not include the keyword.

14. The system of claim 13 , wherein the machine learning algorithm is an untrained machine learning algorithm.

15. The system of claim 13 , wherein the machine learning algorithm understands ambiguous language.

16. The system of claim 13 , wherein the machine learning algorithm comprises a bidirectional encoder representations from transformers (BERT) model.

17. The system of claim 13 , wherein the system in operable for storing a portion of the electronic communications provided to the machine learning algorithm in a first vector and storing the keyword provided to the machine learning algorithm in a second vector.

18. The system of claim 13 , wherein the electronic communications comprise one or more of an email, a chat and a text.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 26, 2023
From: LENOVO (UNITED STATES) INC.
To: LENOVO (SINGAPORE) PTE. LTD.
Reel/Frame 066140/0696 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 16, 2023
From: YOKOYAMA, OJURO; SUMI, HIROYUKI; YOSHIYAMA, NORITOSHI; MARKAS, ANASTASSIOS
To: LENOVO (UNITED STATES) INC.
Reel/Frame 065233/0466 →
Continuity (1)
Related Publication 20250124111A1 · Apr 17, 2025
References Cited (9)
US 5377354A · Scannell · 1994 [cited by examiner]
US 8903752B1 · Pacovsk · 2014 [cited by examiner]
US 10621507B2 · Venkataraman · 2020 [cited by examiner]
US 11368423B1 · Plater-Zyberk · 2022 [cited by examiner]
US 11720842B2 · Galginaitis · 2023 [cited by examiner]
US 20050060643A1 · Glass · 2005 [cited by examiner]
US 20180300315A1 · Leal · 2018 [cited by examiner]
Pelleg, Dan, “X-means: Extending K-means with Efficient Estimation of the Number of Clusters”, 8 pgs. [cited by applicant]
Peng, Sancheng, “A survey on deep learning for textual emotion analysis in social networks”, Digital Communications and Networks 8 (2022) 745-762, (Oct. 14, 2021), 18 pgs. [cited by applicant]