IP Library › Granted Patent US 12,387,059
Granted Patent B2
US 12,387,059 · App. 17/554,606 · Granted Aug 12, 2025

Identifying zones of interest in text transcripts using deep learning

Inventors: Akshay Aravindakshan Thoniparambil (Bangalore, IN); Naman Gupta (Jaipur, IN); Manish Agarwal (Bangalore, IN); Sourav Choudhary (Samastipur, IN)
Assignee: Capital One Services, LLC
G06F40/58G06F16/3344G06F16/93G06F40/216G06N3/08
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,387,059
App. No.
17/554,606
Granted
Aug 12, 2025
Kind
B2
Abstract

Systems, methods, and computer program products for identifying zones of interest in text transcripts. An application may receive input specifying a text statement type and determine a plurality of heuristics for identifying statements of the statement type in transcripts. The application may determine, based on a first heuristic, a first text statement of the statement type. The application may generate, based on a clustering algorithm, a plurality of additional statements of the statement type. The application may receive a first text transcript. The application may identify, based on a second heuristic, a first text statement in the first text transcript, where the first text transcript statement is of the statement type. The application may generate a graphical indication that the first text transcript statement is of the statement type, and display the first transcript statement and the graphical indication on a display.

Claims (70)

1. A method, comprising:

receiving, by an application executing on a processor, input specifying a text statement type;

determining, by the application based on the text statement type, a plurality of heuristics for identifying text statements of the text statement type in a plurality of text transcripts;

determining, by the application based on a first heuristic of the plurality of heuristics, a first text statement of the text statement type;

generating, by the application based on a clustering algorithm and the first text statement, a plurality of additional text statements of the text statement type;

assigning, by a classification model, a classification tag to each of the plurality of additional text statements, the classification tag indicating that each of the plurality of additional text statements is of the text statement type, wherein the classification model is trained based on a training dataset comprising a plurality of training terms, wherein one or more of the plurality of training terms is masked in the training dataset based on a respective frequency each training term appears among a plurality of terms in the plurality of text transcripts and whether the respective frequency exceeds a threshold;

receiving, by the application based on the first text statement and the plurality of additional text statements, a first text transcript of the plurality of text transcripts;

identifying, by the application based on a second heuristic of the plurality of heuristics, a first text transcript statement in the first text transcript, wherein the first text transcript statement is of the text statement type;

generating a graphical indication that the first text transcript statement is of the text statement type;

displaying at least the first text transcript statement and the graphical indication on a display;

the method further comprising:

determining, by the application based on a third heuristic of the plurality of heuristics, a plurality of seed words related to the text statement type; and

generating, by the application based on the clustering algorithm and the plurality of seed words, the plurality of additional text statements of the text statement type; and

the method further comprising, prior to receiving the input:

receiving, by the application, the training dataset comprising the plurality of training terms;

determining, by the application, the respective frequency each training term appears among the plurality of terms in the plurality of text transcripts;

classifying, by the application, a first subset of the plurality of training terms as being associated with the text statement type based on the respective frequencies of the first subset exceeding the threshold;

classifying, by the application, a second subset of the plurality of training terms as not being associated with the text statement type based on the respective frequencies of the second subset not exceeding the threshold; and

masking the first subset of the plurality of training terms from the training dataset.

2. The method of claim 1 , wherein the first text transcript statement is further identified based on one or more of: (i) a length of the first text transcript statement, (ii) whether a predetermined phrase is present in the first text transcript statement, (iii) whether the first text transcript statement is associated with a customer or a customer support agent, (iv) a sentiment score computed for the first text transcript statement, and (v) whether a term of the first text transcript statement is associated with a negative sentiment.

3. The method of claim 1 , wherein the text statement types comprise one or more of: (i) a reason a customer associated with the text transcript contacted a customer support agent, (ii) a call resolution associated with the customer support agent, (iii) a personal story associated with the customer, and (iv) a negative statement made by the customer.

4. The method of claim 1 , wherein the first text transcript is identified based on one or more statements of the first text transcript matching the first text statement or at least one of the plurality of additional text statements, wherein the graphical indication comprises one or more of: (i) a symbol, (ii) a highlighting effect applied to the first text transcript statement, (iii) a bold effect applied to the first text transcript statement, or (iv) an italics effect applied to the first text transcript statement.

5. The method of claim 1 , wherein the text statement type is defined as being separate from a statement in the plurality of text transcripts.

6. A non-transitory computer-readable storage medium, the computer-readable storage medium including instructions that when executed by a processor, cause the processor to:

receive, by an application, input specifying a text statement type;

determine, by the application based on the text statement type, a plurality of heuristics for identifying statements of the text statement type in a plurality of text transcripts;

determine, by the application based on a first heuristic of the plurality of heuristics, a first text statement of the text statement type;

generate, by the application based on a clustering algorithm and the first text statement, a plurality of additional text statements of the text statement type;

assign, by a classification model, a classification tag to each of the plurality of additional text statements, the classification tag indicating that each of the plurality of additional text statements is of the text statement type, wherein the classification model is trained based on a training dataset comprising a plurality of training terms, wherein one or more of the plurality of training terms is masked in the training dataset based on a respective frequency each training term appears among a plurality of terms in the plurality of text transcripts and whether the respective frequency exceeds a threshold;

receive, by the application based on the first text statement and the plurality of additional text statements, a first text transcript of the plurality of text transcripts;

identify, by the application based on a second heuristic of the plurality of heuristics, a first text transcript statement in the first text transcript, wherein the first text transcript statement is of the text statement type;

generate a graphical indication that the first text transcript statement is of the text statement type;

display at least the first text transcript statement and the graphical indication on a display;

the instructions further causing the processor to:

determine, by the application based on a third heuristic of the plurality of heuristics, a plurality of seed words related to the text statement type; and

generate, by the application based on the clustering algorithm and the plurality of seed words, the plurality of additional text statements of the text statement type; and

the instructions further causing the processor to, prior to receiving the input:

receive, by the application, the training dataset comprising the plurality of training terms;

determine, by the application, the respective frequency each training term appears among a plurality of terms in the plurality of text transcripts;

classify, by the application, a first subset of the plurality of training terms as being associated with the text statement type based on the respective frequencies of the first subset exceeding the threshold;

classify, by the application, a second subset of the plurality of training terms as not being associated with the text statement type based on the respective frequencies of the second subset not exceeding the threshold; and

mask the first subset of the plurality of training terms from the training dataset.

7. The computer-readable storage medium of claim 6 , wherein the first text transcript statement is further identified based on one or more of: (i) a length of the first text transcript statement, (ii) whether a predetermined phrase is present in the first text transcript statement, (iii) whether the first text transcript statement is associated with a customer or a customer support agent, (iv) a sentiment score computed for the first text transcript statement, and (v) whether a term of the first text transcript statement is associated with a negative sentiment.

8. The computer-readable storage medium of claim 6 , wherein the text statement types comprise one or more of: (i) a reason a customer associated with the text transcript contacted a customer support agent, (ii) a call resolution associated with the customer support agent, (iii) a personal story associated with the customer, and (iv) a negative statement made by the customer.

9. The computer-readable storage medium of claim 6 , wherein the first text transcript is identified based on one or more statements of the first text transcript matching the first text statement or at least one of the plurality of additional text statements, wherein the graphical indication comprises one or more of: (i) a symbol, (ii) a highlighting effect applied to the first text transcript statement, (iii) a bold effect applied to the first text transcript statement, or (iv) an italics effect applied to the first text transcript statement.

10. The computer-readable storage medium of claim 6 , wherein the text statement type is defined as being separate from a statement in the plurality of text transcripts.

11. A computing apparatus comprising:

a processor; and

a memory storing instructions that, when executed by the processor, cause the processor to:

receive, by an application, input specifying a text statement type;

determine, by the application based on the text statement type, a plurality of heuristics for identifying statements of the text statement type in a plurality of text transcripts;

determine, by the application based on a first heuristic of the plurality of heuristics, a first text statement of the text statement type;

generate, by the application based on a clustering algorithm and the first text statement, a plurality of additional text statements of the text statement type;

assign, by a classification model, a classification tag to each of the plurality of additional text statements, the classification tag indicating that each of the plurality of additional text statements is of the text statement type, wherein the classification model is trained based on a training dataset comprising a plurality of training terms, wherein one or more of the plurality of training terms is masked in the training dataset based on a respective frequency each training term appears among a plurality of terms in the plurality of text transcripts and whether the respective frequency exceeds a threshold;

receive, by the application based on the first text statement and the plurality of additional text statements, a first text transcript of the plurality of text transcripts;

identify, by the application based on a second heuristic of the plurality of heuristics, a first text transcript statement in the first text transcript, wherein the first text transcript statement is of the text statement type;

generate a graphical indication that the first text transcript statement is of the text statement type;

display at least the first text transcript statement and the graphical indication on a display;

the instructions further causing the processor to:

determine, by the application based on a third heuristic of the plurality of heuristics, a plurality of seed words related to the text statement type; and

generate, by the application based on the clustering algorithm and the plurality of seed words, the plurality of additional text statements of the text statement type; and

the instructions further causing the processor to, prior to receiving the input:

receive, by the application, the training dataset comprising the plurality of training terms;

determine, by the application, a respective frequency each training term appears among the plurality of terms in the plurality of text transcripts;

classify, by the application, a first subset of the plurality of training terms as being associated with the text statement type based on the respective frequencies of the first subset exceeding the threshold;

classify, by the application, a second subset of the plurality of training terms as not being associated with the text statement type based on the respective frequencies of the second subset not exceeding the threshold; and

mask the first subset of the plurality of training terms from the training dataset.

12. The computing apparatus of claim 11 , wherein the first text transcript statement is further identified based on one or more of: (i) a length of the first text transcript statement, (ii) whether a predetermined phrase is present in the first text transcript statement, (iii) whether the first text transcript statement is associated with a customer or a customer support agent, (iv) a sentiment score computed for the first text transcript statement, and (v) whether a term of the first text transcript statement is associated with a negative sentiment.

13. The computing apparatus of claim 11 , wherein the text statement types comprise one or more of: (i) a reason a customer associated with the text transcript contacted a customer support agent, (ii) a call resolution associated with the customer support agent, (iii) a personal story associated with the customer, and (iv) a negative statement made by the customer, wherein the first text transcript is identified based on one or more statements of the first text transcript matching the first text statement or at least one of the plurality of additional text statements, wherein the graphical indication comprises one or more of: (i) a symbol, (ii) a highlighting effect applied to the first text transcript statement, (iii) a bold effect applied to the first text transcript statement, or (iv) an italics effect applied to the first text transcript statement.

14. The computing apparatus of claim 11 , wherein the text statement type is defined as being separate from a statement in the plurality of text transcripts.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 17, 2021
From: THONIPARAMBIL, AKSHAY ARAVINDAKSHAN; GUPTA, NAMAN; AGARWAL, MANISH; CHOUDHARY, SOURAV
To: CAPITAL ONE SERVICES, LLC
Reel/Frame 058419/0832 →
Continuity (1)
Related Publication 20230196035A1 · Jun 22, 2023
References Cited (35)
US 7606714B2 · Williams · 2009 [cited by examiner]
US 9477752B1 · Romano · 2016 [cited by examiner]
US 9767165B1 · Tacchi · 2017 [cited by applicant]
US 9779081B2 · Simard et al. · 2017 [cited by applicant]
US 9922025B2 · Cross et al. · 2018 [cited by applicant]
US 10754883B1 · Kannu · 2020 [cited by examiner]
US 11023675B1 · Neervannan · 2021 [cited by applicant]
US 11227183B1 · Connors et al. · 2022 [cited by applicant]
US 11294974B1 · Shukla et al. · 2022 [cited by applicant]
US 20040210443A1 · Kuhn · 2004 [cited by examiner]
US 20120259801A1 · Ji · 2012 [cited by examiner]
US 20130346424A1 · Zhang et al. · 2013 [cited by applicant]
US 20170353605A1 · Dumaine · 2017 [cited by applicant]
US 20180121539A1 · Ciulla · 2018 [cited by applicant]
US 20190139551A1 · Steelberg · 2019 [cited by examiner]
US 20190163817A1 · Milenova · 2019 [cited by applicant]
US 20200302011A1 · Mishra · 2020 [cited by applicant]
US 20200410012A1 · Moon et al. · 2020 [cited by applicant]
US 20210089971A1 · Grabau et al. · 2021 [cited by applicant]
US 20210097472A1 · Inamdar et al. · 2021 [cited by applicant]
US 20210157990A1 · Lima · 2021 [cited by examiner]
US 20210158234A1 · Sivasubramanian · 2021 [cited by examiner]
US 20220230116A1 · Dubey · 2022 [cited by applicant]
US 20220293107A1 · Leaman · 2022 [cited by examiner]
US 20220318485A1 · Narayanan · 2022 [cited by applicant]
US 20220365955A1 · Ramamohan · 2022 [cited by applicant]
Gildea et al. “Topic-based Language Models using EM”. Sixth European Conference on Speech Communication and Technology, 1999 (Year: 1999). [cited by examiner]
Levine et al. “PMI-Masking: Principled Masking of Correlated Spans”. arXiv:2010.01825v1 [cs.LG] Oct. 5, 2020 (Year: 2020). [cited by examiner]
Devlin et al., “BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding”, Cornell University [online], Submitted Oct. 11, 2018, 16 pages. [cited by applicant]
Author Unknown, “Okapi BM25”, Wikipedia the Free Encyclopedia [online], Retrieved from the Internet URL: <https://en.wikipedia.org/wiki/Okapi_BM25>, Retrieved on May 12, 2021, 4 pages. [cited by applicant]
Pagliardini et al., “Unsupervised Learning of Sentence Embeddings using Compositional n-Gram Features”, NAACL 2018—Conference of the North American Chapter of the Association for Computational Linguistics, pp. 528-540. [cited by applicant]
Author Unknown, “DBSCAN”, Wikipedia [online], Retrieved from Internet URL:<https://en.wikipedia.org/wiki/DBSCAN>, Retrieved on Dec. 14, 2021, 8 pages. [cited by applicant]
Author Unknown, “Sentiment analysis”, Wikipedia [online], Retrieved from Internet URL:<https://en.wikipedia.org/wiki/Sentiment_analysis', Retrieved on Dec. 14, 2021, 16 pages. [cited by applicant]
Author Unknown, “How to Search Chat History”, Published Jun. 3, 2021, Retrieved from Internet URL:<https://help.livehelpnow.net/1/kb/article/1584/how-to-search-chat-history>, Retrieved on Dec. 14, 2021, 2 pages. [cited by applicant]
Author Unkown, “Transcripts—Accessing transcripts”, Olark [online], Retrieved from Internet URL:<https://www.olark.com/help/view-transcripts/>, Retrieved on Dec. 14, 2021, 8 pages. [cited by applicant]