IP Library Granted Patent US 12,413,664
Granted Patent B2
US 12,413,664 · App. 18/223,815 · Granted Sep 9, 2025

Identification and prevention of sensitive information exposure in telephonic conversations

Inventors: Manu K M (Bangalore, IN); Balaji Sankar Kumar (Bangalore, IN); Vidya Chandrashekar (Bangalore, IN); vamshi dondapati (Hyderabad, IN); Arun Aravind (Mahe, IN); Akshat Dixit (Lucknow, IN); Arun Sabaresh Anantha Narayanan (Bangalore, IN)
Assignee: INTERNATIONAL BUSINESS MACHINES CORPORATION
H04M3/2281G10L15/1815H04L9/088H04L9/30H04M3/2218H04M2201/40
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,413,664
App. No.
18/223,815
Granted
Sep 9, 2025
Kind
B2
Abstract

An embodiment converts, by a voice-to-text converter, voice data to text data. The embodiment extracts, by an extractor, an intent and an entity from the text data. The embodiment predicts, by a predictor, based on the intent and the entity, a sensitive information. The embodiment compares, by an intersector, the text data to the predicted sensitive information. The embodiment determines, by the intersector, responsive to the comparing, whether the text data includes the predicted sensitive information. The embodiment intersects, by the intersector, responsive to a determination that the text data includes the predicted sensitive information, the voice data.

Claims (60)

1. A computer-implemented method comprising:

converting to text data in real-time, by a voice-to-text converter, voice data of a conversation occurring in real-time between a first party and a second party;

extracting, by an extractor, an intent and an entity from the text data, wherein the intent comprises a request by the first party to the second party, and wherein the entity comprises a category of information being requested;

predicting, by a predictor, based on the request and the category of information being requested, a sensitive information that is likely to be disclosed in the real-time conversation;

comparing, by an intersector, the text data in real-time to the predicted sensitive information;

determining, by the intersector, responsive to the comparing, whether the text data includes the predicted sensitive information; and

intersecting, by the intersector, responsive to a determination that the text data includes the predicted sensitive information, the voice data.

2. The method of claim 1 , wherein the voice data is associated with a speaker of a telephonic conversation, further comprising:

intersecting the voice data includes preventing a transmission of the voice data to a listener in the telephonic conversation.

3. The method of claim 1 , further comprising:

retrieving a key corresponding to an identifier associated with the sensitive information, and a value corresponding to the sensitive information; and

predicting the sensitive information by determining whether the text data matches at least one of the key or the value.

4. The method of claim 3 , further comprising:

decomposing the predicted sensitive information into a key-value pair; and

storing the key-value pair.

5. The method of claim 1 , further comprising generating a sensitive information warning.

6. The method of claim 1 , further comprising generating a speaker agreement.

7. The method of claim 1 , further comprising:

determining a risk score associated with a telephonic conversation; and

generating a risk warning, responsive to a determination that the risk score meets a predetermined threshold.

8. The method of claim 7 , wherein the risk score is based on at least one of a call history, a call frequency, and a call time.

9. The method of claim 1 , further comprising:

deactivating, responsive to a user selection, at least one of the voice-to-text converter, the extractor, the predictor, and the intersector.

10. The method of claim 1 , further comprising:

comparing the entity to a key in a first key-value pair, wherein the entity is extracted from a first portion of voice data;

anticipating, responsive to the entity matching the key, that a potential voice response to the first portion of the voice data has a likelihood of including a first sensitive information corresponding to the first key-value pair; and

causing, responsive to the anticipating, a manipulation of a communication channel carrying the voice data.

11. The method of claim 1 , further comprising:

comparing a first portion of the text data with a first key-value pair;

anticipating, responsive to the first portion of the text data at least partially matching a portion of the first key-value pair, that an upcoming portion of the voice data has a likelihood of including a first sensitive information corresponding to the first key-value pair; and

causing a response to the anticipation, a manipulation of a communication channel carrying voice data.

12. A computer program product comprising one or more computer readable storage media, and program instructions collectively stored on the one or more computer readable storage media, the program instructions executable by a processor to cause the processor to perform operations comprising:

converting to text data in real-time, by a voice-to-text converter, voice data of a conversation occurring in real-time between a first party and a second party;

extracting, by an extractor, an intent and an entity from the text data, wherein the intent comprises a request by the first party to the second party, and wherein the entity comprises a category of information being requested;

predicting, by a predictor, based on the request and the category of information being requested, a sensitive information that is likely to be disclosed in the real-time conversation;

comparing, by an intersector, the text data in real-time to the predicted sensitive information;

determining, by the intersector, responsive to the comparing, whether the text data includes the predicted sensitive information; and

intersecting, by the intersector, responsive to a determination that the text data includes the predicted sensitive information, the voice data.

13. The computer program product of claim 12 , wherein the voice data is associated with a speaker of a telephonic conversation, further comprising:

intersecting the voice data includes preventing a transmission of the voice data to a listener in the telephonic conversation.

14. The computer program product of claim 12 , further comprising:

retrieving a key corresponding to an identifier associated with the sensitive information, and a value corresponding to the sensitive information; and

predicting the sensitive information by determining whether the text data matches at least one of the key or the value.

15. The computer program product of claim 14 , further comprising:

decomposing the predicted sensitive information into a key-value pair; and

storing the key-value pair.

16. The computer program product of claim 12 , further comprising generating a sensitive information warning.

17. The computer program product of claim 12 , further comprising generating a speaker agreement.

18. A computer system comprising a processor and one or more computer readable storage media, and program instructions collectively stored on the one or more computer readable storage media, the program instructions executable by the processor to cause the processor to perform operations comprising:

converting to text data in real-time, by a voice-to-text converter, voice data of a conversation occurring in real-time between a first party and a second party;

extracting, by an extractor, an intent and an entity from the text data, wherein the intent comprises a request by the first party to the second party, and wherein the entity comprises a category of information being requested;

predicting, by a predictor, based on the request and the category of information being requested, a sensitive information that is likely to be disclosed in the real-time conversation;

comparing, by an intersector, the text data in real-time to the predicted sensitive information;

determining, by the intersector, responsive to the comparing, whether the text data includes the predicted sensitive information; and

intersecting, by the intersector, responsive to a determination that the text data includes the predicted sensitive information, the voice data.

19. The computer system of claim 18 , wherein the voice data is associated with a speaker of a telephonic conversation, further comprising:

intersecting the voice data includes preventing a transmission of the voice data to a listener in the telephonic conversation.

20. The computer system of claim 18 , further comprising:

retrieving a key corresponding to an identifier associated with the sensitive information, and a value corresponding to the sensitive information; and

predicting the sensitive information by determining whether the text data matches at least one of the key or the value.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jul 19, 2023
From: K M, MANU; KUMAR, BALAJI SANKAR; CHANDRASHEKAR, VIDYA; DONDAPATI, VAMSHI; ARAVIND, ARUN; DIXIT, AKSHAT; ANANTHA NARAYANAN, ARUN SABARESH
To: INTERNATIONAL BUSINESS MACHINES CORPORATION
Reel/Frame 064315/0055 →
Continuity (1)
Related Publication 20250030795A1 · Jan 23, 2025
References Cited (25)
US 8706486B1 · Devarajan et al. · 2014 [cited by applicant]
US 9443005B2 · Khandekar · 2016 [cited by applicant]
US 10469663B2 · Milstein et al. · 2019 [cited by applicant]
US 10747894B1 · Cline · 2020 [cited by examiner]
US 10778839B1 · Newstadt et al. · 2020 [cited by applicant]
US 11423018B1 · Paiz · 2022 [cited by examiner]
US 11755756B1 · Cline · 2023 [cited by examiner]
US 20090254971A1 · Herz · 2009 [cited by examiner]
US 20130266127A1 · Schachter et al. · 2013 [cited by applicant]
US 20160219024A1 · Verzun et al. · 2016 [cited by applicant]
US 20190026494A1 · Smith · 2019 [cited by examiner]
US 20210241607A1 · Rhoads et al. · 2021 [cited by applicant]
US 20220122628A1 · McCloskey et al. · 2022 [cited by applicant]
US 20220164472A1 · Cannon · 2022 [cited by examiner]
US 20220350825A1 · van de Nieuwegiessen · 2022 [cited by examiner]
US 20220366904A1 · Martinson et al. · 2022 [cited by applicant]
US 20220399009A1 · Okada et al. · 2022 [cited by applicant]
CN 111740951A · 2020 [cited by applicant]
TW 202307644A · 2023 [cited by applicant]
TW 202509910A · 2025 [cited by applicant]
WO 2020117504A1 · 2020 [cited by applicant]
Petracca et al., AuDroid: Preventing Attacks on Audio Channels in Mobile Devices, Apr. 1, 2016. [cited by applicant]
ip.com, Flexible and Effective Method for Sensitive Information Detection for Both Structured and Unstructured Data Using Two Pipelined Models, Jan. 10, 2023. [cited by applicant]
ip.com, Detection and Warning System For Sensitive Information Revealed During Phone Calls, Aug. 19, 2020. [cited by applicant]
Taiwan Patent Office, “First Office Action,” Apr. 10, 2025, 16 Pages, TW Application No. 113126517. [cited by applicant]
Cited By (1)
US 12,614,034