SYSTEMS AND METHODS FOR REDACTION OF SENSITIVE INFORMATION FROM VOICE COMMUNICATIONS
The redaction system and methods use primary triggers and secondary triggers to initiate redaction capabilities. A primary trigger is initiated when sensitive data is first detected and acts forward in time from the detection event. A secondary trigger is initiated at the end or during the timing of the first trigger and acts both backwards and forwards in time from its initiation. Combining the detection results from these two triggers allows confidence that all sensitive data is removed while avoiding false positives.
1 . A method of redacting voice communications, comprising:
receiving a recorded voice communication at a redaction system;
analyzing the recorded voice communication using a plurality of primary triggers to identify a first data set of sensitive information and a corresponding first set of time slices;
analyzing the recorded voice communication using a plurality of secondary triggers to identify a second data set of sensitive information and a corresponding second set of time slices;
combining the first set of time slices and the second set of time to determine a combined set of time slices;
redacting any audio data from the recorded voice communication occurring within the combined set of time slices; and
storing the redacted voice communication in a voice communication database of the redaction system.
2 . The method according to claim 1 , further comprising;
generating a text transcript of the recorded voice communication;
redacting any text in the text transcript corresponding to the audio data from the recorded voice communication occurring within the combined set of time slices; and
storing the text transcript in a transcript database.
3 . The method according to claim 1 , wherein the text transcript comprises an identifier to identify the redacted voice communication stored in the voice communication database.
4 . The method according to claim 1 , wherein the redacting comprises blanking, obfuscating, or cutting audio data from the recorded voice communication occurring within the combined set of time slices.
5 . The method according to claim 1 , wherein at least one secondary trigger of the plurality of secondary triggers is activated only if a corresponding primary trigger of the plurality of primary triggers fails to identify sensitive content.
6 . The method according to claim 1 , wherein the plurality of secondary triggers are different than the plurality of primary triggers.
7 . The method according to claim 1 , wherein each primary trigger of the plurality of primary triggers checks for sensitive content by identifying a first detection event and analyzing audio data in the recorded voice communication for a first predetermined time period after the first detection event.
8 . The method according to claim 7 , wherein audio data occurring before the detection event is not checked for each primary trigger of the plurality of primary triggers.
9 . The method according to claim 7 , wherein each secondary trigger of the plurality of second triggers checks for sensitive content by identifying a second detection event and analyzing audio data in the recorded voice communication for a second predetermined time period after the second detection event and analyzing audio data in the recorded voice communication for a third predetermined time period before the second detection event.
10 . The method according to claim 7 , wherein the first detection event is recognition of a start of audio data associated with sensitive content.
11 . The method according to claim 10 , wherein the first detection event is an identification of a phrase or number combination associated with sensitive content.
12 . The method according to claim 11 , wherein a time slice associated with the first detection event is set to a beginning of the first detection event.
13 . The method according to claim 11 , wherein a time slice associated with the detection event is set to an end of the first detection event.
14 . The method according to claim 1 , further comprising:
encrypting the redacted voice communication prior to storage.
15 . The method according to claim 1 , further comprising:
storing an encrypted version of the recorded voice communication in the voice communication database in association with the redacted voice communication database.
16 . The method according to claim 1 , further comprising:
analyzing audio data within the combined set of time slices to identify a third set of time slices comprising non-sensitive data; and
removing the third set of time slices from the combined set of time slices prior to redacting the recorded voice communication.