IP Library Granted Patent US 11,114,092
Granted Patent B2
US 11,114,092 · App. 16/845,300 · Granted Sep 7, 2021

Real-time voice processing systems and methods

Inventor: Romain Sambarino (Paris, FR)
Assignee: Groupe Allo Media SAS
G10L15/183G10L15/26G10L15/32H04M3/5175
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,114,092
App. No.
16/845,300
Granted
Sep 7, 2021
Kind
B2
Abstract

A computer-implemented method and supporting system transcribes spoken words being monitored from a telephonic interaction among two or more individuals. Telephonic interactions among the individuals are monitored, and at least two of the individuals are each assigned to a separate channel. While still being monitored, each of the channels is assigned a context-based speech recognition models, and in substantially real-time, the monitored telephonic interaction is transcribed from speech to text based on the different assigned models.

Claims (31)

1. A computer-implemented method for protecting access to

information while transcribing spoken words to text, the method comprising:

electronically monitoring a telephonic interaction, the interaction comprising at least two different channels;

while monitoring the telephonic interaction, assigning one of a plurality of context-based speech recognition models to at least one of the two different channels wherein the one of the plurality of context-based speech recognition models is trained to recognize protected information and the plurality of context-based speech recognition models are organized in a multi-level hierarchical structure;

for at least one of the two different channels, swapping the assigned context-based speech recognition model with an alternate context-based speech recognition model while still monitoring the telephonic interaction and, wherein the speech from one of the other of the two different channels comprises a question, and wherein the alternate context-based speech recognition model comprises a fixed set of possible answers for the question; and

as the text is being transcribed, transcribing the monitored telephonic interaction from speech to text based on the assigned model such that the protected information is obfuscated.

2. The method of claim 1 , wherein one of the two different channels represents a service consumer and the other one of the two different channels represents a service provider.

3. The method of claim 1 , wherein the protected information is subject to legal protections under one or more statutes.

4. The method of claim 3 , wherein the one or more statutes comprise the General Data Protection Regulation (“GDPR”), the Health Insurance Portability and Accountability Act (“HIPAA”), or the Family Educational Rights and Privacy Act (“FERPA”).

5. The method of claim 1 , wherein a subset of the plurality of context-based speech recognition models include role-specific variants such that the assignment of one of the plurality of context-based speech recognition models to one of the channels is based in part on a transaction role of one of the channels.

6. The method of claim 1 , wherein the protected information comprises financial account data.

7. The method of claim 1 , wherein the protected information comprises healthcare data.

8. The method of claim 1 , wherein the alternate context-based speech recognition model assigned to at least one of the two different channels is based at least in part on speech from one of the other of the two different channels.

9. The method of claim 1 , further comprising presenting, via an electronic user interface and while monitoring the telephonic interaction, the transcribed text to a user operating as one of the channels such that the protected information cannot be seen by the user.

10. A system for protecting information while transcribing spoken words to text, the system comprising:

at least one memory for storing computer-executable instructions; and

at least one processor for executing the instructions stored on the memory, wherein execution of the instructions programs the at least one processor to perform operations comprising:

electronically monitoring a telephonic interaction, the interaction comprising at least two different channels;

while monitoring the telephonic interaction, assigning one of a plurality of context-based speech recognition models to at least one of the two different channels wherein the one of the plurality of context-based speech recognition models is trained to recognize protected information and the plurality of context-based speech recognition models are organized in a multi-level hierarchical structure;

for at least one of the two different channels, swapping the assigned context-based speech recognition model with an alternate context-based speech recognition model while still monitoring the telephonic interaction, and wherein the speech from one of the other of the two different channels comprises a question, and wherein the alternate context-based speech recognition model comprises a fixed set of possible answers for the question; and

as the text is being transcribed, transcribing the monitored telephonic interaction from speech to text based on the assigned model such that the protected information is obfuscated.

11. The system of claim 10 , further comprising a data storage module for storing the transcribed text.

12. The system of claim 10 , further comprising a user interface for presenting the transcribed text to a user while the user participates in the telephonic interaction.

13. The system of claim 10 , further comprising a model storage module for storing the plurality of context-based speech recognition models.

14. The system of claim 10 , wherein one of the two different channels represents a service consumer and the other one of the two different channels represents a service provider.

15. The system of claim 10 , wherein the protected information is subject to legal protections under one or more statutes.

16. The system of claim 15 , wherein the one or more statutes comprise the General Data Protection Regulation (“GDPR”), the Health Insurance Portability and Accountability Act (“HIPAA”), or the Family Educational Rights and Privacy Act (“FERPA”).

17. The system of claim 10 , wherein a subset of the plurality of context-based speech recognition models include role-specific variants such that the assignment of one of the plurality of context-based speech recognition models to one of the channels is based in part on a transaction role of one of the channels.

18. The system of claim 10 , wherein the protected information comprises financial account data.

19. The system of claim 10 , wherein the protected information comprises healthcare data.

20. The system of claim 10 , wherein the alternate context-based speech recognition model assigned to at least one of the two different channels is based at least in part on speech from one of the other of the at least two different channels.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 13, 2020
From: SAMBARINO, ROMAIN
To: GROUPE ALLO MEDIA SAS
Reel/Frame 052376/0390 →
Continuity (3)
Continuation 16692000 · Nov 22, 2019
Continuation 16272495 · Feb 11, 2019
Related Publication 20200258505A1 · Aug 13, 2020