IP Library Granted Patent US 11,862,169
Granted Patent B2
US 11,862,169 · App. 17/017,941 · Granted Jan 2, 2024

Multilingual transcription at customer endpoint for optimizing interaction results in a contact center

Inventors: Valentine C. Matula (Granville, OH); Pushkar Yashavant Deole (Pune, IN); Sandesh Chopdekar (Pune, IN); Navin Daga (Silapathar, IN)
Assignee: Avaya Management L.P.
G10L15/26G10L15/22G10L15/30G10L2015/223
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,862,169
App. No.
17/017,941
Granted
Jan 2, 2024
Kind
B2
Abstract

Providing speech-to-text (STT) transcription by a user endpoint device includes initiating an audio communication between an enterprise server and the user endpoint device, the audio communication comprising a voice interaction between a user associated with the user endpoint device and an agent associated with an agent device to which the enterprise server routes the audio communication; performing a first STT of at least a portion of the voice interaction to produce a first transcribed speech in a first language; concurrent with performing the first STT, performing, by the user endpoint device, a second STT of the at least the portion of the voice interaction to produce a second transcribed speech in a second language different than the first language, and transmitting the at least the portion of the voice interaction and at least the first transcribed speech from the user endpoint device to the enterprise server.

Claims (52)

1. A processor-based method for providing speech-to-text (STT) transcription, comprising:

initiating, by a user endpoint device, an audio communication between an enterprise server and the user endpoint device, the audio communication comprising a voice interaction between a user associated with the user endpoint device and an agent associated with an agent device to which the enterprise server routes the audio communication;

determining, by the user endpoint device, whether the user endpoint device has the computational capability to perform STT of at least a portion of the voice interaction;

transmitting, by the user endpoint device to the enterprise server, the determination as to whether the user endpoint device has the computational capability to perform STT of the at least the portion of the voice interaction;

performing, by the user endpoint device during the audio communication, a first STT of the at least the portion of the voice interaction to produce a first transcribed speech in a first language;

concurrent with performing the first STT, performing, by the user endpoint device, a second STT of the at least the portion of the voice interaction to produce a second transcribed speech in a second language different than the first language, and

transmitting, by the user endpoint device during the audio communication, the at least the portion of the voice interaction and at least the first transcribed speech from the user endpoint device to the enterprise server.

2. The method of claim 1 , further comprising:

transmitting, by the user endpoint device during the audio communication, the second transcribed speech from the user endpoint device to the enterprise server.

3. The method of claim 1 , wherein the at least the portion of the voice interaction comprises first speech provided by the user and second speech provided by the agent.

4. The method of claim 3 , wherein the second transcribed speech comprises a transcription of the second speech provided by the agent, and further comprising:

displaying, by the user endpoint device, the transcription of the second speech provided by the agent on a display apparatus of the user endpoint device during the audio communication.

5. The method of claim 1 , wherein the at least the portion of the voice interaction and the first transcribed speech are transmitted substantially concurrently with each other, wherein the first transcribed speech is transmitted via a digital channel and the at least the portion of the voice interaction is transmitted via a voice channel.

6. The method of claim 1 , wherein the first language is a default system language defined by the enterprise server.

7. The method of claim 1 , wherein:

the first language is a related language of the agent, and

the second language is a related language of a user.

8. The method of claim 7 , further comprising:

sending, by the user endpoint device, an inquiry to the enterprise server requesting an identification of a more-related language of the agent; and

in response to sending the inquiry, receiving, by the user endpoint device from the enterprise server, the identification of the more-related language of the agent.

9. The method of claim 1 , wherein:

initiating the audio communication, performing the first STT, and transmitting the at least the portion of the voice interaction and the first transcribed speech are performed by one of

an audio-capable application executing on the user endpoint device,

a plug-in of a web browser executing on the user endpoint device, or

a customized client application executing on the user endpoint device.

10. A user endpoint device for providing speech-to-text (STT) transcription, comprising:

a memory device storing executable instructions; and

a processor in communication with the memory device, wherein the processor when executing the executable instructions:

initiates an audio communication between an enterprise server and the user endpoint device, the audio communication comprising a voice interaction between a user associated with the user endpoint device and an agent associated with an agent device to which the enterprise server routes the audio communication;

determines whether the user endpoint device has the computational capability to perform STT of at least a portion of the voice interaction;

transmits to the enterprise server, the determination as to whether or not the user endpoint device has the computational capability to perform STT of the at least the portion of the voice interaction;

performs during the audio communication, a first STT of the at least the portion of the voice interaction to produce a first transcribed speech in a first language;

concurrent with performing the first STT, performs a second STT of the at least the portion of the voice interaction to produce a second transcribed speech in a second language different than the first language, and

transmits, during the audio communication, the at least the portion of the voice interaction and at least the first transcribed speech from the user endpoint device to the enterprise server.

11. The system of claim 10 , wherein the processor when executing the executable instructions:

transmits, during the audio communication, the second transcribed speech from the user endpoint device to the enterprise server.

12. The system of claim 10 , wherein the at least the portion of the voice interaction comprises first speech provided by the user and second speech provided by the agent.

13. The system of claim 12 , wherein the second transcribed speech comprises a transcription of the second speech provided by the agent, and wherein the processor when executing the executable instructions:

displays the transcription of the second speech provided by the agent on a display apparatus of the user endpoint device during the audio communication.

14. The system of claim 10 , wherein the at least the portion of the voice interaction and the first transcribed speech are transmitted substantially concurrently with each other, wherein the first transcribed speech is transmitted via a digital channel and the at least the portion of the voice interaction is transmitted via a voice channel.

15. The system of claim 10 , wherein the first language is a default system language defined by the enterprise server.

16. The system of claim 10 , wherein:

the first language is a related language of the agent, and

the second language is a related language of a user.

17. The system of claim 16 , wherein the processor when executing the executable instructions:

sends an inquiry to the enterprise server requesting an identification of a more-related language of the agent; and

in response to sending the inquiry, receives from the enterprise server, the identification of the more-related language of the agent.

18. The system of claim 10 , wherein:

initiating the audio communication, performing the first STT, and transmitting the at least the portion of the voice interaction and the first transcribed speech are performed by one of

an audio-capable application executing on the user endpoint device,

a plug-in of a web browser executing on the user endpoint device, or

a customized client application executing on the user endpoint device.

Assignments (10)
INTELLECTUAL PROPERTY SECURITY AGREEMENT – SUPPLEMENT NO. 5 Recorded Jun 3, 2024
From: AVAYA LLC; AVAYA MANAGEMENT L.P.
To: WILMINGTON SAVINGS FUND SOCIETY, FSB, AS COLLATERAL AGENT
Reel/Frame 067606/0438 →
INTELLECTUAL PROPERTY SECURITY AGREEMENT – SUPPLEMENT NO. 3 Recorded May 29, 2024
From: AVAYA LLC (FORMERLY KNOWN AS AVAYA INC.); AVAYA MANAGEMENT L.P.
To: WILMINGTON SAVINGS FUND SOCIETY, FSB, AS COLLATERAL AGENT
Reel/Frame 067559/0295 →
INTELLECTUAL PROPERTY SECURITY AGREEMENT Recorded Oct 2, 2023
From: AVAYA LLC (F/K/A AVAYA INC.); AVAYA MANAGEMENT L.P.
To: CITIBANK, N.A., AS COLLATERAL AGENT
Reel/Frame 065093/0584 →
RELEASE OF SECURITY INTEREST IN PATENTS (REEL/FRAME 53955/0436) Recorded May 18, 2023
From: WILMINGTON TRUST, NATIONAL ASSOCIATION, AS NOTES COLLATERAL AGENT
To: AVAYA MANAGEMENT L.P.; AVAYA INC.; INTELLISIST, INC.; AVAYA INTEGRATED CABINET SOLUTIONS LLC
Reel/Frame 063705/0023 →
RELEASE OF SECURITY INTEREST IN PATENTS (REEL/FRAME 61087/0386) Recorded May 18, 2023
From: WILMINGTON TRUST, NATIONAL ASSOCIATION, AS NOTES COLLATERAL AGENT
To: AVAYA MANAGEMENT L.P.; AVAYA INC.; INTELLISIST, INC.; AVAYA INTEGRATED CABINET SOLUTIONS LLC
Reel/Frame 063690/0359 →
INTELLECTUAL PROPERTY SECURITY AGREEMENT Recorded May 4, 2023
From: AVAYA INC.; AVAYA MANAGEMENT L.P.; INTELLISIST, INC.
To: CITIBANK, N.A., AS COLLATERAL AGENT
Reel/Frame 063542/0662 →
INTELLECTUAL PROPERTY SECURITY AGREEMENT Recorded May 3, 2023
From: AVAYA MANAGEMENT L.P.; AVAYA INC.; INTELLISIST, INC.; KNOAHSOFT INC.
To: WILMINGTON SAVINGS FUND SOCIETY, FSB [COLLATERAL AGENT]
Reel/Frame 063742/0001 →
INTELLECTUAL PROPERTY SECURITY AGREEMENT Recorded Aug 5, 2022
From: AVAYA INC.; INTELLISIST, INC.; AVAYA MANAGEMENT L.P.; AVAYA CABINET SOLUTIONS LLC
To: WILMINGTON TRUST, NATIONAL ASSOCIATION, AS COLLATERAL AGENT
Reel/Frame 061087/0386 →
SECURITY INTEREST Recorded Sep 25, 2020
From: AVAYA INC.; AVAYA MANAGEMENT L.P.; INTELLISIST, INC.; AVAYA INTEGRATED CABINET SOLUTIONS LLC
To: WILMINGTON TRUST, NATIONAL ASSOCIATION
Reel/Frame 053955/0436 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Sep 14, 2020
From: MATULA, VALENTINE C.; DEOLE, PUSHKAR YASHAVANT; CHOPDEKAR, SANDESH; DAGA, NAVIN
To: AVAYA MANAGEMENT L.P.
Reel/Frame 053760/0495 →
Continuity (1)
Related Publication 20220084523A1 · Mar 17, 2022