IP Library Granted Patent US 12,243,551
Granted Patent B2
US 12,243,551 · App. 18/503,597 · Granted Mar 4, 2025

Performing artificial intelligence sign language translation services in a video relay service environment

Inventor: Conrad Maxwell (Herriman, UT)
Assignee: Sorenson IP Holdings, LLC
G10L21/10G06T13/205G06T13/40G06V40/28G10L13/027H04N7/147H04N7/15G09B21/009G10L2021/065
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,243,551
App. No.
18/503,597
Granted
Mar 4, 2025
Kind
B2
Abstract

Video relay services, communication systems, non-transitory machine-readable storage media, and methods are disclosed herein. A video relay service may include at least one server configured to receive a video stream including sign language content from a video communication device during a real-time communication session. The server may also be configured to automatically translate the sign language content into a verbal language translation during the real-time communication session without assistance of a human sign language interpreter. Further, the server may be configured to transmit the verbal language translation during the real-time communication session.

Claims (43)

1. A system for providing automated translation services during a communication session, the system comprising:

a video relay system configured to connect to a first device over a first connection and to connect to a second device over a second connection to facilitate a communication session between the first device and the second device;

at least one processor in operable communication with the video relay system, the at least one processor configured to:

obtain an audio stream from one or more of the first device and the second device during the communication session;

analyze the audio stream;

automatically translate audio content from the audio stream into sign language content during the communication session without assistance from a human sign language interpreter; and

transmit the sign language content to one or more of the first device and the second device during the communication session.

2. The system of claim 1 , wherein the translation of the audio content from the audio stream into the sign language content is performed using a convolutional neural network or a deep neural network.

3. The system of claim 1 , wherein the at least one processor is further configured to translate the audio content into text.

4. The system of claim 3 , wherein the at least one processor is further configured to transmit the text to one or more of the first device and the second device during the communication session.

5. The system of claim 1 , wherein the audio content is automatically translated into sign language content using video files stored in an AI translation database.

6. The system of claim 1 , wherein the at least one processor is further configured to:

obtain a video stream from one or more of the first device and the second device during the communication session, the video stream including other sign language content;

automatically translate the other sign language content from the video stream into other media content during the communication session; and

transmit the other media content to one or more of the first device and the second device during the communication session.

7. The system of claim 6 , wherein the other media content is text content or audio content that is a translation of the other sign language content.

8. A method comprising:

obtaining an audio stream from one or more of a first device and a second device during a communication session between the first device and the second device, the communication session facilitated by a video relay system configured to connect to the first device over a first connection and to connect to the second device over a second connection;

analyzing the audio stream;

automatically translating audio content from the audio stream into sign language content during the communication session without assistance from a human sign language interpreter; and

transmitting the sign language content to one or more of the first device and the second device during the communication session.

9. The method of claim 8 , wherein the translation of the audio content from the audio stream into the sign language content is performed using a convolutional neural network or a deep neural network.

10. The method of claim 8 , further comprising translating the audio content into text.

11. The method of claim 10 , further comprising transmitting the text to one or more of the first device and the second device during the communication session.

12. The method of claim 8 , wherein the audio content is automatically translated into sign language content using video files stored in an AI translation database.

13. The method of claim 8 , further comprising:

obtaining a video stream from one or more of the first device and the second device during the communication session, the video stream including other sign language content;

automatically translating the other sign language content from the video stream into other media content during the communication session; and

transmitting the other media content to one or more of the first device and the second device during the communication session.

14. The method of claim 13 , wherein the other media content is text content or audio content that is a translation of the other sign language content.

15. One or more non-transitory computer-readable media configured to store instructions, which when executed by a device or system, cause or direct performance of operations, the operations comprising:

obtaining an audio stream from one or more of a first device and a second device during a communication session between the first device and the second device, the communication session facilitated by a video relay system configured to connect to the first device over a first connection and to connect to the second device over a second connection;

analyzing the audio stream;

automatically translating audio content from the audio stream into sign language content during the communication session without assistance from a human sign language interpreter; and

transmitting the sign language content to one or more of the first device and the second device during the communication session.

16. The computer-readable media of claim 15 , wherein the translation of the audio content from the audio stream into the sign language content is performed using a convolutional neural network or a deep neural network.

17. The computer-readable media of claim 15 , wherein the operations further comprise translating the audio content into text.

18. The computer-readable media of claim 17 , wherein the operations further comprise transmitting the text to one or more of the first device and the second device during the communication session.

19. The computer-readable media of claim 15 , wherein the audio content is automatically translated into sign language content using video files stored in an AI translation database.

20. The computer-readable media of claim 15 , wherein the operations further comprise:

obtaining a video stream from one or more of the first device and the second device during the communication session, the video stream including other sign language content;

automatically translating the other sign language content from the video stream into other media content during the communication session; and

transmitting the other media content to one or more of the first device and the second device during the communication session.

Assignments (2)
SECURITY INTEREST Recorded Apr 23, 2024
From: SORENSON COMMUNICATIONS, LLC; INTERACTIVECARE, LLC; CAPTIONCALL, LLC
To: OAKTREE FUND ADMINISTRATION, LLC, AS COLLATERAL AGENT
Reel/Frame 067573/0201 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 8, 2023
From: MAXWELL, CONRAD A.
To: SORENSON IP HOLDINGS LLC
Reel/Frame 065499/0920 →