SIGN LANGUAGE PROCESSING
A method may include in response to a communication session not being established between a first communication device and a second communication device, obtaining an audio message from the first communication device. In these and other embodiments, the method may include storing the audio message and after storing the audio message, generating, by an automated generation system that includes one or more first machine learning models, video that includes sign language content corresponding to the audio message. The method may also include storing the video.
1 . A method comprising:
in response to a communication session not being established between a first communication device and a second communication device, obtaining an audio message from the first communication device;
storing the audio message;
after storing the audio message, generating, by an automated generation system that includes one or more first machine learning models, video that includes sign language content corresponding to the audio message;
storing the video; and
after storing the video, training one or more second machine learning models of an automated recognition system configured to translate sign language into language data using the video and language data from the audio message.
2 . The method of claim 1 , further comprising before training the automated recognition system, obtaining consent from at least one of a first user associated with the first communication device and a second user associated with the second communication device.
3 . The method of claim 1 , further comprising after storing the video, training one or more of the first machine learning models of the automated generation system using the video and the language data.
4 . The method of claim 1 , further comprising before obtaining the audio message, directing second audio to the first communication device.
5 . The method of claim 4 , wherein second audio is generated via an automated system to interact with a user of the first communication device.
6 . The method of claim 1 , wherein the video is generated in response to a request from a user associated with the second communication device to view the video.
7 . The method of claim 1 , further comprising transcribing the audio message using automated speech recognition to generate text corresponding to the sign language content, wherein the text is used to train the one or more second machine learning models.
8 . The method of claim 1 , further comprising:
storing a plurality of audio messages and corresponding videos that include the audio message and the video, each of the audio messages generated from a different communication session not being established;
determining which of the plurality of audio messages and corresponding videos is usable for training; and
training the one or more second machine learning models of the automated recognition system using the audio messages and corresponding videos determined to be usable for training.
9 . The method of claim 8 , wherein a first audio message and corresponding first video is determined to be usable for training based on obtaining consent from one or more users associated with the first audio message.
10 . At least one non-transitory computer-readable media configured to store one or more instructions that, in response to being executed by a system, cause or direct the system to perform the method of claim 1 .
11 . A method comprising:
in response to a communication session not being established between a first communication device and a second communication device, obtaining an audio message from the first communication device;
storing the audio message;
after storing the audio message, generating, by an automated generation system that includes one or more first machine learning models, video that includes sign language content corresponding to the audio message; and
storing the video.
12 . A system comprising:
one or more computer readable mediums including instructions;
one or more computing systems coupled to the one or more computer readable mediums and configured to execute the instructions to cause or direct the system to perform operations, the operations comprising:
in response to a communication session not being established between a first communication device and a second communication device, obtaining an audio message from the first communication device;
storing the audio message;
after storing the audio message, generating, by an automated generation system that includes one or more first machine learning models, video that includes sign language content corresponding to the audio message;
storing the video; and
after storing the video, training one or more second machine learning models of an automated recognition system configured to translate sign language into language data using the video and language data from the audio message.
13 . The system of claim 12 , wherein the operations further include before training the automated recognition system, obtaining consent from at least one of a first user associated with the first communication device and a second user associated with the second communication device.
14 . The system of claim 12 , wherein the operations further include after storing the video, training one or more of the first machine learning models of the automated generation system using the video and the language data.
15 . The system of claim 12 , wherein the operations further include before obtaining the audio message, directing second audio to the first communication device.
16 . The system of claim 15 , wherein second audio is generated via an automated system to interact with a user of the first communication device.
17 . The system of claim 12 , wherein the video is generated in response to a request from a user associated with the second communication device to view the video.
18 . The system of claim 12 , wherein the operations further include transcribing the audio message using automated speech recognition to generate text corresponding to the sign language content, wherein the text is used to train the one or more second machine learning models.
19 . The system of claim 12 , wherein the operations further include:
storing a plurality of audio messages and corresponding videos that include the audio message and the video, each of the audio messages generated from a different communication session not being established;
determining which of the plurality of audio messages and corresponding videos is usable for training; and
training the one or more second machine learning models of the automated recognition system using the audio messages and corresponding videos determined to be usable for training.
20 . The system of claim 19 , wherein a first audio message and corresponding first video is determined to be usable for training based on obtaining consent from one or more users associated with the first audio message.