SYSTEMS AND METHODS FOR IMPROVED AUDIO/VIDEO CONFERENCES
Systems and methods for efficient management of an audio/video conference are disclosed. The methods comprise recording voice data of a first user connected to a conference while the user is in a first state, determining the first user is talking while in the first state, and initiating playback of the recorded voice data of the first user to a plurality of other users, wherein a playback rate of the recorded voice data is variable.
1 . A method comprising:
receiving voice data of a first user via a user device connected to a conference session;
based on determining that the user device is in a mute state in the conference session, displaying a selectable indicator on a user interface of the user device indicating that the voice data is not being transmitted to other user devices connected to the conference session; and
in response to receiving a selection of the selectable indicator:
identifying speech fillers in the voice data;
modifying the voice data by removing the identified speech fillers; and
automatically transmitting the modified voice data of the first user to the other user devices connected to the conference session.
2 . The method of claim 1 , wherein the speech fillers comprise at least one of:
pauses, hesitations, stutters, filler words, or discourse markers.
3 . The method of claim 1 , wherein the speech fillers are identified in the voice data by analyzing the voice data in real-time using a natural language processing algorithm.
4 . The method of claim 1 , further comprising:
based on determining that the user device is in the mute state in the conference session, storing the voice data in a buffer until a storage threshold of the buffer is reached.
5 . The method of claim 4 , wherein when the storage threshold of the buffer is reached the method further comprises:
ceasing display of the selectable indicator; and
removing the voice data from the buffer.
6 . The method of claim 4 , wherein the selectable indicator indicates how much storage space in the buffer is left before the stored voice data reaches the storage threshold of the buffer.
7 . The method of claim 1 , further comprising, in response to receiving the selection of the selectable indicator, switching the user device to an unmuted state.
8 . The method of claim 1 , wherein the modifying the voice data, further comprises, adjusting a playback speed of the voice data.
9 . The method of claim 1 , wherein the user device is a first user device and wherein the method further comprises:
determining that voice data of a second user at a second user device is being received when automatically transmitting the modified voice data of the first user;
determining that the voice data from the second user has a higher-priority value than the voice data from the first user; and
interrupting transmission of the modified voice data of the first user and transmitting the voice data of the second user.
10 . The method of claim 9 , further comprising:
reinitiating transmission of the modified voice data of the first user after the transmission of the voice data of the second user has completed.
11 . A system comprising:
input/output circuitry;
control circuitry configured to:
receive voice data of a first user via the input/output circuitry of a user device connected to a conference session;
based on determining that the user device is in a mute state in the conference session, display a selectable indicator on a user interface of the user device indicating that the voice data is not being transmitted to other user devices connected to the conference session; and
in response to receiving a selection of the selectable indicator:
identify speech fillers in the voice data;
modify the voice data by removing the identified speech fillers; and
automatically transmit the modified voice data of the first user to the other user devices connected to the conference session.
12 . The system of claim 11 , wherein the speech fillers comprise at least one of:
pauses, hesitations, stutters, filler words, or discourse markers.
13 . The system of claim 11 , wherein the control circuitry is configured to identify the speech fillers in the voice data by analyzing the voice data in real-time using a natural language processing algorithm.
14 . The system of claim 11 , wherein, based on determining that the user device is in the mute state in the conference session, the control circuitry is further configured to:
store the voice data in a buffer until a storage threshold of the buffer is reached.
15 . The system of claim 14 , wherein when the storage threshold of the buffer is reached, the control circuitry is further configured to:
cease display of the selectable indicator; and
remove the voice data from the buffer.
16 . The system of claim 14 , wherein the selectable indicator indicates how much storage space in the buffer is left before the stored voice data reaches the storage threshold of the buffer.
17 . The system of claim 11 , wherein the control circuitry is further configured to:
in response to receiving the selection of the selectable indicator, switching the user device to an unmuted state.
18 . The system of claim 11 , wherein the control circuitry configured to modify the voice data, is further configured to:
adjust a playback speed of the voice data.
19 . The system of claim 11 , wherein the user device is a first user device and wherein the control circuitry is further configured to:
determine that voice data of a second user at a second user device is being received when automatically transmitting the modified voice data of the first user;
determine that the voice data from the second user has a higher-priority value than the voice data from the first user; and
interrupt transmission of the modified voice data of the first user and transmitting the voice data of the second user.
20 . The system of claim 19 , wherein the control circuitry is further configured to:
reinitiate transmission of the modified voice data of the first user after the transmission of the voice data of the second user has completed.