SYSTEMS AND METHODS FOR TEAM COOPERATION WITH REAL-TIME RECORDING AND TRANSCRIPTION OF CONVERSATIONS AND/OR SPEECHES
Methods and systems for team cooperation with real-time recording of one or more moment-associating elements. For example, a method includes: delivering, in response to an instruction, an invitation to each member of one or more members associated with a workspace; granting, in response to acceptance of the invitation by one or more subscribers of the one or more members, subscription permission to the one or more subscribers; receiving the one or more moment-associating elements; transforming the one or more moment-associating elements into one or more pieces of moment-associating information; and transmitting at least one piece of the one or more pieces of moment-associating information to the one or more subscribers.
1 . A computer-implemented method for team cooperation with real-time recording of one or more moment-associating elements, the method comprising:
delivering, in response to an instruction, an invitation to each member of one or more members associated with a workspace;
granting, in response to acceptance of the invitation by one or more subscribers of the one or more members, subscription permission to the one or more subscribers;
receiving the one or more moment-associating elements;
transforming the one or more moment-associating elements into one or more pieces of moment-associating information; and
transmitting at least one piece of the one or more pieces of moment-associating information to the one or more subscribers;
wherein the transforming the one or more moment-associating elements into one or more pieces of moment-associating information includes:
segmenting the one or more moment-associating elements into a plurality of moment-associating segments;
assigning a segment speaker for each segment of the plurality of moment-associating segments;
transcribing the plurality of moment-associating segments into a plurality of transcribed segments; and
generating the one or more pieces of moment-associating information based at least in part on the plurality of transcribed segments and the segment speaker assigned for each segment of the plurality of moment-associating segments.
2 . The computer-implemented method of claim 1 , further comprising receiving event information associated with an event, the event information includes at least one of:
one or more speaker names;
one or more speech titles;
one or more starting times;
one or more end times;
a custom vocabulary;
location information; and
attendee information;
wherein the transforming the one or more moment-associating elements into one or more pieces of moment-associating information includes transforming the one or more moment-associating elements into one or more pieces of moment-associating information based at least in part on the event information.
3 . The computer-implemented method of claim 2 , further comprising:
connecting with one or more calendar systems containing the event information; and
receiving the event information from the one or more calendar systems.
4 . The computer-implemented method of claim 2 , wherein the transforming the one or more moment-associating elements into one or more pieces of moment-associating information based at least in part on the event information includes:
creating a custom language model based at least in part on the event information; and
transcribing the plurality of moment-associating segments into a plurality of transcribed segments based at least in part on the custom language model.
5 . The computer-implemented method of claim 1 , wherein the receiving the one or more moment-associating elements includes assigning a timestamp associated with each element of the one or more moment-associating elements.
6 . The computer-implemented method of claim 1 , wherein the receiving the one or more moment-associating elements includes at least one selected from receiving one or more audio elements, receiving one or more visual elements, and receiving one or more environmental elements.
7 . The computer-implemented method of claim 6 , wherein the receiving one or more audio elements includes at least one selected from receiving one or more voice elements of one or more voice-generating sources and receiving one or more ambient sound elements.
8 . The computer-implemented method of claim 6 , wherein the receiving one or more visual elements includes at least one selected from receiving one or more pictures, receiving one or more images, receiving one or more screenshots, receiving one or more video frames, receiving one or more projections, and receiving one or more holograms.
9 . The computer-implemented method of claim 6 , wherein the receiving one or more environmental elements includes at least one selected from receiving one or more global positions, receiving one or more location types, and receiving one or more moment conditions.
10 . The computer-implemented method of claim 6 , wherein the receiving one or more environmental elements includes at least one selected from receiving a longitude, receiving a latitude, receiving an altitude, receiving a country, receiving a city, receiving a street, receiving a location type, receiving a temperature, receiving a humidity, receiving a movement, receiving a velocity of a movement, receiving a direction of a movement, receiving an ambient noise level, and receiving one or more echo properties.
11 . The computer-implemented method of claim 6 , wherein the transforming the one or more moment-associating elements into one or more pieces of moment-associating information includes:
segmenting the one or more audio elements into a plurality of audio segments;
assigning a segment speaker for each segment of the plurality of audio segments;
transcribing the plurality of audio segments into a plurality of text segments; and
generating the one or more pieces of moment-associating information based at least in part on the plurality of text segments and the segment speaker assigned for each segment of the plurality of audio segments.
12 . The computer-implemented method of claim 11 , wherein the transcribing the plurality of audio segments into a plurality of text segments includes transcribing two or more segments of the plurality of audio segments in conjunction with each other.
13 . The computer-implemented method of claim 1 , and further comprising:
receiving one or more voice elements of one or more voice-generating sources; and
receiving one or more voiceprints corresponding to the one or more voice-generating sources respectively.
14 . The computer-implemented method of claim 13 , wherein the transforming the one or more moment-associating elements into one or more pieces of moment-associating information further includes:
segmenting the one or more moment-associating elements into the plurality of moment-associating segments based at least in part on the one or more voiceprints;
assigning a segment speaker for each segment of the plurality of moment-associating segments based at least in part on the one or more voiceprints; and
transcribing the plurality of moment-associating segments into the plurality of transcribed segments based at least in part on the one or more voiceprints.
15 . The computer-implemented method of claim 13 , wherein the receiving one or more voiceprints corresponding to the one or more voice-generating sources respectively includes at least one of:
receiving one or more acoustic models corresponding to the one or more voice-generating sources respectively; and
receiving one or more language models corresponding to the one or more voice-generating sources respectively.
16 . The computer-implemented method of claim 1 , wherein the transcribing the plurality of moment-associating segments into a plurality of transcribed segments includes:
transcribing a first segment of the plurality of moment-associating segments into a first transcribed segment of the plurality of transcribed segments;
transcribing a second segment of the plurality of moment-associating segments into a second transcribed segment of the plurality of transcribed segments; and
correcting the first transcribed segment based at least in part on the second transcribed segment.
17 . The computer-implemented method of claim 1 , wherein the segmenting the one or more moment-associating elements into a plurality of moment-associating segments includes:
determining one or more speaker-change timestamps, each timestamp of the one or more speaker-change timestamps corresponding to a timestamp when a speaker change occurs;
determining one or more sentence-change timestamps, each timestamp of the one or more sentence-change timestamps corresponding to a timestamp when a sentence change occurs; and
determining one or more topic-change timestamps, each timestamp of the one or more topic-change timestamps corresponding to a timestamp when a topic change occurs.
18 . The computer-implemented method of claim 17 , wherein the segmenting the one or more moment-associating elements into a plurality of moment-associating segments is performed based at least in part on one of:
the one or more speaker-change timestamps;
the one or more sentence-change timestamps; and
the one or more topic-change timestamps.
19 . A system for team cooperation with real-time recording of one or more moment-associating elements, the system comprising:
an invitation delivering module configured to deliver, in response to an instruction, an invitation to each member of one or more members associated with a workspace;
a permission module configured to grant, in response to acceptance of the invitation by one or more subscribers of the one or more members, subscription permission to the one or more subscribers;
a receiving module configured to receive the one or more moment-associating elements;
a transforming module configured to transform the one or more moment-associating elements into one or more pieces of moment-associating information; and
a transmitting module configured to transmit at least one piece of the one or more pieces of moment-associating information to the one or more subscribers;
wherein the transforming module is further configured to:
segment the one or more moment-associating elements into a plurality of moment-associating segments;
assign a segment speaker for each segment of the plurality of moment-associating segments;
transcribe the plurality of moment-associating segments into a plurality of transcribed segments; and
generate the one or more pieces of moment-associating information based at least in part on the plurality of transcribed segments and the segment speaker assigned for each segment of the plurality of moment-associating segments.
20 . A non-transitory computer-readable medium with instructions stored thereon, that when executed by a processor, perform the processes comprising:
delivering, in response to an instruction, an invitation to each member of one or more members associated with a workspace;
granting, in response to acceptance of the invitation by one or more subscribers of the one or more members, subscription permission to the one or more subscribers;
receiving the one or more moment-associating elements;
transforming the one or more moment-associating elements into one or more pieces of moment-associating information; and
transmitting at least one piece of the one or more pieces of moment-associating information to the one or more subscribers;
wherein the transforming the one or more moment-associating elements into one or more pieces of moment-associating information includes:
segmenting the one or more moment-associating elements into a plurality of moment-associating segments;
assigning a segment speaker for each segment of the plurality of moment-associating segments;
transcribing the plurality of moment-associating segments into a plurality of transcribed segments; and
generating the one or more pieces of moment-associating information based at least in part on the plurality of transcribed segments and the segment speaker assigned for each segment of the plurality of moment-associating segments.