IP Library Granted Patent US 12682905
Granted Patent B2
US 12682905 · App. 18/405,092 · Granted Jul 14, 2026

Voice data transmission method and apparatus

Inventors: Jung Ho Kim (Seoul, KR); Young Kwang Kim (Seoul, KR); Soo Hwan Park (Seoul, KR); Sang Wook Lee (Seoul, KR); Dong Ho Cha (Seoul, KR); Jun Ho Kang (Seoul, KR); Hee Tae Yoon (Seoul, KR)
Assignee: SAMSUNG SDS CO., LTD.
G10L17/02G10L25/51
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12682905
App. No.
18/405,092
Granted
Jul 14, 2026
Kind
B2
Abstract

A method for transmitting voice data is provided. The method may include receiving voice data of a particular speaker from a voice data collection server; determining whether destination terminals are first-type terminals; and based on a determination that the destination terminals are the first-type terminals, transmitting the received voice data to the destination terminals through a channel corresponding to the particular speaker, from among a plurality of predefined channels for the destination terminals.

Claims (56)

1 . A voice data transmission method performed by at least one processor, comprising:

receiving voice data of a plurality of speakers including voice data of a particular speaker from a voice data collection server;

determining whether destination terminals are first-type terminals; and

based on a determination that the destination terminals are the first-type terminals, transmitting the received voice data to the destination terminals through a channel corresponding to the particular speaker, from among a plurality of predefined channels for the destination terminals

wherein the transmitting the received voice data to the destination terminals further comprises:

based on a determination that the destination terminals are second-type terminals, transmitting the voice data of the plurality of speakers to the destination terminals through a single channel,

wherein the first-type terminals are terminals capable of identifying a speaker of first voice data based on a channel through which the first voice data has been received, and

wherein the second-type terminals are terminals capable of identifying a speaker of second voice data by analyzing a source of the second voice data, regardless of a channel through which the second voice data has been received.

2 . The voice data transmission method of claim 1 , wherein a number of the plurality of predefined channels is greater than a number of voice data collected by the voice data collection server.

3 . The voice data transmission method of claim 2 , wherein the number of the plurality of predefined channels is determined based on a number of the destination terminals.

4 . The voice data transmission method of claim 1 , wherein the transmitting the received voice data to the destination terminals comprises:

determining whether the particular speaker of the received voice data is a new speaker; and

based on a determination that the particular speaker of the received voice data is not the new speaker, transmitting the received voice data to the destination terminals through a channel that is previously allocated to previously-received voice data.

5 . The voice data transmission method of claim 1 , wherein the transmitting the received voice data to the destination terminals comprises:

determining whether the particular speaker of the received voice data is a new speaker; and

based on a determination that the particular speaker of the received voice data is the new speaker, transmitting the received voice data to the destination terminals through a least recently used channel among the plurality of predefined channels.

6 . The voice data transmission method of claim 1 , wherein the transmitting the received voice data comprises:

determining whether the particular speaker of the received voice data is a new speaker based on a mapping table between channel identifiers (IDs) and speaker IDs;

based on a determination that the particular speaker of the received voice data is the new speaker, replacing a speaker ID in an entry of the mapping table that corresponds to a speaker of oldest voice data with a speaker ID of the new speaker, and transmitting the received voice data to the destination terminals via a channel having a channel ID corresponding to the speaker ID of the new speaker.

7 . The voice data transmission method of claim 1 , wherein the received voice data includes a predefined number or less of voice data having vocal intensities higher than a reference level, as detected during a measurement period, that are selected, among a plurality of voice data, in a descending order of the vocal intensities.

8 . The voice data transmission method of claim 1 , wherein

the particular speaker is one of a plurality of speakers, of which respective voice data are received from the voice data collection server, and

a number of the plurality of predefined channels is constant regardless of a number of the plurality of speakers.

9 . A voice data transmission system, comprising:

a voice data collection server configured to select one or more voice data based on periodic utterance quantities of voice data received from user terminals; and

a channel control server configured to control one or more channels through which the selected voice data are transmitted to the user terminals,

wherein the channel control server is configured to:

generate channels for transmitting voice data between the user terminals and the channel control server;

identify a type of the user terminals and one or more speakers of the selected voice data, and set one or more channels corresponding to the identified one or more speakers, among the generated channels; and

transmit the selected voice data through the set one or more channels

wherein, when voice data of a plurality of speakers are received and the user terminals are determined to be first-type terminals, the channel control server transmits voice data of each speaker through a respective channel corresponding to each speaker from among the generated channels,

wherein, when voice data of a plurality of speakers are received and the user terminals are determined to be second-type terminals, the channel control server transmits the voice data of the plurality of speakers to the user terminals through a single channel selected from among the generated channels,

wherein the first-type terminals are terminals capable of identifying a speaker of first voice data based on a channel through which the first voice data has been received, and

wherein the second-type terminals are terminals capable of identifying a speaker of second voice data by analyzing a source of the second voice data, regardless of a channel through which the second voice data has been received.

10 . The voice data transmission system of claim 9 , wherein the voice data collection server is configured to analyze vocal intensities of voice data collected from the user terminals and select voice data of a particular speaker based on a result of analysis.

11 . A voice data transmission apparatus, comprising:

a processor; and

a memory configured to store instructions,

wherein the instructions, when executed by the processor, cause the processor to:

receive voice data of a plurality of speakers including voice data of a particular speaker from a voice data collection server;

determine whether destination terminals are first-type terminals; and

based on a determination that the destination terminals are the first-type terminals, transmit the received voice data to the destination terminals through a channel corresponding to the particular speaker, from among a plurality of predefined channels for the destination terminals,

wherein the instructions further cause the processor to, based on a determination that the destination terminals are second-type terminals, transmit the voice data of the plurality of speakers to the destination terminals through a single channel,

wherein the first-type terminals are terminals capable of identifying a speaker of first voice data based on a channel through which the first voice data has been received, and

wherein the second-type terminals are terminals capable of identifying a speaker of second voice data by analyzing a source of the second voice data, regardless of a channel through which the second voice data has been received.

12 . The voice data transmission apparatus of claim 11 , wherein a number of the plurality of predefined channels is greater than a number of voice data collected by the voice data collection server.

13 . The voice data transmission apparatus of claim 12 , wherein the number of the plurality of predefined channels is determined based on a number of destination terminals.

14 . The voice data transmission apparatus of claim 11 , wherein the instructions further cause the processor to:

determine whether the particular speaker of the received voice data is a new speaker; and

based on a determination that the particular speaker of the received voice data is not the new speaker, transmit the received voice data to the destination terminals through a channel that is previously allocated to previously-received voice data.

15 . The voice data transmission apparatus of claim 11 , wherein the instructions further cause the processor to:

determine whether the particular speaker of the received voice data is a new speaker; and

based on a determination that the particular speaker of the received voice data is the new speaker, transmit the received voice data to the destination terminals through a least recently used channel among the plurality of predefined channels.

16 . The voice data transmission apparatus of claim 11 , wherein

the particular speaker is one of a plurality of speakers, of which respective voice data are received from the voice data collection server, and

a number of the plurality of predefined channels is constant regardless of a number of the plurality of speakers.