IP Library › Granted Patent US 10,873,661
Granted Patent B2
US 10,873,661 · App. 16/388,827 · Granted Dec 22, 2020

Voice communication method, voice communication apparatus, and voice communication system

Inventors: Qinghe Wang (Beijing, CN); Dongfang Wang (Beijing, CN); Tongshang Su (Beijing, CN); Leilei Cheng (Beijing, CN); Wei Song (Beijing, CN); Yang Zhang (Beijing, CN); Ning Liu (Beijing, CN); Haitao Wang (Beijing, CN); Jun Wang (Beijing, CN); Guangyao Li (Beijing, CN)
Assignees: HEFEI XINSHENG OPTOELECTRONICS TECHNOLOGY CO., LTD.; BOE TECHNOLOGY GROUP CO., LTD.
H04M3/22G06K9/00288G10L17/00H04M3/20
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,873,661
App. No.
16/388,827
Granted
Dec 22, 2020
Kind
B2
Abstract

A voice communication method, a voice communication apparatus, and a voice communication system are disclosed. The method includes: at a transmitting side, obtaining voice information; determining whether the voice information is uttered by a preset user, and transmitting the voice information to a peer device if it is determined that the voice information is uttered by the preset user, and prohibiting the transmission of the voice information otherwise; and at a receiving side, receiving voice information transmitted from a peer device; collecting a first environmental information, and determining whether the first environmental information meets a voice output condition; outputting the voice information if it is determined that the first environmental information meets the voice output condition, and prohibiting the output of the voice information otherwise.

Claims (82)

1. A method for improving security of voice communication at a voice communication apparatus on a transmitting side, the method comprising:

obtaining voice information;

determining whether the voice information is uttered by a preset user;

in response to determining that the voice information is uttered by the preset user:

determining a voice environment in which the voice communication apparatus on the transmitting side is located;

performing optimization processing on the voice information according to the voice environment; and

transmitting the optimized voice information to a peer device; and

in response to determining that the voice information is not uttered by the preset user, prohibiting the transmission of the voice information,

wherein the step of performing optimization processing on the voice information according to the voice environment comprises:

determining whether there is noise in the vicinity of the voice communication apparatus on the transmitting side according to the obtained voice environment;

determining the volume of the noise when there is noise; and

performing a noise reduction processing on the voice information and increasing the volume of the voice information when the volume of the noise is greater than a threshold, and

wherein the step of determining whether the voice information is uttered by a preset user comprises one or more of the steps of:

obtaining facial features of a person who utters the voice information, and determining whether the facial features are consistent with facial features of the preset user; or

obtaining action features of the person who utters the voice information, and determining whether the action features are consistent with action features of the preset user.

2. The method of claim 1 , wherein the step of determining whether the voice information is uttered by a preset user further comprises:

determining whether audio features of the voice information are consistent with audio features of the preset user.

3. A method for improving security of voice communication at a voice communication apparatus on a receiving side, the method comprising:

receiving voice information transmitted from a peer device;

collecting first environmental information of the voice communication apparatus on the receiving side;

determining whether the first environmental information meets a voice output condition;

in response to determining that the first environmental information meets the voice output condition, outputting the voice information; and

in response to determining that the first environmental information does not meet the voice output condition, prohibiting the output of the voice information,

wherein the step of determining whether the first environmental information meets the voice output condition comprises one or more of the steps of:

determining whether there is only one voice recipient;

determining whether facial features of the voice recipient are consistent with facial features of a preset receiving user; or

determining whether distances from other users than the voice recipient to the voice recipient exceed a preset threshold.

4. The method of claim 3 , wherein the step of outputting the voice information comprises:

collecting second environment information when the voice information is output; and

in response to determining that the second environment information does not meet a voice output condition, stopping outputting the voice information or switching to outputting a preset voice.

5. The method of claim 4 , wherein the step of outputting the voice information comprises:

in response to determining that the second environmental information does not meet the voice output condition, outputting an interference superimposed voice, the interference superimposed voice being a voice in which interfering audio is superimposed on the voice information.

6. The method of claim 3 , wherein the step of outputting the voice information comprises:

adjusting output volume of the voice information according to the first environment information, and outputting the adjusted voice information.

7. A voice communication apparatus on a receiving side for improving security of voice communication, the apparatus comprising:

a processor;

a memory storing instructions which, when executed by the processor, cause the processor to perform the method of claim 3 .

8. The voice communication apparatus of claim 7 , wherein the instructions, when executed by the processor, further cause the processor to collect a second environment information when the voice information is output; and:

in response to determining that the second environment information does not meet a voice output condition, stop outputting the voice information and switch to outputting a preset voice; or

in response to determining that the second environmental information does not meet the speech output condition, output an interference superimposed voice, the interference superimposed voice being a voice in which interfering audio is superimposed on the voice information.

9. The voice communication apparatus of claim 7 , wherein the instructions, when executed by the processor, further cause the processor to:

adjust output volume of the voice information according to the first environment information; and

output the adjusted voice information.

10. A system for improving security of voice communication, the system comprising:

one or more apparatuses for improving security of voice communication comprising:

a processor; and

a memory storing instructions which, when executed by the processor,

cause the processor to:

obtain voice information from a sound sensor;

determine whether the voice information is uttered by a preset user;

in response to determining that the voice information is uttered by the preset user

determine a voice environment in which the voice communication apparatus on the transmitting side is located;

perform optimization processing on the voice information according to the voice environment; and

transmit, via a communicator, the optimized voice information to a peer device; and

in response to determining that the voice information is not uttered by the preset user, prohibit the transmission of the voice information; and,

wherein the instructions which, when executed by the processor, further cause the processor to:

determine whether there is noise in the vicinity of the voice communication apparatus on the transmitting side according to the obtained voice environment;

determine the volume of the noise when there is noise; and

perform a noise reduction processing on the voice information and increase the volume of the voice information when the volume of the noise is greater than a threshold, and

wherein the instructions which, when executed by the processor, further cause the processor to:

obtain facial features of a person who utters the voice information, and determine whether the facial features are consistent with facial features of the preset user; or

obtain action features of the person who utters the voice information, and determining whether the action features are consistent with action features of the preset user; and

one or more apparatuses for improving security of voice communication according to claim 7 .

11. A voice communication apparatus on a transmitting side for improving security of voice communication, the apparatus comprising:

a processor; and

a memory storing instructions which, when executed by the processor, cause the processor to:

obtain voice information from a sound sensor;

determine whether the voice information is uttered by a preset user;

in response to determining that the voice information is uttered by the preset user;

determine a voice environment in which the voice communication apparatus on the transmitting side is located;

perform optimization processing on the voice information according to the voice environment; and

transmit, via a communicator, the optimized voice information to a peer device; and

in response to determining that the voice information is not uttered by the preset user, prohibit the transmission of the voice information,

wherein the instructions which, when executed by the processor, further cause the processor to:

determine whether there is noise in the vicinity of the voice communication apparatus on the transmitting side according to the obtained voice environment;

determine the volume of the noise when there is noise; and

perform a noise reduction processing on the voice information and increase the volume of the voice information when the volume of the noise is greater than a threshold, and

wherein the instructions which, when executed by the processor, further cause the processor to:

obtain facial features of a person who utters the voice information, and determine whether the facial features are consistent with facial features of the preset user; or

obtain action features of the person who utters the voice information, and determining whether the action features are consistent with action features of the preset user.

12. The voice communication apparatus of claim 11 , wherein the instructions, when executed by the processor, further cause the processor to perform the operations of:

determining whether audio features of the voice information are consistent with audio features of the preset user.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 18, 2019
From: WANG, QINGHE; WANG, DONGFANG; SU, TONGSHANG; CHENG, LEILEI; SONG, WEI; ZHANG, YANG; LIU, NING; WANG, HAITAO; WANG, JUN; LI, GUANGYAO
To: HEFEI XINSHENG OPTOELECTRONICS TECHNOLOGY CO., LTD.; BOE TECHNOLOGY GROUP CO., LTD.
Reel/Frame 048933/0533 →
Priority Claims (1)
CN 2018 1 1160149 · Sep 30, 2018 · national
Continuity (1)
Related Publication 20200106879A1 · Apr 2, 2020