IP Library Patent Application 16425513
Patent Application
App. No. 16/425,513

METHOD, DEVICE AND COMPUTER STORAGE MEDIUM FOR SPEECH INTERACTION

Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US None
App. No.
16/425,513
Abstract

A method, a device and a computer storage medium for speech interaction are disclosed. The method includes: receiving speech data transmitted by a first terminal device; obtaining a speech recognition result and a voiceprint recognition result of the speech data; obtaining a response text for the speech recognition result, and performing speech conversion for the response text with the voiceprint recognition result; and transmitting audio data obtained from the conversion to the first terminal device. Speech self-adaptation of human-machine interaction may be achieved, and the real feeling and interest of human-machine speech interaction may be enhanced and improved, respectively.

Claims (33)

1 . A method for speech interaction, comprising:

receiving speech data transmitted by a first terminal device;

obtaining a speech recognition result and a voiceprint recognition result of the speech data;

obtaining a response text for the speech recognition result, and performing speech conversion for the response text with the voiceprint recognition result; and

transmitting audio data obtained from the conversion to the first terminal device.

2 . The method according to claim 1 , wherein the voiceprint recognition result comprises at least one kind of identity information of user's gender, age, region and occupation.

3 . The method according to claim 1 , wherein the obtaining a response text for the speech recognition result comprises:

performing searching and matching with the speech recognition result to obtain at least one of a text search result and a prompt text corresponding to the speech recognition result.

4 . The method according to claim 3 , further comprising:

under the condition that an audio search result is obtained by performing searching and matching with the speech recognition result, transmitting the audio search result to the first terminal device.

5 . The method according to claim 1 , wherein the obtaining a response text for the speech recognition result comprises:

performing searching and matching with the speech recognition result and the voiceprint recognition result to obtain at least one of a text search result and a prompt text corresponding to the speech recognition result and the voiceprint recognition result.

6 . The method according to claim 1 , wherein the performing speech conversion for the response text with the voiceprint recognition result comprises:

determining a voice synthesis parameter corresponding to the voiceprint recognition result according to a correspondence relationship between preset identity information and the voice synthesis parameter; and

performing the speech conversion for the response text with the determined voice synthesis parameter.

7 . The method according to claim 6 , further comprising:

receiving and storing the correspondence relationship set by a second terminal device.

8 . The method according to claim 1 , wherein before performing speech conversion for the response text with the voiceprint recognition result, the method further comprises:

judging whether the first terminal device is set as a self-adaptive speech response, under the condition that the first terminal device is set as a self-adaptive speech response, continuing to perform speech conversion for the response text with the voiceprint recognition result; and

under the condition that the first terminal device is not set as a self-adaptive speech response, performing speech conversion for the response text with a preset or default voice synthesis parameter.

9 . A device, comprising:

one or more processors;

a storage for storing one or more programs,

said one or more programs are executed by said one or more processors to enable said one or more processors to implement a method for speech interaction, wherein the method comprises:

receiving speech data transmitted by a first terminal device;

obtaining a speech recognition result and a voiceprint recognition result of the speech data;

obtaining a response text for the speech recognition result, and performing speech conversion for the response text with the voiceprint recognition result; and

transmitting audio data obtained from the conversion to the first terminal device.

10 . A storage medium comprising computer-executable instructions, when the computer-executable instructions are executed by a computer processor, the computer-executable instructions being used to implement a method for speech interaction, wherein the method comprises:

receiving speech data transmitted by a first terminal device;

obtaining a speech recognition result and a voiceprint recognition result of the speech data;

obtaining a response text for the speech recognition result, and performing speech conversion for the response text with the voiceprint recognition result; and

transmitting audio data obtained from the conversion to the first terminal device.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded May 30, 2021
From: BAIDU ONLINE NETWORK TECHNOLOGY (BEIJING) CO., LTD.
To: BAIDU ONLINE NETWORK TECHNOLOGY (BEIJING) CO., LTD.; SHANGHAI XIAODU TECHNOLOGY CO. LTD.
Reel/Frame 056811/0772 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded May 29, 2019
From: CHANG, XIANTANG
To: BAIDU ONLINE NETWORK TECHNOLOGY (BEIJING) CO., LTD.
Reel/Frame 049310/0846 →