IP Library › Granted Patent US 12,211,508
Granted Patent B2
US 12,211,508 · App. 17/788,591 · Granted Jan 28, 2025

Server-side processing method and server for actively initiating dialogue, and voice interaction system capable of initiating dialogue

Inventors: Weisi Shi (Suzhou, CN); Hongbo Song (Suzhou, CN); Chengya Zhu (Suzhou, CN); Shuai Fan (Suzhou, CN)
Assignee: AI SPEECH CO., LTD.
G10L15/30G10L15/22H04L67/141
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,211,508
App. No.
17/788,591
Granted
Jan 28, 2025
Kind
B2
Abstract

A server-side processing method for implementing an active initiation of a dialogue is disclosed, comprising: establishing a communication connection with a voice client, in response to a received request for establishing a connection from the voice client; receiving an information stream sent by the voice client through the communication connection; performing a dialogue decision-making process according to the information stream, obtaining and outputting an adapted dialogue content to the voice client upon determining that it is an active dialogue scenario. A server and a system for implementing an active initiation of a dialogue are also provided. The disclosed solutions realize intelligent decision-making for voice interaction, and can actively initiate a dialogue based on server-side decision-making, improving interaction experience and realizing intelligent interaction.

Claims (28)

1. A server-side processing method for actively initiating a dialogue, comprising:

establishing a communication connection with a voice client, in response to a received request for connecting from the voice client;

receiving an information stream sent by the voice client through the communication connection;

performing a dialogue decision-making process according to the information stream, obtaining and outputting an adapted dialogue content to the voice client upon determining that it is an active dialogue scenario,

wherein the performing a dialogue decision-making process according to the information stream, obtaining and outputting an adapted dialogue content to the voice client upon determining that it is an active dialogue scenario comprises

configuring a trigger condition of an active dialogue scenario and storage of dialogue content associated with the trigger condition;

determining whether it is an active dialogue scenario according to the information stream and the configured trigger condition of the active dialogue scenario, obtaining and outputting dialogue content stored in association with a current trigger condition to the voice client when the active dialogue scenario is determined,

wherein the information stream comprises audio information picked up by the voice client, the triggering condition comprises that recognition content is contained and the recognition content is invalid semantics, and the determining whether it is an active dialogue scenario according to the information stream and the configured trigger condition of the active dialogue scenario comprises

recognizing the audio information to obtain a recognition result;

determining whether the recognition result contains a recognition content, performing semantic parsing on the recognition content when the recognition content is contained, and when the semantic parsing result is invalid semantics, an active dialogue scenario is determined.

2. The server-side processing method according to claim 1 , wherein the triggering condition further comprises that recognition content is not contained while a corresponding context status exists, and the determining whether it is an active dialogue scenario according to the information stream and the configured trigger condition of the active dialogue scenario further comprises

determining whether the recognition result contains the recognition content, acquiring a context status of a voice interaction scenario for determination when the recognition content is not contained, and determining that the active dialogue scenario exists when the acquired context status of the voice interaction scenario is a corresponding context status in the trigger condition.

3. The server-side processing method according to claim 2 , wherein the corresponding context status in the trigger condition comprises waiting for inquiry and breaking silence.

4. The server-side processing method according to claim 1 , wherein the communication connection is a long connection of duplex communication.

5. An electronic device, comprising at least one processor and a memory communicatively connected to the at least one processor, wherein the memory stores instructions executable by the at least one processor, which are executed by the at least one processor to enable the at least one processor to perform steps of the method of claim 1 .

6. A non-transitory computer-readable_storage medium storing a computer program, wherein the program implements steps of the method of claim 1 when executed by a processor.

7. A server for implementing an active initiation of a dialogue, which is configured with

a communication module for establishing a communication connection with a voice client, in response to a received request for connecting from the voice client;

an information receiving module for receiving an information stream sent by the voice client through the communication connection;

a dialogue decision-making module for performing a dialogue decision-making process according to the information stream, obtaining and outputting an adapted dialogue content to the voice client when an active dialogue scenario is determined; and

a configuration module for configuring a trigger condition of an active dialogue scenario and storing dialogue content associated with the trigger condition;

the dialogue decision-making module comprises

a condition determination unit for determining whether it is an active dialogue scenario according to the information stream and the configured trigger condition of the active dialogue scenario, and calling a dialogue initiation unit when the active dialogue scenario is determined; and

the dialogue initiating unit for acquiring and outputting dialogue content stored in association with a current trigger condition to the voice client,

wherein the information stream comprises audio information picked up by the voice client, and the trigger condition comprises that the audio information contains recognition content and the recognition content is invalid semantics, and the audio information does not contain recognition content and has corresponding context status comprising waiting for inquiry and breaking silence.

8. A voice interaction system capable of actively initiating a dialogue, which comprises a voice client and a voice server-side, wherein,

the voice client is configured to initiate a connection request with the voice server-side, output collected audio information to the voice server-side in real time through the established communication connection after the communication connection is established, and play dialogue content sent by the voice server-side upon receiving the same; and

the voice server-side is the server for implementing the active initiation of the dialogue according to claim 7 .

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jun 23, 2022
From: SHI, WEISI; SONG, HONGBO; ZHU, CHENGYA; FAN, SHUAI
To: AI SPEECH CO., LTD.
Reel/Frame 060296/0618 →
Priority Claims (1)
CN 201911364247.0 · Dec 26, 2019 · national
Continuity (1)
Related Publication 20230037913A1 · Feb 9, 2023
References Cited (41)
US 6496799B1 · Pickering · 2002 [cited by applicant]
US 10331791B2 · Anbazhagan · 2019 [cited by examiner]
US 10891152B2 · Anbazhagan · 2021 [cited by examiner]
US 11423911B1 · Fu · 2022 [cited by examiner]
US 20040230434A1 · Galanes · 2004 [cited by examiner]
US 20050080629A1 · Attwater · 2005 [cited by examiner]
US 20050091059A1 · Lecoeuche · 2005 [cited by examiner]
US 20050154591A1 · Lecoeuche · 2005 [cited by examiner]
US 20150339745A1 · Peter · 2015 [cited by examiner]
US 20170110129A1 · Gelfenbeyn · 2017 [cited by examiner]
US 20170125008A1 · Maisonnier · 2017 [cited by examiner]
US 20170263269A1 · Kuo · 2017 [cited by examiner]
US 20180232436A1 · Elson et al. · 2018 [cited by applicant]
US 20180260856A1 · Balasubramanian · 2018 [cited by examiner]
US 20190115016A1 · Seok · 2019 [cited by examiner]
US 20190139547A1 · Wu · 2019 [cited by examiner]
US 20190251965A1 · Dharne · 2019 [cited by examiner]
US 20190347067A1 · Jolfaei · 2019 [cited by examiner]
CN 105975511A · 2016 [cited by applicant]
CN 106020488A · 2016 [cited by applicant]
CN 107004410A · 2017 [cited by applicant]
CN 108446286A · 2018 [cited by applicant]
CN 109036388A · 2018 [cited by applicant]
CN 109543010A · 2019 [cited by applicant]
CN 109658928A · 2019 [cited by applicant]
CN 110209792A · 2019 [cited by applicant]
CN 110211573A · 2019 [cited by applicant]
CN 110265009A · 2019 [cited by applicant]
CN 110442701A · 2019 [cited by applicant]
CN 111107156A · 2020 [cited by applicant]
JP 2016206469A · 2016 [cited by applicant]
JP 2017067849A · 2017 [cited by applicant]
KR 102047385B1 · 2019 [cited by applicant]
WO 2018151766A1 · 2018 [cited by applicant]
WO 2021129262A1 · 2021 [cited by applicant]
Foreign Communication from Related Application—First Chinese Office Action with English Translation, CN Patent Application No. 201911364247.0 filed Dec. 26, 2019, 17 pages. [cited by applicant]
Foreign Communication from Related Application—Second Chinese Office Action with English Translation, CN Patent Application No. 201911364247.0 filed Dec. 26, 2019, 19 pages. [cited by applicant]
Foreign Communication from Related Application—International Search Report and Written Opinion of the International Searching Authority, International Patent Application No. PCT/CN2020/130325 dated Feb. 18, 2021, with E… [cited by applicant]
Foreign Communication from Related Application—Communication Pursuant to Article 94(3) EPC, issued Apr. 29, 2024, EP Patent Application No. 20907823.7 filed Nov. 20, 2020, 5 pages. [cited by applicant]
Foreign Communication from Related Application—Supplementary European Search Report, issued May 19, 2023, EP Patent Application No. 20907823.7 filed Nov. 20, 2020, 8 pages. [cited by applicant]
Foreign Communication from Related Application—Notice of Reasons for Refusal with English Translation, issued May 9, 2023, JP Patent Application No. 2022538904 filed Nov. 20, 2020, 6 pages. [cited by applicant]