IP Library Granted Patent US 8,321,223
Granted Patent B2
US 8,321,223 · App. 12/472,724 · Granted Nov 27, 2012

Method and system for speech synthesis using dynamically updated acoustic unit sets

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 8,321,223
App. No.
12/472,724
Granted
Nov 27, 2012
Kind
B2
Abstract

A method for performing speech synthesis on textual content at a client. The method includes the steps of: performing speech synthesis on the textual content based on a current acoustical unit set S current in a corpus at the client; analyzing the textual content and generating a list of target units with corresponding context features, selecting multiple acoustical unit candidates for each target unit according to the context features based on an acoustical unit set S total that is more plentiful than the current acoustical unit set S current in the corpus at the client, and determining acoustical units suitable for speech synthesis for the textual content according to the multiple unit candidates; and updating the current acoustical unit set S current in the corpus at the client based on the determined acoustical units.

Claims (33)

1. A method for performing speech synthesis of textual content at a client data processing system, the method comprising:

performing speech synthesis of textual content based on a current acoustical unit set S current in a corpus at said client;

analyzing said textual content and generating a list of target units with corresponding context features;

selecting multiple acoustical unit candidates for each target unit according to said context features based on an acoustical unit set S total that is more plentiful than the current acoustical unit set S current in the corpus at said client;

determining acoustical units suitable for speech synthesis for said textual content according to said multiple unit candidates; and

updating the current acoustical unit set S current in the corpus at said client based on the determined acoustical units,

wherein at least one of the performing, analyzing, selecting, determining and updating is performed by a processing device.

2. The method according to claim 1 , further comprising the step of:

downloading a set S 0 of a small number of acoustical units, which can perform speech synthesis to all kinds of textual contents and which can ensure an acceptable speech synthesis quality, as an initial current acoustical unit set in the corpus on said client to make S current =S 0 .

3. The method according to claim 1 , wherein said step of determining acoustical units further comprises:

ranking said multiple acoustical unit candidates to determine, according to importance for the textual content, an acoustical unit set for updating the current acoustical unit set in the corpus at said client.

4. The method according to claim 3 , further comprising the step of:

downloading into said client an acoustical unit set S Δ which (i) belongs to the acoustical unit set used for update and (ii) is not included in the current acoustical unit set in the corpus at said client; and

wherein the current acoustical unit set S current in the corpus on said client is updated by making S current =S current +S Δ in said updating step.

5. The method according to claim 3 , wherein the unit candidates are ranked based on how many times each unit candidate has been selected.

6. The method according to claim 5 , wherein multiple acoustical unit candidates of different target units are ranked together.

7. The method according to claim 5 , wherein multiple acoustical unit candidates of each target unit are ranked separately.

8. A system having at least one processing device for enabling speech synthesis of textual content at a client data processing system, the system comprising:

speech synthesis means configured to perform speech synthesis of textual content based on a current acoustical unit set S current in a corpus on said client;

analysis means configured to analyze said textual content and generate a list of target units with corresponding context features;

selection means configured to select multiple acoustical unit candidates for each target unit according to said context features based on an acoustical unit set S total that is more plentiful than the current acoustical unit set S current in the corpus at said client;

determining means configured to determine acoustical units suitable for speech synthesis for said textual content according to said multiple unit candidates; and

update means configured to update the current acoustical unit set S current in the corpus on said client based at the determined acoustical units.

9. The system according to claim 8 , further comprising:

means configured to download a set S 0 of a small number of acoustical units which can perform speech synthesis to all kinds of textual contents and which can ensure an acceptable speech synthesis quality, as an initial current acoustical unit set in the corpus on said client to make S current =S 0 .

10. The system according to claim 8 , wherein said determining means comprises:

means to rank said multiple acoustical unit candidates to determine, according to importance for the textual content, an acoustical unit set for updating the current acoustical unit set in the corpus at said client.

11. The system according to claim 10 , wherein said determining means further comprises:

means for determining an acoustical unit set S Δ which (i) belongs to the acoustical unit set used for update and (ii) is not included in the current acoustical unit set in the corpus at said client; and

wherein said update means is configured to update the current acoustical unit set S current in the corpus on said client by making S current =S current +S Δ .

12. The system according to claim 10 , wherein said determining means is configured to rank the unit candidates based on how many times each unit candidate has been selected.

13. The system according to claim 12 , wherein said determining means comprises means to rank multiple acoustical unit candidates of different target units together.

14. The system according to claim 12 , wherein said determining means comprises means to separately rank multiple acoustical unit candidates of each target unit.

Assignments (9)
RELEASE (REEL 052935 / FRAME 0584) Recorded Jan 2, 2025
From: WELLS FARGO BANK, NATIONAL ASSOCIATION
To: CERENCE OPERATING COMPANY
Reel/Frame 069797/0818 →
CORRECTIVE ASSIGNMENT TO CORRECT THE REPLACE THE CONVEYANCE DOCUMENT WITH THE NEW ASSIGNMENT PREVIOUSLY RECORDED AT REEL: 050836 FRAME: 0191. ASSIGNOR(S) HEREBY CONFIRMS THE ASSIGNMENT. Recorded Apr 19, 2022
From: NUANCE COMMUNICATIONS, INC.
To: CERENCE OPERATING COMPANY
Reel/Frame 059804/0186 →
SECURITY AGREEMENT Recorded Jun 15, 2020
From: CERENCE OPERATING COMPANY
To: WELLS FARGO BANK, N.A.
Reel/Frame 052935/0584 →
RELEASE OF SECURITY INTEREST Recorded Jun 12, 2020
From: BARCLAYS BANK PLC
To: CERENCE OPERATING COMPANY
Reel/Frame 052927/0335 →
SECURITY AGREEMENT Recorded Nov 7, 2019
From: CERENCE OPERATING COMPANY
To: BARCLAYS BANK PLC
Reel/Frame 050953/0133 →
CORRECTIVE ASSIGNMENT TO CORRECT THE ASSIGNEE NAME PREVIOUSLY RECORDED AT REEL: 050836 FRAME: 0191. ASSIGNOR(S) HEREBY CONFIRMS THE INTELLECTUAL PROPERTY AGREEMENT. Recorded Oct 29, 2019
From: NUANCE COMMUNICATIONS, INC.
To: CERENCE OPERATING COMPANY
Reel/Frame 050871/0001 →
INTELLECTUAL PROPERTY AGREEMENT Recorded Oct 23, 2019
From: NUANCE COMMUNICATIONS, INC.
To: CERENCE INC.
Reel/Frame 050836/0191 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded May 9, 2013
From: INTERNATIONAL BUSINESS MACHINES CORPORATION
To: NUANCE COMMUNICATIONS, INC.
Reel/Frame 030381/0123 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 5, 2009
From: MENG, FAN PING; QIN, YONG; SHI, QIN; SHUANG, ZHIWEI
To: INTERNATIONAL BUSINESS MACHINES CORPORATION
Reel/Frame 023053/0429 →