IP Library Granted Patent US 7,401,020
Granted Patent B2
US 7,401,020 · App. 10/306,950 · Granted Jul 15, 2008

Application of emotion-based intonation and prosody to speech in text-to-speech systems

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 7,401,020
App. No.
10/306,950
Granted
Jul 15, 2008
Kind
B2
Abstract

A text-to-speech system that includes an arrangement for accepting text input, an arrangement for providing synthetic speech output, and an arrangement for imparting emotion-based features to synthetic speech output. The arrangement for imparting emotion-based features includes an arrangement for accepting instruction for imparting at least one emotion-based paradigm to synthetic speech output, as well as an arrangement for applying at least one emotion-based paradigm to synthetic speech output.

Claims (14)

1. A method of converting text to speech, said method comprising the steps of:

accepting text input;

providing synthetic speech output corresponding to the text input;

imparting emotion-based features to synthetic speech output;

said step of imparting emotion-based features comprising:

accepting instruction for imparting at least one emotion-based paradigm to synthetic speech output, wherein said step of accepting instruction further comprises accepting emotion-based commands from a user interface; and

applying at least one emotion-based paradigm to synthetic speech output, said step of applying at least one emotion-based paradigm to synthetic speech output comprising:

altering at least one segment to be used in synthetic speech output, whereby emotion in speech is reflected in how individual words or syllables are stressed;

altering at least one prosodic pattern to be used in synthetic speech output, whereby emotion in speech is reflected in prosodic patterns; and

selectably applying a single emotion-based paradigm over a single utterance of synthetic speech output; or

applying a variable emotion-based paradigm over individual segments of an utterance of synthetic speech output.

2. The method according to claim 1 , wherein said step of accepting instruction comprises accepting commands from an emotion-based markup language associated with the user interface.

3. The method according to claim 1 , wherein said step of applying at least one emotion-based paradigm comprises altering at least one of: prosody, intonation, and intonation intensity in synthetic speech output.

4. The method according to claim 1 , wherein said step of applying at least one emotion-based paradigm comprises altering at least one of speed and amplitude in order to affect prosody, intonation and intonation intensity in synthetic speech output.

Assignments (4)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 14, 2023
From: NUANCE COMMUNICATIONS, INC.
To: MICROSOFT TECHNOLOGY LICENSING, LLC
Reel/Frame 065578/0676 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 6, 2009
From: INTERNATIONAL BUSINESS MACHINES CORPORATION
To: NUANCE COMMUNICATIONS, INC.
Reel/Frame 022354/0566 →
RECORD TO CORRECT TITLE OF INVENTION ON AN ASSIGNMENT PREVIOUSLY RECORDED ON REEL 013547 FRAME 0621. (ASSIGNMENT OF ASSIGNOR'S INTEREST) Recorded Jul 21, 2003
From: EIDE, ELLEN M.
To: IBM CORPORATION
Reel/Frame 014296/0425 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 29, 2002
From: EIDE, ELLEN M.
To: IBM CORPORATION
Reel/Frame 013547/0621 →