IP Library Granted Patent US 8,428,952
Granted Patent B2
US 8,428,952 · App. 13/494,164 · Granted Apr 23, 2013

Text-to-speech user's voice cooperative server for instant messaging clients

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 8,428,952
App. No.
13/494,164
Granted
Apr 23, 2013
Kind
B2
Abstract

A system and method to allow an author of an instant message to enable and control the production of audible speech to the recipient of the message. The voice of the author of the message is characterized into parameters compatible with a formative or articulative text-to-speech engine such that upon receipt, the receiving client device can generate audible speech signals from the message text according to the characterization of the author's voice. Alternatively, the author can store samples of his or her actual voice in a server so that, upon transmission of a message by the author to a recipient, the server extracts the samples needed only to synthesize the words in the text message, and delivers those to the receiving client device so that they are used by a client-side concatenative text-to-speech engine to generate audible speech signals having a close likeness to the actual voice of the author.

Claims (31)

1. A method comprising:

analyzing text within a body of a first user's text instant message to determine text-to-speech synthesis control parameters that are to be used to produce a synthesized audible representation of the text within the body of the text instant message;

extracting, from text-to-speech synthesis control parameters that are associated with the first user and comprise one or more voice synthesis control parameters which determine distinctive intelligible characteristics representative of the first user, a subset of the text-to-speech synthesis control parameters associated with the first user, the subset corresponding to the text-to-speech synthesis control parameters determined during the analyzing as those that are to be used to produce the synthesized audible representation of the text within the body of the text instant message;

sending the text instant message along with the subset of text-to-speech synthesis control parameters to a second user's device, the subset of text-to-speech synthesis control parameters being attached to the text instant message;

receiving the text instant message along with the subset of text-to-speech synthesis control parameters by the second user's device; and

at the second user's device, performing text-to-speech synthesis of the text instant message according to the subset of text-to-speech synthesis control parameters to produce the synthesized audible representation of the text within the body of the text instant message having the distinctive intelligible characteristics representative of the first user.

2. The method of claim 1 , further comprising establishing the text to speech synthesis control parameters associated with the first user, and wherein the step of establishing text-to-speech synthesis control parameters associated with the first user comprises establishing one or more voice characteristic parameters compatible with an articulative text-to-speech engine.

3. The method of claim 1 , further comprising establishing the text to speech synthesis control parameters associated with the first user, and wherein the step of establishing text-to-speech synthesis control parameters associated with the first user comprises establishing one or more phoneme samples of the first user's actual voice, the one or more phoneme samples being stored by a server and being compatible with a concatenative text-to-speech engine.

4. The method of claim 1 , wherein the first user is a sender of the text instant message.

5. The method of claim 1 , wherein the first user is an author of the text instant message.

6. The method of claim 1 , wherein sending the text instant message along with the subset of text-to-speech synthesis control parameters comprises sending the text instant message and the subset of text-to-speech synthesis control parameters from an authoring device.

7. The method of claim 1 , wherein sending the text instant message along with the subset of text-to-speech synthesis control parameters comprises sending the text instant message and the subset of text-to-speech synthesis control parameters from a server.

8. A method comprising:

analyzing text within a body of a first user's text instant message to determine text-to-speech synthesis control parameters that are to be used to produce a synthesized audible representation of the text within the body of the text instant message;

extracting, from text-to-speech synthesis control parameters that are associated with the first user and comprise one or more voice synthesis control parameters which determine distinctive intelligible characteristics representative of the first user, a subset of the text-to-speech synthesis control parameters associated with the first user, the subset corresponding to the text-to-speech synthesis control parameters determined during the analyzing as those that are to be used to produce the synthesized audible representation of the text within the body of the text instant message; and

sending the text instant message along with the subset of text-to-speech synthesis control parameters to a second user's device, the subset of text-to-speech synthesis control parameters being attached to the text instant message.

9. The method of claim 8 , wherein the first user is a sender of the text instant message.

10. The method of claim 8 , wherein the first user is an author of the text instant message.

11. The method of claim 8 , wherein sending the text instant message along with the subset of text-to-speech synthesis control parameters comprises sending the text instant message and the subset of text-to-speech synthesis control parameters from an authoring device.

12. The method of claim 8 , wherein sending the text instant message along with the subset of text-to-speech synthesis control parameters comprises sending the text instant message and the subset of text-to-speech synthesis control parameters from a server.

13. The method of claim 8 , wherein analyzing the text within the body of the text instant message to determine the text-to-speech synthesis control parameters that are to be used to produce the synthesized audible representation of the text within the body of the text instant message comprises analyzing the text within the body of the text instant message to determine which phonemes of a set of possible phonemes are to be used to produce the synthesized audible representation of the text within the body of the text instant message.

14. At least one computer readable storage device encoded with computer-readable instructions which, when executed, perform a method, the method comprising:

analyzing text within a body of a first user's text instant message to determine text-to-speech synthesis control parameters that are to be used to produce a synthesized audible representation of the text within the body of the text instant message; and

extracting, from text-to-speech synthesis control parameters that are associated with the first user and comprise one or more voice synthesis control parameters which determine distinctive intelligible characteristics representative of the first user, a subset of the text-to-speech synthesis control parameters associated with the first user, the subset corresponding to the text-to-speech synthesis control parameters determined during the analyzing as those that are to be used to produce the synthesized audible representation of the text within the body of the text instant message; and

sending the text instant message along with the subset of text-to-speech synthesis control parameters to a second user's device, the subset of text-to-speech synthesis control parameters being attached to the text instant message.

15. The at least one computer readable storage device of claim 14 , wherein the first user is a sender of the text instant message.

16. The at least one computer readable storage device of claim 14 , wherein the first user is an author of the text instant message.

17. The at least one computer readable storage device of claim 14 , wherein sending the text instant message along with the subset of text-to-speech synthesis control parameters comprises sending the text instant message and the subset of text-to-speech synthesis control parameters from a server.

18. The at least one computer readable storage device of claim 14 , wherein analyzing the text within the body of the text instant message to determine the text-to-speech synthesis control parameters that are to be used to produce the synthesized audible representation of the text within the body of the text instant message comprises analyzing the text within the body of the text instant message to determine which phonemes of a set of possible phonemes are to be used to produce the synthesized audible representation of the text within the body of the text instant message.

19. The at least one computer readable storage device of claim 14 , wherein sending the text instant message along with the subset of text-to-speech synthesis control parameters comprises sending the text instant message and the subset of text-to-speech synthesis control parameters from an authoring device.

20. The at least one computer readable storage device of claim 14 , wherein the method further comprises establishing the text to speech synthesis control parameters associated with the first user, and wherein the step of establishing text-to-speech synthesis control parameters associated with the first user comprises establishing one or more phoneme samples of the first user's actual voice, the one or more phoneme samples being stored by a server and being compatible with a concatenative text-to-speech engine.

Assignments (9)
RELEASE (REEL 052935 / FRAME 0584) Recorded Jan 2, 2025
From: WELLS FARGO BANK, NATIONAL ASSOCIATION
To: CERENCE OPERATING COMPANY
Reel/Frame 069797/0818 →
CORRECTIVE ASSIGNMENT TO CORRECT THE REPLACE THE CONVEYANCE DOCUMENT WITH THE NEW ASSIGNMENT PREVIOUSLY RECORDED AT REEL: 050836 FRAME: 0191. ASSIGNOR(S) HEREBY CONFIRMS THE ASSIGNMENT. Recorded Apr 19, 2022
From: NUANCE COMMUNICATIONS, INC.
To: CERENCE OPERATING COMPANY
Reel/Frame 059804/0186 →
SECURITY AGREEMENT Recorded Jun 15, 2020
From: CERENCE OPERATING COMPANY
To: WELLS FARGO BANK, N.A.
Reel/Frame 052935/0584 →
RELEASE OF SECURITY INTEREST Recorded Jun 12, 2020
From: BARCLAYS BANK PLC
To: CERENCE OPERATING COMPANY
Reel/Frame 052927/0335 →
SECURITY AGREEMENT Recorded Nov 7, 2019
From: CERENCE OPERATING COMPANY
To: BARCLAYS BANK PLC
Reel/Frame 050953/0133 →
CORRECTIVE ASSIGNMENT TO CORRECT THE ASSIGNEE NAME PREVIOUSLY RECORDED AT REEL: 050836 FRAME: 0191. ASSIGNOR(S) HEREBY CONFIRMS THE INTELLECTUAL PROPERTY AGREEMENT. Recorded Oct 29, 2019
From: NUANCE COMMUNICATIONS, INC.
To: CERENCE OPERATING COMPANY
Reel/Frame 050871/0001 →
INTELLECTUAL PROPERTY AGREEMENT Recorded Oct 23, 2019
From: NUANCE COMMUNICATIONS, INC.
To: CERENCE INC.
Reel/Frame 050836/0191 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jun 27, 2012
From: NIEMEYER, TERRY WADE; OROZCO, LILIANA
To: INTERNATIONAL BUSINESS MACHINES CORPORATION
Reel/Frame 028449/0250 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jun 27, 2012
From: INTERNATIONAL BUSINESS MACHINES CORPORATION
To: NUANCE COMMUNICATIONS, INC.
Reel/Frame 028449/0268 →