IP Library Granted Patent US 9,026,445
Granted Patent B2
US 9,026,445 · App. 13/847,850 · Granted May 5, 2015

Text-to-speech user's voice cooperative server for instant messaging clients

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 9,026,445
App. No.
13/847,850
Granted
May 5, 2015
Kind
B2
Abstract

A system and method to allow an author of an instant message to enable and control the production of audible speech to the recipient of the message. The voice of the author of the message is characterized into parameters compatible with a formative or articulative text-to-speech engine such that upon receipt, the receiving client device can generate audible speech signals from the message text according to the characterization of the author's voice. Alternatively, the author can store samples of his or her actual voice in a server so that, upon transmission of a message by the author to a recipient, the server extracts the samples needed only to synthesize the words in the text message, and delivers those to the receiving client device so that they are used by a client-side concatenative text-to-speech engine to generate audible speech signals having a close likeness to the actual voice of the author.

Claims (24)

1. A method comprising:

analyzing text within a body of a first user's text instant message to determine text-to-speech synthesis control parameters that are to be used to produce a synthesized audible representation of the text within the body of the text instant message; and

extracting, from text-to-speech synthesis control parameters that are associated with the first user and comprise one or more voice synthesis control parameters which determine distinctive intelligible characteristics representative of the first user, a subset of the text-to-speech synthesis control parameters associated with the first user;

wherein the text to speech synthesis control parameters are compatible with a Local Frequency Oscillator (LFO) method of voice synthesis and are to be used to produce the synthesized audible representation of the text within the body of the text instant message.

2. The method of claim 1 , further comprising sending the text instant message and the subset of text-to-speech synthesis control parameters, attached to the text instant message, to a second user's device;

receiving the text instant message along with the subset of text-to-speech synthesis control parameters by the second user's device; and

at the second user's device, performing text-to-speech synthesis of the text instant message implementing the subset of text-to-speech synthesis control parameters to produce the synthesized audible representation of the text within the body of the text instant message having the distinctive intelligible characteristics representative of the first user.

3. The method of claim 2 , wherein receiving the text instant message along with the subset of text-to-speech synthesis control parameters by the second user's device comprises receiving the text instant message along with the subset of text-to-speech synthesis control parameters by a portable device.

4. The method of claim 1 , wherein the first user is an author of the text instant message.

5. The method of claim 1 , wherein extracting, from text-to-speech synthesis control parameters that are associated with the first user and comprise one or more voice synthesis control parameters which determine distinctive intelligible characteristics representative of the first user, a subset of the text-to-speech synthesis control parameters associated with the first user comprises extracting the subset from a server.

6. At least one computer-readable storage device encoded with computer-readable instructions which, when executed, causes performance of a method, the method comprising:

analyzing text within a body of a first user's text instant message to determine text-to-speech synthesis control parameters that are to be used to produce a synthesized audible representation of the text within the body of the text instant message; and

extracting, from text-to-speech synthesis control parameters that are associated with the first user and comprise one or more voice synthesis control parameters which determine distinctive intelligible characteristics representative of the first user, a subset of the text-to-speech synthesis control parameters associated with the first user;

wherein the text to speech synthesis control parameters are compatible with a Local Frequency Oscillator (LFO) method of voice synthesis and are to be used to produce the synthesized audible representation of the text within the body of the text instant message.

7. The at least one computer-readable storage device of claim 6 , wherein the method further comprises sending the text instant message and the subset of text-to-speech synthesis control parameters, attached to the text instant message, to a second user's device.

8. The at least one computer-readable storage device of claim 7 , wherein sending the text instant message and the subset of text-to-speech synthesis control parameters, attached to the text instant message, to a second user's device comprises sending the text instant message from a portable device.

9. The at least one computer-readable storage device of claim 6 , wherein the first user is an author of the text instant message.

10. The at least one computer-readable storage device of claim 6 , wherein extracting, from text-to-speech synthesis control parameters that are associated with the first user and comprise one or more voice synthesis control parameters which determine distinctive intelligible characteristics representative of the first user, a subset of the text-to-speech synthesis control parameters associated with the first user comprises extracting the subset from a server.

11. A method, comprising:

receiving with a receiving device a text instant message together with one or more text-to-speech synthesis control parameters including one or more voice synthesis control parameters which determine distinctive intelligible characteristics representative of an author of the text instant message, the one or more text-to-speech synthesis control parameters representing a subset of a larger set of text-to-speech synthesis control parameters associated with the author and determining the distinctive intelligible characteristics representative of the author of the text instant message,

wherein receiving with a receiving device a text instant message together with one or more text-to-speech synthesis control parameters including one or more voice synthesis control parameters comprises receiving parameters compatible with a Local Frequency Oscillator (LFO) method of voice synthesis.

12. The method of claim 11 , further comprising performing text-to-speech synthesis on the text instant message with the receiving device by using the one or more text-to-speech synthesis control parameters to produce a synthesized audible representation of the text instant message having the distinctive intelligible characteristics of the author.

13. The method of claim 12 , further comprising deleting the one or more text-to-speech synthesis control parameters from the receiving device subsequent to performing the text-to-speech synthesis.

14. The method of claim 11 , further comprising temporarily storing the one or more text-to-speech synthesis control parameters on the receiving device.

Assignments (9)
RELEASE (REEL 052935 / FRAME 0584) Recorded Jan 2, 2025
From: WELLS FARGO BANK, NATIONAL ASSOCIATION
To: CERENCE OPERATING COMPANY
Reel/Frame 069797/0818 →
CORRECTIVE ASSIGNMENT TO CORRECT THE REPLACE THE CONVEYANCE DOCUMENT WITH THE NEW ASSIGNMENT PREVIOUSLY RECORDED AT REEL: 050836 FRAME: 0191. ASSIGNOR(S) HEREBY CONFIRMS THE ASSIGNMENT. Recorded Apr 19, 2022
From: NUANCE COMMUNICATIONS, INC.
To: CERENCE OPERATING COMPANY
Reel/Frame 059804/0186 →
SECURITY AGREEMENT Recorded Jun 15, 2020
From: CERENCE OPERATING COMPANY
To: WELLS FARGO BANK, N.A.
Reel/Frame 052935/0584 →
RELEASE OF SECURITY INTEREST Recorded Jun 12, 2020
From: BARCLAYS BANK PLC
To: CERENCE OPERATING COMPANY
Reel/Frame 052927/0335 →
SECURITY AGREEMENT Recorded Nov 7, 2019
From: CERENCE OPERATING COMPANY
To: BARCLAYS BANK PLC
Reel/Frame 050953/0133 →
CORRECTIVE ASSIGNMENT TO CORRECT THE ASSIGNEE NAME PREVIOUSLY RECORDED AT REEL: 050836 FRAME: 0191. ASSIGNOR(S) HEREBY CONFIRMS THE INTELLECTUAL PROPERTY AGREEMENT. Recorded Oct 29, 2019
From: NUANCE COMMUNICATIONS, INC.
To: CERENCE OPERATING COMPANY
Reel/Frame 050871/0001 →
INTELLECTUAL PROPERTY AGREEMENT Recorded Oct 23, 2019
From: NUANCE COMMUNICATIONS, INC.
To: CERENCE INC.
Reel/Frame 050836/0191 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded May 30, 2013
From: NIEMEYER, TERRY WADE; OROZCO, LILIANA
To: INTERNATIONAL BUSINESS MACHINES CORPORATION
Reel/Frame 030515/0320 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded May 30, 2013
From: INTERNATIONAL BUSINESS MACHINES CORPORATION
To: NUANCE COMMUNICATIONS, INC.
Reel/Frame 030515/0342 →