IP Library Granted Patent US 9,607,609
Granted Patent B2
US 9,607,609 · App. 14/496,832 · Granted Mar 28, 2017

Method and apparatus to synthesize voice based on facial structures

Inventors: Shamim Begum (Beaverton, OR); Alexander A. Oganezov (Portland, OR)
Assignee: INTEL CORPORATION
G10L13/027G06K9/00315G10L13/047
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 9,607,609
App. No.
14/496,832
Granted
Mar 28, 2017
Kind
B2
Abstract

Disclosed are embodiments for use in an articulatory-based text-to-speech conversion system configured to establish an articulatory speech synthesis model of a person's voice based on facial characteristics defining exteriorly visible articulatory speech synthesis model parameters of the person's voice and on a predefined articulatory speech synthesis model selected from among stores of predefined models.

Claims (10)

1. An apparatus for use in an articulatory-based text-to-speech conversion system to establish an articulatory speech synthesis model of a first person's voice, the apparatus comprising:

a facial structure input device to acquire image data representing a visage of the first person, in which the visage includes facial characteristics defining exteriorly visible articulatory speech synthesis model parameters of the first person's voice;

a facial characteristics matching system to select a predefined articulatory speech synthesis model from among stores of predefined models modeling voices of different persons, each of the predefined models comprising exteriorly visible and interiorly concealed articulatory speech synthesis model parameters associated with a respective person that is different than the first person, the selection based at least in part on one or both of the facial characteristics or the exteriorly visible articulatory speech synthesis model parameters, associated with the first person, corresponding to a selected predefined articulatory speech synthesis model that is associated with one of the predefined models modeling a second person's voice, the second person being different than the first person; and

an articulatory modeling system to establish the articulatory speech synthesis model of the first person's voice by including in it the interiorly concealed articulatory speech synthesis model parameters available from the selected predefined articulatory speech synthesis model.

2. The apparatus of claim 1 , in which the selection is based on a measure of a face-matching correlation between the facial characteristics of the visage of the first person and facial characteristics defining the exteriorly visible articulatory speech synthesis model parameters of the predefined models.

3. The apparatus of claim 2 , in which the measure of face-matching correlation is derived using a hidden Markovian model.

4. The apparatus of claim 1 , in which the facial structure input device is configured to acquire the image data by capturing an image with an imager in a user equipment device.

5. The apparatus of claim 1 , in which the facial characteristics matching system is configured to select by comparing the one or both of the facial characteristics or the exteriorly visible articulatory speech synthesis model parameters to those of the predefined models.

6. The apparatus of claim 1 , in which the facial characteristics matching system is configured to select by communicating the image data from a user equipment device to a server for initiating a comparison of the one or both of the facial characteristics or the exteriorly visible articulatory speech synthesis model parameters to those of the predefined models.

7. The apparatus of claim 1 , in which the articulatory system is configured to synthesize speech based on the articulatory speech synthesis model of the first person's voice.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Sep 26, 2014
From: BEGUM, SHAMIM; OGANEZOV, ALEXANDER A.
To: INTEL CORPORATION
Reel/Frame 033830/0260 →
Continuity (1)
Related Publication 20160093284A1 · Mar 31, 2016