IP Library Granted Patent US 10,388,269
Granted Patent B2
US 10,388,269 · App. 15/583,068 · Granted Aug 20, 2019

System and method for intelligent language switching in automated text-to-speech systems

Inventors: Gregory Pulz (Cranbury, NJ); Harry E. Blanchard (Rumson, NJ); Lan Zhang (Malvern, PA)
Assignee: AT&T INTELLECTUAL PROPERTY I, L.P.
G10L13/086G06F17/289G10L13/047
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,388,269
App. No.
15/583,068
Granted
Aug 20, 2019
Kind
B2
Abstract

Systems, methods, and computer-readable storage media for providing for intelligent switching of languages and/or pronunciations in a text-to-speech system. As the system receives text, the text is analyzed to identify portions which should have speech constructed using a pronunciation distinct from the remaining portions of the text. The text-to-speech system uses multiple pronunciation dictionaries to generate and produce speech corresponding to the text, where the identified portions of the text are in a different language or have a different accent from the remainder of the text. Having generated speech corresponding to the text in multiple languages, accents, or dialects, the system combines the portions, then communicates the speech to the text recipient.

Claims (43)

1. A method comprising:

selecting, via a speech processing system, a first language for a first part of a text and a second language for a second part of the text;

generating, via the speech processing system and based on a first location of a device, first speech comprising a first portion corresponding to at least the first part of the text and a second portion corresponding to at least the second part of the text, the first portion in the first language and the second portion in the second language;

communicating the first speech to the device; and

when the device is at a second location:

generating, via the speech processing system, second speech from the text wherein the second speech comprises the first portion and the second portion both being in a same language; and

communicating the second speech to the device.

2. The method of claim 1 , wherein the first language is a primary language of a recipient and the second language is selected based on an original pronunciation of the second part of the text.

3. The method of claim 2 , wherein the first part of the text is an address number and the second part of the text is a street name.

4. The method of claim 1 , wherein the first language and the second language correspond to distinct regional accents of a single language.

5. The method of claim 1 , wherein one of the first language and the second language is selected based on one of an age, an ethnicity, and a language of a sender of the text.

6. The method of claim 1 , further comprising:

receiving, from a recipient, input indicating a category corresponding to one of the first part of the text and the second part of the text.

7. The method of claim 1 , wherein the generating of the first speech occurs on a mobile device.

8. The method of claim 1 , further comprising identifying the first portion and the second portion using a first language pronunciation database corresponding to the first language and a second language pronunciation database corresponding to the second language.

9. The method of claim 1 , wherein the first location differs from the second location.

10. A speech processing system comprising:

a processor; and

a computer-readable storage medium having instructions stored which, when executed by the processor, cause the processor to perform operations comprising:

selecting a first language for a first part of a text and a second language for a second part of the text;

generating, based on a first location of a device, first speech comprising a first portion corresponding to at least the first part of the text and a second portion corresponding to at least the second part of the text, the first portion in the first language and the second portion in the second language;

communicating the first speech to the device; and

when the device is at a second location:

generating second speech from the text wherein the second speech comprises the first portion and the second portion both being in a same language; and

communicating the second speech to the device.

11. The speech processing system of claim 10 , wherein the first language is a primary language of a recipient, and the second language is selected based on an original pronunciation of the second part of the text.

12. The speech processing system of claim 11 , wherein the first part of the text is an address number and the second part of the text is a street name.

13. The speech processing system of claim 10 , wherein the first language and the second language correspond to distinct regional accents of a single language.

14. The speech processing system of claim 10 , wherein one of the first language and the second language is selected based on one of an age, an ethnicity, and a language of a sender of the text.

15. The speech processing system of claim 10 , wherein the computer-readable storage medium stores additional instructions which, when exceeded by the processor, cause the processor to perform operations further comprising:

receiving, from a recipient, input indicating a category corresponding to one of the first part of the text and the second part of the text.

16. The speech processing system of claim 10 , wherein the generating of the first speech occurs on a mobile device.

17. The speech processing system of claim 10 , wherein the computer-readable storage medium stores additional instructions which, when exceeded by the processor, cause the processor to perform operations further comprising:

identifying the first portion and the second portion using a first language pronunciation database corresponding to the first language and a second language pronunciation database corresponding to the second language.

18. The speech processing system of claim 10 , wherein the first location differs from the second location.

19. A computer-readable storage device having instructions stored which, when executed by a speech processing system, cause the speech processing system to perform operations comprising:

selecting a first language for a first part of a text and a second language for a second part of the text;

generating, based on a first location of a device, first speech comprising a first portion corresponding to at least the first part of the text and a second portion corresponding to at least the second part of the text, the first portion in the first language and the second portion in the second language;

communicating the first speech to the device; and

when the device is at a second location:

generating second speech from the text wherein the second speech comprises the first portion and the second portion both being in a same language; and

communicating the second speech to the device.

20. The computer-readable storage device of claim 19 , wherein the first language is a primary language of a recipient and the second language is selected based on an original pronunciation of the second part of the text.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 17, 2021
From: AT&T INTELLECTUAL PROPERTY I, L.P.
To: HYUNDAI MOTOR COMPANY; KIA CORPORATION
Reel/Frame 058135/0446 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 7, 2018
From: PULZ, GREGORY; BLANCHARD, HARRY E.; ZHANG, LAN
To: AT&T INTELLECTUAL PROPERTY I, L.P.
Reel/Frame 045134/0686 →
Continuity (2)
Continuation 14022991 · Sep 10, 2013
Related Publication 20170236509A1 · Aug 17, 2017