IP Library Granted Patent US 8,370,151
Granted Patent B2
US 8,370,151 · App. 12/687,196 · Granted Feb 5, 2013

Systems and methods for multiple voice document narration

Inventors: Raymond C. Kurzweil (Newton, MA); Paul Albrecht (Bedford, MA); Peter Chapman (Bedford, MA)
Assignee: K-NFB Reading Technology, Inc.
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 8,370,151
App. No.
12/687,196
Granted
Feb 5, 2013
Kind
B2
Abstract

Disclosed are techniques and systems to provide a narration of a text in multiple different voices where the portions of the text narrated using the different voices are selected by a user.

Claims (66)

1. A computer implemented method, comprising:

displaying on a display device text from an electronic document that has a sequence of words the sequence of words rendered by one or more computing devices in a user interface rendered on a display device;

applying, by one or more computer devices, in response to a user-based selection of a first portion of words in the sequence of words, a first indicium to the user-selected first portion of words in the sequence of words;

associating, by the one or more computer devices, a first narration voice to the first portion of words in the sequence of words; and

associating, by the one or more computers, a second, different narration voice to a second portion of words in the sequence of words, the second portion of the words in the sequence of words being different from the first portion of words in the sequence of words.

2. The method of claim 1 wherein the second portion of words in the sequence of words are user selected, the method further comprises:

applying, in response to the user selected second portion of words in the sequence of words, a second indicium.

3. The method of claim 1 wherein the second portion of words in the sequence of words do not have an assigned indicium.

4. The method of claim 3 , further comprising:

applying, by the one or more computing devices in response to a user-based selection of a third portion of words in the sequence of words, a second indicium to the user-selected third portion of words in the sequence of words; and

associating a third narration voice to the third portion of words in the sequence of words.

5. The method of claim 1 wherein the first and second narration voices are selected from the group consisting of text-to-speech synthesized voice, an audio recording of speech and a voice model that defines and controls features of the voice.

6. The method of claim 1 wherein the first and second narration voices are audio recordings of speech.

7. The method of claim 1 further comprising: receiving the user selection of one of a plurality of narration voices to associate with the selected first portion of words in response to a user selecting a voice model from a drop down menu.

8. The method of claim 1 , further comprising:

generating, by the one or more computer systems, an audible output corresponding to the words in the sequence of words, with the words in the first portion of words narrated using the first narration voice and the words in the second portion of words remaining unselected and being narrated using the second narration voice.

9. The method of claim 8 , wherein the second narration voice comprises a default narration voice.

10. The method of claim 1 , wherein the first indicium comprises a highlighting applied over the user-based selection of the first portion of words in the sequence of words.

11. The method of claim 1 , further comprising:

modifying one or both of the first and second narration voices by at least one of modifying a reading speed associated with the voice model, modifying a volume associated with the voice model, modifying the gender of the character associated with the voice model, modifying the age of the character and modifying a language of the voice model.

12. The method of claim 1 , further comprising:

automatically identifying by the one or more computers third portions of text in the sequence of words; and

automatically associating by the one or more computers the third portions of text with a third, different narration voice model.

13. A computer program product residing on a computer readable medium, the computer program product comprising instructions for causing a processor to:

display an electronic document having sequence of words in a user interface rendered on a display device;

apply in response to a user-based selection of a first portion of words in the sequence of words, a first indicium to the user-selected first portion of words in the sequence of words;

associate a first narration voice to the first portion of words in the sequence of words; and

associate a second, different narration voice to a second portion of words in the sequence of words, the second portion of the words in the sequence of words being different from the first portion of words in the sequence of words.

14. The computer program product of claim 13 wherein the second portion of words in the sequence of words are user selected and the computer program product further comprises instructions for causing the processor to:

apply, in response to the user selected second portion of words in the sequence of words, a second indicium.

15. The computer program product of claim 13 , wherein the second portion of words in the sequence of words do not have an assigned indicium.

16. The computer program product of claim 15 , wherein the computer program product further comprises instructions for causing the processor to:

apply, in response to a user-based selection of a third portion of words in the sequence of words, a second indicium to the user-selected third portion of words in the sequence of words; and

associate a third narration voice to the third portion of words in the sequence of words.

17. The computer program product of claim 13 wherein the computer program product further comprises instructions for causing the processor to:

display a drop down menu including a plurality of narration voice models; and

receive the user selection of one of a plurality of narration voice models to associate with the selected first portion of words.

18. The computer program product of claim 13 , wherein the computer program product further comprises instructions for causing the processor to:

generate an audible output corresponding to the words in the sequence of words, with the words in the first portion of words narrated using the first narration voice and the words in the second portion of words remaining unselected and being narrated using the second narration voice.

19. The computer program product of claim 13 , wherein the computer program product further comprises instructions for causing the processor to:

modify one or both of the first and second narration voices by at least one of modifying a reading speed associated with the voice model, modifying a volume associated with the voice model, modifying the gender of the character associated with the voice model, modifying the age of the character and modifying a language of the voice model.

20. The computer program product of claim 13 , further comprising instructions to:

automatically identify third portions of text in the sequence of words; and

automatically associate the third portions of text with a third, different narration voice model.

21. A system comprising:

a display device;

a memory; and

a computing device configured to:

access an electronic version of a document;

display a sequence of words from the electronic document in a user interface rendered on the display device;

apply in response to a user-based selection of a first portion of words in the sequence of words from the document, a first indicium to the user-selected first portion of words in the sequence of words;

associate a first narration voice to the first portion of words in the sequence of words; and

associate a second narration voice to a second portion of words in the sequence of words, the second portion of the words in the sequence of words being different from the first portion of words in the sequence of words.

22. The system of claim 21 wherein the second portion of words in the sequence of words are user selected and the computing device is further configured to:

apply, in response to the user selected second portion of words in the sequence of words, a second indicium.

23. The system of claim 21 wherein the second portion of words in the sequence of words do not have an assigned indicium and the system automatically associates the second narration voice with a default narration voice.

24. The system of claim 21 , wherein the computing device is further configured to:

apply, in response to a user-based selection of a third portion of words in the sequence of words, a second indicium to the user-selected third portion of words in the sequence of words; and

associate a third narration voice to the third portion of words in the sequence of words.

25. The system of claim 21 , wherein the computing device is further configured to:

generate an audible output corresponding to the words in the sequence of words, with the words in the first portion of words narrated using the first narration voice and the words in the second portion of words remaining unselected and being narrated using the second narration voice.

26. The system of claim 21 , wherein the computing device is further configured to:

modify one or both of the first and second narration voices by at least one of modifying a reading speed associated with the voice model, modifying a volume associated with the voice model, modifying the gender of the character associated with the voice model, modifying the age of the character and modifying a language of the voice model.

27. The system of claim 21 , further configured to:

automatically identify third portions of text in the sequence of words; and

automatically associate the third portions of text with a third, different narration voice model.

Assignments (8)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 12, 2021
From: EM ACQUISITION CORP., INC.
To: T PLAY HOLDINGS LLC
Reel/Frame 055896/0481 →
RELEASE OF SECURITY INTEREST Recorded Sep 17, 2015
From: FISH & RICHARDSON P.C.
To: DIMENSIONAL STACK ASSETS LLC
Reel/Frame 036629/0762 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Sep 17, 2015
From: DIMENSIONAL STACK ASSETS, LLC
To: EM ACQUISITION CORP., INC.
Reel/Frame 036593/0328 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded May 1, 2015
From: K-NFB READING TECHNOLOGY, INC.
To: DIMENSIONAL STACK ASSETS LLC
Reel/Frame 035546/0205 →
LIEN Recorded Dec 30, 2014
From: K-NFB HOLDING TECHNOLOGY, IMC.
To: FISH & RICHARDSON P.C.
Reel/Frame 034599/0860 →
CHANGE OF NAME Recorded Mar 21, 2013
From: K-NFB READING TECHNOLOGY, INC.
To: K-NFB HOLDING TECHNOLOGY, INC.
Reel/Frame 030058/0669 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 21, 2013
From: K-NFB HOLDING TECHNOLOGY, INC.
To: K-NFB READING TECHNOLOGY, INC.
Reel/Frame 030059/0351 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Sep 1, 2010
From: KURZWEIL, RAYMOND C.; ALBRECHT, PAUL; CHAPMAN, PETER
To: K-NFB READING TECHNOLOGY, INC.
Reel/Frame 024921/0268 →
Continuity (2)
Provisional Application 61144947 · Jan 15, 2009
Related Publication 20100318362A1 · Dec 16, 2010