IP Library Granted Patent US 8,972,259
Granted Patent B2
US 8,972,259 · App. 12/878,402 · Granted Mar 3, 2015

System and method for teaching non-lexical speech effects

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 8,972,259
App. No.
12/878,402
Granted
Mar 3, 2015
Kind
B2
Abstract

A method and system for teaching non-lexical speech effects includes delexicalizing a first speech segment to provide a first prosodic speech signal and data indicative of the first prosodic speech signal is stored in a computer memory. The first speech segment is audibly played to a language student and the student is prompted to recite the speech segment. The speech uttered by the student in response to the prompt, is recorded.

Claims (57)

1. A method, comprising:

delexicalizing a speech segment to generate a first prosodic speech signal;

encoding the first prosodic speech signal to generate a musically encoded prosodic speech signal;

storing data indicative of the musically encoded prosodic speech signal in a computer memory;

audibly playing the musically encoded prosodic speech signal to a language student;

prompting, in response to the audibly playing, the student to recite the speech segment from which the musically encoded prosodic speech signal originated;

recording audible speech uttered by the student in response to the prompt;

delexicalizing the audible speech to generate a second prosodic speech signal;

calculating at least one error signal based on a difference between:

a difference between a duration of a first syllable in the first prosodic speech signal and a duration of a second syllable in the first prosodic speech signal, the first syllable in the first prosodic speech signal being non-adjacent to the second syllable in the first prosodic speech signal; and

a difference between a duration of a first syllable in the second prosodic speech signal and a duration of a second syllable in the second prosodic speech signal, the first syllable in the second prosodic speech signal being non-adjacent to the second syllable in the second prosodic speech signal.

2. The method of claim 1 , further comprising:

determining prosodic characteristics of the first prosodic speech signal and the second prosodic speech signal, the calculating the at least one error signal being based on the prosodic characteristics of the first prosodic speech signal and the second prosodic speech signal.

3. The method of claim 1 , further comprising:

determining prosodic characteristics of the first prosodic speech signal and the second prosodic speech signal, the calculating the at least one error signal being based on the prosodic characteristics of the first prosodic speech signal and the second prosodic speech signal, the prosodic characteristics including rhythm, the prosodic characteristics further including at least one of prosody; duration in time; a total number of syllables; or a pitch level of at least one syllable.

4. The method of claim 1 , wherein

the at least one error signal is further based on at least one of a difference in prosody between the first prosodic speech signal and the second prosodic speech signal; a difference in a total number of syllables between the first prosodic speech signal and the second prosodic speech signal; a difference in duration in units of time between the first prosodic speech signal and the second prosodic speech signal; a difference in pitch between respective syllables between the first prosodic speech signal and the second prosodic speech signal; or a difference in intonation between the first prosodic speech signal and the second prosodic speech signal.

5. The method of claim 1 , wherein the encoding includes implementing a musical instrument digital interface (MIDI) framework.

6. The method of claim 1 , wherein the encoding includes converting a frequency of the first prosodic speech signal to a musical instrument digital interface (MIDI) pitch note number.

7. The method of claim 1 , wherein the calculating the at least one error signal is further based on a sum of the difference between:

the difference between the duration of the first syllable in the first prosodic speech signal and the duration of the second syllable in the first prosodic speech signal, the first syllable in the first prosodic speech signal being non-adjacent to the second syllable in the first prosodic speech signal; and

the difference between the duration of the first syllable in the second prosodic speech signal and the duration of the second syllable in the second prosodic speech signal, the first syllable in the second prosodic speech signal being non-adjacent to the second syllable in the second prosodic speech signal.

8. The method of claim 7 , wherein the calculating the at least one error signal is further based on a normalization of the sum of the difference.

9. A method, comprising:

audibly presenting a non-lexicalized musically encoded phrase in a target language to a student;

receiving, at a computer, a lexicalized version of the non-lexicalized musically encoded phrase in the target language from the student;

comparing, at the computer, non-lexical properties of the non-lexicalized musically encoded phrase with non-lexical properties of the lexicalized version of the non-lexicalized musically encoded phrase to calculate at least one error signal based on a difference between:

a difference between a duration of a first syllable in the non-lexicalized musically encoded phrase and a duration of a second syllable in the non-lexicalized musically encoded phrase, the first syllable in the non-lexicalized musically encoded phrase being non-adjacent to the second syllable in the non-lexicalized musically encoded phrase; and

a difference between a duration of a first syllable in the lexicalized version of the non-lexicalized musically encoded phrase and a duration of a second syllable in the lexicalized version of the non-lexicalized musically encoded phrase, the first syllable in the lexicalized version of the non-lexicalized musically encoded phrase being non-adjacent to the second syllable in the lexicalized version of the non-lexicalized musically encoded phrase; and

adjusting a future lesson for the student based upon the error signal.

10. The method of claim 9 , wherein the non-lexical properties of the non-lexicalized musically encoded phrase and the lexicalized version of the non-lexicalized musically encoded phrase are visually presented to a user.

11. The method of claim 9 , wherein the at least one error signal is based on a difference in rhythm between the non-lexicalized musically encoded phrase and the lexicalized version of the non-lexicalized musically encoded phrase, the at least one error signal further based on at least one of a difference in prosody between the non-lexicalized musically encoded phrase and the lexicalized version of the non-lexicalized musically encoded phrase; a difference in the total number of syllables between the non-lexicalized musically encoded phrase and the lexicalized version of the non-lexicalized musically encoded phrase; a difference in duration in units of time between the non-lexicalized musically encoded phrase and the lexicalized version of the non-lexicalized musically encoded phrase; a difference in pitch between respective ones of the syllables between the non-lexicalized musically encoded phrase and the lexicalized version of the non-lexicalized musically encoded phrase; or a difference in intonation between the non-lexicalized musically encoded phrase and the lexicalized version of the non-lexicalized musically encoded phrase.

12. The method of claim 9 , wherein the non-lexical properties of the non-lexicalized musically encoded phrase and the non-lexical properties of the lexicalized version of the non-lexicalized musically encoded phrase include rhythm, and further include at least one of prosody; duration in time; a total number of syllables; or a pitch level of at least one syllable.

13. The method of claim 9 , wherein the at least one error signal further based on a sum of the difference between:

the difference between the duration of the first syllable in the non-lexicalized musically encoded phrase and the duration of the second syllable in the non-lexicalized musically encoded phrase, the first syllable in the non-lexicalized musically encoded phrase being non-adjacent to the second syllable in the non-lexicalized musically encoded phrase; and

the difference between the duration of the first syllable in the lexicalized version of the non-lexicalized musically encoded phrase and the duration of the second syllable in the lexicalized version of the non-lexicalized musically encoded phrase, the first syllable in the lexicalized version of the non-lexicalized musically encoded phrase being non-adjacent to the second syllable in the lexicalized version of the non-lexicalized musically encoded phrase.

14. A non-transitory processor-readable medium storing code representing instructions to be executed by a processor, the code comprising code to cause the processor to:

encode a first prosodic speech signal to generate a musically encoded first prosodic speech signal by mapping each syllable of the first prosodic speech signal to a musical note;

store the musically encoded first prosodic speech signal;

audibly play the musically encoded first prosodic speech signal to a language student;

prompt the student to recite the speech segment from which the musically encoded first prosodic speech signal originated;

record an utterance from the language student in response to the prompt;

delexicalize the utterance to generate a second prosodic speech signal; and

calculate at least one error signal based on a difference between:

a difference between a duration of a first syllable in the first prosodic speech signal and a duration of a second syllable in the first prosodic speech signal, the first syllable in the first prosodic speech signal being non-adjacent to the second syllable in the first prosodic speech signal; and

a difference between a duration of a first syllable in the second prosodic speech signal and a duration of a second syllable in the second prosodic speech signal, the first syllable in the second prosodic speech signal being non-adjacent to the second syllable in the second prosodic speech signal.

15. The non-transitory processor-readable medium of claim 14 , further comprising code to cause the processor to:

determine prosodic characteristics of the first prosodic speech signal and the second prosodic speech signal, the code to cause the processor to calculate includes code to cause the processor to calculate the at least one error signal based on the prosodic characteristics of the first prosodic speech signal and the second prosodic speech signal.

16. The non-transitory processor-readable medium of claim 14 , further comprising code to cause the processor to:

determine prosodic characteristics of the first prosodic speech signal and the second prosodic speech signals, the code to cause the processor to calculate includes code to cause the processor to calculate the at least one error signal based on the prosodic characteristics of the first prosodic speech signal and the second prosodic speech signal, the prosodic characteristics including rhythm, the prosodic characteristics further including at least one of prosody; duration in time; a total number of syllables; or a pitch level of at least one syllable.

17. The non-transitory processor-readable medium of claim 14 , wherein

the at least one error signal is further based on at least one of a difference in prosody between the first prosodic speech signal and the second prosodic speech signal; a difference in a total number of syllables between the first prosodic speech signal and the second prosodic speech signal; a difference in duration in units of time between the first prosodic speech signal and the second prosodic speech signal; a difference in pitch between respective syllables between the first prosodic speech signal and the second prosodic speech signal; or a difference in intonation between the first prosodic speech signal and the second prosodic speech signal.

18. The non-transitory processor-readable medium of claim 14 , wherein the musically encoded first prosodic speech signal is encoded in a musical instrument digital interface (MIDI) framework.

19. The non-transitory processor-readable medium of claim 14 , wherein the code to cause the processor to encode includes code to cause the processor to encode the first prosodic speech signal at least partially by converting a frequency of the first prosodic speech signal to a musical instrument digital interface (MIDI) pitch note number.

20. The non-transitory processor-readable medium of claim 14 , wherein the at least one error signal further based on a sum of the difference between:

the difference between the duration of the first syllable in the first prosodic speech signal and the duration of the second syllable in the first prosodic speech signal, the first syllable in the first prosodic speech signal being non-adjacent to the second syllable in the first prosodic speech signal; and

the difference between the duration of the first syllable in the second prosodic speech signal and the duration of the second syllable in the second prosodic speech signal, the first syllable in the second prosodic speech signal being non-adjacent to the second syllable in the second prosodic speech signal.

Assignments (11)
RELEASE OF SECURITY INTEREST IN SPECIFIED PATENTS Recorded Mar 1, 2023
From: JPMORGAN CHASE BANK, N.A.
To: IXL LEARNING, INC.; THINKMAP, INC.; WYZANT, INC.; ROSETTA STONE LLC
Reel/Frame 062904/0514 →
SECURITY INTEREST Recorded Mar 1, 2023
From: IXL LEARNING, INC.; THINKMAP, INC.; WYZANT, INC.; ROSETTA STONE LLC; TEACHER SYNERGY LLC; EMMERSION LEARNING, INC.
To: JPMORGAN CHASE BANK, N.A.
Reel/Frame 062846/0032 →
CHANGE OF NAME Recorded May 17, 2021
From: ROSETTA STONE LTD.
To: ROSETTA STONE LLC
Reel/Frame 056256/0603 →
GRANT OF PATENT SECURITY INTEREST Recorded Mar 12, 2021
From: IXL LEARNING INC.; THINKMAP, INC.; WYZANT, INC.; ROSETTA STONE LLC (F/K/A ROSETTA STONE LTD.)
To: JPMORGAN CHASE BANK, N.A., AS ADMINISTRATIVE AGENT
Reel/Frame 055581/0469 →
RELEASE OF SECURITY INTEREST IN PATENTS AT REEL/FRAME NO. 54085/0920 Recorded Mar 12, 2021
From: ROYAL BANK OF CANADA, AS COLLATERAL AGENT
To: ROSETTA STONE LTD.
Reel/Frame 055583/0555 →
RELEASE OF SECURITY INTEREST IN PATENTS AT REEL/FRAME NO. 54085/0934 Recorded Mar 12, 2021
From: ROYAL BANK OF CANADA, AS COLLATERAL AGENT
To: ROSETTA STONE LTD.
Reel/Frame 055583/0562 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 15, 2020
From: SILICON VALLEY BANK
To: ROSETTA STONE, LTD; LEXIA LEARNING SYSTEMS LLC
Reel/Frame 054086/0105 →
FIRST LIEN PATENT SECURITY AGREEMENT Recorded Oct 15, 2020
From: ROSETTA STONE LTD.; LEXIA LEARNING SYSTEMS LLC
To: ROYAL BANK OF CANADA, AS COLLATERAL AGENT
Reel/Frame 054085/0920 →
SECOND LIEN PATENT SECURITY AGREEMENT Recorded Oct 15, 2020
From: ROSETTA STONE LTD.; LEXIA LEARNING SYSTEMS LLC
To: ROYAL BANK OF CANADA, AS COLLATERAL AGENT
Reel/Frame 054085/0934 →
SECURITY AGREEMENT Recorded Oct 30, 2014
From: ROSETTA STONE, LTD.; LEXIA LEARNING SYSTEMS LLC
To: SILICON VALLEY BANK
Reel/Frame 034105/0733 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Sep 14, 2011
From: TEPPERMAN, JOSEPH; STANLEY, THEBAN; HACIOGLU, KADRI
To: ROSETTA STONE, LTD.
Reel/Frame 026904/0355 →