IP Library Granted Patent US 10,068,569
Granted Patent B2
US 10,068,569 · App. 13/932,506 · Granted Sep 4, 2018

Generating acoustic models of alternative pronunciations for utterances spoken by a language learner in a non-native language

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,068,569
App. No.
13/932,506
Granted
Sep 4, 2018
Kind
B2
Abstract

A non-transitory processor-readable medium storing code representing instructions to be executed by a processor includes code to cause the processor to receive acoustic data representing an utterance spoken by a language learner in a non-native language in response to prompting the language learner to recite a word in the non-native language and receive a pronunciation lexicon of the word in the non-native language. The pronunciation lexicon includes at least one alternative pronunciation of the word based on a pronunciation lexicon of a native language of the language learner. The code causes the processor to generate an acoustic model of the at least one alternative pronunciation in the non-native language and identify a mispronunciation of the word in the utterance based on a comparison of the acoustic data with the acoustic model. The code causes the processor to send feedback related to the mispronunciation of the word to the language learner.

Claims (19)

1. A non-transitory processor-readable medium storing code representing instructions to be executed by a processor, the code comprising code to cause the processor to:

receive acoustic data representing an utterance spoken by a language learner in a non-native language in response to prompting the language learner to recite a word in the non-native language;

receive a pronunciation lexicon of the word in the non-native language, the pronunciation lexicon of the word including at least one alternative pronunciation of the word determined based on a pronunciation lexicon of a native language of the language learner, the at least one alternative pronunciation of the word being a phonological error in the non-native language and having a probability greater than a threshold level of being spoken by the language learner when the language learner attempts to recite the word in the non-native language;

generate an acoustic model of the at least one alternative pronunciation of the word from the pronunciation lexicon of the word in the non-native language;

identify a mispronunciation of the word in the utterance based on a comparison of the acoustic data with the at least one alternative pronunciation of the word that is included in the acoustic model; and

send feedback related to the mispronunciation of the word to the language learner, wherein

the acoustic data is first acoustic data, the code to cause the processor to generate the acoustic model includes code to cause the processor to:

generate a maximum likelihood native model based on native data, the native data including a corpus of audio data provided by at least one native speaker,

generate a mix-up native triphone model based on the maximum likelihood native model, and

generate a maximum likelihood non-native model based at least on the mix-up native triphone model and second acoustic data.

2. A non-transitory processor-readable medium storing code representing instructions to be executed by a processor, the code comprising code to cause the processor to:

receive acoustic data representing an utterance spoken by a language learner in a non-native language in response to prompting the language learner to recite a word in the non-native language;

receive a pronunciation lexicon of the word in the non-native language, the pronunciation lexicon of the word including at least one alternative pronunciation of the word determined based on a pronunciation lexicon of a native language of the language learner, the at least one alternative pronunciation of the word being a phonological error in the non-native language and having a probability greater than a threshold level of being spoken by the language learner when the language learner attempts to recite the word in the non-native language;

generate an acoustic model of the at least one alternative pronunciation of the word from the pronunciation lexicon of the word in the non-native language;

identify a mispronunciation of the word in the utterance based on a comparison of the acoustic data with the at least one alternative pronunciation of the word that is included in the acoustic model; and

send feedback related to the mispronunciation of the word to the language learner, wherein

the acoustic data is first acoustic data, the code to cause the processor to generate the acoustic model includes code to cause the processor to:

generate a maximum likelihood non-native model based at least in part on second acoustic data and a mix-up native triphone model generated by a maximum likelihood native model training process, and

generate the acoustic model as a minimum phone error trained acoustic model based on the maximum likelihood non-native model.

Assignments (14)
SECURITY INTEREST Recorded Mar 1, 2023
From: IXL LEARNING, INC.; THINKMAP, INC.; WYZANT, INC.; ROSETTA STONE LLC; TEACHER SYNERGY LLC; EMMERSION LEARNING, INC.
To: JPMORGAN CHASE BANK, N.A.
Reel/Frame 062846/0032 →
RELEASE OF SECURITY INTEREST IN SPECIFIED PATENTS Recorded Mar 1, 2023
From: JPMORGAN CHASE BANK, N.A.
To: IXL LEARNING, INC.; THINKMAP, INC.; WYZANT, INC.; ROSETTA STONE LLC
Reel/Frame 062904/0514 →
CHANGE OF NAME Recorded May 17, 2021
From: ROSETTA STONE LTD.
To: ROSETTA STONE LLC
Reel/Frame 056256/0603 →
RELEASE OF SECURITY INTEREST IN PATENTS AT REEL/FRAME NO. 54085/0920 Recorded Mar 12, 2021
From: ROYAL BANK OF CANADA, AS COLLATERAL AGENT
To: ROSETTA STONE LTD.
Reel/Frame 055583/0555 →
GRANT OF PATENT SECURITY INTEREST Recorded Mar 12, 2021
From: IXL LEARNING INC.; THINKMAP, INC.; WYZANT, INC.; ROSETTA STONE LLC (F/K/A ROSETTA STONE LTD.)
To: JPMORGAN CHASE BANK, N.A., AS ADMINISTRATIVE AGENT
Reel/Frame 055581/0469 →
RELEASE OF SECURITY INTEREST IN PATENTS AT REEL/FRAME NO. 54085/0934 Recorded Mar 12, 2021
From: ROYAL BANK OF CANADA, AS COLLATERAL AGENT
To: ROSETTA STONE LTD.
Reel/Frame 055583/0562 →
SECOND LIEN PATENT SECURITY AGREEMENT Recorded Oct 15, 2020
From: ROSETTA STONE LTD.; LEXIA LEARNING SYSTEMS LLC
To: ROYAL BANK OF CANADA, AS COLLATERAL AGENT
Reel/Frame 054085/0934 →
FIRST LIEN PATENT SECURITY AGREEMENT Recorded Oct 15, 2020
From: ROSETTA STONE LTD.; LEXIA LEARNING SYSTEMS LLC
To: ROYAL BANK OF CANADA, AS COLLATERAL AGENT
Reel/Frame 054085/0920 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 15, 2020
From: SILICON VALLEY BANK
To: ROSETTA STONE, LTD; LEXIA LEARNING SYSTEMS LLC
Reel/Frame 054086/0105 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jul 14, 2017
From: SIIVOLA, VESA
To: ROSETTA STONE LTD.
Reel/Frame 043004/0384 →
CERTIFICATE OF AMENDMENT Recorded Jun 19, 2017
From: FAIRFIELD & SONS, LTD.
To: ROSETTA STONE LTD.
Reel/Frame 042895/0056 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jun 19, 2017
From: HACIOGLU, KADRI
To: FAIRFIELD & SONS, LTD. D/B/A FAIRFIELD LANGUAGE TECHNOLOGIES
Reel/Frame 042895/0148 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jun 19, 2017
From: STANLEY, THEBAN
To: FAIRFIELD & SONS, LTD. D/B/A FAIRFIELD LANGUAGE TECHNOLOGIES INC.
Reel/Frame 042749/0248 →
SECURITY AGREEMENT Recorded Oct 30, 2014
From: ROSETTA STONE, LTD.; LEXIA LEARNING SYSTEMS LLC
To: SILICON VALLEY BANK
Reel/Frame 034105/0733 →