AUTOMATED SPEECH RECOGNITION SYSTEM
There is provided an automated speech recognition system that applies weights to grapheme-to-phoneme models, and interpolates pronunciations from combinations of the models, to recognize utterances of foreign named entities for naive, informed, and in-between pronunciations.
1 . An automated speech recognition (ASR) system, comprising:
a microphone;
a recognition dictionary storage that contains:
(a) a first recognition dictionary that stores a first pronunciation of a token that was generated from a first grapheme-to-phoneme model (G2P) for said token; and
(b) a second recognition dictionary that stores a second interpretation of said token that was generated from a second G2P model for said token;
a G2P weight storage that contains:
(a) a first G2P weight that is applicable to said first G2P model to yield said first pronunciation for said token; and
(b) a second G2P weight that is applicable to said second G2P model to yield said second pronunciation for said token;
a processor that receives an utterance containing a spoken form of said token from said microphone; and
a memory that contains instructions that are readable by said processor to control said processor to:
obtain metadata concerning said token;
modify said first G2P weight and said second G2P weight based on said metadata, thus yielding a first weighted G2P model and a second weighted G2P model;
interpolate said first weighted G2P model and said second weighted G2P model to yield a resultant pronunciation for said token; and
provide an output based on said resultant pronunciation.
2 . The ASR system of claim 1 ,
wherein said utterance is spoken by a user, and
wherein said metadata identifies a characteristic of said user.
3 . The ASR system of claim 2 , wherein said characteristic of said user is a native language of said user.
4 . The ASR system of claim 1 , further comprising:
a user device; and
a global positioning system that identifies a present location of said user device,
wherein said metadata comprises said present location.
5 . The ASR system of claim 1 , wherein said output comprises a signal to control a device.