Systems and methods for machine-learning based multi-lingual pronunciation generation
Systems and methods for machine-learning based multi-lingual pronunciation generation are disclosed. A method for machine-learning based multi-lingual pronunciation generation may include: (1) training a language origin prediction machine learning model; (2) training a pronunciation generator machine learning model; (3) receiving, by a pronunciation computer program, a word for pronunciation guidance; (4) predicting, by the pronunciation computer program and using the trained language origin prediction machine learning model, a language origin of the word; (5) predicting, by the pronunciation computer program and using the trained pronunciation generator machine learning model and the language origin, a syllable-by-syllable pronunciation for the word; and (6) returning, by the pronunciation computer program, the syllable-by-syllable pronunciation.
1 . A method for machine-learning based multi-lingual pronunciation generation, comprising:
training a language origin prediction machine learning model;
training a plurality of pronunciation generator machine learning models, wherein each of the trained pronunciation generator machine learning model is specific to a language origin;
receiving, by a pronunciation computer program, a word for pronunciation guidance;
predicting, by the pronunciation computer program and using the trained language origin prediction machine learning model, a language origin of the word;
selecting, by the pronunciation computer program, one of the plurality of trained pronunciation generator machine learning models for the predicted language origin;
predicting, by the pronunciation computer program and using the selected trained pronunciation generator machine learning model, a syllable-by-syllable pronunciation for the word; and
returning, by the pronunciation computer program, the syllable-by-syllable pronunciation.
2 . The method of claim 1 , wherein the word is a name.
3 . The method of claim 1 , wherein the language origin prediction machine learning model and/or the pronunciation generator machine learning model are trained using supervised learning.
4 . The method of claim 1 , further comprising:
receiving, by the pronunciation computer program, feedback; and
re-training the language origin prediction machine learning model and/or the pronunciation generator machine learning model using the feedback.
5 . The method of claim 1 , wherein the word is received at an application executed by a user electronic device that is in communication with the pronunciation computer program.
6 . The method of claim 5 , wherein the application outputs audio of the syllable-by-syllable pronunciation.
7 . The method of claim 5 , wherein the application outputs text of the syllable-by-syllable pronunciation.
8 . The method of claim 1 , wherein the pronunciation computer program is integrated into a videoconferencing computer program.
9 . A system, comprising:
a trained language origin prediction machine learning model;
a plurality of trained pronunciation generator machine learning models, wherein each of the trained pronunciation generator machine learning model is specific to a language origin; and
an electronic device executing a pronunciation computer program that is configured to receive a word for pronunciation guidance, to predict using the trained language origin prediction machine learning model, a language origin of the word, to select one of the plurality of trained pronunciation generator machine learning models for the predicted language origin, to predict, using the selected trained pronunciation generator machine learning model, a syllable-by-syllable pronunciation of the word, and to output the syllable-by-syllable pronunciation.
10 . The system of claim 9 , wherein the word is a name.
11 . The system of claim 9 , wherein the trained language origin prediction machine learning model and/or the trained pronunciation generator machine learning model are trained using supervised learning.
12 . The system of claim 9 , wherein the pronunciation computer program is further configured to receive feedback and to retrain the trained language origin prediction machine learning model and/or the trained pronunciation generator machine learning model using the feedback.
13 . The system of claim 9 , further comprising:
a user electronic device executing an application, wherein application is configured to receive the word.
14 . The system of claim 13 , wherein the application outputs audio of the syllable-by-syllable pronunciation on a speaker.
15 . The system of claim 13 , wherein the application outputs text of the syllable-by-syllable pronunciation on a display.
16 . The system of claim 9 , wherein the pronunciation computer program is integrated into a videoconferencing computer program.