IP Library › Granted Patent US 9,390,710
Granted Patent B2
US 9,390,710 · App. 14/604,991 · Granted Jul 12, 2016

Method for reranking speech recognition results

Inventors: Jung Yun Seo (Seoul, KR); Sang Woo Kang (Seoul, KR); Myoung-Wan Koo (Seoul, KR); Hark Soo Kim (Seoul, KR); Hyeok Ju Ahn (Wonju-si, KR); Yeong Kil Song (Chuncheon-si, KR); Maeng Sik Choi (Samcheok-si, KR)
Assignees: SOGANG UNIVERSITY RESEARCH FOUNDATION; KANGWON NATIONAL UNIVERSITY UNIVERSITY-INDUSTRY COOPERATION FOUNDATION
G10L15/063G10L2015/0635
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 9,390,710
App. No.
14/604,991
Granted
Jul 12, 2016
Kind
B2
Abstract

Provided is a speech recognition method using machine learning, including: receiving a speech signal as an input, performing speech recognition to generate speech recognition result information including multiple candidate sentences and ranks of the respective candidate sentences; processing the multiple candidate sentences included in the speech recognition result information according to a machine learning model which is learned in advance and changing the ranks of the multiple candidate sentences to re-rank the multiple candidate sentences; and selecting the highest-rank candidate sentence among the re-ranked multiple candidate sentences as a speech recognition result. Particularly, the machine learning model is generated by: receiving the speech signal and a correct answer sentence as inputs; generating the speech recognition result information and a correct answer set; generating learning data by using the correct answer set; and performing the machine learning of changing the ranks of the candidate sentences.

Claims (16)

1. A speech recognition method using machine learning, comprising:

receiving a speech signal as an input, performing speech recognition to generate speech recognition result information including multiple candidate sentences and ranks of the respective candidate sentences;

processing the multiple candidate sentences included in the speech recognition result information according to a machine learning model which is learned in advance and changing the ranks of the multiple candidate sentences to re-rank the multiple candidate sentences; and

selecting a highest-rank candidate sentence among the re-ranked multiple candidate sentences as a speech recognition result,

wherein the machine learning model is generated by:

receiving the speech signal and a correct answer sentence as inputs;

performing the speech recognition on the speech signal to generate the speech recognition result information including the multiple candidate sentences and sentence scores representing the ranks of the respective candidate sentences;

adding the correct answer sentence to the speech recognition result information to generate a correct answer set;

extracting features of the candidate sentences and the correct answer sentence included in the correct answer set to generate learning data; and

performing the machine learning of changing the ranks of the candidate sentences according to differences between the features of the candidate sentences and the features of the correct answer sentence based on the learning data, and

wherein the features include speech recognition ranks, a sentence score of the highest-rank candidate sentence, a morpheme bigram, a POS (part of speech) bigram, the number of domain dictionary unregistered words, morphemes/POSs of domain dictionary unregistered words, the number of general dictionary unregistered words, and morphemes/POSs of general dictionary unregistered words.

2. The speech recognition method according to claim 1 , wherein the machine learning is a Rank SVM.

3. The speech recognition method according to claim 1 ,

wherein the correct answer set includes a portion of the multiple candidate sentences, and

wherein the portion of the multiple candidate sentences are candidate sentences of which sentence scores are equal to or higher than a predetermined sentence score among the candidate sentences included in the speech recognition result information.

4. The speech recognition method according to claim 1 , wherein the speech recognition result information is transmitted from a predetermined external speech recognition service server.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 27, 2015
From: SEO, JUNG YUN; KANG, SANG WOO; KOO, MYOUNG-WAN; KIM, HARK SOO; AHN, HYEOK JU; SONG, YEONG KIL; CHOI, MAENG SIK
To: SOGANG UNIVERSITY RESEARCH FOUNDATION; KANGWON NATIONAL UNIVERSITY UNIVERSITY-INDUSTRY COOPERATION FOUNDATION
Reel/Frame 035273/0251 →
Priority Claims (1)
KR 10-2014-0138383 · Oct 14, 2014 · national
Continuity (1)
Related Publication 20160104478A1 · Apr 14, 2016