IP Library Granted Patent US 7,454,347
Granted Patent B2
US 7,454,347 · App. 10/920,454 · Granted Nov 18, 2008

Voice labeling error detecting system, voice labeling error detecting method and program

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 7,454,347
App. No.
10/920,454
Granted
Nov 18, 2008
Kind
B2
Abstract

A labeling part 3 analyzes the character string data to produce a phoneme label and a prosody label, partition the voice data stored in a voice database 1 into phonemic data, and label the phonemic data, employing the phoneme label and the like. A phoneme segmenting part 4 connects the voice data labeled with the same kind of phonemic data, and a formant extracting part 5 specifies the frequency of formant of each piece of phonemic data. A processing part 6 decides an evaluation value for each phonemic data based on the frequency of formant, and an error detection part 7 detects the phonemic data of which a deviation of the evaluation value within a set of phonemic data reaches a predetermined amount.

Claims (79)

1. A voice labeling error detecting system comprising:

data acquisition means for acquiring waveform data representing a waveform of a unit voice and labeling data for identifying a kind of said unit voice;

classification means for classifying the waveform data acquired by said data acquisition means into the kinds of unit voice, based on the labeling data acquired by said data acquisition means;

evaluation value decision means for specifying a frequency of a formant of each unit voice represented by the waveform data acquired by said data acquisition means and determining an evaluation value of said waveform data based on the specified frequency; and

error detection means for detecting the waveform data from among a set of waveform data classified into a same kind, for which a deviation of evaluation value within said set reaches a predetermined amount, and outputting the data representing said detected waveform data, as waveform data having a labeling error, and

wherein said evaluation value H is calculated by the following formula representing a linear combination of values {|f(k)−F(k)|}:

H

=

k

=

1

n

{

f

(

k

)

-

F

(

k

)

·

W

(

k

)

}

wherein F(k) is a frequency of the k-th formant of a unit voice indicated by the waveform data to calculate the evaluation value, and f(k) is an average value of the frequency of the k-th formant of the unit voice indicated by each waveform data classified into the same kind as said waveform data, W(k) is a weighting factor and n is the order of formant of the phoneme having the highest frequency.

2. The voice labeling error detecting system according to claim 1 , characterized in that said evaluation value is a linear combination of plural frequencies of formants in a spectrum of acquired waveform data.

3. The voice labeling error detecting system according to claim 1 or 2 , characterized in that said evaluation value deciding means deals with a frequency at a maximal value of a spectrum in the waveform data as the frequency of formant of unit voice indicated by said waveform data.

4. The voice labeling error detecting system according to any one of claim 1 or 2 , characterized in that said evaluation value deciding means specifies an order of formant used to decide the evaluation value of the waveform data as the kind of unit voice indicated by said waveform data, corresponding to the kind of labeling data.

5. The voice labeling error detecting system according to any one of claim 1 or 2 , characterized in that said error detection means detects the waveform data associated with the labeling data indicating a voiceless state at which a magnitude of voice represented by said waveform data reaches a predetermined amount as the waveform data in which the labeling has an error.

6. The voice labeling error detecting system according to claim 1 or 2 , characterized in that said classification means comprises means for concatenating each waveform data classified into the same kind in the form in which two adjacent pieces of waveform data sandwiches data indicating a voiceless state therebetween.

7. A voice labeling error detecting method comprising the steps of:

acquiring waveform data representing a waveform of a unit voice and labeling data for identifying a kind of said unit voice;

classifying said acquired waveform data into the kinds of unit voice, based on said acquired labeling data;

specifying a frequency of a formant of each unit voice represented by the waveform data and deciding an evaluation value of said waveform data based on the specified frequency; and

detecting the waveform data having a labeling error, from among a set of waveform data classified into a same kind, in which a deviation of evaluation value within said set reaches a predetermined amount and outputting data representing said detected waveform data,

wherein said evaluation value H is calculated by the following formula representing a linear combination of values {|f(k)−F(k)|}:

H

=

k

=

1

n

{

f

(

k

)

-

F

(

k

)

·

W

(

k

)

}

wherein F(k) is a frequency of the k-th formant of a unit voice indicated by the waveform data to calculate the evaluation value, and f(k) is an average value of the frequency of the k-th formant of the unit voice indicated by each waveform data classified into the same kind as said waveform data, W(k) is a weighting factor and n is the order of formant of the phoneme having the highest frequency.

Assignments (5)
CORRECTIVE ASSIGNMENT TO CORRECT THE REMOVE PATENT NUMBERS 10342096;10671117; 10716375; 10716376;10795407;10795408; AND 10827591 PREVIOUSLY RECORDED AT REEL: 58314 FRAME: 657. ASSIGNOR(S) HEREBY CONFIRMS THE ASSIGNMENT. Recorded Feb 29, 2024
From: RAKUTEN, INC.
To: RAKUTEN GROUP, INC.
Reel/Frame 068066/0103 →
CHANGE OF NAME Recorded Dec 6, 2021
From: RAKUTEN, INC.
To: RAKUTEN GROUP, INC.
Reel/Frame 058314/0657 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 1, 2015
From: JVC KENWOOD CORPORATION
To: RAKUTEN, INC.
Reel/Frame 037179/0777 →
MERGER Recorded Apr 6, 2012
From: KENWOOD CORPORATION
To: JVC KENWOOD CORPORATION
Reel/Frame 028001/0636 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 23, 2004
From: KOYAMA, RIKA
To: KABUSHIKI KAISHA KENWOOD
Reel/Frame 016012/0295 →