IP Library Granted Patent US 10,249,306
Granted Patent B2
US 10,249,306 · App. 14/760,617 · Granted Apr 2, 2019

Speaker identification device, speaker identification method, and recording medium

Inventors: Masahiro Tani (Tokyo, JP); Takafumi Koshinaka (Tokyo, JP); Yoshifumi Onishi (Tokyo, JP); Shigeru Sawada (Tokyo, JP)
Assignee: NEC CORPORATION
G10L17/12G10L17/04G10L17/06
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,249,306
App. No.
14/760,617
Granted
Apr 2, 2019
Kind
B2
Abstract

A speaker identification device includes: a primary speaker identification unit that computes, for each pre-stored registered speaker, a score that indicates the similarity between input speech and speech of the registered speakers; a similar speaker selection unit that selects a plurality of the registered speakers as similar speakers according to the height of the scores thereof; a learning unit that creates a classifier for each similar speaker by sorting the speech of a certain similar speaker among the similar speakers as a positive instance and the speech of the other similar speakers as negative instances; and a secondary speaker identification unit that computes, for each classifier, a score of the classifier with respect to the input speech, and outputs an identification result.

Claims (23)

1. A speaker identification device comprising:

a primary speaker identification unit which computes, for each registered speaker stored in advance, a score that indicates similarity between input speech and speech of the registered speakers;

a similar speaker selection unit which selects a plurality of the registered speakers as similar speakers according to height of the scores;

a learning unit which creates a plurality of classifiers, each classifier corresponding to a different speaker of similar speakers,

wherein for each classifier, the classifier corresponds speech of the different speaker to which the classifier corresponds as a positive instance and speech of other speakers of the similar speakers as negative instances; and

a secondary speaker identification unit which computes, for each classifier, a score of the classifier with respect to the input speech and outputs an identification result.

2. The speaker identification device according to claim 1 , wherein

the learning unit stores in advance pairs of the similar speakers selected by the similar speaker selection unit in the past and classifiers created by the learning unit in the past as history, and creates a classifier only when there is a difference between the similar speakers in the history and the similar speakers selected by the similar speaker selection unit.

3. The speaker identification device according to claim 1 , wherein the similar speaker selection unit selects a preset number of the similar speakers.

4. The speaker identification device according to claim 1 , wherein the similar speaker selection unit selects the similar speakers based on a preset score threshold.

5. The speaker identification device according to claim 1 , wherein the classifier is a Support Vector Machine (SVM), and the score of the classifier is a distance from a feature point of the input speech to a classification plane.

6. A speaker identification method comprising:

computing, for each registered speaker stored in advance, a score that indicates similarity between input speech and speech of the registered speakers;

selecting a plurality of the registered speakers as similar speakers according to height of the scores;

creating a plurality of classifiers, each classifier corresponding to a different speaker of similar speakers;

for each classifier, corresponding, by the classifier, speech of the different speaker to which the classifier corresponds as a positive instance and speech of other speakers of the similar speakers as negative instances; and

computing, for each classifier, a score of the classifier with respect to the input speech and outputting an identification result.

7. A non-transitory computer readable medium that stores therein a speaker identification program that causes a computer to execute:

primary speaker identification processing of computing, for each registered speaker stored in advance, a score that indicates similarity between input speech and speech of the registered speakers;

similar speaker selection processing of selecting a plurality of the registered speakers as similar speakers according to height of the scores;

learning processing of creating a plurality of classifiers, each classifier corresponding to a different speaker of similar speakers;

for each classifier, corresponding, by the classifier, speech of the different speaker to which the classifier corresponds as a positive instance and speech of other speakers of the similar speakers as negative instances; and

secondary speaker identification processing of computing, for each classifier, a score of the classifier with respect to the input speech and outputs an identification result.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jul 13, 2015
From: TANI, MASAHIRO; KOSHINAKA, TAKAFUMI; ONISHI, YOSHIFUMI; SAWADA, SHIGERU
To: NEC CORPORATION
Reel/Frame 036071/0402 →
Priority Claims (1)
JP 2013-006350 · Jan 17, 2013 · national
Continuity (1)
Related Publication 20150356974A1 · Dec 10, 2015