IP Library Granted Patent US 8,849,391
Granted Patent B2
US 8,849,391 · App. 13/304,789 · Granted Sep 30, 2014

Speech sound intelligibility assessment system, and method and program therefor

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 8,849,391
App. No.
13/304,789
Granted
Sep 30, 2014
Kind
B2
Abstract

The speech sound intelligibility assessment system includes: an output section for presenting a speech sound to a user; a biological signal measurement section for measuring an electroencephalogram signal of the user; a positive component determination section for determining presence/absence of a positive component of an event-related potential in the electroencephalogram signal in a zone from 600 ms to 800 ms from a starting point, which is a point in time at which the output section presents a speech sound; a negative component determination section for determining presence/absence of a negative component of an event-related potential in the electroencephalogram signal in a zone from 100 ms to 300 ms from the same starting point; and an assessment section for evaluating whether the user has clearly aurally comprehended the presented speech sound or not based on the results of determination of presence/absence of the positive and negative components, respectively.

Claims (96)

1. A speech sound intelligibility assessment system comprising:

a speech sound database retaining a plurality of speech sounds;

a presented-speech sound control section for determining a speech sound to be presented by referring to the speech sound database;

an output section for presenting the determined speech sound to a user;

a biological signal measurement section for measuring an electroencephalogram signal of the user;

a positive component determination section for determining presence or absence of a positive component of an event-related potential in the electroencephalogram signal in a zone from 600 ms to 800 ms from a starting point, the starting point being a point in time at which the output section presents a speech sound;

a negative component determination section for determining presence or absence of a negative component of an event-related potential in the electroencephalogram signal in a zone from 100 ms to 300 ms from a starting point, the starting point being a point in time at which the output section presents a speech sound; and

a speech sound intelligibility assessment section for evaluating whether the user has clearly aurally comprehended the presented speech sound or not, based on a result of determination as to presence or absence of the positive component acquired from the positive component determination section and a result of determination as to presence or absence of the negative component acquired from the negative component determination section;

wherein the speech sound intelligibility assessment section:

makes an evaluation that the user has clearly aurally comprehended the presented speech sound when the result of determination by the positive component determination section indicates that the positive component is absent;

makes an evaluation that the user has not clearly aurally comprehended the presented speech sound due to an insufficient overall sound pressure when the result of determination by the positive component determination section indicates that the positive component is present and the result of determination by the negative component determination section indicates that the negative component is absent; or

makes an evaluation that the user has not clearly aurally comprehended the presented speech sound due to an insufficient sound pressure of a consonant frequency when the result of determination by the positive component determination section indicates that the positive component is present and the result of determination by the negative component determination section indicates that the negative component is present.

2. The speech sound intelligibility assessment system of claim 1 , wherein

the positive component determination section compares between a predetermined threshold value and a zone average potential of an event-related potential in a zone from 600 ms to 800 ms from a starting point, the starting point being a point in time at which the output section presents a speech sound, and

determines that the positive component is present when the zone average potential is equal to or greater than the threshold value, or

determines that the positive component is absent when the zone average potential is smaller than the threshold value.

3. The speech sound intelligibility assessment system of claim 1 , wherein

the negative component determination section compares between a predetermined threshold value and an absolute value of a negative peak value of an event-related potential in a zone from 100 ms to 300 ms from a starting point, the starting point being a point in time at which the output section presents a speech sound, and

determines that the negative component is present when the absolute value of the peak value is equal to or greater than the threshold value, or

determines that the negative component is absent when the absolute value of the peak value is smaller than the threshold value.

4. The speech sound intelligibility assessment system of claim 1 , wherein, in the speech sound database, a speech sound type, consonant information type, and a group concerning probability of confusion are associated with each of the plurality of speech sounds retained therein.

5. The speech sound intelligibility assessment system of claim 1 , wherein

the speech sound database further retains gain information defining a gain for each of frequency bands concerning the plurality of speech sounds, the speech sound intelligibility assessment system further comprising:

a stimulation speech sound gain adjustment section for, with respect to any speech sound that is evaluated by the speech sound intelligibility assessment section as not being clearly aurally comprehended by the user due to an insufficient overall sound pressure, updating the frequency-by-frequency gain information concerning speech sounds that is retained in the speech sound database so as to increase the gain for the entire frequency band, and with respect to any speech sound that is evaluated by the speech sound intelligibility assessment section as not being clearly aurally comprehended by the user due to an insufficient sound pressure of a consonant frequency, calculating a consonant frequency band of the speech sound and updating the frequency-by-frequency gain information concerning speech sounds that is retained in the speech sound database so as to increase the gain for the consonant frequency band.

6. A speech sound intelligibility assessment comprising:

a speech sound database retaining a plurality of speech sounds;

a presented-speech sound control section for determining a speech sound to be presented by referring to the speech sound database;

an output section for presenting the determined speech sound to a user;

a biological signal measurement section for measuring an electroencephalogram signal of the user;

a positive component determination section for determining presence or absence of a positive component of an event-related potential in the electroencephalogram signal in a zone from 600 ms to 800 ms from a starting point, the starting point being a point in time at which the output section presents a speech sound;

a negative component determination section for determining presence or absence of a negative component of an event-related potential in the electroencephalogram signal in a zone from 100 ms to 300 ms from a starting point, the starting point being a point in time at which the output section presents a speech sound; and

a speech sound intelligibility assessment section for evaluating whether the user has clearly aurally comprehended the presented speech sound or not, based on a result of determination as to presence or absence of the positive component acquired from the positive component determination section and a result of determination as to presence or absence of the negative component acquired from the negative component determination section;

wherein, in the speech sound database, a speech sound type, consonant information type, and a group concerning probability of confusion are associated with each of the plurality of speech sounds retained therein; and

further comprising an event-related potential processing section for referring to association between the speech sound type, the consonant information type, and the group concerning probability of confusion stored in the speech sound database, and generating electroencephalogram data by taking an arithmetic mean of event-related potentials corresponding to the presented speech sound, with respect to each of the speech sound type, the consonant information type, and the group concerning probability of confusion.

7. The speech sound intelligibility assessment system of claim 6 , wherein,

the output section presents a plurality of speech sounds;

the positive component determination section and the negative component determination section receive electroencephalogram data obtained by taking an arithmetic mean of event-related potentials with respect to each speech sound type, each consonant type, or each group concerning probability of confusion in connection with the presented plurality of speech sounds;

based on the electroencephalogram data, the positive component determination section determines presence or absence of the positive component of the event-related potential with respect to each speech sound type, each consonant type, or each group concerning probability of confusion; and

based on the electroencephalogram data, the negative component determination section determines presence or absence of the negative component of the event-related potential with respect to each speech sound type, each consonant type, or each group concerning probability of confusion.

8. A speech sound intelligibility assessment method comprising the steps of:

providing a speech sound database retaining a plurality of speech sounds;

determining a speech sound to be presented by referring to the speech sound database;

presenting the determined speech sound to a user;

measuring an electroencephalogram signal of the user;

determining presence or absence of a positive component of an event-related potential in the electroencephalogram signal in a zone from 600 ms to 800 ms from a starting point, the starting point being a point in time at which the output section presents a speech sound;

determining presence or absence of a negative component of an event-related potential in the electroencephalogram signal in a zone from 100 ms to 300 ms from a starting point, the starting point being a point in time at which the output section presents a speech sound; and

evaluating whether the user has clearly aurally comprehended the presented speech sound or not based on a result of determination of presence or absence of the positive component and a result of determination of presence or absence of the negative component;

wherein the evaluating step:

makes an evaluation that the user has clearly aurally comprehended the presented speech sound when a result of determination of presence or absence of the positive component indicates that the positive component is absent;

makes an evaluation that the user has not clearly aurally comprehended the presented speech sound due to an insufficient overall sound pressure when a result of determination of presence or absence of the positive component indicates that the positive component is present and a result of determination of presence or absence of the negative component indicates that the negative component is absent; or

makes an evaluation that the user has not clearly aurally comprehended the presented speech sound due to an insufficient sound pressure of a consonant frequency when a result of determination of presence or absence of the positive component indicates that the positive component is present and a result of determination of presence or absence of the negative component indicates that the negative component is present.

9. A computer program, stored on a non-transitory computer-readable medium, to be executed by a computer mounted in a speech sound intelligibility assessment system including a speech sound database retaining a plurality of speech sounds, wherein

the computer program causes the computer in the speech sound intelligibility assessment system to execute the steps of:

determining a speech sound to be presented by referring to the speech sound database;

presenting the determined speech sound to a user;

measuring an electroencephalogram signal of the user;

determining presence or absence of a positive component of an event-related potential in the electroencephalogram signal in a zone from 600 ms to 800 ms from a starting point, the starting point being a point in time at which the output section presents a speech sound;

determining presence or absence of a negative component of an event-related potential in the electroencephalogram signal in a zone from 100 ms to 300 ms from a starting point, the starting point being a point in time at which the output section presents a speech sound; and

evaluating a speech sound intelligibility indicating whether the user has clearly aurally comprehended the presented speech sound or not based on a result of determination of presence or absence of the positive component and a result of determination of presence or absence of the negative component;

wherein the evaluating step:

makes an evaluation that the user has clearly aurally comprehended the presented speech sound when a result of determination of presence or absence of the positive component indicates that the positive component is absent;

makes an evaluation that the user has not clearly aurally comprehended the presented speech sound due to an insufficient overall sound pressure when a result of determination of presence or absence of the positive component indicates that the positive component is present and a result of determination of presence or absence of the negative component indicates that the negative component is absent; or

makes an evaluation that the user has not clearly aurally comprehended the presented speech sound due to an insufficient sound pressure of a consonant frequency when a result of determination of presence or absence of the positive component indicates that the positive component is present and a result of determination of presence or absence of the negative component indicates that the negative component is present.

10. A speech sound intelligibility assessment apparatus comprising:

a presented-speech sound control section for determining a speech sound to be presented by referring to a speech sound database retaining a plurality of speech sounds;

a positive component determination section for determining presence or absence of a positive component of an event-related potential in an electroencephalogram signal of a user measured by a biological signal measurement section in a zone from 600 ms to 800 ms from a starting point, the starting point being a point in time at which the speech sound is presented;

a negative component determination section for determining presence or absence of a negative component in an event-related potential in the electroencephalogram signal in a zone from 100 ms to 300 ms from a starting point, the starting point being a point in time at which a speech sound is presented by an output section; and

a speech sound intelligibility assessment section for evaluating whether the user has clearly aurally comprehended the presented speech sound or not based on a result of determination as to presence or absence of the positive component acquired from the positive component determination section and a result of determination as to presence or absence of the negative component acquired from the negative component determination section;

wherein the speech sound intelligibility assessment section:

makes an evaluation that the user has clearly aurally comprehended the presented speech sound when the result of determination by the positive component determination section indicates that the positive component is absent;

makes an evaluation that the user has not clearly aurally comprehended the presented speech sound due to an insufficient overall sound pressure when the result of determination by the positive component determination section indicates that the positive component is present and the result of determination by the negative component determination section indicates that the negative component is absent; or

makes an evaluation that the user has not clearly aurally comprehended the presented speech sound due to an insufficient sound pressure of a consonant frequency when the result of determination by the positive component determination section indicates that the positive component is present and the result of determination by the negative component determination section indicates that the negative component is present.

11. A method of operating a speech sound intelligibility assessment system, comprising:

a step where a presented-speech sound control section determines a speech sound to be presented by referring to a speech sound database retaining a plurality of speech sounds;

a step where an output section presents the determined speech sound to a user;

a step where an electroencephalogram measurement section measures an electroencephalogram signal of the user;

a step where a positive component determination section determines presence or absence of a positive component of an event-related potential in the electroencephalogram signal in a zone from 600 ms to 800 ms from a starting point, the starting point being a point in time at which the output section presents a speech sound;

a step where a negative component determination section determines presence or absence of a negative component of an event-related potential in the electroencephalogram signal in a zone from 100 ms to 300 ms from a starting point, the starting point being a point in time at which the output section presents a speech sound; and

a step where a speech sound intelligibility assessment section evaluates whether the user has clearly aurally comprehended the presented speech sound or not based on a result of determination of presence or absence of the positive component and a result of determination of presence or absence of the negative component;

wherein the evaluating step by the speech sound intelligibility assessment section:

makes an evaluation that the user has clearly aurally comprehended the presented speech sound when a result of determination of presence or absence of the positive component indicates that the positive component is absent;

makes an evaluation that the user has not clearly aurally comprehended the presented speech sound due to an insufficient overall sound pressure when a result of determination of presence or absence of the positive component indicates that the positive component is present and a result of determination of presence or absence of the negative component indicates that the negative component is absent; or

makes an evaluation that the user has not clearly aurally comprehended the presented speech sound due to an insufficient sound pressure of a consonant frequency when a result of determination of presence or absence of the positive component indicates that the positive component is present and a result of determination of presence or absence of the negative component indicates that the negative component is present.

12. A speech sound intelligibility assessment system comprising:

a speech sound database retaining a plurality of speech sounds;

a presented-speech sound control section for determining a speech sound to be presented by referring to the speech sound database;

an output section for presenting the determined speech sound to a user;

a biological signal measurement section for measuring an electroencephalogram signal of the user;

a positive component determination section for determining presence or absence of a positive component of an event-related potential in the electroencephalogram signal in a zone from 600 ms to 800 ms from a starting point, the starting point being a point in time at which the output section presents a speech sound;

a negative component determination section for determining presence or absence of a negative component of an event-related potential in the electroencephalogram signal in a zone from 100 ms to 300 ms from a starting point, the starting point being a point in time at which the output section presents a speech sound; and

a speech sound intelligibility assessment section for evaluating whether the user has clearly aurally comprehended the presented speech sound or not, based on a result of determination as to presence or absence of the positive component acquired from the positive component determination section and a result of determination as to presence or absence of the negative component acquired from the negative component determination section; and

a stimulation speech sound gain adjustment section for, when the positive component is absent and the negative component is absent, increasing the gain for the entire frequency band; or when the positive component is absent and the negative component is present, calculating a consonant frequency band of the speech sound and increasing the gain for the consonant frequency band;

wherein the speech sound intelligibility assessment section:

makes an evaluation that the user has clearly aurally comprehended the presented speech sound when the result of determination by the positive component determination section indicates that the positive component is absent;

makes an evaluation that the user has not clearly aurally comprehended the presented speech sound due to an insufficient overall sound pressure when the result of determination by the positive component determination section indicates that the positive component is present and the result of determination by the negative component determination section indicates that the negative component is absent; or

makes an evaluation that the user has not clearly aurally comprehended the presented speech sound due to an insufficient sound pressure of a consonant frequency when the result of determination by the positive component determination section indicates that the positive component is present and the result of determination by the negative component determination section indicates that the negative component is present.

Assignments (2)
CORRECTIVE ASSIGNMENT TO CORRECT THE ERRONEOUSLY FILED APPLICATION NUMBERS 13/384239, 13/498734, 14/116681 AND 14/301144 PREVIOUSLY RECORDED ON REEL 034194 FRAME 0143. ASSIGNOR(S) HEREBY CONFIRMS THE ASSIGNMENT. Recorded Dec 24, 2020
From: PANASONIC CORPORATION
To: PANASONIC INTELLECTUAL PROPERTY MANAGEMENT CO., LTD.
Reel/Frame 056788/0362 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 10, 2014
From: PANASONIC CORPORATION
To: PANASONIC INTELLECTUAL PROPERTY MANAGEMENT CO., LTD.
Reel/Frame 034194/0143 →