SIGNAL PROCESSING DEVICE, SIGNAL PROCESSING METHOD, AND SIGNAL PROCESSING PROGRAM
A signal processing device ( 10 ) includes: a speech enhancement unit ( 11 ) that generates, from an observation signal, an enhancement signal in which a voice of a speaker is enhanced; an original sound addition unit ( 12 ) that adds the observation signal to the enhancement signal; and a speech recognition unit ( 13 ) that performs speech recognition on the enhancement signal to which the observation signal is added by the original sound addition unit ( 12 ).
1 . A signal processing device comprising:
a speech enhancement unit that generates, from an observation signal, an enhancement signal in which a voice of a speaker is enhanced;
an addition unit that adds the observation signal to the enhancement signal; and
a speech recognition unit that performs speech recognition on the enhancement signal to which the observation signal is added by the addition unit.
2 . The signal processing device according to claim 1 , wherein the observation signal is an audio signal recorded by a single microphone.
3 . The signal processing device according to claim 1 , wherein the addition unit adjusts a weight of an observation signal to be added to the enhancement signal according to a ratio of a noise signal included in the observation signal.
4 . The signal processing device according to claim 3 , wherein the addition unit weights only an observation signal to be added to the enhancement signal, or weights both the observation signal and the observation signal to be added to the enhancement signal in a relationship in which a sum of a weight of the observation signal and a weight of the observation signal to be added to the enhancement signal is 1.
5 . A signal processing method executed by a signal processing device, the signal processing method comprising:
a process of generating, from an observation signal, an enhancement signal in which a voice of a speaker is enhanced;
a process of adding the observation signal to the enhancement signal; and
a process of performing speech recognition on the enhancement signal to which the observation signal is added in the adding process.
6 . A signal processing program for causing a computer to execute:
generating, from an observation signal, an enhancement signal in which a voice of a speaker is enhanced;
adding the observation signal to the enhancement signal; and
performing speech recognition on the enhancement signal to which the observation signal is added in the adding step.
7 . The signal processing method according to claim 5 , wherein the observation signal is an audio signal recorded by a single microphone.
8 . The signal processing method according to claim 5 , wherein a weight of an observation signal is adjusted to be added to the enhancement signal according to a ratio of a noise signal included in the observation signal.
9 . The signal processing method according to claim 5 , wherein an observation signal is weighted and added to the enhancement signal, or weights both the observation signal and the observation signal to be added to the enhancement signal in a relationship in which a sum of a weight of the observation signal and a weight of the observation signal to be added to the enhancement signal is 1.
10 . The signal processing program according to claim 6 , wherein the observation signal is an audio signal recorded by a single microphone.
11 . The signal processing program according to claim 6 , wherein a weight of an observation signal is adjusted to be added to the enhancement signal according to a ratio of a noise signal included in the observation signal.
12 . The signal processing program according to claim 6 , wherein an observation signal is weighted and added to the enhancement signal, or weights both the observation signal and the observation signal to be added to the enhancement signal in a relationship in which a sum of a weight of the observation signal and a weight of the observation signal to be added to the enhancement signal is 1.
13 . The signal processing device according to claim 1 , wherein the speech recognition unit performs speech recognition on modification emphasis signal and outputs a speech recognition result obtained by converting a message signal into text.
14 . The signal processing device according to claim 1 , wherein the speech recognition unit performs speech enhancement processing using a trained deep learning model.
15 . The signal processing method according to claim 5 , wherein speech recognition is performed on modification emphasis signal and a speech recognition result obtained by converting a message signal into text is output.
16 . The signal processing method according to claim 5 , wherein speech enhancement processing is performed using a trained deep learning model.
17 . The signal processing program according to claim 6 , wherein speech recognition is performed on modification emphasis signal and a speech recognition result obtained by converting a message signal into text is output.
18 . The signal processing program according to claim 6 , wherein speech enhancement is performed processing using a trained deep learning model.