SOUND ESTIMATING MODEL ACQUIRING APPARATUS, SOUND ESTIMATING APPARATUS, SOUND ESTIMATING MODEL ACQUIRING METHOD, SOUND ESTIMATING METHOD AND PROGRAM
According to an aspect of the present invention, there is provided a sound estimation model acquisition device including a model acquisition unit configured to acquire a mathematical model that estimates an estimated time series that is a time series satisfying a predetermined estimation condition from one or a plurality of time series on a basis of a first sound time series that is a time series indicating a first sound, a second sound time series that is a time series indicating a second sound, and first difference information indicating at least a partial difference between the first sound and the second sound, in which the mathematical model is a mathematical model that estimates the estimated time series on the basis of an input time series that is a time series indicating an input sound and second difference information that is information indicating at least a partial difference between the input time series and the estimated time series, and the estimation condition is a condition that a difference between a difference from the input time series and a difference indicated by the second difference information is smaller than a predetermined difference.
1 . A sound estimation model acquisition device comprising:
a processor; and
a storage medium having computer program instructions stored thereon, wherein the computer program instruction, when executed by the processor, perform processing of:
acquiring a mathematical model that estimates an estimated time series that is a time series satisfying a predetermined estimation condition from one or a plurality of time series on a basis of a first sound time series that is a time series indicating a first sound, a second sound time series that is a time series indicating a second sound, and first difference information indicating at least a partial difference between the first sound and the second sound,
wherein the mathematical model is a mathematical model that estimates the estimated time series on a basis of an input time series that is a time series indicating an input sound and second difference information that is information indicating at least a partial difference between the input time series and the estimated time series, and
the estimation condition is a condition that a difference between a difference from the input time series and a difference indicated by the second difference information is smaller than a predetermined difference.
2 . The sound estimation model acquisition device according to claim 1 ,
wherein the first difference information and the second difference information are pieces of text data.
3 . The sound estimation model acquisition device according to claim 1 ,
wherein the first difference information and the second difference information are pieces of image data.
4 . A sound estimation device comprising:
a processor; and
a storage medium having computer program instructions stored thereon, wherein the computer program instruction, when executed by the processor, perform processing of:
acquiring query information including a sound time series that is a time series indicating a sound and second difference information indicating at least a partial difference between an estimated time series that is a time series satisfying a predetermined estimation condition and the sound time series; and
estimating a time series satisfying the estimation condition from one or a plurality of time series on a basis of the query information, by using a mathematical model acquired by a sound estimation model acquisition device comprising: a processor; and a storage medium having computer program instructions stored thereon, wherein the computer program instruction, when executed by the processor, perform processing of: acquiring the mathematical model that estimates a time series satisfying the estimation condition from one or a plurality of time series on a basis of a first sound time series that is a time series indicating a first sound, a second sound time series that is a time series indicating a second sound, and first difference information indicating at least a partial difference between the first sound and the second sound, the mathematical model being a mathematical model that estimates a time series satisfying the estimation condition on a basis of an input time series that is a time series indicating an input sound and second difference information that is information indicating at least a partial difference between the input time series and a time series satisfying the estimation condition, and the estimation condition being a condition that a difference between a difference from the input time series and a difference indicated by the second difference information is smaller than a predetermined difference.
5 . A sound estimation model acquisition method comprising
acquiring a mathematical model that estimates an estimated time series that is a time series satisfying a predetermined estimation condition from one or a plurality of time series on a basis of a first sound time series that is a time series indicating a first sound, a second sound time series that is a time series indicating a second sound, and first difference information indicating at least a partial difference between the first sound and the second sound,
wherein the mathematical model is a mathematical model that estimates the estimated time series on a basis of an input time series that is a time series indicating an input sound and second difference information that is information indicating at least a partial difference between the input time series and the estimated time series, and
the estimation condition is a condition that a difference between a difference from the input time series and a difference indicated by the second difference information is smaller than a predetermined difference.
6 . A sound estimation method comprising:
acquiring query information including a sound time series that is a time series indicating a sound and second difference information indicating at least a partial difference from an estimated time series that is a time series satisfying a predetermined estimation condition; and
estimating a time series satisfying the estimation condition from one or a plurality of time series on a basis of the query information acquired, by using a mathematical model acquired by a sound estimation model acquisition device comprising: a processor; and a storage medium having computer program instructions stored thereon, wherein the computer program instruction, when executed by the processor, perform processing of: acquiring the mathematical model that estimates a time series satisfying the estimation condition from one or a plurality of time series on a basis of a first sound time series that is a time series indicating a first sound, a second sound time series that is a time series indicating a second sound, and first difference information indicating at least a partial difference between the first sound and the second sound, the mathematical model being a mathematical model that estimates a time series satisfying the estimation condition on a basis of an input time series that is a time series indicating an input sound and second difference information that is information indicating at least a partial difference between the input time series and a time series satisfying the estimation condition, and the estimation condition being a condition that a difference between a difference from the input time series and a difference indicated by the second difference information is smaller than a predetermined difference.
7 . A non-transitory computer readable medium which stores a program for causing a computer to function as the sound estimation model acquisition device according to claim 1 .
8 . A non-transitory computer readable medium which stores a program for causing a computer to function as the sound estimation device according to claim 4 .
9 . The sound estimation model acquisition device according to claim 1 ,
wherein the mathematical model is also acquired on a basis of a first sound event class label indicating a classification result of the first sound time series and a second sound event class label indicating a classification result of the second sound time series.
10 . The sound estimation device according to claim 4 ,
wherein the mathematical model is also acquired on a basis of a first sound event class label indicating a classification result of the first sound time series and a second sound event class label indicating a classification result of the second sound time series.