Apparatus and method for encoding/decoding audio signal using information of previous frame
Disclosed is an apparatus and method for encoding/decoding an audio signal using information of a previous frame. An audio signal encoding method includes: generating a current latent vector by reducing dimension of a current frame of an audio signal; generating a concatenation vector by concatenating a previous latent vector generated by reducing dimension of a previous frame of the audio signal with the current latent vector; and encoding and quantizing the concatenation vector.
1. A method for encoding audio signal, comprising:
generating a current latent vector by reducing a dimension of a current frame of an audio signal;
generating a concatenation vector by concatenating a previous latent vector generated by reducing a dimension of a previous frame of the audio signal with the current latent vector; and
encoding and quantizing the concatenation vector to output a bit stream, and
wherein the generating the current latent vector reduces the dimension of the current frame of the audio signal using a neural network;
wherein the neural network is trained according to a loss function of the current latent vector calculated by setting the previous latent vector as a conditional probability.
2. The method of claim 1 ,
wherein the neural network is trained according to an entropy of the current latent vector calculated by setting the previous latent vector as a conditional probability.