Device for detecting music data from video contents, and method for controlling same
A data processing method according to the present invention comprises the steps of: receiving an input of video contents including a video stream and an audio stream; detecting music data from the audio stream; and filtering the audio stream so that the music data detected from the audio stream is removed.
1. A data processing method comprising:
receiving video content including a video stream and an audio stream;
detecting music data from the audio stream; and
filtering the audio stream to remove the music data detected from the audio stream,
wherein the detecting of the music data from the audio stream comprises a division operation of dividing the audio stream into music data and voice data and a detection operation of detecting a section in which the music data exists from the audio stream,
wherein the division operation is performed by a first artificial intelligence (AI) model which is trained in advance,
wherein the first AI model, which is composed of an artificial neural network that performs deep learning or machine learning, is configured to perform learning using training data labeled as music or voice,
wherein the first AI model is configured to output a probability that each preset unit section of the audio stream corresponds to the music data and a probability that each preset unit section of the audio stream corresponds to the voice data,
wherein the detection operation is performed by a second artificial intelligence (AI) model which is trained in advance, and
wherein the second AI model is configured to perform learning using training data identified in advance as including music or not.
2. The data processing method of claim 1 , wherein the filtering of the audio stream comprises:
determining whether the detected music data is a copyrighted work on the basis of copyright information of the detected music data; and
filtering the audio stream in accordance with whether the detected music data is a copyrighted work.
3. The data processing method of claim 1 , further comprising changing the detected music data with substitute music data which is different from the music data in the audio stream.