IP Library Granted Patent US 11,462,210
Granted Patent B2
US 11,462,210 · App. 16/593,488 · Granted Oct 4, 2022

Data collecting method and system

Inventors: Jung Woo Ha (Seongnam-si, KR); Jung Myung Kim (Seongnam-si, KR); Jang Yeon Park (Seongnam-si, KR); Chanju Kim (Seongnam-si, KR); Dong Won Kim (Seongnam-si, KR)
Assignees: NAVER CORPORATION; LINE CORPORATION
G10L15/16G06N3/08G06N20/00G10L25/51
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,462,210
App. No.
16/593,488
Granted
Oct 4, 2022
Kind
B2
Abstract

A method of determining a highlight section of a sound source includes obtaining a sound source and classification information of the sound source, and learning a neural network by using the sound source and the classification information. The neural network includes an input layer including a node corresponding to a feature value of each of a plurality of sections obtained by splitting the sound source according to a time axis, an output layer including a node corresponding to the classification information, a hidden layer defined between the input layer and the output layer, a first function between the input layer and the hidden layer, and a second function between the hidden layer and the output layer, wherein the first function includes an attention model for calculating a weighted sum of the feature value of each section. The highlight section of the sound source is determined based on weight information of a feature value node of each section included in the first function.

Claims (20)

1. A method of determining a highlight section of a sound source, the method comprising:

obtaining, by a processor, a sound source and classification information of the sound source, wherein the classification information is in a form of a vector indicating a probability of corresponding to each classification;

learning, by the processor, a neural network by using the sound source and the classification information, the neural network comprising an input layer including a node corresponding to a feature value of each of a plurality of sections obtained by splitting the sound source into said plurality of sections according to a time axis, an output layer including a node corresponding to the classification information, a hidden layer defined between the input layer and the output layer, a first function between the input layer and the hidden layer, and a second function between the hidden layer and the output layer, wherein the first function comprises an attention model for calculating a weighted sum of the feature value of each section; and

determining, by the processor, the highlight section of the sound source, based on weight information of a feature value node of each section included in the first function, wherein the weight information of the feature value node of each section indicates a degree of contribution of each section to an estimation of the classification information of the sound source.

2. The method of claim 1 , wherein the hidden layer comprises a node corresponding to an integrated feature value of the sound source obtained from the feature value of each section, according to the first function.

3. The method of claim 1 , wherein the first function further comprises a 1-1st function calculating a similarity of an output value of the attention model and an output value of a recurrent neural network (RRN) model, wherein the hidden layer comprises a node of the similarity.

4. The method of claim 1 , further comprising, after the obtaining of the sound source and the classification information,

converting the sound source comprising sound data regarding the time axis to include energy data regarding the time axis,

wherein the plurality of sections are obtained by splitting the converted sound source according to the time axis.

5. The method of claim 4 , wherein the determining of the highlight section comprises determining the highlight section based on the weight information of the feature value node of each section and the energy data of each section.

6. The method of claim 1 , wherein the determining of the highlight section comprises determining an important section based on the weight information of the feature value node of each section and determining the highlight section among a plurality of sections of the sound source by referring to energy data within a section of a pre-set range before and after the important section.

7. The method of claim 6 , wherein the determining of the highlight section comprises determining the highlight section in response to a point of time when momentum of the energy data is greatest within the section of the pre-set range.

8. A non-transitory tangible computer readable recording medium storing a computer program for determining a highlight section of a sound source, the program when executed by a computer performing the method of claim 1 .

9. An apparatus for determining a highlight section of a sound source, the apparatus comprising:

a processor configured to include a plurality of functional units for performing a plurality predetermined functions, the functional units including,

a sound source obtainer configured to obtain a sound source and classification information of the sound source, wherein the classification information is in a form of a vector indicating a probability of corresponding to each classification;

a neural network processor configured to learn a neural network by using the sound source and the classification information, the neural network comprising an input layer including a node corresponding to a feature value of each of a plurality of sections obtained by splitting the sound source into said plurality of sections according to a time axis, an output layer including a node corresponding to the classification information, a hidden layer defined between the input layer and the output layer, a first function between the input layer and the hidden layer, and a second function between the hidden layer and the output layer, wherein the first function comprises an attention model for calculating a weighted sum of the feature value of each section; and

a highlight determiner configured to determine a highlight section of the sound source, based on weight information of a feature value node of each section included in the first function, wherein the weight information of the feature value node of each section indicates a degree of contribution of each section to an estimation of the classification information of the sound source.

10. The method of claim 1 , wherein the classification information includes at least one of a genre, a mood, a preferred age group, a theme and an atmosphere of the sound source.

11. The apparatus of claim 9 , wherein the classification information includes at least one of a genre, a mood, a preferred age group, a theme and an atmosphere of the sound source.

Assignments (7)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 12, 2024
From: Z INTERMEDIATE GLOBAL CORPORATION
To: LY CORPORATION
Reel/Frame 067091/0109 →
CHANGE OF NAME Recorded Apr 10, 2024
From: LINE CORPORATION
To: Z INTERMEDIATE GLOBAL CORPORATION
Reel/Frame 067069/0467 →
CORRECTIVE ASSIGNMENT TO CORRECT THE SPELLING OF THE ASSIGNEES CITY IN THE ADDRESS SHOULD BE TOKYO, JAPAN PREVIOUSLY RECORDED AT REEL: 058597 FRAME: 0303. ASSIGNOR(S) HEREBY CONFIRMS THE ASSIGNMENT. Recorded Jan 17, 2023
From: A HOLDINGS CORPORATION
To: LINE CORPORATION
Reel/Frame 062401/0490 →
CORRECTIVE ASSIGNMENT TO CORRECT THE THE CITY SHOULD BE SPELLED AS TOKYO PREVIOUSLY RECORDED AT REEL: 058597 FRAME: 0141. ASSIGNOR(S) HEREBY CONFIRMS THE ASSIGNMENT. Recorded Jan 17, 2023
From: LINE CORPORATION
To: A HOLDINGS CORPORATION
Reel/Frame 062401/0328 →
CHANGE OF NAME Recorded Dec 28, 2021
From: LINE CORPORATION
To: A HOLDINGS CORPORATION
Reel/Frame 058597/0141 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 28, 2021
From: A HOLDINGS CORPORATION
To: LINE CORPORATION
Reel/Frame 058597/0303 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 4, 2019
From: HA, JUNG WOO; KIM, JUNG MYUNG; PARK, JANG YEON; KIM, CHANJU; KIM, DONG WON
To: NAVER CORPORATION; LINE CORPORATION
Reel/Frame 050629/0109 →
Priority Claims (1)
KR 10-2017-0045391 · Apr 7, 2017 · national
Continuity (2)
Continuation PCTKR2018004061 · Apr 6, 2018
Related Publication 20200035225A1 · Jan 30, 2020
Cited By (1)
US 12,676,161