IP Library Granted Patent US 12,437,185
Granted Patent B2
US 12,437,185 · App. 17/564,912 · Granted Oct 7, 2025

Apparatus and method for artificial intelligence neural network based on co-evolving neural ordinary differential equations

Inventors: No Seong Park (Seoul, KR); Sheo Yon Jhin (Goyang-si, KR); Min Ju Jo (Seoul, KR); Tae Yong Kong (Seoul, KR); Jin Sung Jeon (Seoul, KR)
Assignee: UIF (UNIVERSITY INDUSTRY FOUNDATION), YONSEI UNIVERSITY
G06N3/045G06F17/13G06N3/048
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,437,185
App. No.
17/564,912
Granted
Oct 7, 2025
Kind
B2
Abstract

An apparatus for an artificial intelligence neural network based on co-evolving neural ordinary differential equations (NODEs) includes a main NODE module configured to provide a downstream machine learning task; and an attention NODE module configured to receive the downstream machine learning task and provide attention to the main NODE module, in which the main NODE module and the attention NODE module may influence each other over time so that the main NODE module outputs a multivariate time-series value at a given time for an input sample x.

Claims (30)

1. An apparatus for an artificial intelligence neural network based on co-evolving neural ordinary differential equations (NODEs), the apparatus comprising:

at least one processor and memory storing instructions performed by the at least one processor;

a main NODE module configured to provide a downstream machine learning task at an initial time; and

an attention NODE module configured to receive the downstream machine learning task from the main NODE module and provide attention to the main NODE module based on the downstream machine learning task,

wherein the main NODE module receives the attention provided from the attention NODE module and provides the downstream machine learning task at a next time after the initial time such that the main NODE module and the attention NODE module influence each other over time so that the main NODE module outputs a multivariate time-series value at a given time for an input sample x,

wherein the attention is used in a feature extraction layer before a NODE layer and does not introduce a new NODE model that is internally combined with the attention,

wherein Explainable Tensorized Neural (ETN)-ODE uses the attention to derive a correlation matrix in the feature extraction layer and then evolve the correlation matrix using the NODE layer, and

wherein the main NODE module and the attention NODE module are each implemented via the at least one processor.

2. The apparatus of claim 1 , further comprising: a feature extraction module configured to extract a feature for the input sample x to generate an initial feature vector at the initial time, and provide the initial feature vector to the main NODE module,

wherein the feature extraction module is implemented via the at least one processor.

3. The apparatus of claim 2 , further comprising: an initial attention generating module configured to receive the initial feature vector to generate initial attention, and provide the initial attention to the attention NODE module to support calculation of a multivariate time-series value at a next time,

wherein the initial attention generating module is implemented via the at least one processor.

4. The apparatus of claim 1 , wherein the main NODE module performs integration of an ordinary differential equation (ODE) function with respect to the ODE function for time, and the ODE function receives 1) the multivariate time-series value and 2) the attention at the given time.

5. The apparatus of claim 4 , wherein the attention NODE module performs element-wise multiplication between the multivariate time-series value and a sigmoid activation function of a time-evolving matrix as the ODE function.

6. The apparatus of claim 4 , wherein the attention NODE module performs integration of an attention generation function for the ODE function for time, and the attention generation function receives 1) a time evolution matrix representing a logit value of the attention and 2) the multivariate time-series value is input at the given time.

7. The apparatus of claim 1 , further comprising: a classification module configured to receive the multivariate time-series value at the given time and performs prediction on the input sample x,

wherein the classification module is implemented via the at least one processor.

8. A method for an artificial intelligence neural network based on co-evolving neural ordinary differential equations (NODEs), the method comprising:

providing, by a main NODE module, a downstream machine learning task at an initial time; and

receiving, by an attention NODE module, the downstream machine learning task from the main NODE module and providing attention to the main NODE module based on the downstream machine learning task,

wherein the main NODE module receives the attention provided from the attention NODE module and provides the downstream machine learning task at a next time such that the main NODE module and the attention NODE module influence each other over time so that the main NODE module outputs a multivariate time-series value at a given time for an input sample x,

wherein the attention is used in a feature extraction layer before a NODE layer and does not introduce a new NODE model that is internally combined with the attention,

wherein Explainable Tensorized Neural (ETN)-ODE uses the attention to derive a correlation matrix in the feature extraction layer and then evolve the correlation matrix using the NODE layer.

9. The method of claim 8 , further comprising:

extracting, by a feature extraction module, a feature for the input sample x to generate an initial feature vector at the initial time, and providing the initial feature vector to the main NODE module.

10. The method of claim 9 , further comprising:

receiving, by an initial attention generating module, the initial feature vector to generate initial attention, and providing the initial attention to the attention NODE module to support calculation of a multivariate time-series value at a next time.

11. The method of claim 8 , wherein the main NODE module performs integration of an ordinary differential equation (ODE) function with respect to adjoint times, and the ODE function receives 1) the multivariate time-series value and 2) the attention at the given time.

12. The method of claim 8 , further comprising:

receiving, by a classification module, the multivariate time-series value at the given time and performing prediction on the input sample x.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 29, 2021
From: PARK, NO SEONG; JHIN, SHEO YON; JO, MIN JU; KONG, TAE YONG; JEON, JIN SUNG
To: UIF (UNIVERSITY INDUSTRY FOUNDATION), YONSEI UNIVERSITY
Reel/Frame 058501/0942 →
Priority Claims (1)
KR 10-2021-0181699 · Dec 17, 2021 · national
Continuity (1)
Related Publication 20230196071A1 · Jun 22, 2023
References Cited (7)
KR 1020210031197A · 2021 [cited by applicant]
Xia, Hedi, et al. “Heavy ball neural ordinary differential equations.” Advances in Neural Information Processing Systems 34 (2021): 18646-18659. (Year: 2021). [cited by examiner]
Li, Bei, et al. “ODE Transformer: An Ordinary Differential Equation-Inspired Model for Sequence Generation.” (Year: 2021). [cited by examiner]
Zhang, Jing, et al. “Continuous self-attention models with neural ODE networks.” Proceedings of the AAAI Conference on Artificial Intelligence. vol. 35. No. 16. 2021. (Year: 2021). [cited by examiner]
Baier-Reinio, Aaron, and Hans De Sterck. “N-ode transformer: A depth-adaptive variant of the transformer using neural ordinary differential equations.” arXiv preprint arXiv:2010.11358 (2020). (Year: 2020). [cited by examiner]
Gao, Penglei, et al. “Explainable tensorized neural ordinary differential equations for arbitrary-step time series prediction.” IEEE Transactions on Knowledge and Data Engineering 35.6 (2022): 5837-5850. (Year: 2022). [cited by examiner]
Guo, Dan, et al. “Iterative context-aware graph inference for visual dialog.” Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition. 2020. (Year: 2020). [cited by examiner]