IP Library Granted Patent US 12,346,809
Granted Patent B2
US 12,346,809 · App. 17/480,999 · Granted Jul 1, 2025

Method, device, and storage medium for deep learning based domain adaptation with data fusion for aerial image data analysis

Inventors: Jingyang Lu (Germantown, MD); Erik Blasch (Arlington, VA); Roman Ilin (Dayton, OH); Hua-mei Chen (Germantown, MD); Dan Shen (Germantown, MD); Nichole Sullivan (Germantown, MD); Genshe Chen (Germantown, MD)
Assignee: INTELLIGENT FUSION TECHNOLOGY, INC.
G06N3/08G06F18/214G06F18/253G06V20/00
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,346,809
App. No.
17/480,999
Granted
Jul 1, 2025
Kind
B2
Abstract

Embodiments of the present disclosure provide a method, a device, and a storage medium for domain adaptation for efficient learning fusion (DAELF). The method includes acquiring data from a plurality of data sources of a plurality of sensors; for each of the plurality of sensors, training an auxiliary classifier generative adversarial network (AC-GAN) by a hardware processor with data from each data source of the plurality of data sources, thereby obtaining a trained feature extraction network and a trained label prediction network for each data source; forming a decision-level fusion network or a feature-level fusion network; and training the decision-level fusion network or the feature-level fusion network with a source-only mode or a generate to adapt (GTA) mode; and applying the trained decision-level fusion network or the trained feature-level fusion network to detect a target of interest.

Claims (48)

1. A domain adaptation for efficient learning fusion (DAELF) method, comprising:

acquiring data from a plurality of data sources of a plurality of sensors;

for each of the plurality of sensors, training an auxiliary classifier generative adversarial network (AC-GAN) by a hardware processor, wherein the AC-GAN includes a feature extraction network, a label prediction network, a generator network, and a discriminator network, with data from each data source of the plurality of data sources, thereby obtaining a trained feature extraction network and a trained label prediction network for each data source;

using the trained feature extraction network and the trained label prediction network for each data source on a sensor side, and a corresponding centralized fusion network on a fusion center side to form a decision-level fusion network; or using the trained feature extraction network for each data source on the sensor side and a corresponding centralized fusion network on the fusion center side to form a feature-level fusion network;

training the decision-level fusion network or the feature-level fusion network with a source-only mode or a generate to adapt (GTA) mode, wherein:

at the source-only mode, the trained feature extraction network for each data source and the corresponding centralized fusion network are trained with labeled source data, and

at the GTA mode, the trained feature extraction network for each data source and the corresponding centralized fusion network are trained separately, wherein the trained feature extraction network for each data source is trained with the labeled source data and unlabeled target data; and the corresponding centralized fusion network is trained with the labeled source data only; and

applying the trained decision-level fusion network or the trained feature-level fusion network to detect a target of interest.

2. The method according to claim 1 , wherein training the AC-GAN includes:

inputting a source sample of the data from each data source into the feature extraction network for each data source to generate an embedding feature used by both the label prediction network and the generator network for each data source; and

inputting a target sample of the data from each data source into the feature extract network for each data source to generate an embedding feature only used by the generator network for each data source.

3. The method according to claim 1 , wherein:

at a training phase, for each data source, the AC-GAN has a stream 1 , including the feature extraction network and the label prediction network, and a stream 2 , including the feature extraction network, the generator network, and the discriminator network.

4. The method according to claim 1 , further including:

displaying the target of interest detected by the trained decision-level fusion network or the trained feature-level fusion network.

5. A domain adaptation for efficient learning fusion (DAELF) device, comprising:

a memory, configured to store program instructions for performing a DAELF method; and

a processor, coupled with the memory and, when executing the program instructions, configured for:

acquiring data from a plurality of data sources of a plurality of sensors;

for each of the plurality of sensors, training an auxiliary classifier generative adversarial network (AC-GAN) by a hardware processor, wherein the AC-GAN includes a feature extraction network, a label prediction network, a generator network, and a discriminator network, with data from each data source of the plurality of data sources, thereby obtaining a trained feature extraction network and a trained label prediction network for each data source;

using the trained feature extraction network and the trained label prediction network for each data source on a sensor side, and a corresponding centralized fusion network on a fusion center side to form a decision-level fusion network; or

using the trained feature extraction network for each data source on the sensor side and a corresponding centralized fusion network on the fusion center side to form a feature-level fusion network;

training the decision-level fusion network or the feature-level fusion network with a source-only mode or a generate to adapt (GTA) mode, wherein:

at the source-only mode, the trained feature extraction network for each data source and the corresponding centralized fusion network are trained with labeled source data, and

at the GTA mode, the trained feature extraction network for each data source and the corresponding centralized fusion network are trained separately, wherein the trained feature extraction network for each data source is trained with the labeled source data and unlabeled target data; and the corresponding centralized fusion network is trained with the labeled source data only; and

applying the trained decision-level fusion network or the trained feature-level fusion network to detect a target of interest.

6. The device according to claim 5 , wherein training the AC-GAN includes:

inputting a source sample of the data from each data source into the feature extraction network for each data source to generate an embedding feature used by both the label prediction network and the generator network for each data source; and

inputting a target sample of the data from each data source into the feature extract network for each data source to generate an embedding feature only used by the generator network for each data source.

7. The device according to claim 5 , wherein:

at a training phase, for each data source, the AC-GAN has a stream 1 , including the feature extraction network and the label prediction network, and a stream 2 , including the feature extraction network, the generator network, and the discriminator network.

8. The device according to claim 5 , wherein the method further includes:

displaying the target of interest detected by the trained decision-level fusion network or the trained feature-level fusion network.

9. A non-transitory computer-readable storage medium, containing program instructions for, when being executed by a processor, performing a domain adaptation for efficient learning fusion (DAELF) method, the method comprising:

acquiring data from a plurality of data sources of a plurality of sensors;

for each of the plurality of sensors, training an auxiliary classifier generative adversarial network (AC-GAN) by a hardware processor, wherein the AC-GAN includes a feature extraction network, a label prediction network, a generator network, and a discriminator network, with data from each data source of the plurality of data sources, thereby obtaining a trained feature extraction network and a trained label prediction network for each data source;

using the trained feature extraction network and the trained label prediction network for each data source on a sensor side, and a corresponding centralized fusion network on a fusion center side to form a decision-level fusion network; or using the trained feature extraction network for each data source on the sensor side and a corresponding centralized fusion network on the fusion center side to form a feature-level fusion network;

training the decision-level fusion network or the feature-level fusion network with a source-only mode or a generate to adapt (GTA) mode, wherein:

at the source-only mode, the trained feature extraction network for each data source and the corresponding centralized fusion network are trained with labeled source data, and

at the GTA mode, the trained feature extraction network for each data source and the corresponding centralized fusion network are trained separately, wherein the trained feature extraction network for each data source is trained with the labeled source data and unlabeled target data; and the corresponding centralized fusion network is trained with the labeled source data only; and

applying the trained decision-level fusion network or the trained feature-level fusion network to detect a target of interest.

10. The storage medium according to claim 9 , wherein training the AC-GAN includes:

inputting a source sample of the data from each data source into the feature extraction network for each data source to generate an embedding feature used by both the label prediction network and the generator network for each data source; and

inputting a target sample of the data from each data source into the feature extract network for each data source to generate an embedding feature only used by the generator network for each data source.

11. The storage medium according to claim 9 , wherein:

at a training phase, for each data source, the AC-GAN has a stream 1 , including the feature extraction network and the label prediction network, and a stream 2 , including the feature extraction network, the generator network, and the discriminator network.

12. The storage medium according to claim 9 , wherein the method further includes:

displaying the target of interest detected by the trained decision-level fusion network or the trained feature-level fusion network.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Sep 23, 2021
From: LU, JINGYANG; BLASCH, ERIK; ILIN, ROMAN; CHEN, HUA-MEI; SHEN, DAN; SULLIVAN, NICHOLE; CHEN, GENSHE
To: INTELLIGENT FUSION TECHNOLOGY, INC.
Reel/Frame 057574/0095 →
Continuity (2)
Provisional Application 63081036 · Sep 21, 2020
Related Publication 20220092420A1 · Mar 24, 2022
References Cited (15)
US 11210570B2 · Shen · 2021 [cited by examiner]
US 11580405B2 · Santillán Perez et al. · 2023 [cited by examiner]
Decision-Level Fusion Method for Emotion Recognition using Multimodal Emotion Recognition Information, Kyu-Seob Song, 2018 15th International Conference on Ubiquitous Robots (UR) Hawaii Convention Center, Hawai'i, USA, … [cited by examiner]
Generate to Adapt: Aligning Domains using Generative Adversarial Networks, Swami Sankaranarayanan et al., arXiv:1704.01705v4 [cs.CV] Apr. 12, 2018 (Year: 2018). [cited by examiner]
Multi-Source Domain Adaptation in the Deep Learning Era, Sicheng Zhao et al., arXiv:2002.12169v1 [cs.LG] Feb. 26, 2020 (Year: 2020). [cited by examiner]
Multilevel Features Convolutional Neural Network for Multifocus Image Fusion, Yong Yang et al., IEEE Transactions on Computational Imaging, vol. 5, No. 2, Jun. 2019 (Year: 2019). [cited by examiner]
Object Detection in Aerial Images Using Feature Fusion Deep Networks, Hao Long et al., DOI 10.1109/ACCESS.2019.2903422 (Year: 2019). [cited by examiner]
Self-Paced Collaborative and Adversarial Network for Unsupervised Domain Adaptation, Weichen Zhang et al., IEEE Transactions on Pattern Analysis and Machine Intelligence, vol. 43, No. 6, Jun. 2021 (Year: 2021). [cited by examiner]
Adaptive Region Aggregation Network: Unsupervised Domain Adaptation With Adversarial Training for ECG Delineation, Ming Chen et al., IEEE 978-1-5090-6631-5/20 (Year: 2020). [cited by examiner]
Classification of Multisource and Hyperspectral Data Based on Decision Fusion, Jon Atli Benediktsson, IEEE Transactions on Geoscience and Remote Sensing, vol. 37, No. 3, May 1999 (Year: 1999). [cited by examiner]
CLDA: an adversarial unsupervised domain adaptation method with classifier-level adaptation, Zhihai He et al., Multimedia Tools and Applications (2020) 79:33973-33991 https://doi.org/10.1007/s11042-020-08877-8 (Year: 20… [cited by examiner]
Domain-Invariant Adversarial Learning for Unsupervised Domain Adaption, Yexun Zhang et al., arXiv:1811.12751v1 [cs. CV] Nov. 30, 2018 (Year: 2018). [cited by examiner]
Improved Techniques for Adversarial Discriminative Domain Adaptation, Aaron Chadha, IEEE Transactions on Image Processing, vol. 29, 2020 (Year: 2020). [cited by examiner]
Learning a Recurrent Residual Fusion Network for Multimodal Matching, Yu Liu et al., Proceedings of the IEEE International Conference on Computer Vision (ICCV), 2017, pp. 4107-4116 (Year: 2017). [cited by examiner]
Multi-source remote sensing data fusion: status and trends, Jixian Zhang, International Journal of Image and Data Fusion, 1:1, 5-24, DOI: 10.1080/19479830903561035 (Year: 2010). [cited by examiner]