IP Library Granted Patent US 12,417,391
Granted Patent B2
US 12,417,391 · App. 17/844,701 · Granted Sep 16, 2025

Modular adaptation for cross-domain few-shot learning

Inventors: Xiao Lin (Union City, CA); Meng Ye (Lawrenceville, NJ); Yunye Gong (West Windsor, NJ); Giedrius T. Burachas (Doylestown, PA); Ajay Divakaran (Monmouth Junction, NJ); Yi Yao (Princeton, NJ); Nikoletta Basiou (Palo Alto, CA)
Assignee: SRI International
G06N3/088G06F18/211G06N3/047
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,417,391
App. No.
17/844,701
Granted
Sep 16, 2025
Kind
B2
Abstract

A method, apparatus and system for adapting a pre-trained network for application to a different dataset includes arranging at least two different types of active adaptation modules in a pipeline configuration, wherein an output of a previous active adaptation module produces an input for a next active adaptation module in the pipeline in the form of adapted network data until a last active adaptation module, and wherein each of the at least two different types of adaptation modules can be switched on or off, determining at least one respective hyperparameter for each of the at least two different types of active adaptation modules, and applying the at least one respective determined hyperparameter to each of the at least two different types of active adaptation modules for processing received data from the pretrained network to determine an adapted network.

Claims (34)

1. A method for adapting a pre-trained network for application to a different dataset, comprising:

arranging at least two different types of active adaptation modules in a pipeline configuration, wherein an output of a previous active adaptation module produces an input for a next active adaptation module in the pipeline in the form of adapted network data until a last active adaptation module, and wherein each of the at least two different types of adaptation modules can be switched on or off;

determining at least one respective hyperparameter for each of the at least two different types of active adaptation modules; and

applying the at least one respective determined hyperparameter to each of the at least two different types of active adaptation modules for processing received data from the pretrained network to determine an adapted network.

2. The method of claim 1 , further comprising:

applying an adaptive learning process to an output of at least one of the at least two different types of active adaptation modules.

3. The method of claim 1 , wherein the arranging, the determining and the applying comprise an iterative process in which at least one of the at least two different types of active adaptation modules in the pipeline or a respective hyperparameter is changed in each subsequent iteration.

4. The method of claim 1 , wherein at least one of the at least two different types of the active adaptation modules or the respective hyperparameters are selected from a collection of historically well-functioning adaptation modules or hyperparameters stored in a storage device.

5. The method of claim 1 , wherein at least one of the at least two different types of the active adaptation modules or the respective hyperparameters are selected based on at least one a user input or an input from a machine learning process.

6. The method of claim 5 , wherein the machine learning process is trained to determine at least one of an adaptation module for the pipeline or a hyperparameter for an adaptation module based on a target dataset.

7. The method of claim 1 , wherein data from the pre-trained network comprises at least one of classification data or data regarding an embedding space.

8. A non-transitory machine-readable medium having stored thereon at least one program, the at least one program including instructions which, when executed by a processor, cause the processor to perform a method in a processor based system for adapting a pre-trained network for application to a different dataset, comprising:

arranging at least two different types of active adaptation modules in a pipeline configuration, wherein an output of a previous active adaptation module produces an input for a next active adaptation module in the pipeline in the form of adapted network data until a last active adaptation module, and wherein each of the at least two different types of adaptation modules can be switched on or off;

determining at least one respective hyperparameter for each of the at least two different types of active adaptation modules; and

applying the at least one respective determined hyperparameter to each of the at least two different types of active adaptation modules for processing received data from the pretrained network to determine an adapted network.

9. The non-transitory machine-readable medium of claim 8 , further comprising:

applying an adaptive learning process to an output of at least one of the at least two different types of active adaptation modules.

10. The non-transitory machine-readable medium of claim 8 , in which the arranging, the determining and the applying steps comprise an iterative process in which at least one of an order of the at least two different types of active adaptation modules in the pipeline or a respective hyperparameter is changed in each subsequent iteration.

11. The non-transitory machine-readable medium of claim 8 , wherein at least one of the at least two different types of the active adaptation modules or the respective hyperparameters are selected from a collection of historically well-functioning adaptation modules or hyperparameters stored in a storage device.

12. The non-transitory machine-readable medium of claim 8 , wherein at least one of the at least two different types of the active adaptation modules or the respective hyperparameters are selected based on at least one a user input or an input from a machine learning process.

13. The non-transitory machine-readable medium of claim 12 , wherein the machine learning process is trained to determine at least one of an adaptation module for the pipeline or a hyperparameter for an adaptation module based on a target dataset.

14. The non-transitory machine-readable medium of claim 8 , wherein data from the pre-trained network comprises at least one of classification data or data regarding an embedding space.

15. A system for adapting a pre-trained network for application to a different dataset, comprising:

a storage device; and

a computing device comprising a processor and a memory having stored therein at least one program, the at least one program including instructions which, when executed by the processor, cause the computing device to perform a method, comprising:

arranging at least two different types of active adaptation modules in a pipeline configuration, wherein an output of a previous active adaptation module produces an input for a next active adaptation module in the pipeline in the form of adapted network data until a last active adaptation module, and wherein each of the at least two different types of adaptation modules can be switched on or off;

determining at least one respective hyperparameter for each of the at least two different types of active adaptation modules; and

applying the at least one respective determined hyperparameter to each of the at least two different types of active adaptation modules for processing received data from the pretrained network to determine an adapted network.

16. The system of claim 15 , further comprising:

applying an adaptive learning process to an output of at least one of the at least two different types of active adaptation modules.

17. The system of claim 15 , in which the arranging, the determining and the applying steps comprise an iterative process in which at least one of an order of the at least two different types of active adaptation modules in the pipeline or a respective hyperparameter is changed in each subsequent iteration.

18. The system of claim 15 , wherein at least one of the at least two different types of the active adaptation modules or the respective hyperparameters are selected from a collection of historically well-functioning adaptation modules or hyperparameters stored in the storage device.

19. The system of claim 15 , wherein at least one of the at least two different types of the active adaptation modules or the respective hyperparameters are selected based on at least one a user input or an input from a machine learning process.

20. The method of claim 19 , wherein the machine learning process is trained to determine at least one of an adaptation module for the pipeline or a hyperparameter for an adaptation module based on a known target dataset.

Assignments (2)
CONFIRMATORY LICENSE Recorded Jul 13, 2022
From: SRI INTERNATIONAL
To: GOVERNMENT OF THE UNITED STATES AS REPRESENTED BY THE SECRETARY OF THE AIR FORCE
Reel/Frame 060649/0922 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jul 12, 2022
From: LIN, XIAO; YE, MENG; GONG, YUNYE; BURACHAS, GIEDRIUS T.; DIVAKARAN, AJAY; YAO, YI; BASIOU, NIKOLETTA
To: SRI INTERNATIONAL
Reel/Frame 060483/0909 →
Continuity (2)
Provisional Application 63214128 · Jun 23, 2021
Related Publication 20220414476A1 · Dec 29, 2022
References Cited (15)
Chen, et al., A simple framework for contrastive learning of visual representations. In ICML, pp. 1597-1607. PMLR, 2020. [cited by applicant]
Finn, et al., Model-agnostic meta-learning for fast adaptation of deep networks. In [cited by applicant]
Guo, et al., Spottune: transfer learning through adaptive fine-tuning. In [cited by applicant]
Helber, et al., Eurosat: A novel dataset and deep learning benchmark for land use and land cover classification. IEEE Journal of Selected Topics in Applied Earth Observations and Remote Sensing. 12(7): 2217-2226, 2019. [cited by applicant]
Li, et al., Rethinking the hyperparameters for fine-tuning. ICLR, 2020, pp. 1-20. [cited by applicant]
Li, et al., Explicit inductive bias for transfer learning with convolutional networks. ICML, 2018, 12 pages. [cited by applicant]
Liu, et al., Learning to propagate labels: Transductive propagation network for few-shot learning. ICLR, 2019, pp. 1-14. [cited by applicant]
Perrone, et al., Scalable hyperparameter transfer learning. In NeurIPS, pp. 6846-6856, 2018. [cited by applicant]
Zammir, et al., Disentangling task transfer learning. In Proceedings of the IEEE, conference on computer vision and pattern recognition, pp. 3712-3722, 2018. [cited by applicant]
Zhai, et al., The visual task adaptation benchmark. 2019, 33 pages. [cited by applicant]
Zoph, et al., Neural architecture search with reinforcement learning. ICLR, 2017, pp. 1-16. [cited by applicant]
Hu, et al., Leveraging the feature distribution in transfer-based few-shot learning. arXiv preprint arXiv: 2006.03806, 2020. [cited by applicant]
Mangla, et al., Charting the right manifold: Manifold mixup for few-shot learning. In WAVC, pp. 2218-2227, 2020. [cited by applicant]
Qiao, et al., Few-shot image recognition by predicting parameters from activations. In CVPR, 2018, 10 pages. [cited by applicant]
Tseng, et al., Cross-domain few-shot classification via learned feature-wise transformation. In ICLR, 2020, pp. 1-18. [cited by applicant]