IP Library Granted Patent US 12,469,315
Granted Patent B2
US 12,469,315 · App. 17/676,134 · Granted Nov 11, 2025

Systems, methods, and apparatuses for implementing transferable visual words by exploiting the semantics of anatomical patterns for self-supervised learning

Inventors: Fatemeh Haghighi (Tempe, AZ); Mohammad Reza Hosseinzadeh Taher (Tempe, AZ); Zongwei Zhou (Tempe, AZ); Jianming Liang (Scottsdale, AZ)
Assignee: Arizona Board of Regents on behalf of Arizona State University
G06V20/70G06V10/26G06V10/761G06V10/764G06V10/774G06V10/82G06V2201/03
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,469,315
App. No.
17/676,134
Granted
Nov 11, 2025
Kind
B2
Abstract

Described herein are means for the generation of Transferable Visual Word (TransVW) models through self-supervised learning in the absence of manual labeling, in which the trained TransVW models are then utilized for the processing of medical imaging. For instance, an exemplary system is specially configured to perform self-supervised learning for an AI model in the absence of manually labeled input, by performing the following operations: receiving medical images as input; performing a self-discovery operation of anatomical patterns by building a set of the anatomical patterns from the medical images received at the system, performing a self-classification operation of the anatomical patterns; performing a self-restoration operation of the anatomical patterns within cropped and transformed 2D patches or 3D cubes derived from the medical images received at the system by recovering original anatomical patterns to learn different sets of visual representation; and providing a semantics-enriched pre-trained AI model having a trained encoder-decoder structure with skip connections in between based on the performance of the self-discovery operation, the self-classification operation, and the self-restoration operation. Other related embodiments are disclosed.

Claims (48)

1 . A system comprising:

a memory to store instructions;

a processor to execute the instructions stored in the memory;

a receive interface to receive a plurality of unlabeled medical images obtained from a plurality of human patients;

wherein the system is specially configured to perform self-supervised learning for an artificial intelligence (AI) model in the absence of manually labeling the plurality of unlabeled medical images, by executing instructions via the processor for:

selecting a subset of the plurality of unlabeled medical images, corresponding to similar patients, based on deep latent features therein;

extracting two-dimensional (2D) patches or three-dimensional (3D) cubes each representing an appearance of an anatomical pattern reoccurring at fixed coordinates across each of the selected subset of the plurality of unlabeled medical images, and assigning one of a plurality of labels to each of the 2D patches or 3D cubes based on their fixed coordinates;

transforming each of the 2D patches or 3D cubes to generate transformed 2D patches or transformed 3D cubes, respectively

by perturbing each of the plurality of anatomical patterns resulting in a plurality of perturbed anatomical patterns;

performing, via an encoder-decoder network having skip connections in between and a classification head at an end of the encoder, a self-classification operation on the plurality of perturbed anatomical patterns by formulating a multi-class classification task that discriminates among the plurality of perturbed anatomical patterns based on their respective label to learn anatomical pattern semantics from the plurality of perturbed anatomical patterns and generate a corresponding plurality of latent representations;

performing a self-restoration operation that receives the plurality of latent representation and recovers the corresponding plurality of anatomical patterns from the plurality of perturbed anatomical patterns to learn image representation from multiple perspectives and encode anatomical diversity in the plurality of anatomical patterns.

2 . The system of claim 1 , wherein transforming each of the 2D patches or 3D cubes to generate transformed 2D patches or transformed 3D cubes comprises applying one or more of the following transformations:

a non-linear transformation;

a local-shuffling transformation;

an out-painting transformation; and

an in-painting transformation.

3 . The system of claim 1 , wherein each of the assigned labels carry local information defining anatomical features selected from a group consisting of:

anterior ribs 2 through 4;

anterior ribs 1 through 3;

right pulmonary artery; and

Left Ventricle (LV).

4 . The system of claim 1 :

wherein the system further comprises an auto-encoder network;

wherein a classification branch of the auto-encoder network encodes the plurality of anatomical patterns into a latent space followed by a sequence of fully-connected (fc) layers; and

wherein the classification branch predicts the label assigned to each of the plurality of anatomical patterns.

5 . The system of claim 1 , wherein the classification branch classifies the plurality of anatomical patterns by applying a categorical cross-entropy loss function.

6 . The system of claim 1 :

wherein the system further comprises an auto-encoder network;

wherein a restoration branch of the auto-encoder network encodes a transformed anatomical pattern from the transformed 2D patches or transformed 3D cubes into a latent space; and

wherein the restoration branch decodes the transformed 2D patches or transformed 3D cubes to an original resolution from the latent space to recover each of the plurality of anatomical patterns from a corresponding transformed anatomical pattern.

7 . A non-transitory computer-readable storage media having instructions stored thereupon that, when executed by a system having at least a processor and a memory therein, cause the system to perform operations including:

receiving a plurality of unlabeled medical images;

wherein the system is specially configured to perform self-supervised learning for an artificial intelligence (AI) model in the absence of manually labeling the plurality of unlabeled images, by performing the following operations:

selecting a subset of the plurality of unlabeled medical images, corresponding to similar patients based on deep latent features therein;

extracting 2D patches or 3D cubes each representing an appearance of an anatomical pattern reoccurring at fixed coordinates across each of the selected subset of the plurality of unlabeled medical images, and assigning one of a plurality of labels to each of the 2D patches or 3D cubes based on their fixed coordinates;

transforming each of the 2D patches or 3D cubes to generate transformed 2D patches or transformed 3D cubes, respectively

by perturbing each of the plurality of anatomical patterns resulting in a plurality of perturbed anatomical patterns;

performing, via an encoder-decoder network having skip connections in between and a classification head at an end of the encoder, a self-classification operation on the plurality of perturbed anatomical patterns by formulating a multi-class classification task that discriminates among the plurality of perturbed anatomical patterns based on their respective label to learn anatomical pattern semantics from the plurality of perturbed anatomical patterns and generate a corresponding plurality of latent representations;

performing a self-restoration operation that receives the plurality of latent representations and recovers the corresponding plurality of anatomical patterns from the plurality of perturbed anatomical patterns to learn image representation from multiple perspectives and encode anatomical diversity in the plurality of anatomical patterns.

8 . A method performed by a system having at least a processor and a memory therein to execute instructions, wherein the method comprises:

receiving a plurality of unlabeled medical images;

wherein the system is specially configured to perform self-supervised learning for an artificial intelligence (AI) model in the absence of manually labeling the plurality of unlabeled medical images, by performing the following operations:

selecting a subset of the plurality of unlabeled medical images, corresponding to similar patients based on deep latent features therein;

extracting two-dimensional (2D) patches or three-dimensional (3D) cubes each representing an appearance of an anatomical pattern reoccurring at fixed coordinates across each of the selected subset of the plurality of unlabeled medical images, and assigning one of a plurality of labels to each of the 2D patches or 3D cubes based on their fixed coordinates;

transforming each of the 2D patches or 3D cubes to generate transformed 2D patches or transformed 3D cubes, respectively

by perturbing each of the plurality of anatomical patterns resulting in a plurality of perturbed anatomical patterns;

performing, via an encoder-decoder network having skip connections in between and a classification head at an end of the encoder, a self-classification operation on the plurality of perturbed anatomical patterns by formulating a multi-class classification task that discriminates among the plurality of perturbed anatomical patterns based on their respective label to learn anatomical pattern semantics from the plurality of perturbed anatomical patterns and generate a corresponding plurality of latent representations;

performing a self-restoration operation that receives the plurality of latent representations and recovers the corresponding plurality of anatomical patterns from the plurality of perturbed anatomical patterns to learn image representation from multiple perspectives and encode anatomical diversity in the plurality of anatomical patterns.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Feb 22, 2022
From: HAGHIGHI, FATEMEH; HOSSEINZADEH TAHER, MOHAMMAD REZA; ZHOU, ZONGWEI; LIANG, JIANMING
To: ARIZONA BOARD OF REGENTS ON BEHALF OF ARIZONA STATE UNIVERSITY
Reel/Frame 059066/0163 →
Continuity (5)
Continuation In Part 17246032 · Apr 30, 2021
Provisional Application 63151567 · Feb 19, 2021
Provisional Application 63110265 · Nov 5, 2020
Provisional Application 63018335 · Apr 30, 2020
Related Publication 20220309811A1 · Sep 29, 2022
References Cited (88)
US 8204842B1 · Zhang et al. · 2012 [cited by applicant]
US 9811765B2 · Wang et al. · 2017 [cited by applicant]
US 20180196873A1 · Yerebakan et al. · 2018 [cited by applicant]
US 20190057774A1 · Velez et al. · 2019 [cited by applicant]
US 20200069973A1 · Lou · 2020 [cited by examiner]
US 20210319556A1 · Chauhan · 2021 [cited by examiner]
US 20210343014A1 · Haghighi et al. · 2021 [cited by applicant]
US 20220309811A1 · Haghighi et al. · 2022 [cited by applicant]
Chen et al., Self-supervised learning for medical image analysis using image context restoration, Jul. 2018, IEEE (Year: 2018). [cited by examiner]
Notice of Allowance for U.S. Appl. No. 17/246,032, mailed Sep. 30, 2024, 16 pages. [cited by applicant]
Notice of Allowance for U.S. Appl. No. 17/246,032, mailed Nov. 25, 2024, 12 pages. [cited by applicant]
Raghu, M., et al., “Transfusion: Understanding Transfer Learning for Medical Imaging,” arXiv preprint arXiv:1902.07208, 2019, 22 pages. [cited by applicant]
Ross, T. et al., “Exploiting the potential of unlabeled endoscopic video data with self-supervised learning,” International Journal of Computer Assisted Radiology and Surgery 13, 2018, pp. 925-933. [cited by applicant]
Sabokrou, M. et al., “Self-Supervised Representation Learning via Neighborhood-Relational Encoding,” In Proceedings of IEEE/CVF International Conference on Computer Vision, 2019, pp. 8010-8019. [cited by applicant]
Setio, A.A.A. et al., “Validation, comparison, and combination of algorithms for automatic detection of pulmonary nodules in computed tomography images: the LUNA16 challenge,” Medical Image Analysis 42, 2017, pp. 1-13, … [cited by applicant]
Simonyan, K. et al., “Very Deep Convolutional Networks for Large-Scale Image Recognition,” arXiv preprint arXiv:1409.1556, 2014. [cited by applicant]
Simpson, A.L. et al., “A large annotated medical image dataset for the development and evaluation of segmentation algorithms,” arXiv preprint arXiv:1902.09063, 2019, 15 pages. [cited by applicant]
Sivic, J. et al., “Video google: a text retrieval approach to object matching in videos,” Proceedings ninth IEEE international conference on computer vision, 2003, pp. 1470-1477, IEEE. [cited by applicant]
Taher, M.R.H., et al., “CAiD: A Self-supervised Learning Framework for Empowering Instance Discrimination in Medical Imaging,” Proceedings of Machine Learning Research, 2022, 20 pages. [cited by applicant]
Tajbakhsh, N. et al., “Computer-aided Pulmonary Embolism Detection Using a Novel Vessel-aligned Multi-planar Image Representation and Convolutional Neural Networks,” Medical Image Computing and Computer-Assisted Interve… [cited by applicant]
Tajbakhsh, N. et al., “Convolutional Neural Networks for Medical Image Analysis: Full Training or Fine Tuning?,” IEEE Transactions on Medical Imaging, vol. 35, No. 5, May 2016, pp. 1299-1312. [cited by applicant]
Tajbakhsh, N. et al., “Embracing imperfect datasets: A review of deep learning solutions for medical image segmentation,” Medical Image Analysis, 63, 2020, p. 101693, 30 pages. [cited by applicant]
Van Gansbeke, W., et al, “Scan: Learning to Classify Images Without Labels,” European Conference on Computer Vision, 2020, pp. 268-285, Springer International Publishing. [cited by applicant]
Vincent, P. et al., “Extracting and Composing Robust Features with Denoising Autoencoders,” In Proceedings of the 25th International Conference on Machine Learning, 2008, pp. 1096-1103. [cited by applicant]
Wang, H. et al., “Comparison of machine learning methods for classifying mediastinal lymph node metastasis of non-small cell lung cancer from 18 F-FDG PET/CT images,” EJNMMI research 7(1), 2017, pp. 1-11. [cited by applicant]
Wang, X. et al., “ChestX-ray8: Hospital-scale Chest X-ray Database and Benchmarks on Weakly-Supervised Classification and Localization of Common Thorax Diseases,” Proceedings of the IEEE Conference on Computer Vision an… [cited by applicant]
Wang, Y. et al., “E2-Train: Training State-of-the-Art CNNs with Over 80% Less Energy” Advances in Neural Information Processing Systems, 32, 2019, 13 pages. [cited by applicant]
Wu, B. et al., “Joint Learning for Pulmonary Nodule Segmentation, Attributes and Malignancy Prediction,” In 2018 IEEE 15th International Symposium on Biomedical Imaging (ISBI 2018), 2018, pp. 1109-1113, IEEE. [cited by applicant]
Yan, X. et al., “ClusterFit: Improving Generalization of Visual Representations,” In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, 2020, pp. 6509-6518. [cited by applicant]
Yosinski, J. et al., “How transferable are features in deep neural networks?” Advances in Neural Information Processing Systems, 27 (NIPS 2014), arXiv preprint arXiv:1411.1792. [cited by applicant]
Yue-Hei Ng, J. et al., “Exploiting Local Features from Deep Networks for Image Retrieval,” In Proceedings of IEEE Conference on Computer Vision and Pattern Recognition Workshops, 2015, pp. 53-61. [cited by applicant]
Zhan, X. et al., “Online Deep Clustering for Unsupervised Representation Learning,” In Proceedings of IEEE/CVF Conference on Computer Vision and Pattern Recognition, 2020, pp. 6688-6697. [cited by applicant]
Zhang, L. et al., “AET vs. AED: Unsupervised Representation Learning by Auto-Encoding Transformations Rather than Data,” Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, 2019, pp. 2542-… [cited by applicant]
Zhang, R. et al., “Colorful Image Colorization,” In Computer Vision—ECCV 2016: 14th European Conference, Amsterdam, The Netherlands, Oct. 11-14, 2016, Proceedings, Part III 14, 2016, pp. 649-666, Springer International … [cited by applicant]
Zhou, Z. et al., “Fine-tuning Convolutional Neural Networks for Biomedical Image Analysis: Actively and Incrementally,” In 30th IEEE Conference on Computer Vision and Pattern Recognition, CVPR 2017, pp. 4761-4772. [cited by applicant]
Zhou, Z. et al., “Models genesis,” Medical Image Analysis, vol. 67, 2021, p. 101840, 23 pages. [cited by applicant]
Zhou, Z. et al., “Models Genesis: Generic Autodidactic Models for 3D Medical Image Analysis,” Medical Image Computing and Computer Assisted Intervention—MICCAI 2019: 22nd International Conference, Shenzhen, China, Oct. … [cited by applicant]
Zhuang, X. et al., “Self-supervised Feature Learning for 3d Medical Images by Playing a Rubik's Cube,” Medical Image Computing and Computer Assisted Intervention—MICCAI 2019: 22nd International Conference, Shenzhen, Chi… [cited by applicant]
“SIIM-ACR Pneumothorax Segmentation” (Online) (2019), https://www.kaggle.com/c/siim-acr-pneumothorax-segmentation/, 26 pages. [cited by applicant]
Advisory Action for U.S. Appl. No. 17/246,032, dated May 6, 2024, 3 pages. [cited by applicant]
Alex, V. et al., “Semisupervised learning using denoising autoencoders for brain lesion detection and segmentation,” Journal of Medical Imaging 4(4), 2017, 041311-041311. [cited by applicant]
Arandjelovic, R. et al., “NetVLAD: CNN architecture for weakly supervised place recognition,” In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, 2016, pp. 5297-5307. [cited by applicant]
Ardila, D. et al., “End-to-end lung cancer screening with three-dimensional deep learning on low-dose chest computed tomography,” Nature Medicine 25(6), 2019, pp. 954-961. [cited by applicant]
Armato III, S.G. et al., “The lung image database consortium (LIDC) and image database resource initiative (IDRI): a completed reference database of lung nodules on CT scans,” Medical Physics 38(2), 2011, pp. 915-931. [cited by applicant]
Bai, W. et al., “Self-Supervised Learning for Cardiac MR Image Segmentation by Anatomical Position Prediction,” In Medical Image Computing and Computer Assisted Intervention—MICCAI 2019: 22nd International Conference, S… [cited by applicant]
Bakas, S. et al., “Identifying the Best Machine Learning Algorithms for Brain Tumor Segmentation, Progression Assessment, and Overall Survival Prediction in the Brats Challenge,” arXiv preprint arXiv:1811.02629, 2018, 2… [cited by applicant]
Bengio, Y. “Learning Deep Architectures for AI”, Foundations and Trends in Machine Learning, vol. 2, No. 1, 2009. 35 pages. [cited by applicant]
Bilic, P. et al., “The Liver Tumor Segmentation Benchmark (LiTS),” Medical Image Analysis, 84, 102680, arXiv preprint arXiv:1901.04056, 2019. [cited by applicant]
Caron, M. et al., “Deep Clustering for Unsupervised Learning of Visual Features,” In Proceedings of the European Conference on Computer Vision (ECCV), 2018, pp. 132-149. [cited by applicant]
Caron, M. et al., “Unsupervised Pre-Training of Image Features on Non-Curated Data,” Proceedings of the IEEE International Conference on Computer Vision, 2019, pp. 2959-2968. [cited by applicant]
Carreira, J. et al., “Quo Vadis, Action Recognition? A New Model and the Kinetics Dataset,” In 2017 IEEE Conference on Computer Vision and Pattern Recognition (CVPR), (Jul. 2017), pp. 6299-6308. [cited by applicant]
Chaitanya, K. et al., “Contrastive learning of global and local features for medical image segmentation with limited annotations,” Advances in Neural Information Processing Systems vol. 33, 2020, pp. 12546-12558. [cited by applicant]
Chen, L. et al., “Self-supervised learning for medical image analysis using image context restoration,” Medical Image Analysis, vol. 58, 2019, 12 pages, https://doi.org/10.1016/j.media.2019.101539. [cited by applicant]
Chen, S. et al., Med3d: Transfer Learning for 3d Medical Image Analysis. arXiv preprint arXiv:1904.00625, 2019, 12 pages. [cited by applicant]
Deshpande, A. et al., “Learning Large-Scale Automatic Image Colorization,” In Proceedings of the IEEE International Conference on Computer Vision, 2015, pp. 567-575. [cited by applicant]
Doersch, C. et al., “Unsupervised Visual Representation Learning by Context Prediction,” Proceedings of the IEEE International Conference on Computer Vision, 2015, pp. 1422-1430, https://doi.org/10.1109/ICCV.2015.167. [cited by applicant]
Fan, Z., et al., “A Generic Unified Deep Model for Learning from Multiple Tasks,” 34 pages. [cited by applicant]
Feng, Z. et al., “Self-Supervised Representation Learning by Rotation Feature Decoupling,” In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, 2019, pp. 10364-10374. [cited by applicant]
Final Office Action for U.S. Appl. No. 17/246,032 dated Feb. 23, 2024, 15 pages. [cited by applicant]
Gibson, E. et al., “Automatic Multi-Organ Segmentation on Abdominal CT with Dense V-Networks,” IEEE Transactions on Medical Imaging, vol. 37, No. 8, Aug. 2018, pp. 1822-1834. [cited by applicant]
Gibson, E. et al., “Niftynet: a deep-learning platform for medical imaging,” Computer Methods and Programs in Biomedicine 158, 2018, pp. 113-122. [cited by applicant]
Gidaris, S. et al., “Learning Representations by Predicting Bags of Visual Words,” In Proceedings of IEEE/CVF Conference on Computer Vision and Pattern Recognition, 2020, pp. 6926-6936. [cited by applicant]
Gidaris, S. et al., “Unsupervised Representation Learning by Predicting Image Rotations,” In ICLR 2018, 17 pages. [cited by applicant]
Gong, Y. et al., “Multi-Scale Orderless Pooling of Deep Convolutional Activation Features,” In Computer Vision—ECCV 2014: 13th European Conference, Zurich, Switzerland, Sep. 6-12, 2014, Proceedings, Part VII 13, 2014, p… [cited by applicant]
Goyal, P. et al., “Scaling and Benchmarking Self-Supervised Visual Representation Learning,” Proceedings of the IEEE/CVF International Conference on Computer Vision, 2019, pp. 6390-6399. [cited by applicant]
Guo, Z., et al., “Discriminative, Restorative, and Adversarial Learning: Stepwise Incremental Pretraining,” 15 pages. [cited by applicant]
Haghighi, F. et al., “Learning semantics-enriched representation via Self-discovery, Self-classification, and Self-restoration,” Medical Image Computing and Computer Assisted Intervention—MICCAI 2020: 23rd International… [cited by applicant]
Haghighi, F. et al., “Transferable visual words: Exploiting the Semantics of Anatomical Patterns for Self-supervised Learning,” IEEE Transactions on Medical Imaging, 40(10), 2021, pp. 2857-2868, https://doi.org/10.1109/… [cited by applicant]
Haghighi, F., et al., “DiRA: Discriminative, Restorative, and Adversarial Learning for Self-supervised Medical Image Analysis,” In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (2022)… [cited by applicant]
Haghighi, F., et al., Transferable Visual Worlds, IEEE Transactions on Medical Imaging, 2020, 21 pages. [cited by applicant]
Hendrycks, D. et al., “Using Self-Supervised Learning Can Improve Model Robustness and Uncertainty,” Advances in Neural Information Processing Systems, 32, 2019,12 pages. [cited by applicant]
Huh, M. et al., “What makes ImageNet good for transfer learning?,” arXiv preprint arXiv:1608.08614 (2016), 10 pages. [cited by applicant]
Isensee, F. et al., “Automated Design of Deep Learning Methods for Biomedical Image Segmentation,” arXiv preprint arXiv:1904.08128, 2020, 55 pages. [cited by applicant]
Jenni, S. et al., “Steering Self-Supervised Feature Learning Beyond Local Pixel Statistics,” Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, 2020, pp. 6408-6417. [cited by applicant]
Johnson, T.B. et al., “Training Deep Models Faster with Robust, Approximate Importance Sampling,” Advances in Neural Information Processing Systems 31, 2018, pp. 7265-7275. [cited by applicant]
Kim, D. et al., “Learning Image Representations by Completing Damaged Jigsaw Puzzles,” In 2018 IEEE Winter Conference on Applications of Computer Vision (WACV), 2018, pp. 793-802, IEEE. [cited by applicant]
Kornblith, S. et al., “Do Better ImageNet Models Transfer Better?” In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, 2019, pp. 2656-2666. [cited by applicant]
Larsson, G. et al., “Colorization as a Proxy Task for Visual Understanding,” In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, 2017, pp. 6874-6883. [cited by applicant]
Larsson, G. et al., “Learning Representations for Automatic Colorization,” In Computer Vision—ECCV 2016: 14th European Conference, Amsterdam, The Netherlands, Oct. 11-14, 2016, Proceedings, Part IV, 14, 2016, pp. 577-59… [cited by applicant]
Lowe, D. G., “Distinctive Image Features from Scale-Invariant Keypoints,” International Journal of Computer Vision, 60(2), 2004, pp. 91-110. [cited by applicant]
Minderer, M. et al., “Automatic Shortcut Removal for Self-Supervised Representation Learning,” In International Conference on Machine Learning. PMLR, 2020, pp. 6927-6937. [cited by applicant]
Mundhenk, T. N. et al., “Improvements to Context Based Self-Supervised Learning.” 2018 IEEE/CVF Conference on Computer Vision and Pattern Recognition, 2018, pp. 9339-9348, https://doi.org/10.1109/CVPR.2018.00973. [cited by applicant]
Newell, A. et al., “How Useful is Self-Supervised Pretraining for Visual Tasks?” In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, 2020, pp. 7343-7352. [cited by applicant]
Neyshabur, B. et al., “What is being transferred in transfer learning?” Advances in Neural Information Processing Systems 33, 2020, pp. 512-523. [cited by applicant]
Non-final Office Action for U.S. Appl. No. 17/246,032, dated Jul. 5, 2023, 15 pages. [cited by applicant]
Noroozi, M. et al., “Unsupervised Learning of Visual Representations by Solving Jigsaw Puzzles,” European Conference on Computer Vision, 2016, pp. 69-84, Springer International Publishing, arXiv preprint arXiv:1603.0924… [cited by applicant]
Notice of Allowance for U.S. Appl. No. 17/246,032, dated Aug. 1, 2024, 12 pages. [cited by applicant]
Pathak, D. et al., “Context Encoders: Feature Learning by Inpainting,” 2016 Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR), pp. 2536-2544, https://doi.org/10.1109/CVPR.2016.278. [cited by applicant]