IP Library Granted Patent US 12,205,294
Granted Patent B2
US 12,205,294 · App. 17/688,575 · Granted Jan 21, 2025

Methods and systems for authentication of a physical document

Inventors: Daniele Pizzocchero (London, GB); Jimmy Moore (London, GB); Zhiyuan Shi (London, GB); Christos Sagonas (London, GB); Mohan Mahadevan (London, GB); Yuanwei Li (London, GB)
Assignee: Onfido Ltd.
G06T7/11G06F2218/12G06T2207/10004G06T2207/10016G06T2207/20081G06T2207/20084G06T2207/30176
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,205,294
App. No.
17/688,575
Granted
Jan 21, 2025
Kind
B2
Abstract

Described herein are computerized methods and systems for authentication of a physical document. An image capture device coupled to a mobile device captures images of a physical document, during which the mobile device adjusts operational parameters of the image capture device, resulting in a sequence of images captured using different capture settings. The mobile device partitions the sequence of images into subsets of images, wherein each subset comprises images with a similar alignment of the physical document and captured using the same capture settings. The mobile device processes the subsets of images to identify a region of interest in each image. The mobile device generates a representation of the identified region of interest using the processed images, generates an authentication score for the document using the representation of the identified region of interest, and determines whether the physical document is authentic based upon the authentication score.

Claims (64)

1. A system for authentication of a physical document, the system comprising a mobile computing device coupled to an image capture device, the mobile computing device configured to:

capture, using the image capture device, images of a physical document in a scene, during which the mobile computing device adjusts one or more operational parameters of the image capture device, resulting in a sequence of images captured using different capture settings;

partition the sequence of images into one or more subsets of images, wherein each subset comprises images with a similar alignment of the physical document and captured using the same capture settings;

process the subsets of images to identify a region of interest in each image, the region of interest corresponding to one or more visual features of the physical document;

generate a representation of the identified region of interest using the processed images;

generate an authentication score for the document using the representation of the identified region of interest; and

determine whether the physical document is authentic based upon the authentication score.

2. The system of claim 1 , wherein the one or more operational parameters comprise one or more of shutter speed, ISO speed, gain and offset, aperture, flash intensity, flash duration, or light balance.

3. The system of claim 1 , wherein the physical document is stationary during capture of the images by the mobile computing device.

4. The system of claim 1 , wherein the physical document remains in a stationary position relative to the image capture device during capture of the images by the mobile computing device.

5. The system of claim 1 , wherein prior to capturing a first image of the physical document in the scene, the mobile computing device generates baseline operational parameters of the image capture device based upon one or more imaging conditions associated with the physical document.

6. The system of claim 5 , wherein adjusting one or more operational parameters of the image capture device comprises adjusting the baseline operational parameters between capturing each image in the sequence of images.

7. The system of claim 6 , wherein adjusting the baseline operational parameters between capturing each image comprises receiving operational parameters used for the previous image and using the received operational parameters to adjust the baseline operational parameters as part of a dynamic feedback loop.

8. The system of claim 1 , wherein the mobile computing device preprocesses the sequence of images received from the image capture device prior to partitioning the sequence of images.

9. The system of claim 8 , wherein preprocessing the sequence of images comprises one or more of: assessing video quality metrics for the entire sequence of images, detecting a location of the physical document in each image of the sequence of images, and determining one or more quality metrics for each image in the sequence of images.

10. The system of claim 9 , wherein the video quality metrics comprise a length of the sequence of images, a frames-per-second (FPS) value associated with the sequence of images, and an image resolution associated with the sequence of images.

11. The system of claim 9 , wherein the one or more quality metrics comprise (i) global image quality metrics including one or more of: glare, blur, white balance, or sensor noise characteristics, (ii) local image quality metrics including one or more of: blur, sharpness, text region confidence, character confidence, or edge detection, or (iii) both the global image quality metrics and the local image quality metrics.

12. The system of claim 11 , wherein the sensor noise characteristics comprise one or more of: blooming, readout noise, or custom calibration variations.

13. The system of claim 1 , wherein processing the selected images to identify a region of interest in each image comprises normalizing an image signal of each image.

14. The system of claim 13 , wherein normalizing an image signal of each image comprises amplifying the image signal associated with a region of interest on the physical document and reducing the image signal associated with a background of the physical document.

15. The system of claim 1 , wherein generating a representation of the identified region of interest comprises executing one or more of a robust principal component analysis (PCA) algorithm or a learned alternative mapping on the image to reconstruct the region of interest.

16. The system of claim 1 , wherein generating an authentication score for the document using the reconstructed region of interest comprises executing one or more machine learning classification models using one or more features of the reconstructed region of interest as input to generate a classification value for the document.

17. The system of claim 16 , wherein the one or more machine learning classification models comprise one or more of: deep learning models, Random Forest algorithms, Support Vector Machines, neural networks, or ensembles thereof.

18. The system of claim 16 , wherein the classification value comprises at least one of a probability that the document is authentic, a confidence score metric that indicates whether the document is authentic, or a similarity metric that indicates whether the document is authentic.

19. The system of claim 17 , wherein at least one of the one or more machine learning classification models is a convolutional neural network.

20. The system of claim 17 , wherein the one or more machine learning classification models is an ensemble classifier comprised of a plurality of convolutional neural networks.

21. The system of claim 16 , wherein one or more interpretable methods are used to validate the classification value.

22. The system of claim 21 , wherein the one or more interpretable methods comprise occlusion of at least a portion of the document, perturbation of at least a portion of the document, or analysis of a heatmap of at least a portion of the document.

23. The system of claim 22 , wherein an output of the one or more interpretable methods comprises an identification of the reconstructed region of interest that represents proof of the document being genuine or fraudulent.

24. The system of claim 16 , wherein the one or more machine learning classification models are trained using a plurality of genuine documents, a plurality of fraudulent documents, or both.

25. The system of claim 24 , wherein the classification value generated by the one or more machine learning classification models is a measure of similarity between one or more of the plurality of genuine documents, one or more of the plurality of fraudulent documents, or both.

26. The system of claim 1 , wherein the images of the physical document comprise one of: images of a front side of the physical document or images of a back side of the physical document.

27. A computerized method of authentication of a physical document, the method comprising:

capturing, using an image capture device coupled to a mobile computing device, images of a physical document in a scene, during which the mobile computing device adjusts one or more operational parameters of the image capture device, resulting in a sequence of images captured using different capture settings;

partitioning, by the mobile computing device, the sequence of images into one or more subsets of images, wherein each subset comprises images with a similar alignment of the physical document and captured using the same capture settings;

processing, by the mobile computing device, the subsets of images to identify a region of interest in each image, the region of interest corresponding to one or more visual features of the physical document;

generating, by the mobile computing device, a representation of the identified region of interest using the processed images;

generating, by the mobile computing device, an authentication score for the document using the representation of the identified region of interest; and

determining, by the mobile computing device, whether the physical document is authentic based upon the authentication score.

28. The method of claim 27 , wherein the one or more operational parameters comprise one or more of shutter speed, ISO speed, gain and offset, aperture, flash intensity, flash duration, or light balance.

29. The method of claim 27 , wherein the physical document is stationary during capture of the images by the mobile computing device.

30. The method of claim 27 , wherein the physical document remains in a stationary position relative to the image capture device during capture of the images by the mobile computing device.

31. The method of claim 27 , wherein prior to capturing a first image of the physical document in the scene, the mobile computing device generates baseline operational parameters of the image capture device based upon one or more imaging conditions associated with the physical document.

32. The method of claim 31 , wherein adjusting one or more operational parameters of the image capture device comprises adjusting the baseline operational parameters between capturing each image in the sequence of images.

33. The method of claim 32 , wherein adjusting the baseline operational parameters between capturing each image comprises receiving operational parameters used for the previous image and using the received operational parameters to adjust the baseline operational parameters as part of a dynamic feedback loop.

34. The method of claim 27 , wherein the mobile computing device preprocesses the sequence of images received from the image capture device prior to partitioning the sequence of images.

35. The method of claim 34 , wherein preprocessing the sequence of images comprises one or more of: assessing video quality metrics for the entire sequence of images, detecting a location of the physical document in each image of the sequence of images, and determining one or more quality metrics for each image in the sequence of images.

36. The method of claim 35 , wherein the video quality metrics comprise a length of the sequence of images, a frames-per-second (FPS) value associated with the sequence of images, and an image resolution associated with the sequence of images.

37. The method of claim 35 , wherein the one or more quality metrics comprise (i) global image quality metrics including one or more of: glare, blur, white balance, or sensor noise characteristics, (ii) local image quality metrics including one or more of: blur, sharpness, text region confidence, character confidence, or edge detection, or (iii) both the global image quality metrics and the local image quality metrics.

38. The method of claim 37 , wherein the sensor noise characteristics comprise one or more of: blooming, readout noise, or custom calibration variations.

39. The method of claim 27 , wherein processing the selected images to identify a region of interest in each image comprises normalizing an image signal of each image.

40. The method of claim 39 , wherein normalizing an image signal of each image comprises amplifying the image signal associated with a region of interest on the physical document and reducing the image signal associated with a background of the physical document.

41. The method of claim 27 , wherein generating a representation of the identified region of interest comprises executing one or more of a robust principal component analysis (PCA) algorithm or a learned alternative mapping on the image to reconstruct the region of interest.

42. The method of claim 27 , wherein generating an authentication score for the document using the reconstructed region of interest comprises executing one or more machine learning classification models using one or more features of the reconstructed region of interest as input to generate a classification value for the document.

43. The method of claim 42 , wherein the one or more machine learning classification models comprise one or more of: deep learning models, Random Forest algorithms, Support Vector Machines, neural networks, or ensembles thereof.

44. The method of claim 42 , wherein the classification value comprises at least one of a probability that the document is authentic, a confidence score metric that indicates whether the document is authentic, or a similarity metric that indicates whether the document is authentic.

45. The method of claim 42 , wherein at least one of the one or more machine learning classification models is a convolutional neural network.

46. The method of claim 42 , wherein the one or more machine learning classification models is an ensemble classifier comprised of a plurality of convolutional neural networks.

47. The method of claim 42 , wherein one or more interpretable methods are used to validate the classification value.

48. The method of claim 47 , wherein the one or more interpretable methods comprise occlusion of at least a portion of the document, perturbation of at least a portion of the document, or analysis of a heatmap of at least a portion of the document.

49. The method of claim 48 , wherein an output of the one or more interpretable methods comprises an identification of the reconstructed region of interest that represents proof of the document being genuine or fraudulent.

50. The method of claim 42 , wherein the one or more machine learning classification models are trained using a plurality of genuine documents, a plurality of fraudulent documents, or both.

51. The method of claim 50 , wherein the classification value generated by the one or more machine learning classification models is a measure of similarity between one or more of the plurality of genuine documents, one or more of the plurality of fraudulent documents, or both.

52. The method of claim 27 , wherein the images of the physical document comprise one of: images of a front side of the physical document or images of a back side of the physical document.

Assignments (5)
SECURITY INTEREST Recorded Jul 25, 2024
From: ONFIDO LTD
To: BMO BANK N.A., AS COLLATERAL AGENT
Reel/Frame 068079/0801 →
RELEASE OF SECURITY INTEREST Recorded Apr 9, 2024
From: HSBC INNOVATION BANK LIMITED (F/K/A SILICON VALLEY BANK UK LIMITED)
To: ONFIDO LTD
Reel/Frame 067053/0607 →
AMENDED AND RESTATED INTELLECTUAL PROPERTY SECURITY AGREEMENT Recorded Dec 21, 2022
From: ONFIDO LTD
To: SILICON VALLEY BANK UK LIMITED
Reel/Frame 062200/0655 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 2, 2022
From: PIZZOCCHERO, DANIELE; MOORE, JIMMY; SHI, ZHIYUAN; SAGONAS, CHRISTOS; MAHADEVAN, MOHAN; LI, YUANWEI
To: ONFIDO LTD.
Reel/Frame 061625/0315 →
SUPPLEMENT NO. 1 TO INTELLECTUAL PROPERTY SECURITY AGREEMENT Recorded Jul 8, 2022
From: ONFIDO LTD
To: SILICON VALLEY BANK
Reel/Frame 060613/0293 →
Continuity (1)
Related Publication 20230281821A1 · Sep 7, 2023
References Cited (97)
US 7590344B2 · Petschnigg · 2009 [cited by applicant]
US 7672475B2 · O'Doherty et al. · 2010 [cited by applicant]
US 8025239B2 · Labrec et al. · 2011 [cited by applicant]
US 8156115B1 · Erol · 2012 [cited by examiner]
US 8627939B1 · Jones · 2014 [cited by examiner]
US 8786767B2 · Rihn et al. · 2014 [cited by applicant]
US 9058535B2 · Guigan · 2015 [cited by applicant]
US 9081988B2 · Dolev · 2015 [cited by applicant]
US 9153005B2 · Tremolada et al. · 2015 [cited by applicant]
US 9418282B2 · Rutz et al. · 2016 [cited by applicant]
US 9779296B1 · Ma · 2017 [cited by examiner]
US 10019626B2 · Schilling et al. · 2018 [cited by applicant]
US 10019627B2 · Kutter et al. · 2018 [cited by applicant]
US 10127454B2 · Pau et al. · 2018 [cited by applicant]
US 10140511B2 · Macciola · 2018 [cited by examiner]
US 10354142B2 · Arlazarov et al. · 2019 [cited by applicant]
US 10402448B2 · Filgueiras de Araujo et al. · 2019 [cited by applicant]
US 10477087B2 · Rivard et al. · 2019 [cited by applicant]
US 10621426B2 · Gaubatz et al. · 2020 [cited by applicant]
US 10789463B2 · Kutter et al. · 2020 [cited by applicant]
US 10872488B2 · Azanza Ladron et al. · 2020 [cited by applicant]
US 11144752B1 · Castelblanco Cruz · 2021 [cited by examiner]
US 11527087B1 · Geusz · 2022 [cited by examiner]
US 20070041628A1 · Fan · 2007 [cited by applicant]
US 20080158258A1 · Lazarus et al. · 2008 [cited by applicant]
US 20120226600A1 · Dolev · 2012 [cited by examiner]
US 20130185618A1 · Macciola · 2013 [cited by examiner]
US 20130262333A1 · Wicker · 2013 [cited by examiner]
US 20140032406A1 · Roach · 2014 [cited by examiner]
US 20140037196A1 · Blair · 2014 [cited by applicant]
US 20140055824A1 · Tremolada et al. · 2014 [cited by applicant]
US 20140105449A1 · Caton · 2014 [cited by examiner]
US 20140355069A1 · Caton · 2014 [cited by examiner]
US 20150093018A1 · Macciola · 2015 [cited by examiner]
US 20150302421A1 · Caton · 2015 [cited by examiner]
US 20150324390A1 · Macciola · 2015 [cited by examiner]
US 20160307035A1 · Schilling et al. · 2016 [cited by applicant]
US 20160350592A1 · Ma · 2016 [cited by examiner]
US 20170132465A1 · Kutter et al. · 2017 [cited by applicant]
US 20170132866A1 · Kuklinski · 2017 [cited by examiner]
US 20170200247A1 · Kuklinski · 2017 [cited by examiner]
US 20190087942A1 · Ma · 2019 [cited by examiner]
US 20190164010A1 · Ma · 2019 [cited by examiner]
US 20190213408A1 · Cali et al. · 2019 [cited by applicant]
US 20190220660A1 · Cali et al. · 2019 [cited by applicant]
US 20190372968A1 · Balogh · 2019 [cited by examiner]
US 20190392196A1 · Sagonas et al. · 2019 [cited by applicant]
US 20200051059A1 · Filler · 2020 [cited by examiner]
US 20200210796A1 · Hsu · 2020 [cited by examiner]
US 20200387700A1 · Wu · 2020 [cited by examiner]
US 20200394763A1 · Ma · 2020 [cited by examiner]
US 20210117520A1 · Ross · 2021 [cited by examiner]
US 20210158036A1 · Huber, Jr. · 2021 [cited by examiner]
US 20210160687A1 · Ross et al. · 2021 [cited by applicant]
US 20210200855A1 · Guigan · 2021 [cited by applicant]
US 20220343483A1 · Desai · 2022 [cited by examiner]
CN 112434727A · 2021 [cited by applicant]
DE 102018102015A1 · 2019 [cited by applicant]
EP 3526781 · 2018 [cited by applicant]
FR 2977349A3 · 2013 [cited by applicant]
WO 2021044082A1 · 2021 [cited by applicant]
A. Hartl et al., “Efficient Verification of Holograms Using Mobile Augmented Reality,” IEEE Transactions on Visualization and Computer Graphics, Nov. 2015, 10 pages. [cited by applicant]
J. Redmon et al., “You Only Look Once: Unified, Real-Time Object Detection,” arXiv:1506.02640v5 [cs.CV] May 9, 2016, available at https://arxiv.org/pdf/1506.02640.pdf, 10 pages. [cited by applicant]
W. Liu et al., “SSD: Single Shot MultiBox Detector,” arXiv:1512.02325v5 [cs.CV] Dec. 29, 2016, available at https://arxiv.org/pdf/1512.02325.pdf, 17 pages. [cited by applicant]
S. Ren et al., “Faster R-CNN: Toward Real-Time Object Detection with Region Proposal Networks,” arXiv:1506.01497v1 [cs.CV] Jun. 4, 2015, available at https://arxiv.org/pdf/1506.01497v1.pdf, 10 pages. [cited by applicant]
A. G. Howard et al., “MobileNets: Efficient Convolutional Neural Networks for Mobile Vision Applications,” arXiv:1704.0486v1 [cs.CV] Apr. 17, 2017, available at https://arxiv.org/pdf/1704.04861.pdf, 9 pages. [cited by applicant]
N. Wojke et al., “Simple Online and Realtime Tracking with a Deep Association Metric,” arXiv:1703.07402v1 [cs.CV] Mar. 21, 2017, available at https://arxiv.org/pdf/1703.07402.pdf, 5 pages. [cited by applicant]
Hartl et al “AR-Based Hologram Detection on Security Documents Using a Mobile Phone” 18 [cited by applicant]
Soukup et al “Mobile Hologram Verification with Deep Learning” 15 [cited by applicant]
P. Bergmann et al., “Tracking without bells and whistles,” arXiv:1903.05625v3 [cs.CV] Aug. 17, 2019, available at https://arxiv.org/pdf/1903.05625.pdf, 16 pages. [cited by applicant]
G. Ciaparrone et al., “Deep Learning in Video Multi-Object Tracking: A Survey,” arXiv:1907.12740v4 [cs.CV] Nov. 19, 2019, available at https://arxiv.org/pdf/1907.12740.pdf, 42 pages. [cited by applicant]
X. Zhou et al., “Tracking Objects as Points,” arXiv:2004.01177v2 [cs.CV] Aug. 21, 2020, available at https://arxiv.org/pdf/2004.01177.pdf, 22 pages. [cited by applicant]
G. Levi and T. Hassner, “LATCH: Learned Arrangements of Three Patch Codes,” 2016 IEEE Winter Conference on Applications of Computer Vision (WACV), Lake Placid, NY, Mar. 7-10, 2016, available at https://talhassner.github… [cited by applicant]
K. Simonyan and A. Zisserman, “Very deep convolutional networks for large-scale image recognition,” arXiv:1409.1556v6 [cs.CV], Apr. 10, 2015, available at https://arxiv.org/pdf/1409.1556v6.pdf, 14 pages. [cited by applicant]
B. Lakshminarayanan et al., “Simple and Scalable Predictive Uncertainty Estimation using Deep Ensembles,” arXiv:1612.01474v3 [stat.ML] Nov. 4, 2017, available at https://arxiv.org/pdf/1612.01474v3.pdf, 15 pages. [cited by applicant]
R. Rahaman and A.H. Thiery, “Uncertainty Quantification and Deep Ensembles,” arXiv:2007.08792v4 [stat.ML] Nov. 2, 2021, available at https://arxiv.org/pdf/2007.08792.pdf, 16 pages. [cited by applicant]
R. Selvaraju et al., “Grad-CAM: Visual Explanations from Deep Networks via Gradient-based Localization,” arXiv:1610.02391 [cs.CV] Dec. 3, 2019, available at https://arxiv.org/pdf/1610.02391.pdf, 23 pages. [cited by applicant]
G. Montavon et al., “Layer-Wise Relevance Propagation: An Overview,” Explainable AI: Interpreting, Explaining and Visualizing Deep Learning, Lecture Notes in Computer Science, vol. 11700, pp. 193-209, Sep. 10, 2019, Spr… [cited by applicant]
M. Sundararajan et al., “Axiomatic Attribution for Deep Networks,” arXiv:1703.01365v2 [cs.LG] Jun. 13, 2017, available at https://arxiv.org/pdf/1703.01365.pdf, 11 pages. [cited by applicant]
P. Kindermans et al., “Learning How to Explain Neural Networks: PatternNet and PatternAttribution,” arXiv:1705.05598v2 [stat.ML] Oct. 24, 2017, available at https://arxiv.org/pdf/1705.05598.pdf, 12 pages. [cited by applicant]
K. Simonyan et al., “Deep Inside Convolutional Networks: Visualising Image Classification Models and Saliency Maps,” arXiv:1312.6034v2 [cs.CV] Apr. 19, 2014, available at https://arxiv.org/pdf/1312.6034.pdf, 8 pages. [cited by applicant]
B. Gao and M.W. Spratling, “Robust Template Matching via Hierarchical Convolutional Features from a Shape Based CNN,” arXiv:2007.15817v3 [cs.CV] May 7, 2021, available at https://arxiv.org/pdf/2007.15817.pdf, 11 pages. [cited by applicant]
R. Polvara et al., “Toward End-to-End Control for UAV Autonomous Landing via Deep Reinforcement Learning,” 2018 International Conference on Unmanned Aircraft Systems (ICUAS), Jun. 12-15, 2018, DOI: 10.1109/ICUAS.2018.84… [cited by applicant]
Y. Yoon et al., “Online Multiple Pedestrians Tracking using Deep Temporal Appearance Matching Association,” arXiv:1907.00831v4 [cs.CV] Oct. 9, 2020, available at https://arxiv.org/pdf/1907.00831.pdf, 28 pages. [cited by applicant]
S. Mallick, “Object Tracking using OpenCV (C++ / Python),” Feb. 13, 2017, available at https://learnopencv.com/object-tracking-using-opencv-cpp-python/, 12 pages. [cited by applicant]
K. He et al., “Deep Residual Learning for Image Recognition,” arXiv:1512.03385v1 [cs.CV] Dec. 10, 2015, available at https://arxiv.org/pdf/1512.03385v1.pdf, 12 pages. [cited by applicant]
C. Szegedy et al., “Rethinking the Inception Architecture for Computer Vision,” arXiv:1500567v3 [cs.CV] Dec. 11, 2015, available at https://arxiv.org/pdf/1512.00567v3.pdf, 10 pages. [cited by applicant]
M. Tan & Q.V. Lee, “EfficientNet: Rethinking Model Scaling for Convolutional Neural Networks,” arXiv:1905.11946v5 [cs.LG] Sep. 11, 2020, available at https://arxiv.org/pdf/1905.11946.pdf, 11 pages. [cited by applicant]
C. Wang et al., “EfficientNet-eLite: Extremely Lightweight and Efficient CNN Models for Edge Devices by Network Candidate Search,” arXiv:2009.07409v1 [cs.CV] Sep. 16, 2020, available at https://arxiv.org/pdf/2009.07409v… [cited by applicant]
T. Zhang et al., “Spatial-Temporal Recurrent Neural Network for Emotion Recognition,” arXiv:1705.0451v1 [cs.CV] May 12, 2017, available at https://arxiv.org/pdf/1705.04515.pdf, 8 pages. [cited by applicant]
Y. Dong et al., “A Hybrid Spatial-temporal Deep Learning Architecture for Lane Detection,” arXiv:2110.04079 [cs.CV] Oct. 14, 2021, available at https://arxiv.org/ftp/arxiv/papers/2110/2110.04079.pdf, 18 pages. [cited by applicant]
G. Balakrishnan et al., “VoxelMorph: A Learning Framework for Deformable Medical Image Registration,” arXiv:1809.05231v3 [cs.CV] Sep. 1, 2019, available at https://arxiv.org/pdf/1809.05231.pdf, 16 pages. [cited by applicant]
I. Rocco et al., “Convolutional neural network architecture for geometric matching,” arXiv:1703.05593v2 [cs.CV] Apr. 13, 2017, available at https://arxiv.org/pdf/1703.05593.pdf, 15 pages. [cited by applicant]
M. Jadenberg et al., “Spatial Transformer Networks,” arXiv:1506.02025v3 [cs.CV] Feb. 4, 2016, available at https://arxiv.org/pdf/1506.02025.pdf, 15 pages. [cited by applicant]
P. H. Seo et al., “Attentive Semantic Alignment with Offset-Aware Correlation Kernels,” arXiv:1808.02128v2 [cs.CV] Oct. 26, 2018, available at https://arxiv.org/pdf/1808.02128.pdf, 21 pages. [cited by applicant]
R. Chen et al., “Video Foreground Detection Algorithm Based on Fast Principal Component Pursuit and Motion Saliency,” Comput. Intell. Neurosci. 2019, doi: 10.1155/2019/4769185, published Feb. 3, 2019, available at https… [cited by applicant]
E. Candes et al., “Robust Principal Component Analysis?”, arXiv:0912.3599v1 [cs.IT] Dec. 18, 2009, available at https://arxiv.org/pdf/0912.3599.pdf, 39 pages. [cited by applicant]
Cited By (1)
US 12,411,095