IP Library Granted Patent US 12,229,930
Granted Patent B2
US 12,229,930 · App. 17/496,246 · Granted Feb 18, 2025

Distortion-based filtering for image classification

Inventors: Utsav Drolia (Milipitas, CA); Srimat Chakradhar (Manalapan, NJ); Sibendu Paul (West Lafayette, IN)
Assignee: NEC Corporation
G06T7/0002G06N3/04H04N23/64G06T2207/20081G06T2207/20084G06T2207/30168G06T2207/30232
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,229,930
App. No.
17/496,246
Granted
Feb 18, 2025
Kind
B2
Abstract

Methods and systems for image filtering include detecting a distortion level of input images, using a distortion detection model that is trained using confidence values generated by a pre-trained image classifier with a set of distorted training images. An analysis is performed on input images having a detected distortion level that is lower than a threshold, with input images having an above-threshold detected distortion level being filtered out.

Claims (31)

1. A computer-implemented method for image filtering, comprising:

detecting a distortion level of input images, using a distortion detection model that is trained using confidence values generated by a pre-trained image classifier with a set of distorted training images;

performing analysis on input images having a detected distortion level that is lower than a threshold, with input images having an above-threshold detected distortion level being filtered out; and

adjusting an image capture setting responsive to a determination that the input image has an above-threshold distortion level.

2. The computer-implemented method of claim 1 , further comprising training the distortion detection model using the set of distorted training images, each distorted training image having an associated target score that is based on the confidence values generated by the pre-trained image classifier.

3. The computer-implemented method of claim 2 , further comprising generating the set of distorted training images by performing a plurality of distortion operations on original training images.

4. The computer-implemented method of claim 3 , wherein generating the set of distorted training images further includes generating a target score for each respective distorted training image by aggregating confidence scores from a plurality of pre-trained image classifiers using the respective distorted training image as input.

5. The computer-implemented method of claim 1 , further comprising training the distortion detection model using a set of distorted training images, each distorted training image having an associated target score that is based on a distance between a classifier softmax output for the distorted training image and a classifier softmax output for an original undistorted version of the distorted training image.

6. The computer-implemented method of claim 1 , wherein the image capture setting is selected from the group consisting of a camera frame rate and a brightness.

7. The computer-implemented method of claim 1 , wherein the distortion detection model is implemented as a neural network model that includes a feature extractor with a plurality of branches of convolutional layers and a regressor with fully connected layers.

8. The computer-implemented method of claim 7 , wherein the plurality of branches include convolutional layers of different kernel sizes.

9. The computer-implemented method of claim 8 , wherein the feature extractor concatenates outputs of the plurality of branches to generate a feature extractor output.

10. A computer-implemented method for image filtering, comprising:

performing a plurality of distortion operations on a set of original training images to generate distorted training images corresponding to the original training images;

training a neural network of a distortion detection model using the set of original training images and the distorted training images corresponding to the original training images, based on confidence values generated by a pre-trained image classifier;

detecting a distortion level of input images, using the distortion detection model;

performing analysis on input images having a detected distortion level that is lower than a threshold, with input images having an above-threshold detected distortion level being filtered out; and

adjusting an image capture setting responsive to a determination that the input image has an above-threshold distortion level.

11. A system for image filtering, comprising:

a hardware processor; and

a memory that stores a computer program, which, when executed by the hardware processor, causes the hardware processor to:

detect a distortion level of input images, using a distortion detection model that is trained using confidence values generated by a pre-trained image classifier with a set of distorted training images;

perform analysis on input images having a detected distortion level that is lower than a threshold, with input images having an above-threshold detected distortion level being filtered out; and

adjust an image capture setting responsive to a determination that the input image has an above-threshold distortion level.

12. The system of claim 11 , wherein the image capture setting is selected from the group consisting of a camera frame rate and a brightness.

13. The system of claim 11 , wherein the computer program further causes the hardware processor to train distortion detection model using the set of distorted training images, each distorted training image having an associated target score that is based on the confidence values generated by the pre-trained image classifier.

14. The system of claim 13 , wherein the computer program further causes the hardware processor to generate the set of distorted training images by performing a plurality of distortion operations on original training images.

15. The system of claim 14 , wherein the computer program further causes the hardware processor to generate a target score for each respective distorted training image by aggregating confidence scores from a plurality of pre-trained image classifiers using the respective distorted training image as input.

16. The system of claim 11 , wherein the computer program further causes the hardware processor to train the distortion detection model using a set of distorted training images, each distorted training image having an associated target score that is based on a distance between a classifier softmax output for the distorted training image and a classifier softmax output for an original undistorted version of the distorted training image.

17. The system of claim 11 , wherein the distortion detection model is implemented as a neural network model that includes a feature extractor with a plurality of branches of convolutional layers and a regressor with fully connected layers.

18. The system of claim 17 , wherein the plurality of branches include convolutional layers of different kernel sizes.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 10, 2024
From: NEC LABORATORIES AMERICA, INC.
To: NEC CORPORATION
Reel/Frame 069540/0269 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 7, 2021
From: DROLIA, UTSAV; CHAKRADHAR, SRIMAT; PAUL, SIBENDU
To: NEC LABORATORIES AMERICA, INC.
Reel/Frame 057731/0004 →
Continuity (3)
Provisional Application 63089681 · Oct 9, 2020
Provisional Application 63089570 · Oct 9, 2020
Related Publication 20220114717A1 · Apr 14, 2022
References Cited (10)
US 20190303720A1 · Karam · 2019 [cited by examiner]
US 20210343030A1 · Sagonas · 2021 [cited by examiner]
US 20220148146A1 · Morzos · 2022 [cited by examiner]
Mittal, Anish, et al. “No-reference image quality assessment in the spatial domain”, IEEE Transactions on image processing, vol. 21, No. 12. Dec. 2012, pp. 4695-4708. [cited by applicant]
Szegedy, Christian, et al. “Going deeper with convolutions”, InProceedings of the IEEE conference on computer vision and pattern recognition. Jun. 1, 2015, pp. 1-9. [cited by applicant]
Deng, Jia, et al. “Imagenet: A large-scale hierarchical image database”, In2009 IEEE conference on computer vision and pattern recognition Jun. 20, 2009, pp. 248-255. [cited by applicant]
He, Kaiming, et al. “Deep residual learning for image recognition”, InProceedings of the IEEE conference on computer vision and pattern recognition. Jun. 2016, pp. 770-778. [cited by applicant]
Talebi, Hossein, et al. “NIMA: Neural Image Assessment”, IEEE Transactions on Image Processing, vol. 27, No. 8. Aug. 2018, pp. 3998-4011. [cited by applicant]
Szegedy, Christian, et al. “Rethinking the inception architecture for computer vision”, InProceedings of the IEEE conference on computer vision and pattern recognition. Jun. 28, 2016, pp. 2818-2826. [cited by applicant]
Li, Yuanqi, et al. “Reducto: On-camera filtering for resource-efficient real-time video analytics”, InProceedings of the Annual conference of the ACM Special Interest Group on Data Communication on the applications, tec… [cited by applicant]