IP Library › Granted Patent US 10,586,312
Granted Patent B2
US 10,586,312 · App. 15/899,331 · Granted Mar 10, 2020

Method for image processing and video compression with sparse zone salient features

Inventor: Christiaan Erik Rijnders (Rome, IT)
Assignee: COGISEN S.R.L.
G06T5/10G06T5/002H04N19/139H04N19/149G06T2207/20212
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,586,312
App. No.
15/899,331
Filed
Feb 19, 2018
Granted
Mar 10, 2020
Kind
B2
Art Unit
2485
USPC
375/240.18
Abstract

A method for video compression through image processing and object detection, based on images or a digital video stream of images, to enhance and isolate frequency domain signals representing content to be identified, and decrease or ignore frequency domain noise with respect to the content. A digital image or sequence of digital images defined in a spatial domain are obtained. One or more pairs of sparse zones are selected, each pair generating a selected feature, each zone defined by two sequences of spatial data. The selected features are transformed into frequency domain data. The transfer function, shape and direction of the frequency domain data are varied for each zone, thus generating a normalized complex vector for each feature. The normalized complex vectors are then combined to define a model of the content to be identified.

Claims (15)

1. A method for video compression through image processing and object detection, to be carried out by an electronic processing unit, based either on images or on a digital video stream of images, the images being defined by a single frame or by sequences of frames of said video stream, with the aim of enhancing and then isolating frequency domain signals representing a content to be identified, and decreasing or ignoring frequency domain noise with respect to the content within the images or the video stream, comprising the steps of:

obtaining a digital image or a sequence of digital images from either a corresponding single frame or a corresponding sequence of frames of said video stream, all the digital images being defined in a spatial domain;

selecting one or more pairs of sparse zones, each covering at least a portion of said single frame or at least two frames of said sequence of frames, each pair of sparse zones generating a selected feature, each zone being defined by two sequences of spatial data;

transforming the selected features into frequency domain data by combining, for each zone, said two sequences of spatial data through a 2D variation of an L-transformation, varying a respective transfer function, shape and direction of the frequency domain data for each zone, thus generating a normalized complex vector for each of said selected features;

combining all said normalized complex vectors to define a model of the content to be identified; and

inputting that model from said selected features in a classifier, therefore obtaining the data for object detection or visual saliency to use for video compression.

2. The method for video compression as defined in claim 1 , wherein the step of transforming the selected features into frequency domain data uses spatial data from a varying number and/or choice of frames.

3. The method of video compression according to claim 1 , wherein a search logic is used on the full input image to generate an input frame where said sparse zones are identified.

4. The method of video compression according to claim 1 , wherein said sparse zones are grouped together, either possibly partially overlapping each other or placed side-to-side, to increase a local resolution of said digital image at said sparse zones.

5. The method of video compression according to claim 1 , wherein the transforming the selected features into frequency domain data is carried out in parallel with respect to said two sequences of spatial data.

6. The method of video compression according to claim 1 , wherein, in the transforming step, first 1D Göertzel calculations are performed by rows and then the results are used for a second step wherein 1D Göertzel calculations are performed by columns, or vice versa.

7. The method of video compression according to claim 1 , wherein, for each sparse zone of a pair, different target frequencies are chosen.

8. The method of video compression according to claim 1 , wherein input cells of digital images for the step of transforming into the frequency domain data are only taken around a position for which a transforming computing is needed.

9. The method of video compression according to claim 8 , wherein the transforming computing of the position is taken by separately calculating the 1D output for the row and column at the position and then combining this into a single value.

10. The method of video compression according to claim 1 , wherein the transfer function is chosen separately for each input of a sparse zone, so that the first input and second input have different discrete transfer function settings.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 2, 2021
From: COGISEN S.R.L.
To: INTEL CORPORATION
Reel/Frame 058274/0696 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 20, 2018
From: RIJNDERS, CHRISTIAAN ERIK
To: COGISEN S.R.L.
Reel/Frame 045286/0536 →
Priority Claims (1)
EP 17156726 · Feb 17, 2017 · regional
Continuity (1)
Related Publication 20180240221A1 · Aug 23, 2018
Cited By (2)
US 12,333,777 US 12,340,274