IP Library › Granted Patent US 11,375,202
Granted Patent B2
US 11,375,202 · App. 17/277,460 · Granted Jun 28, 2022

Translational and affine candidates in a unified list

Inventors: Franck Galpin (Thorigne-Fouillard, FR); Tangi Poirier (Thorigne-Fouillard, FR); Ya Chen (Rennes, FR)
Assignee: InterDigital VC Holdings, Inc.
H04N19/139H04N19/105H04N19/159H04N19/176
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,375,202
App. No.
17/277,460
Granted
Jun 28, 2022
Kind
B2
Abstract

At least a method and an apparatus are presented for efficiently encoding or decoding video. For example, one or more prediction models are determined respectively for one or more prediction candidates used for the video encoding based on one or more control point motion vectors of a current block of the video being encoded or decoded. It is determined from the one or more control point motion vectors that a first prediction model of the one or more prediction models may be a translational prediction model. It is also determined from the one or more control point motion vectors that a second prediction model of the one or more prediction models is to be an affine prediction model. The video is encoded or decoded based on a candidate list comprising the one or more prediction candidates determined respectively from the one or more prediction models.

Claims (36)

1. A method for video decoding, comprising:

determining one or more prediction models respectively for one or more prediction candidates used for the video decoding based on one or more control point motion vectors of a current block of the video being decoded;

determining from the one or more control point motion vectors that a first prediction model of the one or more prediction models is to be a translational prediction model, wherein the first prediction model of a prediction candidate is a translational prediction model when a relative motion between two of the one or more control point motion vectors is less than or equal to a first value or higher than a second value depending on a height and a width of the current block;

determining from the one or more control point motion vectors that a second prediction model of the one or more prediction models is to be an affine prediction model; and

decoding the video based on a candidate list comprising the one or more prediction candidates determined respectively from the one or more prediction models.

2. The method of claim 1 , wherein the determining from the one or more control point motion vectors that a first prediction model of the one or more prediction models is to be a translational prediction model, is based on the one or more control motion vectors being in inter mode.

3. The method of claim 1 , wherein the second prediction model of the one or more prediction models is to be an affine prediction model when a relative motion between two of the one or more control point motion vectors is higher than the first value and lower than or equal to the second value depending on a height and a width of the current block.

4. The method of claim 1 , wherein the affine prediction model is a 4-parameter affine prediction model or a 6-parameter affine prediction model.

5. The method of claim 1 , wherein the one or more prediction models comprise an AMVP mode or a merge mode.

6. An apparatus for video decoding, comprising one or more processors, wherein the one or more processors are configured to:

determine one or more prediction models respectively for one or more prediction candidates used for the video decoding based on one or more control point motion vectors of a current block of the video being decoded;

determine from the one or more control point motion vectors that a first prediction model of the one or more prediction models is to be a translational prediction model, wherein the first prediction model of a prediction candidate is a translational prediction model when a relative motion between two of the one or more control point motion vectors is less than or equal to a first value or higher than a second value depending on a height and a width of the current block;

determine from the one or more control point motion vectors that a second prediction model of the one or more prediction models is to be an affine prediction model; and

decode the video based on a candidate list comprising the one or more prediction candidates determined respectively from the one or more prediction models.

7. The apparatus of claim 6 , wherein the determining from the one or more control point motion vectors that a first prediction model of the one or more prediction models is to be a translational prediction model, is based on the one or more control motion vectors being in inter mode.

8. The apparatus of claim 6 , wherein the second prediction model of the one or more prediction models is to be an affine prediction model when a relative motion between two of the one or more control point motion vectors is higher than the first value and lower than or equal to the second value depending on a height and a width of the current block.

9. The apparatus of claim 6 , wherein the affine prediction model is a 4-parameter affine prediction model or a 6-parameter affine prediction model.

10. The apparatus of claim 6 , wherein the one or more prediction models comprise an AMVP mode or a merge mode.

11. A method for video encoding, comprising:

determining one or more prediction models respectively for one or more prediction candidates used for the video encoding based on one or more control point motion vectors of a current block of the video being encoded;

determining from the one or more control point motion vectors that a first prediction model of the one or more prediction models is to be a translational prediction model, wherein the first prediction model of a prediction candidate is a translational prediction model when a relative motion between two of the one or more control point motion vectors is less than or equal to a first value or higher than a second value depending on a height and a width of the current block;

determining from the one or more control point motion vectors that a second prediction model of the one or more prediction models is to be an affine prediction model; and

encoding the video based on a candidate list comprising the one or more prediction candidates determined respectively from the one or more prediction models.

12. The method of claim 11 , wherein the determining from the one or more control point motion vectors that a first prediction model of the one or more prediction models is to be a translational prediction model, is based on the one or more control motion vectors being in inter mode.

13. The method of claim 11 , wherein the second prediction model of the one or more prediction models is to be an affine prediction model when a relative motion between two of the one or more control point motion vectors is higher than the first value and lower than or equal to the second value depending on a height and a width of the current block.

14. The method of claim 11 , wherein the affine prediction model is a 4-parameter affine prediction model or a 6-parameter affine prediction model.

15. The method of claim 11 , wherein the one or more prediction models comprise an AMVP mode or a merge mode.

16. An apparatus for video encoding, comprising one or more processors, wherein the one or more processors are configured to:

determine one or more prediction models respectively for one or more prediction candidates used for the video encoding based on one or more control point motion vectors of a current block of the video being encoded;

determine from the one or more control point motion vectors that a first prediction model of the one or more prediction models is to be a translational prediction model, wherein the first prediction model of a prediction candidate is a translational prediction model when a relative motion between two of the one or more control point motion vectors is less than or equal to a first value or higher than a second value depending on a height and a width of the current block;

determine from the one or more control point motion vectors that a second prediction model of the one or more prediction models is to be an affine prediction model; and

encode the video based on a candidate list comprising the one or more prediction candidates determined respectively from the one or more prediction models.

17. The apparatus of claim 16 , wherein the determining from the one or more control point motion vectors that a first prediction model of the one or more prediction models is to be a translational prediction model, is based on the one or more control motion vectors being in inter mode.

18. The apparatus of claim 16 , wherein the second prediction model of the one or more prediction models is to be an affine prediction model when a relative motion between two of the one or more control point motion vectors is higher than the first value and lower than or equal to the second value depending on a height and a width of the current block.

19. The apparatus of claim 16 , wherein the affine prediction model is a 4-parameter affine prediction model or a 6-parameter affine prediction model.

20. The apparatus of claim 16 , wherein the one or more prediction models comprise an AMVP mode or a merge mode.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 4, 2023
From: INTERDIGITAL VC HOLDINGS, INC.
To: INTERDIGITAL MADISON PATENT HOLDINGS, SAS
Reel/Frame 062291/0394 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 19, 2021
From: GALPIN, FRANCK; POIRIER, TANGI; CHEN, YA
To: INTERDIGITAL VC HOLDINGS, INC.
Reel/Frame 055650/0061 →
Priority Claims (1)
EP 18306232 · Sep 21, 2018 · regional
Continuity (1)
Related Publication 20210352294A1 · Nov 11, 2021