IP Library Granted Patent US 12,041,244
Granted Patent B2
US 12,041,244 · App. 18/098,453 · Granted Jul 16, 2024

Affine model-based image encoding/decoding method and device

Inventor: Yong Jo Ahn (Seoul, KR)
Assignee: INTELLECTUAL DISCOVERY CO., LTD.
H04N19/159H04N19/105H04N19/176H04N19/52H04N19/96
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,041,244
App. No.
18/098,453
Granted
Jul 16, 2024
Kind
B2
Abstract

In an image encoding/decoding method and device according to the present invention, a candidate list for motion information prediction of a current block is generated, a control point vector of the current block is derived on the basis of the candidate list and a candidate index, a motion vector of the current block is derived on the basis of the control point vector of the current block, and inter-prediction with respect to the current block can be performed by means of the motion vector.

Claims (45)

1. A method of decoding an image, comprising:

generating a candidate list including merge candidates for motion information prediction of a current block in the image, wherein the merge candidates comprise at least one of a plurality of affine candidates or a subblock-based temporal candidate;

deriving, in units of subblocks of the current block, a motion vector of the current block based on the candidate list and a candidate index, the candidate index specifying one of the merge candidates in the candidate list;

generating a prediction block of the current block by performing inter prediction for the current block using the motion vector; and

reconstructing the current block based on the prediction block,

wherein, in response to a case where a merge candidate specified by the candidate index is one of the plurality of the affine candidates, deriving the motion vector comprises:

deriving a control point vector of the current block based on the specified merge candidate; and

deriving the motion vector based on the control point vector of the current block,

wherein, in response to a case where the specified merge candidate is the subblock-based temporal candidate, a motion vector of each of subblocks belonging to the current block is derived using a motion vector of a subblock in a collocated block corresponding to the each subblock in the current block, and

wherein the subblocks belonging to the current block share one reference picture.

2. The method of claim 1 , wherein the affine candidate includes at least one of a spatial candidate or a constructed candidate, and

wherein the spatial candidate is derived from a block encoded with an affine model among spatial neighboring blocks of the current block.

3. The method of claim 2 , wherein the control point vector of the current block includes at least one of a first control point vector corresponding to a top left sample of the current block, a second control point vector corresponding to a top right sample of the current block, or a third control point vector corresponding to a bottom left sample of the current block.

4. The method of claim 3 , wherein the control point vector of the spatial candidate is derived by considering whether a boundary of the current block is located on a boundary of coding tree block (CTU).

5. The method of claim 2 , wherein the constructed candidate is determined based on a combination of motion vectors of neighboring blocks adjacent to the current block.

6. The method of claim 3 , wherein, in response to the case where the specified merge candidate is the one of the plurality of the affine candidates, the motion vector of the current block is derived using at least one of the first control point vector, the second control point vector, a position of the subblock, or the size of the current block.

7. The method of claim 1 , wherein generating the candidate list is selectively performed by considering at least one of a prediction mode of a neighboring block of the current block or a size of the current block.

8. The method of claim 1 , wherein the collocated block belongs to a picture different from the current block, and

wherein the collocated block is representative of a block at a position shifted by a temporal vector from a position of the current block.

9. The method of claim 8 , wherein the temporal vector is determined based on only a left neighboring block among spatial neighboring blocks adjacent to the current block.

10. The method of claim 2 , wherein the candidate list is generated by adding the merge candidates in an order of the subblock-based temporal candidate, the spatial candidate, and the constructed candidate.

11. The method of claim 4 , wherein, in response to a case where the boundary of the current block is not located on the boundary of the CTU, the control point vector of the spatial candidate is derived based on a control point vector of a spatial neighboring block, and

wherein, in response to a case where the boundary of the current block is located on the boundary of the CTU, the control point vector of the spatial candidate is derived based on a motion vector of the spatial neighboring block.

12. A method of encoding an image, comprising:

generating a candidate list including merge candidates for motion information prediction of a current block in the image, wherein the merge candidates comprise at least one of a plurality of affine candidates or a subblock-based temporal candidate;

deriving, in units of subblocks of the current block, a motion vector of the current block based on the candidate list;

generating a prediction block of the current block by performing inter prediction for the current block using the motion vector; and

generating a residual block of the current block based on the prediction block,

wherein a candidate index specifying one of the merge candidates in the candidate list is encoded into a bitstream,

wherein, in response to a case where the motion vector of the current block is derived based on one of the plurality of the affine candidates, deriving the motion vector comprises:

deriving a control point vector of the current block based on the one of the plurality of the affine candidates; and

deriving the motion vector based on the control point vector of the current block,

wherein, in response to a case where the motion vector of the current block is derived based on the subblock-based temporal candidate, a motion vector of each of subblocks belonging to the current block is derived using a motion vector of a subblock in a collocated block corresponding to the each subblock in the current block, and

wherein the subblocks belonging to the current block share one reference picture.

13. A non-transitory computer-readable medium for storing a bitstream generated by an image encoding method, the image encoding method comprising:

generating a candidate list including merge candidates for motion information prediction of a current block in the image, wherein the merge candidates comprise at least one of a plurality of affine candidates or a subblock-based temporal candidate;

deriving, in units of subblocks of the current block, a motion vector of the current block based on the candidate list;

generating a prediction block of the current block by performing inter prediction for the current block using the motion vector; and

generating a residual block of the current block based on the prediction block,

wherein a candidate index specifying one of the merge candidates in the candidate list is encoded into a bitstream,

wherein, in response to a case where the motion vector of the current block is derived based on one of the plurality of the affine candidates, deriving the motion vector comprises:

deriving a control point vector of the current block based on the one of the plurality of the affine candidates; and

deriving the motion vector based on the control point vector of the current block,

wherein, in response to a case where the motion vector of the current block is derived based on the subblock-based temporal candidate, a motion vector of each of subblocks belonging to the current block is derived using a motion vector of a subblock in a collocated block corresponding to the each subblock in the current block, and

wherein the subblocks belonging to the current block share one reference picture.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jul 29, 2025
From: AHN, YONG JO
To: INTELLECTUAL DISCOVERY CO., LTD.
Reel/Frame 071869/0608 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jul 21, 2025
From: INTELLECTUAL DISCOVERY CO., LTD.
To: DOLBY INTERNATIONAL AB
Reel/Frame 072128/0634 →
Priority Claims (1)
KR 10-2018-0038904 · Apr 3, 2018 · national
Continuity (2)
Continuation 17044765
Related Publication 20230283784A1 · Sep 7, 2023
Cited By (1)
US 12,395,644