IP Library Granted Patent US 12,401,790
Granted Patent B2
US 12,401,790 · App. 18/019,384 · Granted Aug 26, 2025

Combining ABT with VVC sub-block-based coding tools

Inventors: Fabrice Leleannec (Betton, FR); Baptiste Esteban (Grossoeuvre, FR); Karam Naser (Mouazé, FR); Gagan Bihari Rath (Rennes, FR)
Assignee: INTERDIGITAL MADISON PATENT HOLDINGS, SAS
H04N19/119H04N19/176H04N19/52H04N19/96
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,401,790
App. No.
18/019,384
Granted
Aug 26, 2025
Kind
B2
Abstract

Asymmetric binary trees are used in combination with coding tools such as transform unit tiling, affine motion compensation, decoder side motion vector refinement, and bidirectional optical flow. In another embodiment asymmetric binary trees is used in combination with subblock temporal motion vector prediction. Embodiments enable tiling of coding blocks into subblocks in accordance with coding tools such as used in Versatile Video Coding.

Claims (36)

1. A method, comprising:

partitioning a block of a video image into subblocks of size in correspondence with one or more coding tools, wherein the partitioning comprises dividing the block according to a regular grid of sub-blocks of size 2 in a direction where a size of the block is not a power-of-two;

refining inter bi-predicted luminance subblocks based on an optical flow for subblocks up to size 16×16; and

encoding the video block using said one or more coding tools, wherein asymmetric binary tree partitioning is used in combination with subblock temporal motion vector prediction.

2. An apparatus, comprising:

a processor, configured to:

partition a block of a video image into subblocks of size in correspondence with one or more coding tools, wherein the partitioning comprises dividing the block according to a regular grid of sub-blocks of size 2 in a direction where a size of the block is not a power-of-two;

refine inter bi-predicted luminance subblocks based on an optical flow for subblocks up to size 16×16; and

encode the video block using said one or more coding tools, wherein asymmetric binary tree partitioning is used in combination with subblock temporal motion vector prediction.

3. A method, comprising:

parsing a video bitstream to determine subblock sizes, wherein the subblock sizes comprise blocks having been partitioned according to a regular grid of sub-blocks of size 2 in a direction where a size of the block is not a power-of-two;

refining inter bi-predicted luminance subblocks based on an optical flow for subblocks up to size 16×16; and

decoding the video block using one or more decoding tools on said subblocks of said determined sizes, wherein asymmetric binary tree partitioning is used in combination with subblock temporal motion vector prediction.

4. An apparatus, comprising:

a processor, configured to:

parse a video bitstream to determine subblock sizes, wherein the subblock sizes comprise blocks having been partitioned according to a regular grid of sub-blocks of size 2 in a direction where a size of the block is not a power-of-two;

refine inter bi-predicted luminance subblocks based on an optical flow for subblocks up to size 16×16; and

decode the video block using one or more decoding tools on said subblocks of said determined sizes, wherein asymmetric binary tree partitioning is used in combination with subblock temporal motion vector prediction.

5. The method of claim 1 , wherein said video coding or decoding tools comprise transform unit tiling, affine motion compensation, decoder side motion vector refinement, and bi-directional optical flow.

6. The method of claim 1 , wherein subblocks are chosen to be equal in size.

7. The method of claim 1 , wherein coding units are tiled into transform units.

8. The apparatus of claim 2 , wherein resulting subblocks are fully contained within a square area corresponding to a regular grid of said video image with a granularity equal to a maximum transform size in width and height.

9. The apparatus of claim 2 , wherein a prediction unit is divided into sub-prediction units and a motion model is used to assign each sub-prediction unit to a dedicated motion vector.

10. The apparatus of claim 2 , wherein a motion vector and reference picture are identified based on a temporal motion vector in an associated reference picture.

11. The method of claim 3 , wherein asymmetric binary tree partitioning is used in combination with subblock temporal motion vector prediction and motion data prediction for a subblock is determined by scaling a Sub-block Temporal Motion Vector Prediction.

12. A device comprising:

an apparatus according to claim 1 ; and

at least one of (i) an antenna configured to receive a signal, the signal including the video block, (ii) a band limiter configured to limit the received signal to a band of frequencies that includes the video block, and (iii) a display configured to display an output representative of a video block.

13. A non-transitory computer readable medium containing data content generated according to the method of claim 1 , for playback using a processor.

14. A computer code program stored in a non-transitory computer-readable storage medium or computer-readable storage memory which, when the program is executed by a computer, cause the computer to carry out the method of claim 1 .

15. The method of claim 3 , wherein said video coding or decoding tools comprise transform unit tiling, affine motion compensation, decoder side motion vector refinement, and bi-directional optical flow.

16. The method of claim 3 , wherein subblocks are chosen to be equal in size.

17. The method of claim 3 , wherein coding units are tiled into transform units.

18. The apparatus of claim 4 , wherein resulting subblocks are fully contained within a square area corresponding to a regular grid of said video image with a granularity equal to a maximum transform size in width and height.

19. The apparatus of claim 4 , wherein a prediction unit is divided into sub-prediction units and a motion model is used to assign each sub-prediction unit to a dedicated motion vector.

20. The apparatus of claim 4 , wherein a motion vector and reference picture are identified based on a temporal motion vector in an associated reference picture.

Assignments (3)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 6, 2023
From: INTERDIGITAL CE PATENT HOLDINGS, SAS
To: INTERDIGITAL MADISON PATENT HOLDINGS, SAS
Reel/Frame 065465/0293 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jul 26, 2023
From: INTERDIGITAL VC HOLDINGS FRANCE, SAS
To: INTERDIGITAL CE PATENT HOLDINGS, SAS
Reel/Frame 064396/0118 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 3, 2023
From: LELEANNEC, FABRICE; ESTEBAN, BAPTISTE; NASER, KARAM; RATH, GAGAN BIHARI
To: INTERDIGITAL VC HOLDINGS FRANCE, SAS
Reel/Frame 062874/0942 →
Priority Claims (1)
EP 20305913 · Aug 7, 2020 · regional
Continuity (1)
Related Publication 20230336721A1 · Oct 19, 2023
References Cited (14)
US 11146810B2 · Chen · 2021 [cited by examiner]
US 11368703B2 · Lim · 2022 [cited by examiner]
US 20130136175A1 · Wang · 2013 [cited by examiner]
US 20200204820A1 · Zhang · 2020 [cited by examiner]
US 20200288150A1 · Jun · 2020 [cited by examiner]
US 20210029370A1 · Li · 2021 [cited by examiner]
WO 2019245841 · 2019 [cited by applicant]
Zheng, Implicit Transform Block Split Process for Asymmetric Partitions, Joint Collaborative Team on Video Coding (JCT-VC) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, Document: JCTVC-J0364, 10th Meeting, Stockhol… [cited by examiner]
Zheng, Implicit Transform Block Split Process for Asymmetric Partitions, Joint Collaborative Team on Video Coding (JCT-VC) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, Document: JCTVC-J0364, 10th Meeting, Stockhol… [cited by applicant]
Bordes et al., Description of SDR, HDR and 360 Video Coding Technology Proposal by Qualcomm and Technicolor—Medium Complexity Version, 10. JVET Meeting, Oct. 4, 2018-20-40-2018; San Diego, The Joint Video Exploration Te… [cited by applicant]
Leleannec et al., EE2 Related: Asymmetric Binary Tree Splitting on Top of VVC, 22. JVET Meeting, Apr. 20, 2021-Apr. 28, 2021, Teleconference, (The Joint Video Exploration Team of ISO/IEC JTC1/SC29/WG11 and ITU-T SG. 16,… [cited by applicant]
Bross et al., Versatile Video Coding (Draft 10), Document JVET-S2001_vG, Joint Video Exploration Team of ISO/IEC JTC1/SC29/WG11, 19th meeting, Teleconference, (2020). [cited by applicant]
High Efficiency Video Coding, Series H: Audiovisual and Multimedia Systems Infrastructure of Audiovisual Services—Coding of Moving Video, ITU-T Telecommunication Standardization Sector of ITU, H.265 (Apr. 2015), 634 pag… [cited by applicant]
Chen et al., Algorithm Description of Joint Exploration Test Model 3 (JEM3), Joint Video Exploration Team (JVET) of ITU-T SG16 WP3 and ISO/IEC JTC1/SC29/WG11, Document: JVET-C1001 v3, 3rd Meeting, Geneva, Switzerland, M… [cited by applicant]