IP Library › Granted Patent US 12,744,883
Granted Patent B2
US 12,744,883 · App. 18/768,800 · Granted Sep 22, 2026

Subblock based motion vector predictor displacement vector reordering using template matching

Inventors: Han Gao (San Diego, CA); Lien-Fei Chen (Hsinchu, TW); Guichun Li (San Jose, CA); Xin Zhao (San Jose, CA); Shan Liu (San Jose, CA)
Assignee: Tencent America LLC
H04N19/105H04N19/139H04N19/159H04N19/176H04N19/513H04N19/70
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,744,883
App. No.
18/768,800
Granted
Sep 22, 2026
Kind
B2
Abstract

Aspects of the disclosure provide a method and an apparatus for video encoding/decoding. The apparatus includes processing circuitry for: receiving prediction information of a current coding block in a current picture from a coded video bitstream, the prediction information indicating that the current coding block is coded using a subblock-based temporal motion vector prediction (SbTMVP) mode; deriving multiple displacement vector (DV) candidates by applying multiple DV offset candidates to a fixed DV predictor of the current coding block; comparing a template of the current coding block with each of multiple templates, each template of the multiple templates being located at a position specified by a corresponding one of the multiple DV candidates; calculating a cost value associated with each one of the multiple DV offset candidates based on the comparing; and reordering DV offset indices of the multiple DV offset candidates based on their calculated cost values.

Claims (44)

1 . A method of video decoding, the method comprising:

receiving prediction information of a current coding block in a current picture from a coded video bitstream, the prediction information indicating that the current coding block is coded using a subblock-based temporal motion vector prediction (SbTMVP) mode;

comparing a template of the current coding block with each of multiple templates, each template of the multiple templates being located at a position specified by a corresponding one of multiple displacement vector (DV) candidates;

calculating a cost value associated with each one of the multiple DV candidates based on the comparison of the template with each of the multiple templates;

selecting a DV candidate from the multiple DV candidates with a lowest calculated cost value or from a reordered list of the multiple DV candidates based on the calculated cost values; and

predicting the current coding block in the SbTMVP mode based at least on the selected DV candidate.

2 . The method of claim 1 , further comprising:

obtaining multiple DV offsets from the coded video bitstream, each DV offset corresponding to a respective one of the multiple DV candidates; and

deriving the multiple DV candidates by applying the multiple DV offsets to a fixed DV predictor of the current coding block.

3 . The method of claim 1 , wherein the predicting comprises:

predicting the current coding block in the SbTMVP mode based at least on an index of the reordered list of the multiple DV candidates that is signaled in the coded video bitstream, the index indicating which DV candidate is selected from the reordered list of the multiple DV candidates for performing SbTMVP.

4 . The method of claim 1 , wherein the selecting the DV candidate from the reordered list comprises selecting the DV candidate of the multiple DV candidates with a lowest calculated cost value by default for performing SbTMVP.

5 . The method of claim 1 , wherein the cost value is calculated by performing Sum of Absolute Differences (SAD), Sum of Absolute Transformed Differences (SATD), Sum of Squared Error (SSE), sub-sampled SAD, or mean-removed SAD.

6 . The method of claim 1 , wherein the multiple DV candidates comprise Merge with Motion Vector Difference (MMVD) candidates.

7 . The method of claim 6 , wherein the comparing, the calculating, and the selecting are performed only for a subset of the MMVD candidates, wherein a relative order of one or more other ones of the MMVD candidates is kept unchanged.

8 . The method of claim 6 , wherein the comparing, the calculating, and the selecting are performed for all of the MMVD candidates, wherein after reordering only a number N of the MMVD candidates which have lowest cost values are used, wherein the number N is less than or equal to a total number of the MMVD candidates.

9 . The method of claim 1 , wherein the multiple DV candidates include multiple DV predictor candidates.

10 . The method of claim 9 , further comprising:

receiving an index signaled in the coded video bitstream, wherein the index indicates which DV predictor candidate is selected from the reordered list of the multiple DV candidates for performing SbTMVP.

11 . The method of claim 10 , wherein the method further comprises selecting a DV predictor candidate with a lowest calculated cost value by default for performing SbTMVP.

12 . The method of claim 10 , wherein the list of the multiple DV candidates is constructed from spatial neighboring coding units (CUs) or from history-based motion vector prediction (HMVP) candidates.

13 . The method of claim 10 , wherein only a first N number of the multiple DV predictor candidates on the reordered list of the multiple DV candidates are signaled.

14 . The method of claim 1 , wherein the list of the multiple DV candidates includes at least (i) a first DV candidate that is derived by applying a DV offset to a fixed DV predictor of the current coding block and (ii) a second DV candidate that includes a DV predictor candidate.

15 . A method of video encoding, the method comprising:

determining that a current coding block in a current picture is to be coded using a subblock-based temporal motion vector prediction (SbTMVP) mode;

comparing a template of the current coding block with each of multiple templates, each template of the multiple templates being located at a position specified by a corresponding one of multiple displacement vector (DV) candidates;

calculating a cost value associated with each one of the multiple DV candidates based on the comparison of the template with each of the multiple templates;

selecting a DV candidate from the multiple DV candidates with a lowest calculated cost value or from a reordered list of the multiple DV candidates based on the calculated cost values; and

encoding the current coding block in the SbTMVP mode in a bitstream based at least on the selected DV candidate.

16 . The method of claim 15 , further comprising:

determining multiple DV offsets, each DV offset corresponding to a respective one of the multiple DV candidates; and

deriving the multiple DV candidates by applying the multiple DV offsets to a fixed DV predictor of the current coding block.

17 . The method of claim 15 , wherein the selecting the DV candidate from the reordered list comprises selecting the DV candidate of the multiple DV candidates with a lowest calculated cost value by default for performing SbTMVP.

18 . A non-transitory computer-readable storage medium storing instructions which, when executed by a processor, cause the processor to perform a method of encoding a bitstream comprising:

determining that a current coding block in a current picture is to be coded using a subblock-based temporal motion vector prediction (SbTMVP) mode;

comparing a template of the current coding block with each of multiple templates, each template of the multiple templates being located at a position specified by a corresponding one of multiple displacement vector (DV) candidates;

calculating a cost value associated with each one of the multiple DV candidates based on the comparison of the template with each of the multiple templates;

selecting a DV candidate from the multiple DV candidates with a lowest calculated cost value or from a reordered list of the multiple DV candidates based on the calculated cost values;

encoding the current coding block in the SbTMVP mode in the bitstream based at least on the selected DV candidate; and

transmitting the encoded bitstream.

19 . The non-transitory computer-readable storage medium of claim 18 , wherein the method further comprises:

determining multiple DV offsets, each DV offset corresponding to a respective one of the multiple DV candidates; and

deriving the multiple DV candidates by applying the multiple DV offsets to a fixed DV predictor of the current coding block.

20 . The non-transitory computer-readable storage medium of claim 18 , wherein the selecting the DV candidate from the reordered list comprises selecting the DV candidate of the multiple DV candidates with a lowest calculated cost value by default for performing SbTMVP.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 8, 2024
From: GAO, HAN; CHEN, LIEN-FEI; LI, GUICHUN; ZHAO, XIN; LIU, SHAN
To: TENCENT AMERICA LLC
Reel/Frame 068226/0753 →
Continuity (3)
Continuation 17985127 · Nov 10, 2022
Provisional Application 63346283 · May 26, 2022
Related Publication 20240364873A1 · Oct 31, 2024
References Cited (20)
US 11838539B2 · Liu et al. · 2023 [cited by applicant]
US 20190387251A1 · Lin et al. · 2019 [cited by applicant]
US 20220078408A1 · Park et al. · 2022 [cited by applicant]
US 20220264147A1 · Robert et al. · 2022 [cited by applicant]
US 20230104476A1 · Chen et al. · 2023 [cited by applicant]
US 20230109532A1 · Chen et al. · 2023 [cited by applicant]
US 20230336737A1 · Chen et al. · 2023 [cited by applicant]
US 20230388513A1 · Chen et al. · 2023 [cited by applicant]
US 20240007615A1 · Liao et al. · 2024 [cited by applicant]
US 20240015333A1 · Chen et al. · 2024 [cited by applicant]
US 20250211767A1 · Chen · 2025 [cited by examiner]
“Versatile Video Coding”, Telecommunication Standardization Sector of ITU, ITU-T Rec. H.266 and ISO/IEC 23090-3, 2020, 516 pages. [cited by applicant]
Coban et al., “Algorithm description of Enhanced Compression Model 4 (ECM 4)”, Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29 23rd Meeting, by teleconference, JVET-Y2025-v2, Jul. 2021, 32 pa… [cited by applicant]
High Efficiency Video Coding, Telecommunication Standardization Sector of ITU, Rec. Itu-T H.265 v4, Dec. 2016, 664 pages. [cited by applicant]
Yang et al., “Subblock-Based Motion Derivation and Inter Prediction Refinement in Versatile Video Coding Standard”, IEEE Transactions Circuits Syst. Video Technol., vol. 31, No. 10, Oct. 2021, pp. 3862-3877. [cited by applicant]
Chen et al., “EE-2.3: SBTMVP with MMVD”, Tencent, vJoint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29 30th Meeting, Antayla, Apr. 21-28, 2023, Document: JVET-AD0209-v2, Apr. 2023, 4 pages. [cited by applicant]
Chen et al., “Non-EE-2: SbTMVP with MMVD”, Tencent, Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29 30th Meeting, by teleconference, Jan. 11-20, 2023, Document: JVET-AC0213-v2, Jan. 2023, 3 p… [cited by applicant]
Extended European Search Report received for European Patent Application No. 22932484.3, mailed on Mar. 30, 2026, 11 pages. [cited by applicant]
Salehifar et al., “Non-EE-2: Template Matching-based Reordering for Extended MMVD Design”, Bytedance Inc., Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29 24th Meeting, by teleconference, Oct… [cited by applicant]
Zhang et al., “AHG12: Adaptive Reordering of Merge Candidates with Template Matching”, Bytedance Inc., Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29 22nd Meeting, by teleconference, Apr. 20… [cited by applicant]