IP Library › Granted Patent US 12,750,512
Granted Patent B2
US 12,750,512 · App. 18/970,073 · Granted Sep 29, 2026

Search area refinement for decoder motion refinement

Inventors: Zhi Zhang (Munich, DE); Po-Han Lin (Taipei, TW); Jian-Liang Lin (Su'ao Township, TW); Yan Zhang (San Diego, CA); Chun-Chi Chen (San Diego, CA); Han Huang (San Diego, CA); Vadim Seregin (San Diego, CA); Marta Karczewicz (San Diego, CA)
Assignee: QUALCOMM Incorporated
H04N19/44H04N19/105H04N19/176
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,750,512
App. No.
18/970,073
Granted
Sep 29, 2026
Kind
B2
Abstract

An example device includes one or more processors configured to determine a first search area in a first reference picture for a current block of the video data. The one or more processors are configured to determine a first initial reference block in the first search area. The one or more processors are configured to apply a local illumination compensation model to the first search area to generate a refined first search area. The one or more processors are configured to apply template matching to the refined first search area to determine a first candidate motion vector having a lowest template matching cost for the refined first search area. The one or more processors are configured to decode the current block based on the first candidate motion vector.

Claims (55)

1 . A method of decoding video data, the method comprising:

determining a first search area in a first reference picture for a current block of the video data, the first reference picture comprising a list 0 (L0) reference picture;

determining a first initial reference block in the first search area;

determining a second search area in a second reference picture for a current block of the video data, the second reference picture comprising a list 1 (L1) reference picture;

determining a second initial reference block in the second search area;

determining a minimum difference between a L0 template prediction and an L1 template prediction;

determining at least one of a first local illumination compensation (LIC) model or a second LIC model based on the minimum difference;

applying the first LIC model to the first search area to generate a refined first search area;

applying template matching to the refined first search area to determine a first candidate motion vector having a lowest template matching cost for the refined first search area;

applying the second LIC model to the second search area to generate a refined second search area;

applying template matching to the refined second search area to determine a second candidate motion vector having a lowest template matching cost for the refined second search area; and

decoding the current block based on the first candidate motion vector and the second candidate motion vector.

2 . The method of claim 1 , further comprising determining to decode the current block using decoder side motion vector refinement.

3 . The method of claim 1 , further comprising determining the first LIC model based on reconstructed neighbor samples of the current block.

4 . The method of claim 1 , wherein applying template matching comprises determining the first candidate motion vector having the lowest template matching cost for the refined first search area between a first reference template in the first reference picture and a reconstructed template in a current picture, the current picture comprising the current block.

5 . The method of claim 1 , wherein applying template matching comprises determining the first candidate motion vector having the lowest template matching cost for the refined first search area between a first reference template in the first reference picture and a second reference template in a second reference picture.

6 . The method of claim 1 , wherein applying the first LIC model comprises determining a LIC template, the LIC template comprising an above template and a left template.

7 . The method of claim 1 , wherein the applying the first LIC model comprises determining a LIC template, wherein the LIC template is a rectangular block.

8 . The method of claim 1 , wherein the first LIC model comprises a smoothing filter configured to remove high frequencies from the first search area.

9 . The method of claim 1 , further comprising encoding the current block prior to decoding the current block.

10 . A device for decoding video data, the device comprising:

one or more memories configured to store the video data; and

one or more processors, the one or more processors communicatively coupled to the one or more memories and configured to:

determine a first search area in a first reference picture for a current block of the video data, the first reference picture comprising a list 0 (L0) reference picture;

determine a first initial reference block in the first search area;

determine a second search area in a second reference picture for a current block of the video data, the second reference picture comprising a list 1 (L1) reference picture;

determine a second initial reference block in the second search area;

determine a minimum difference between a L0 template prediction and an L1 template prediction;

determine at least one of a first local illumination compensation (LIC) model or a second LIC model based on the minimum difference;

apply the first LIC model to the first search area to generate a refined first search area;

apply template matching to the refined first search area to determine a first candidate motion vector having a lowest template matching cost for the refined first search area;

apply the second LIC model to the second search area to generate a refined second search area;

apply template matching to the refined second search area to determine a second candidate motion vector having a lowest template matching cost for the refined second search area; and

decode the current block based on the first candidate motion vector and the second candidate motion vector.

11 . The device of claim 10 , wherein the one or more processors are further configured to determine to decode the current block using decoder side motion vector refinement.

12 . The device of claim 10 , wherein the one or more processors are further configured to determine the first LIC model based on reconstructed neighbor samples of the current block.

13 . The device of claim 10 , wherein to apply template matching, the one or more processors are configured to determine the first candidate motion vector having the lowest template matching cost for the refined first search area between a first reference template in the first reference picture and a reconstructed template in a current picture, the current picture comprising the current block.

14 . The device of claim 10 , wherein to apply template matching, the one or more processors are configured to determine the first candidate motion vector having the lowest template matching cost for the refined first search area between a first reference template in the first reference picture and a second reference template in a second reference picture.

15 . The device of claim 10 , wherein to apply the first LIC model, the one or more processors are configured to determine a LIC template, the LIC template comprising an above template and a left template.

16 . The device of claim 10 , wherein to apply the first LIC model, the one or more processors are configured to determine a LIC template, wherein the LIC template is a rectangular block.

17 . The device of claim 10 , wherein the first LIC model comprises a smoothing filter configured to remove high frequencies from the first search area.

18 . The device of claim 10 , further comprising a display configured to display decoded video data.

19 . The device of claim 10 , wherein the one or more processors are further configured to encode the current block prior to decoding the current block.

20 . A device for decoding video data, the device comprising:

means for determining a first search area in a first reference picture for a current block of the video data, the first reference picture comprising a list 0 (L0) reference picture;

means for determining a first initial reference block in the first search area;

means for determining a second search area in a second reference picture for a current block of the video data, the second reference picture comprising a list 1 (L1) reference picture;

means for determining a second initial reference block in the second search area;

means for determining a minimum difference between a L0 template prediction and an L1 template prediction;

means for determining at least one of a first local illumination compensation (LIC) model or a second LIC model based on the minimum difference;

means for applying the first LIC model to the first search area to generate a refined first search area;

means for applying template matching to the refined first search area to determine a first candidate motion vector having a lowest template matching cost for the refined first search area;

means for applying the second LIC model to the second search area to generate a refined second search area;

means for applying template matching to the refined second search area to determine a second candidate motion vector having a lowest template matching cost for the refined second search area; and

means for decoding the current block based on the first candidate motion vector and the second candidate motion vector.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 22, 2025
From: ZHANG, ZHI; LIN, PO-HAN; LIN, JIAN-LIANG; ZHANG, YAN; CHEN, CHUN-CHI; HUANG, HAN; SEREGIN, VADIM; KARCZEWICZ, MARTA
To: QUALCOMM INCORPORATED
Reel/Frame 069969/0446 →
Continuity (2)
Provisional Application 63615106 · Dec 27, 2023
Related Publication 20250220208A1 · Jul 3, 2025
References Cited (36)
US 20160366416A1 · Liu · 2016 [cited by examiner]
US 20220329823A1 · Chen · 2022 [cited by examiner]
US 20220417522A1 · Huang · 2022 [cited by examiner]
US 20240364865A1 · Deng · 2024 [cited by examiner]
US 20240380903A1 · Deng · 2024 [cited by examiner]
US 20250142081A1 · Vishwanath · 2025 [cited by examiner]
WO 2016200779A1 · 2016 [cited by applicant]
“Decoder Side Motion Vector Refinement for Versatile Video Coding”—Gao et al., 978-1-7281-1817-8/19/$31.00 c 2019 IEEE (Year: 2019). [cited by examiner]
“Adaptive linear model for intra block copy with local illumination compensation”—Wang et al., 2023 IEEE International Conference on Visual Communications and Image Processing (VCIP) (Year: 2023). [cited by examiner]
“A VVC Proposal With Quaternary Tree Plus Binary-Ternary Tree Coding Block Structure and Advanced Coding Techniques”—Huang et al., IEEE Transactions on Circuits and Systems for Video Technology, vol. 30, No. 5, May 2020… [cited by examiner]
Bross B., et al., “Versatile Video Coding Editorial Refinements on Draft 10”, JVET-T2001-v2, Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 2920th Meeting, by Teleconference, Oct. 7-16, 2020, p… [cited by applicant]
Chen J., et al., “Algorithm Description for Versatile Video Coding and Test Model 10 (VTM 10)”, JVET-S2002-v2, Joint Video Experts Team (JVET)of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11 19th Meeting: by Teleconfer… [cited by applicant]
Chen J., et al., “Algorithm Description for Versatile Video Coding and Test Model 11 (VTM 11)”, JVET-T2002-v2, Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29, 20th Meeting, by teleconference… [cited by applicant]
Coban M., et al., “Algorithm Description of Enhanced Compression Model 10 (ECM 10)”, JVET-AE2025-v1, Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29, 31st Meeting, Geneva, CH, Jul. 11-19, 202… [cited by applicant]
Coban M., et al., “Algorithm Description of Enhanced Compression Model 11 (ECM 11)”, JVET-AF2025, Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29 32nd Meeting, Hannover, DE, Oct. 13-20, 2023,… [cited by applicant]
ITU-T H.265: “Series H: Audiovisual and Multimedia Systems Infrastructure of Audiovisual Services—Coding of Moving Video”, High Efficiency Video Coding, The International Telecommunication Union, Jun. 2019, 696 Pages. [cited by applicant]
ITU-T H.266: “Series H: Audiovisual and Multimedia Systems Infrastructure of Audiovisual Services—Coding of Moving Video”, Versatile Video Coding, The International Telecommunication Union, Aug. 2020, 516 pages. [cited by applicant]
Karczewicz M., et al., “Common Test Conditions and Evaluation Procedures for Enhanced Compression Tool Testing”, JVET-AF2017-v1, Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29 32nd Meeting, … [cited by applicant]
Karczewicz M., et al., “Common Test Conditions and Evaluation Procedures for Enhanced Compression Tool Testing”, JVET-Y2017-v1, Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29, 25th Meeting, … [cited by applicant]
Seregin V., et al., “CE4-3.1a and CE4-3.1b: Unidirectional Local Illumination Compensation with Affine Prediction”, JVET-O0066-v1, Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, 15th … [cited by applicant]
Seregin V., et al., “Exploration Experiment on Enhanced Compression Beyond VVC Capability (EE2)”, Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29, 32nd Meeting, Hannover, DE, Oct. 13-20, 2023… [cited by applicant]
Seregin V., “Merge Branch ‘xlxiangli/ecm16-ahg7cfg’ into ‘master’”, ECM Software, Created on Jul. 22, 2021, pp. 1-3. [cited by applicant]
Sullivan G.J., et al., “Overview of the High Efficiency Video Coding (HEVC) Standard”, IEEE Transactions on Circuits and Systems for Video Technology, IEEE Service Center, Piscataway, NJ, US, vol. 22, No. 12, Dec. 1, 20… [cited by applicant]
Wang Y., et al., “EE2 Test 2.6g, 2.6h, 2.6i, 2.6j: Combination of Tests on LIC Improvement”, JVET-AG0276_r2, Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29, 33rd Meeting, by Teleconference, … [cited by applicant]
Wang Y., et al., “Non-EE2: Extension of Local Illumination Compensation”, JVET-AF0200_r1, Joint Video Experts Team (JVET)of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29, 32nd Meeting, Hannover, DE, Oct. 13-20, 2023, JVET-AF… [cited by applicant]
Wang Y-K., et al., “High Efficiency Video Coding (HEVC) Defect Report”, JCTVC-N1003_v1, Joint Collaborative Team on Video Coding (JCT-VC) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11 14th Meeting: Vienna, AT, Jul 2… [cited by applicant]
Xiu X., et al., “EE2-Test2.7: Improvements on Local Illumination Compensation”, JVET-AD0213-v1, Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29 30th Meeting, Antalya, TR, Apr. 21-28, 2023, pp… [cited by applicant]
Xiu X., et al., “Non-EE2: Enhancements on local illumination compensation”, JVET-AF0191-v2, Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29, JVET-AF0191_r1, 32nd Meeting, Hannover, DE, Oct. 1… [cited by applicant]
Zhang N., et al., “EE2-3.2: LIC flag derivation for merge candidates with template costs”, JVET-AF0128-v2, Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29, 32nd Meeting, Hannover, DE, Oct. 13… [cited by applicant]
Zhang Y., et al., “EE2: Test 2.1, 2.2, 2.5, 2.6a, 2.6e and 2.6f on LIC improvement”, JVET-AG0176-v2, Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29, 33rd Meeting, by teleconference, Jan. 17-… [cited by applicant]
Zhang Y., et al., “EE2-2.1: Regression Based Affine Candidate Derivation”, JVET-AA0107-v2, Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29, 27th Meeting, by teleconference, Jul. 13-22, 2022, … [cited by applicant]
Zhang Y., et al., “Non-EE2: On LIC Flag in Merge Mode”, Doc: JVET-AF0194, Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29, 32nd Meeting, Hannover, DE, Oct. 13-20, 2023, JVET-AF0194-V2, pp. 1-… [cited by applicant]
Coban M., et al., “Algorithm Description of Enhanced Compression Model 9 (ECM 9)”, 142. MPEG Meeting; Apr. 24, 2023-Apr. 28, 2023, Antalya, (Motion Picture Expert Group or Iso/iec Jtc1/sc29/wg11), No. m63602, JVET-AD202… [cited by applicant]
Filippov (Huawei) A., et al., “CE1-related: Simplification of LIC parameter derivation and its unification with CCLM”, JVET-N0410-v2, Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11 14t… [cited by applicant]
Filippov (Ofinno) A., et al., “Non-EE2: LIC Extensions”, JVET-AE0140-v2, Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29, 31st Meeting, Geneva, CH, Jul. 11-19, 2023, No. JVET-AE0140, Jul. 13,… [cited by applicant]
International Search Report and Written Opinion—PCT/US2024/058799—ISA/EPO—Mar. 26, 2025 13 Pages. [cited by applicant]