IP Library › Granted Patent US 12,348,759
Granted Patent B2
US 12,348,759 · App. 18/514,228 · Granted Jul 1, 2025

Concept of interweaved prediction

Inventors: Kai Zhang (San Diego, CA); Li Zhang (San Diego, CA); Hongbin Liu (Beijing, CN); Yue Wang (Beijing, CN)
Assignees: BEIJING BYTEDANCE NETWORK TECHNOLOGY CO., LTD.; BYTEDANCE INC.
H04N19/513H04N19/105H04N19/119H04N19/176H04N19/573
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,348,759
App. No.
18/514,228
Granted
Jul 1, 2025
Kind
B2
Abstract

Methods, systems, and devices related to sub-block based motion prediction in video coding are described. In one representative aspect, a video processing method includes partitioning a video block into a first set of sub-blocks according to a first pattern, partitioning the video block into a second set of sub-blocks according to a second pattern, in which at least one sub-block in the second set has a different size than a sub-block in the first set, and determining a prediction block corresponding to a combination of a first intermediate prediction block that is predictively generated from the first set of sub-blocks and a second intermediate prediction block that is predictively generated from the second set of sub-blocks.

Claims (37)

1. A method of video processing, comprising:

partitioning, for a conversion between a video block of a video and a bitstream of the video, the video block into a first set of sub-blocks according to a first pattern;

partitioning the video block into a second set of sub-blocks according to a second pattern, wherein at least one sub-block in the second set has a different dimension than a sub-block in the first set, and wherein the first pattern and the second pattern are determined based on a dimension of the video block;

determining a prediction block for the video block, wherein the prediction block is a combination of a first intermediate prediction block generated from the first set of sub-blocks and a second intermediate prediction block generated from the second set of sub-blocks; and

performing the conversion based on the determining.

2. The method of claim 1 , wherein the first intermediate prediction block or the second intermediate prediction block is generated using at least one of (1) an affine prediction method, (2) an alternative temporal motion vector prediction method, (3) a spatial-temporal motion vector prediction method, (4) a bi-directional optical flow method, or (5) a frame-rate up conversion method.

3. The method of claim 1 , wherein the sub-blocks in the first or the second set have a rectangular shape.

4. The method of claim 1 , wherein the sub-blocks in the first or the second set have non-uniform shapes.

5. The method of claim 1 , wherein partitioning the video block into the first set of sub-blocks is performed for a motion compensation of the video block based on a reference picture in a first reference picture list, and partitioning the video block into the second set of sub-blocks is performed for a motion compensation of the video block based on a reference picture in a second reference picture list that is different from the first reference picture list.

6. The method of claim 1 , wherein partitioning the video block into the first set of sub-blocks is performed for a motion compensation of the video block based on a reference picture in a first reference picture list, and partitioning the video block into the second set of sub-blocks is performed for a motion compensation of the video block from a reference picture in a second reference picture list that is same as the first reference picture list.

7. The method of claim 6 , wherein the motion compensation of the video block from a reference picture in the second reference picture list is performed by:

partitioning the video block into a third set of sub-blocks according to a third pattern;

generating a third intermediate prediction block based on the third set of sub-blocks;

partitioning the video block into a fourth set of sub-blocks according to a fourth pattern, wherein at least one sub-block in the fourth set has a different dimension than a sub-block in the third set;

generating a fourth intermediate prediction block based on the fourth set of sub-blocks;

determining a second prediction block based on the third intermediate prediction block and the fourth intermediate prediction block; and

determining a third prediction block based on the prediction block and the second prediction block.

8. The method of claim 1 , wherein the prediction block is determined as a weighted combination of the first intermediate prediction block weighted using a first set of weights and the second intermediate prediction block weighted using a second set of weights.

9. The method of claim 8 , wherein the first set of weights or the second set of weights includes fixed-weight values.

10. The method of claim 9 , wherein at least one value in the first set of weights is different than another value in the first set of weights.

11. The method of claim 10 , wherein at least one value in the second set of weights is different than another value in the second set of weights.

12. The method of claim 1 , further comprising performing a prediction for each sub-block in the first set of sub-blocks to determine the first intermediate prediction block.

13. The method of claim 1 , further comprising performing a prediction for each sub-block in the second set of sub-blocks to determine the second intermediate prediction block.

14. The method of claim 1 , wherein the conversion includes encoding the video block into the bitstream.

15. The method of claim 1 , wherein the conversion includes decoding the video block from the bitstream.

16. A video processing apparatus comprising a processor and a non-transitory memory with instructions thereon, wherein the instructions upon execution by the processor, cause the processor to:

partition, for a conversion between a video block of a video and a bitstream of the video, the video block into a first set of sub-blocks according to a first pattern;

partition the video block into a second set of sub-blocks according to a second pattern, wherein at least one sub-block in the second set has a different dimension than a sub-block in the first set, and wherein the first pattern and the second pattern are determined based on a dimension of the video block;

determine a prediction block for the video block, wherein the prediction block is a combination of a first intermediate prediction block generated from the first set of sub-blocks and a second intermediate prediction block generated from the second set of sub-blocks; and

perform the conversion based on the determining.

17. The video processing apparatus of claim 16 , wherein the first intermediate prediction block or the second intermediate prediction block is generated using at least one of (1) an affine prediction method, (2) an alternative temporal motion vector prediction method, (3) a spatial-temporal motion vector prediction method, (4) a bi-directional optical flow method, or (5) a frame-rate up conversion method.

18. The video processing apparatus of claim 16 , wherein the sub-blocks in the first or the second set have a rectangular shape, or the sub-blocks in the first or the second set have non-uniform shapes.

19. A non-transitory computer-readable recording medium storing a bitstream of a video which is generated by a method performed by a video processing apparatus, wherein the method comprises:

partitioning a video block into a first set of sub-blocks according to a first pattern;

partitioning the video block into a second set of sub-blocks according to a second pattern, wherein at least one sub-block in the second set has a different dimension than a sub-block in the first set, and wherein the first pattern and the second pattern are determined based on a dimension of the video block;

determining a prediction block for the video block, wherein the prediction block is a combination of a first intermediate prediction block generated from the first set of sub-blocks and a second intermediate prediction block generated from the second set of sub-blocks; and

generating the bitstream based on the determining.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 11, 2024
From: ZHANG, KAI; ZHANG, LI
To: BYTEDANCE INC.
Reel/Frame 066103/0269 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 11, 2024
From: LIU, HONGBIN; WANG, YUE
To: BEIJING BYTEDANCE NETWORK TECHNOLOGY CO., LTD.
Reel/Frame 066103/0407 →
Priority Claims (1)
WO PCT/CN2018/089242 · May 31, 2018 · international
Continuity (3)
Continuation 16951137 · Nov 18, 2020
Continuation PCTIB2019054467 · May 30, 2019
Related Publication 20240107053A1 · Mar 28, 2024
References Cited (145)
US 6489995B1 · Kok · 2002 [cited by applicant]
US 6807231B1 · Wiegand · 2004 [cited by applicant]
US 7801217B2 · Boyce · 2010 [cited by applicant]
US 8204109B2 · Xiong · 2012 [cited by applicant]
US 9544601B2 · Zhao · 2017 [cited by applicant]
US 9736481B2 · Zhang · 2017 [cited by applicant]
US 9860559B2 · Zhang · 2018 [cited by applicant]
US 9860562B2 · Zhang · 2018 [cited by applicant]
US 9912925B2 · Ye · 2018 [cited by applicant]
US 9986257B2 · Zhang · 2018 [cited by applicant]
US 9998742B2 · Chen · 2018 [cited by applicant]
US 10057578B2 · Rapaka · 2018 [cited by applicant]
US 10277910B2 · Xiu · 2019 [cited by applicant]
US 10375411B2 · Zhao · 2019 [cited by applicant]
US 10440340B2 · Ye · 2019 [cited by applicant]
US 10462439B2 · He · 2019 [cited by applicant]
US 10469847B2 · Xiu · 2019 [cited by applicant]
US 10477214B2 · Zhang · 2019 [cited by applicant]
US 10812835B2 · Wang · 2020 [cited by applicant]
US 20030031258A1 · Wang · 2003 [cited by applicant]
US 20060193388A1 · Woods · 2006 [cited by applicant]
US 20060268166A1 · Bossen · 2006 [cited by examiner]
US 20080101707A1 · Mukherjee · 2008 [cited by applicant]
US 20090232207A1 · Chen · 2009 [cited by applicant]
US 20100118943A1 · Shiodera · 2010 [cited by applicant]
US 20110122942A1 · Kudana · 2011 [cited by applicant]
US 20110200110A1 · Chen · 2011 [cited by applicant]
US 20120082224A1 · Van der Auwera · 2012 [cited by applicant]
US 20120219216A1 · Sato · 2012 [cited by applicant]
US 20130128974A1 · Chien · 2013 [cited by applicant]
US 20140044179A1 · Li · 2014 [cited by examiner]
US 20140192883A1 · Seregin · 2014 [cited by applicant]
US 20150016528A1 · Wang · 2015 [cited by applicant]
US 20150341657A1 · Onno · 2015 [cited by applicant]
US 20150350687A1 · Zhai · 2015 [cited by applicant]
US 20150373343A1 · Hendry · 2015 [cited by applicant]
US 20160100163A1 · Rapaka · 2016 [cited by applicant]
US 20170048552A1 · An · 2017 [cited by applicant]
US 20170168709A1 · Zhong · 2017 [cited by applicant]
US 20170214932A1 · Huang · 2017 [cited by applicant]
US 20170223377A1 · Bankoski · 2017 [cited by applicant]
US 20170332099A1 · Lee · 2017 [cited by applicant]
US 20180014017A1 · Li · 2018 [cited by applicant]
US 20180048889A1 · Zhang · 2018 [cited by applicant]
US 20180070105A1 · Jin · 2018 [cited by applicant]
US 20180131943A1 · Park · 2018 [cited by applicant]
US 20180184117A1 · Chen · 2018 [cited by applicant]
US 20180192069A1 · Chen · 2018 [cited by applicant]
US 20180213239A1 · Mukherjee · 2018 [cited by applicant]
US 20180241998A1 · Chen · 2018 [cited by applicant]
US 20180270500A1 · Li · 2018 [cited by applicant]
US 20180278950A1 · Chen · 2018 [cited by applicant]
US 20190045192A1 · Socek · 2019 [cited by applicant]
US 20190230350A1 · Chen · 2019 [cited by examiner]
US 20190246122A1 · Zhang · 2019 [cited by applicant]
US 20190273943A1 · Zhao · 2019 [cited by applicant]
US 20190379870A1 · Ye · 2019 [cited by applicant]
US 20200112740A1 · Chien · 2020 [cited by applicant]
US 20200221120A1 · Robert · 2020 [cited by applicant]
US 20200228815A1 · Xu · 2020 [cited by applicant]
US 20210029356A1 · Zhang · 2021 [cited by applicant]
US 20210297673A1 · Zhang · 2021 [cited by applicant]
US 20210329250A1 · Zhang · 2021 [cited by applicant]
US 20220312016A1 · Zhang · 2022 [cited by applicant]
CN 101252686A · 2008 [cited by applicant]
CN 101350920A · 2009 [cited by applicant]
CN 101491107A · 2009 [cited by applicant]
CN 101621693A · 2010 [cited by applicant]
CN 101626505A · 2010 [cited by applicant]
CN 101766030A · 2010 [cited by applicant]
CN 101833768A · 2010 [cited by applicant]
CN 102037732A · 2011 [cited by applicant]
CN 102577388A · 2012 [cited by applicant]
CN 104168483A · 2014 [cited by applicant]
CN 104244002A · 2014 [cited by applicant]
CN 104488271A · 2015 [cited by applicant]
CN 105103554A · 2015 [cited by applicant]
CN 105580365A · 2016 [cited by applicant]
CN 105723707A · 2016 [cited by applicant]
CN 105791858A · 2016 [cited by applicant]
CN 106303544A · 2017 [cited by applicant]
CN 106464885A · 2017 [cited by applicant]
CN 106688237A · 2017 [cited by applicant]
CN 106797476A · 2017 [cited by applicant]
CN 107079150A · 2017 [cited by applicant]
CN 107092787A · 2017 [cited by applicant]
CN 107231557A · 2017 [cited by applicant]
CN 108028933A · 2018 [cited by applicant]
CN 108109629A · 2018 [cited by applicant]
CN 108271023A · 2018 [cited by applicant]
CN 108293131A · 2018 [cited by applicant]
CN 108432250A · 2018 [cited by applicant]
CN 108702509A · 2018 [cited by applicant]
CN 108781282A · 2018 [cited by applicant]
CN 113454999B · 2024 [cited by applicant]
CN 113348669B · 2024 [cited by applicant]
CN 113597760B · 2024 [cited by applicant]
EP 1658726A2 · 2006 [cited by applicant]
JP 2010068103A · 2010 [cited by applicant]
KR 20140146541A · 2014 [cited by applicant]
TW 200644644A · 2006 [cited by applicant]
TW 201433153A · 2014 [cited by applicant]
TW 201507443A · 2015 [cited by applicant]
TW 201820872A · 2018 [cited by applicant]
TW 201841505A · 2018 [cited by applicant]
TW 201902214A · 2019 [cited by applicant]
TW I850252B · 2024 [cited by applicant]
WO 2015196126A1 · 2015 [cited by applicant]
WO 2017059926A1 · 2017 [cited by applicant]
WO 2018054286A1 · 2018 [cited by applicant]
WO 2018070152A1 · 2018 [cited by applicant]
WO 2018226015A1 · 2018 [cited by applicant]
WO 2019004283A1 · 2019 [cited by applicant]
WO 2019229705A1 · 2019 [cited by applicant]
WO 2020008325A1 · 2020 [cited by applicant]
Flierl et al. “Multihypothesis Pictures for H.26L,” Proceedings 2001 International Conference on Image Processing, ICIP 2001—Thessaloniki, Greece, Oct. 7-10, 2001, Institute of Electrical and Electronics Engineers, New … [cited by applicant]
Chen et al. “Algorithm Description of Joint Exploration Test Model 7 (JEM 7),” Joint Video Exploration Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, 7th Meeting: Torino, IT, Jul. 13-21, 2017, document J… [cited by applicant]
Zhang et al. “CE4-Related: Interweaved Prediction for Affine Motion Compensation,” Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IECJTC 1/SC 29/WG 11, 11th Meeting, Ljubljana, SI, Jul. 10-18, 2018, documen… [cited by applicant]
Boyce, Jill M. “Weighted Prediction in the H.264/MPEG AVG Video Coding Standard,” Proceedings/ 2004 IEEE Intemational Syposium on Circuits and Systems, May 23-26, 2004, Sheraton Vancouver Wall Centre Hotel, Vancouver Br… [cited by applicant]
Zhang et al. “CE4-Related: Simplified Affine Prediction,” Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, 11th Meeting, Ljubljana, SI, Jul. 10-18, 2018, document JVET-K0103, 2018. [cited by applicant]
Zhang et al. “CE10: Interweaved Prediction for Affine Motion Compensation (Test 10.5.1 and Test 10.5.2),” Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, 12th Meeting, Macao, CN, Oct. … [cited by applicant]
Zhang et al. “Non-CE2: Interweaved Prediction for Affine Motion Compensation,” Joint Video Experts Team (JVET) o ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, 13th Meeting, Marrakech, MA, Jan. 9-18, 2019, document JVE… [cited by applicant]
Luo et al. “CE9: Addressing the Decoding Latency Issue for Decoder-Side Motion Vector Refinement (DMVR),” Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, 12th Meeting: Macau, CN, Oct. … [cited by applicant]
Andersson et al. “CE11: Deblocking of Sub-Block Boundaries for Luma,” Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, 12th Meeting, Macao, CN, Oct. 3-12, 2018, document JVET-L0074, 201… [cited by applicant]
Bordes et al. “CE4-Related: UC with Reduced Memory Buffer,” Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, 12th Meeting, Macao, CN, Oct. 3-12, 2018, document JVET-L0203, 2018. [cited by applicant]
Xiu et al. “CE9-Related: A Simplified Design of Bi-Directional Optical Flow (BIO),” Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, 12th Meeting, Macao, CN, Oct. 3-12, 2018, document J… [cited by applicant]
JVET-K0102-v2, Zhang, K., et al., “CE4-related: Interweaved Prediction for Affine Motion Compensation,” Bytedance, Jul. 7, 2023, 7 pages. [cited by applicant]
Document: JVET-L0265, Zhang, K., et al., “CE4: Affine Prediction with 4x4 Sub-blocks for Chroma Components (Test 4.1.16),” Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11 12th Meeting: … [cited by applicant]
Document: JVET-M0310-v4, “CE2-related: Using shorter-tap filter for 4x4 sized partition,” Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11 13th Meeting: Marrakech, MA, Jan. 9-18, 2019, 7… [cited by applicant]
Non-Final Office Action from U.S. Appl. No. 17/359,890 dated May 25, 2023. [cited by applicant]
International Search Report and Written Opinion from PCT/182019/054466 dated Sep. 9, 2019 (18 pages). [cited by applicant]
International Search Report and Written Opinion from PCT/182019/054467 dated Sep. 9, 2019 (17 pages). [cited by applicant]
International Search Report and Written Opinion from PCT/182019/054505 dated Sep. 9, 2019 (18 pages). [cited by applicant]
International Search Report and Written Opinion from PCT/182019/057399 dated Jan. 8, 2020 (17 pages). [cited by applicant]
International Search Report and Written Opinion from PCT/182019/057400 dated Nov. 6, 2019 (12 pages). [cited by applicant]
International Search Report and Written Opinion from PCT/CN2020/070113 dated Mar. 26, 2020 (9 pages). [cited by applicant]
International Search Report and Written Opinion from PCT/CN2020/070115 dated Mar. 24, 2020 (10 pages). [cited by applicant]
International Search Report and Written Opinion from PCT/CN2020/070119 dated Mar. 26, 2020 (10 pages). [cited by applicant]
International Search Report and Written Opinion from PCT/CN2020/071660 dated Apr. 13, 2020 (11 pages). [cited by applicant]
Non-Final Office Action from U.S. Appl. No. 17/342,900 dated Oct. 25, 2021. [cited by applicant]
Notice of Allowance from U.S. Appl. No. 16/951,137 dated Jul. 24, 2023. [cited by applicant]
Non-Final Office Action from U.S. Appl. No. 16/951,137 dated Oct. 4, 2022. [cited by applicant]
Final Office Action from U.S. Appl. No. 16/951,137 dated Apr. 6, 2023. [cited by applicant]
CHinese Notice of Allowance from Chinese Patent Application No. 202080007867.3 dated May 28, 2024, 6 pages. [cited by applicant]
Chinese Notice of Allowance from Chinese Patent Application No. 202080007864.X dated Apr. 15, 2024, 7 pages. [cited by applicant]