IP Library Granted Patent US 12,563,228
Granted Patent B2
US 12,563,228 · App. 17/435,669 · Granted Feb 24, 2026

Sub-picture bitstream extraction and reposition

Inventor: Yong He (San Diego, CA)
Assignee: InterDigital VC Holdings, Inc.
H04N19/597H04N19/172H04N19/30H04N19/70
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,563,228
App. No.
17/435,669
Granted
Feb 24, 2026
Kind
B2
Abstract

Systems and methods described herein employ a high-level syntax design that supports a sub-picture extraction and reposition process. An input video may be encoded into multiple representations, each representation may be represented as a layer. A layer picture may be partitioned into multiple sub-pictures. Each sub-picture may have its own tile partitioning, resolution, color format and bit depth. Each sub-picture is encoded independently from other sub-pictures of the same layer, but it may be inter-predicted from the corresponding sub-pictures from its dependent layers. Each sub-picture may refer to a sub-picture parameter set where the sub-picture properties such as resolution and coordinate is signaled. Each sub-picture parameter set may refer to a PPS where the resolution of the entire picture is signaled.

Claims (30)

1 . A method comprising:

encoding in a bitstream a video including at least one picture comprising a plurality of sub-pictures, at least one of the sub-pictures being a layered sub-picture encoded using a plurality of layers, each sub-picture having a respective level and each layer of the layered sub-picture having a respective level; and

signaling in the bitstream a data structure, wherein the data structure indicates the level of each of the plurality of respective sub-pictures, including the level of each respective layer of the layered sub-picture;

wherein the level indicates a predefined set of constraints on values of syntax elements of the respective sub-picture or the respective layer of the layered sub-picture.

2 . The method of claim 1 , further comprising signaling, for each of the sub-pictures, information indicating a tier for the respective sub-picture.

3 . The method of claim 1 , further comprising signaling, for each of the sub-pictures, information indicating a profile for the respective sub-picture.

4 . The method of claim 1 , wherein each of the sub-pictures is associated with a layer, and wherein each sub-picture within a layer is encoded independently from other sub-pictures in the same layer.

5 . An apparatus comprising:

a processor configured to perform at least:

encoding in a bitstream a video comprising a plurality of sub-pictures, at least one of the sub-pictures being a layered sub-picture encoded using a plurality of layers, each sub-picture having a respective level and each layer of the layered sub-picture having a respective level; and

signaling in the bitstream a data structure, wherein the data structure indicates the level of each of the plurality of respective sub-pictures, including the level of each respective layer of the layered sub-picture;

wherein the level indicates a predefined set of constraints on values of syntax elements of the respective sub-picture or the respective layer of the layered sub-picture.

6 . The apparatus of claim 5 , wherein each of the sub-pictures is associated with a layer, and wherein each sub-picture within a layer is encoded independently from other sub-pictures in the same layer.

7 . The apparatus of claim 5 , wherein the processor is further configured to signal at least one output sub-picture set in the bitstream, wherein the output sub-picture set identifies at least a subset of the plurality of sub-pictures and includes the level for each of the sub-pictures in the subset.

8 . The apparatus of claim 5 , wherein the processor is further configured to signal at least one output sub-picture set in the bitstream, wherein the output sub-picture set identifies at least a subset of the plurality of sub-pictures and includes position offset information for each of the sub-pictures in the subset.

9 . A method comprising:

decoding from a bitstream a video including at least one picture comprising a plurality of sub-pictures, at least one of the sub-pictures being a layered sub-picture encoded using a plurality of layers, each sub-picture having a respective level and each layer of the layered sub-picture having a respective level; and

decoding from the bitstream a data structure, wherein the data structure indicates the level of each of the plurality of respective sub-pictures, including the level of each respective layer of the layered sub-picture;

wherein the level indicates a predefined set of constraints on values of syntax elements of the respective sub-picture or the respective layer of the layered sub-picture.

10 . The method of claim 9 , further comprising selecting an output sub-picture set of the sub-pictures based at least in part on the level, wherein decoding the video comprises decoding the selected output sub-picture set.

11 . The method of claim 9 , wherein each of the sub-pictures is associated with a layer, and wherein at least one sub-picture within a layer is decoded independently from other sub-pictures in the same layer.

12 . The method of claim 9 , further comprising composing at least one output frame from the decoded plurality of sub-pictures.

13 . An apparatus comprising:

a processor configured to perform at least:

decoding from a bitstream a video including at least one picture comprising a plurality of sub-pictures, at least one of the sub-pictures being a layered sub-picture encoded using a plurality of layers, each sub-picture having a respective level and each layer of the layered sub-picture having a respective level; and

decoding from the bitstream a data structure, wherein the data structure indicates the level of each of the plurality of respective sub-pictures, including the level of each respective layer of the layered sub-picture;

wherein the level indicates a predefined set of constraints on values of syntax elements of the respective sub-picture or the respective layer of the layered sub-picture.

14 . The apparatus of claim 13 , wherein the processor is further configured to select an output sub-picture set of the sub-pictures based at least in part on the level, and wherein the selected output sub-picture set is decoded.

15 . The apparatus of claim 13 , wherein each of the sub-pictures is associated with a layer, and wherein at least one sub-picture within a layer is decoded independently from other sub-pictures in the same layer.

16 . The apparatus of claim 13 , wherein the processor is further configured to compose at least one output frame from the decoded plurality of sub-pictures.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jul 11, 2024
From: VID SCALE, INC.
To: INTERDIGITAL VC HOLDINGS, INC.
Reel/Frame 068284/0031 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 11, 2022
From: HE, YONG
To: VID SCALE, INC.
Reel/Frame 060787/0994 →
Continuity (3)
Provisional Application 62855446 · May 31, 2019
Provisional Application 62816703 · Mar 11, 2019
Related Publication 20220141488A1 · May 5, 2022
References Cited (47)
US 20040008766A1 · Wang · 2004 [cited by examiner]
US 20060256851A1 · Wang · 2006 [cited by examiner]
US 20080095228A1 · Hannuksela · 2008 [cited by applicant]
US 20130188709A1 · Deshpande · 2013 [cited by applicant]
US 20140086344A1 · Wang · 2014 [cited by examiner]
US 20150016547A1 · Tabatabai · 2015 [cited by examiner]
US 20150063453A1 · Kang · 2015 [cited by applicant]
US 20150312580A1 · Hannuksela · 2015 [cited by examiner]
US 20150365702A1 · Deshpande · 2015 [cited by examiner]
US 20150381998A1 · Wang · 2015 [cited by examiner]
US 20160044324A1 · Deshpande · 2016 [cited by examiner]
US 20160165247A1 · Deshpande · 2016 [cited by examiner]
US 20160173887A1 · Deshpande · 2016 [cited by examiner]
US 20160191926A1 · Deshpande · 2016 [cited by examiner]
US 20160255373A1 · Deshpande · 2016 [cited by examiner]
US 20160366428A1 · Deshpande · 2016 [cited by examiner]
US 20170150160A1 · Deshpande · 2017 [cited by examiner]
US 20190058895A1 · Deshpande · 2019 [cited by examiner]
US 20200186833A1 · Oh · 2020 [cited by examiner]
US 20210021814A1 · Wang · 2021 [cited by examiner]
US 20210337228A1 · Wang · 2021 [cited by examiner]
CN 101548548A · 2009 [cited by applicant]
CN 104067620A · 2014 [cited by applicant]
CN 105874804A · 2016 [cited by applicant]
CN 106105210A · 2016 [cited by applicant]
JP 2016518763A · 2016 [cited by applicant]
JP 2016528804A · 2016 [cited by applicant]
JP 2017510100A · 2017 [cited by applicant]
WO 2015102959A1 · 2015 [cited by applicant]
WO 2020146665A1 · 2020 [cited by applicant]
International Telecommunication Union, “High Efficiency Video Coding”. Series H: Audiovisual and Multimedia System; Infrastructure of audiovisual services, Coding of moving video, ITU-T Recommendation H.264, Dec. 2016, … [cited by applicant]
International Search Report and Written Opinion of the International Searching Authority for PCT/US2020/022070 mailed Jun. 24, 2020, 9 pages. [cited by applicant]
International Telecommunication Union, “High Efficiency Video Coding”. Series H: Audiovisual and Multimedia Systems; cc-Coding of Moving Video, Recommendation ITU-T H.265, Telecommunication Standardization Sector of ITU… [cited by applicant]
International Preliminary Report on Patentability for PCT/US2020/022070 issued on Aug. 25, 2021, 6 pages. [cited by applicant]
Hannuksela, Miska M., et al. “AHG12: On grouping of tiles”. Joint Video Experts Team (JVET) of ITU-T SG16 WP3 and ISO/IEC JTC 1/SC29/WG11, Document No. JVET-M0261, 13th Meeting: Marrakech, MA, Jan. 9-18, 2019 (11 pages). [cited by applicant]
Hannuksela, Miska M. “AHG12/AHG17: On merging of MCTSs for viewport-dependent streaming”. Joint Video Experts Team (JVET) of ITU-T SG16 WP3 and ISO/IEC JTC 1/SC29/WG11, Document No. JVET-M0388, 13th Meeting: Marrakech, … [cited by applicant]
Ramasubramonian, Adarsh K., et al. “MV-HEVC/SHVC HLS: Sub-DPB based DPB operations”. Joint Collaborative Team on Video Coding (JCT-VC) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, Document No. JCTVC-O0217, 15th Me… [cited by applicant]
He, Yong, et al. “AHG12/AHG17: On Sub-picture parameter set”. Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, Document No. JVET-N0099, 14th Meeting: Geneva, CH, Mar. 19-27, 2019 (7 pag… [cited by applicant]
Skupin, Robert, et al. “AHG12: conformance of temporally independent tile groups”. Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, Document No. JVET-N0354, 14th Meeting: Geneva, CH, Ma… [cited by applicant]
Wang, Ye-Kui, et al. “AHG12: Harmonized proposal for sub-picture-based coding for VVC”. Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, Document No. JVET-N0826-v1, 14th Meeting: Geneva… [cited by applicant]
Skupin, Robert, et al. “AHG17: Subpicture level info for extraction and merging”. Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, Document No. JVET-P0984-v3, 16th Meeting: Geneva, CH, … [cited by applicant]
“Draft Requirements for Immersive Media Access and Delivery”. ISO/IEC JTC 1/SC 29/WG 11, Document No. N18357, Geneva, CH—Mar. 2019 (18 pages). [cited by applicant]
He, Yong, et al. “AHG12: On picture and sub-picture signaling”. Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, Document No. JVET-O0182, 15th Meeting: Gothenburg, SE, Jul. 3-12, 2019 (… [cited by applicant]
He, Yong, et al. “AHG12: On picture and sub-picture signaling”. Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, Document No. JVET-O0182r1, 15th Meeting: Gothenburg, SE, Jul. 3-12, 2019… [cited by applicant]
He, Yong, et al. “AHG12: On layer-based sub-picture extraction and reposition”. Joint Video Experts Team (JVET) of Itu-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, Document No. JVET-O0183, 15th Meeting: Gothenburg, SE, J… [cited by applicant]
He, Yong, et al. “AHG12: On sub-picture info SEI message”. Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, Document No. JVET-O0700, 15th Meeting: Gothenburg, SE, Jul. 3-12, 2019 (2 pag… [cited by applicant]
Wang, Y.K. et al., ISO/IEC JTC1/SC29/WG11 N18227-v1, “WD 4 of ISO/IEC 23090-2 OMAF 2nd Edition”, Jan. 2019 (227 pages). [cited by applicant]