IP Library Granted Patent US 12,519,947
Granted Patent B2
US 12,519,947 · App. 18/809,072 · Granted Jan 6, 2026

Techniques for bitstream extraction for subpicture in coded video stream

Inventors: Byeongdoo Choi (Palo Alto, CA); Stephan Wenger (Hillsborough, CA); Shan Liu (San Jose, CA)
Assignee: TENCENT AMERICA LLC
H04N19/132H04N19/176H04N19/30H04N19/46
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,519,947
App. No.
18/809,072
Granted
Jan 6, 2026
Kind
B2
Abstract

A method, computer program, and computer system are provided for video coding. The video coding includes decoding one or more previously encoded frames of a video source that are designated as reference frames for the video source, searching the reference frames for one or more candidate pixel blocks for an input frame of the video source, and encoding the input frame based on the one or more candidate pixel blocks.

Claims (38)

1 . A method of video decoding, executable by a processor, comprising:

receiving video data having one or more subpictures;

extracting a sub-bitstream associated with resampling parameters and spatial scalability parameters corresponding to the one or more subpictures, wherein the sub-bitstream is extracted from a bitstream associated with the video data based on an array of index values;

decoding the video data based on the resampling and spatial scalability parameters in the extracted sub-bitstream;

scaling the video data based on the resampling and spatial scalability parameters in the extracted a sub-bitstream; and

updating left and right offset values of a scaling window in a high-level syntax structure of the extracted sub-bitstream based on a left boundary position of the extracted sub-image; and

updating upper and lower offset values of the scaling window signaled in a syntax structure based on an upper boundary position of the extracted sub-bitstream.

2 . The method of claim 1 , further comprising enabling an adaptive resolution change within the received video data based on the resampling parameters.

3 . The method of claim 1 , wherein the resampling parameters correspond to one or more flags signaled in a parameter set associated with the video data.

4 . The method of claim 1 , wherein spatial scalability parameters correspond to one or more flags signaled in a parameter set associated with the video data.

5 . The method of claim 1 , wherein resampling of the video data during decoding is disabled based on the resampling parameters.

6 . The method of claim 1 , wherein the array of index values comprise at least one of a target output layer set index, a target highest temporal identification value, and an array of target subpicture index values associated with the video data.

7 . The method of claim 1 , wherein the high-level syntax structure is one of a Network Abstraction Layer (NAL) unit header, a slice header, a tile group header, a Supplementary Enhancement Information (SEI) message, parameter set, and an Access Unit (AU) delimiter.

8 . A method of video encoding, executable by a processor, comprising:

generating video data having one or more subpictures;

generating a bitstream associated with video data having one or more subpictures, wherein the bitstream comprises a sub-bitstream that includes resampling parameters and spatial scalability parameters corresponding to the one or more subpictures, wherein the sub-bitstream is extracted from a bitstream associated with the video data based on an array of index values;

scaling the video data based on the resampling and spatial scalability parameters in the sub-bitstream; and

encoding the video data based on the resampling and spatial scalability parameters in the extracted sub-bitstream,

wherein left and right offset values of a scaling window are updated in a high-level syntax structure of the extracted sub-bitstream based on a left boundary position of the extracted sub-image, and

wherein upper and lower offset values of the scaling window signaled in a syntax structure are updated based on an upper boundary position of the extracted sub-bitstream.

9 . The method of claim 8 , further comprising enabling an adaptive resolution change within the received video data based on the resampling parameters.

10 . The method of claim 8 , wherein the resampling parameters correspond to one or more flags signaled in a parameter set associated with the video data.

11 . The method of claim 8 , wherein spatial scalability parameters correspond to one or more flags signaled in a parameter set associated with the video data.

12 . The method of claim 8 , wherein resampling of the video data during decoding is disabled based on the resampling parameters.

13 . The method of claim 8 , wherein the array of index values comprise at least one of a target output layer set index, a target highest temporal identification value, and an array of target subpicture index values associated with the video data.

14 . The method of claim 8 , wherein the high-level syntax structure is one of a Network Abstraction Layer (NAL) unit header, a slice header, a tile group header, a Supplementary Enhancement Information (SEI) message, parameter set, and an Access Unit (AU) delimiter.

15 . A method of video decoding, executable by a processor, comprising:

receiving a bitstream associated with video data having one or more subpictures;

wherein a sub-bitstream associated with resampling parameters and spatial scalability parameters corresponding to the one or more subpictures is extracted from the bitstream associated with the video data based on an array of index values,

wherein the video data is decoded based on the resampling and spatial scalability parameters in the extracted sub-bitstream,

wherein the video data is scaled based on the resampling and spatial scalability parameters in the extracted a sub-bitstream,

wherein left and right offset values of a scaling window in a high-level syntax structure of the extracted sub-bitstream are updated based on a left boundary position of the extracted sub-image, and

wherein upper and lower offset values of the scaling window signaled in a syntax structure are updated based on an upper boundary position of the extracted sub-bitstream.

16 . The method of claim 15 , wherein an adaptive resolution change is enabled within the received video data based on the resampling parameters.

17 . The method of claim 15 , wherein the resampling parameters correspond to one or more flags signaled in a parameter set associated with the video data.

18 . The method of claim 15 , wherein spatial scalability parameters correspond to one or more flags signaled in a parameter set associated with the video data.

19 . The method of claim 15 , wherein resampling of the video data during decoding is disabled based on the resampling parameters.

20 . The method of claim 15 , wherein the array of index values comprise at least one of a target output layer set index, a target highest temporal identification value, and an array of target subpicture index values associated with the video data.

Continuity (4)
Continuation 18153131 · Jan 11, 2023
Continuation 17335600 · Jun 1, 2021
Provisional Application 63037202 · Jun 10, 2020
Related Publication 20240414347A1 · Dec 12, 2024
References Cited (23)
US 8665968B2 · Chen et al. · 2014 [cited by applicant]
US 10021392B2 · Puri · 2018 [cited by applicant]
US 20030076858A1 · Deshpande · 2003 [cited by applicant]
US 20130177083A1 · Chen · 2013 [cited by applicant]
US 20130188719A1 · Chen · 2013 [cited by applicant]
US 20140044161A1 · Chen · 2014 [cited by examiner]
US 20140301463A1 · Rusanovskyy · 2014 [cited by examiner]
US 20170353718A1 · Rodriguez et al. · 2017 [cited by applicant]
US 20190297339A1 · Hannuksela et al. · 2019 [cited by applicant]
US 20200404269A1 · Choi et al. · 2020 [cited by applicant]
US 20210392332A1 · Choi et al. · 2021 [cited by applicant]
US 20230275698A1 · Pan · 2023 [cited by applicant]
WO 2019162230A1 · 2019 [cited by applicant]
Communication dated May 16, 2022 from the Russian Patent Office in Russian Application No. 2021132549. [cited by applicant]
International Search Report dated Sep. 15, 2021 in International Application No. PCT/US21/36134. [cited by applicant]
Written Opinion of the International Searching Authority dated Sep. 15, 2021 in International Application No. PCT/US21/36134. [cited by applicant]
Benjamin Bross et al., “Versatile Video Coding (Draft 9)”, Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, JVET-R2001-vA, Apr. 15-24, 2020, 18th Meeting: by teleconference, 525 pages. [cited by applicant]
Hirabayashi Mitsuhiro et al., “AHG8/AHG12 Subpicture-based reference picture resampling signaling”, Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, JVET-Q0232-v1, Jan. 7-17, 2020, 17th… [cited by applicant]
Office Action issued Dec. 12, 2022 in Japanese Application No. 2021-562874. [cited by applicant]
Robert Skupin et al., “AHG12: Sei message handling in subpicture extraction”, Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, JVET-S0099, Jun. 22-Jul. 1, 2020, 19th meeting: by telecon… [cited by applicant]
Yao-Jen Chang et al., “AHG8: On scaling window constraint”, Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, JVET-S0126rl, Jun. 22-Jul. 1, 2020, 19th Meeting: by teleconference (3 pages… [cited by applicant]
Ye-Kui Wang et al., “AHG8/AHG9/AHG12: On the combination of RPR, subpictures, and scalability”, Joint Video Experts Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11, JVET-R0058- v5, Apr. 15-24, 2020, 18th M… [cited by applicant]
Extended European Search Report issued Mar. 14, 2023 in European Application No. 21785738.2. [cited by applicant]