Method for parameter set reference constraints in coded video stream
There is included a method and apparatus comprising computer code configured to cause a processor or processors to perform obtaining video data comprising data of a plurality of semantically independent source pictures, determining, among the video data, whether references are associated with any of a first access unit (AU) and a second AU according to at least one picture order count (POC) signal value included with the video data, and outputting a first quantity of the references set to the first AU and a second quantity of the references set to the second AU based on the at least one POC signal value.
1 . A method for video decoding in a decoder, the method comprising:
obtaining video data comprising data of a plurality of pictures, and of which a same value space of values of adaptation_parameter_set_id and aps_params_type is shared across spatial layers of the video data regardless of nuh_layer_id values, and of which a same value space of sps_video_parameter_set_id is shared in a sequence parameter set;
determining, among the video data, whether references are associated with any of a first access unit (AU) and a second AU based on comparing respective POC values with the POC signal value, determining that the at least first one of the references comprises a first value that is less than the POC signal value, and determining that the at least second one of the references comprises a second value that is equal to or greater than the POC signal value; and
decoding the video data based on at least a first one of the references being set to the first AU and at least a second one of the references being set to the second AU based on a picture order count (POC) signal value.
2 . The method according claim 1 ,
wherein the references comprise at least one of pictures, slices, and tiles of the video data.
3 . The method according to claim 1 ,
wherein the references comprise the slices, and
wherein the POC signal value is included in a slice header of the video data.
4 . The method according to claim 1 , wherein outputting the at least first one of the references set to the first AU and the at least second one of the references set to the second AU based on the POC signal value is further based on whether a quantity of the references comprises one of a plurality of POC values that is less than the POC signal value.
5 . The method according to claim 1 ,
wherein the POC signal value is included in a video parameter set (VPS) of the video data, and
wherein the method further comprises determining whether the VPS data comprises at least one flag indicating whether one or more of the references are divided into a plurality of sub-regions.
6 . The method according to claim 5 , further comprising:
determining, in a case where the at least one flag indicates that the one or more of the references are divided into the plurality of sub-regions, a signaling value, included in the sequence parameter set of the video data, specifying an offset of a portion of at least one of the sub-regions.
7 . The method according to claim 1 ,
wherein the pictures represent a spherical 360 picture.
8 . A method for video encoding, the method comprising:
obtaining video data comprising data of a plurality of pictures; and
encoding the video data such that a same value space of values of adaptation_parameter_set_id and aps_params_type is shared across spatial layers of the video data regardless of nuh_layer_id values, and of which a same value space of sps_video_parameter_set_id is shared in a sequence parameter set, and
encoding the video data further comprises determining, among the video data, whether references are associated with any of a first access unit (AU) and a second AU based on comparing respective POC values with the POC signal value, determining that the at least first one of the references comprises a first value that is less than the POC signal value, and determining that the at least second one of the references comprises a second value that is equal to or greater than the POC signal value, and based on at least a first one of the references being set to the first AU and at least a second one of the references being set to the second AU based on a picture order count (POC) signal value.
9 . The method according claim 8 ,
wherein the references comprise at least one of pictures, slices, and tiles of the video data.
10 . The method according to claim 8 ,
wherein the references comprise the slices, and
wherein the POC signal value is included in a slice header of the video data.
11 . The method according to claim 8 , wherein the at least first one of the references set to the first AU and the at least second one of the references set to the second AU based on the POC signal value is further based on whether a quantity of the references comprises one of a plurality of POC values that is less than the POC signal value.
12 . The method according to claim 8 ,
wherein the POC signal value is included in a video parameter set (VPS) of the video data, and
wherein the method further comprises determining whether the VPS data comprises at least one flag indicating whether one or more of the references are divided into a plurality of sub-regions.
13 . The method according to claim 12 , further comprising:
determining, in a case where the at least one flag indicates that the one or more of the references are divided into the plurality of sub-regions, a signaling value, included in the sequence parameter set of the video data, specifying an offset of a portion of at least one of the sub-regions.
14 . The method according to claim 8 ,
wherein the pictures represent a spherical 360 picture.
15 . A method of encoding visual media data, the method comprising:
generating a bitstream of the visual media data according to an encoding process;
the bitstream of the visual media data comprising encoded video data such that a same value space of values of adaptation_parameter_set_id and aps_params_type is shared across spatial layers of the video data regardless of nuh_layer_id values, and a same value space of sps_video_parameter_set_id is shared in a sequence parameter set, and
the encoding process comprises encoding and transmitting the video data based on determining, among the video data, whether references are associated with any of a first access unit (AU) and a second AU based on comparing respective POC values with the POC signal value, determining that the at least first one of the references comprises a first value that is less than the POC signal value, and determining that the at least second one of the references comprises a second value that is equal to or greater than the POC signal value, and based on at least a first one of the references being set to the first AU and at least a second one of the references being set to the second AU based on a picture order count (POC) signal value, and
transmitting the generated bitstream of the visual media data.
16 . The method according claim 15 ,
wherein the references comprise at least one of pictures, slices, and tiles of the video data.
17 . The method according to claim 15 ,
wherein the references comprise the slices, and
wherein the POC signal value is included in a slice header of the video data.
18 . The method according to claim 15 , wherein the at least first one of the references set to the first AU and the at least second one of the references set to the second AU based on the POC signal value is further based on whether a quantity of the references comprises one of a plurality of POC values that is less than the POC signal value.
19 . The method according to claim 15 ,
wherein the POC signal value is included in a video parameter set (VPS) of the video data, and
wherein the method further comprises determining whether the VPS data comprises at least one flag indicating whether one or more of the references are divided into a plurality of sub-regions.
20 . The method according to claim 19 , wherein he video encoding method further comprises determining, in a case where the at least one flag indicates that the one or more of the references are divided into the plurality of sub-regions, a signaling value, included in the sequence parameter set of the video data, specifying an offset of a portion of at least one of the sub-regions.