IP Library Granted Patent US 11,477,489
Granted Patent B2
US 11,477,489 · App. 17/223,619 · Granted Oct 18, 2022

Apparatus, a method and a computer program for video coding and decoding

Inventors: Kashyap Kammachi Sreedhar (Tampere, FI); Emre Baris Aksu (Tampere, FI); Lukasz Kondrad (Munich, DE)
Assignee: Nokia Technologies Oy
H04N19/70H04N19/162H04N19/184
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,477,489
App. No.
17/223,619
Granted
Oct 18, 2022
Kind
B2
Abstract

A method comprising: writing, in a container file, a first video-based point cloud compression (V-PCC) bitstream and a second V-PCC bitstream, wherein said first and second V-PCC bitstreams are associated with a common group based on at least one logical context; writing, in the container file, an indication about the common group between the first V-PCC bitstream and the second V-PCC bitstream; generating a media presentation description (MPD) file with a first representation belonging to a first adaptation set associated with the first V-PCC bitstream and a second representation belonging to a second adaptation set associated with the second V-PCC bitstream; and writing, in the MPD file, at least one information element describing grouping information of the first representation belonging to the first adaptation set and the second representation belonging to the second adaptation set, wherein said information element is provided with at least one attribute indicating that said first and second V-PCC bitstreams are user-switchable alternatives upon rendering.

Claims (61)

1. A method comprising:

writing, in a container file, a first video-based point cloud compression bitstream and a second video-based point cloud compression bitstream, wherein said first and second video-based point cloud compression bitstreams are associated with a common group based on at least one logical context;

writing, in the container file, an indication about the common group between the first video-based point cloud compression bitstream and the second video-based point cloud compression bitstream;

generating a dynamic adaptive streaming file with a first representation belonging to a first adaptation set associated with the first video-based point cloud compression bitstream corresponding to a first object, and a second representation belonging to a second adaptation set associated with the second video-based point cloud compression bitstream corresponding to a second object; and

writing, in the dynamic adaptive streaming file, at least one information element describing grouping information of the first representation belonging to the first adaptation set and the second representation belonging to the second adaptation set, wherein said information element is provided with at least one attribute indicating that said first and second video-based point cloud compression bitstreams, corresponding respectively to the first object and the second object, are user-switchable alternatives upon rendering.

2. The method according to claim 1 , further comprising

encapsulating the first and the second video-based point cloud compression bitstream in a single-track or in a multi-track container.

3. The method according to claim 1 , wherein said indication is configured to be carried out with a syntax element defining an entity group.

4. The method according to claim 1 , wherein said indication is configured to be carried out with a syntax element defining user data.

5. The method according to claim 1 , wherein said indication is configured to be carried out with a syntax element defining a track group.

6. An apparatus comprising at least one processor and at least one non- transitory memory, said at least one memory stored with computer program code thereon, the at least one memory and the computer program code configured to, with the at least one processor, cause the apparatus at least to perform:

write, in a container file, a first video-based point cloud compression bitstream and a second video-based point cloud compression bitstream, wherein said first and second video-basedpoint cloud compression bitstreams are associated with a common group based on at least one logical context;

write, in the container file, an indication about the common group between the first video-based point cloud compression bitstream and the second video-based point cloud compression bitstream;

generate a dynamic adaptive streaming file with a first representation belonging to a first adaptation set associated with the first video-based point cloud compression bitstream corresponding to a first object, and a second representation belonging to a second adaptation set associated with the second video-based point cloud compression bitstream corresponding to a second object; and

write, in the dynamic adaptive streaming file, at least one information element describing grouping infoimation of the first representation belonging to the first adaptation set and the second representation belonging to the second adaptation set, wherein said information element is provided with at least one attribute indicating that said first and second video-based point cloud compression bitstreams, corresponding respectively to the first object and the second object, are user-switchable alternatives upon rendering.

7. The apparatus according to claim 6 , wherein the computer program code is configured to cause the apparatus to

encapsulate the first video-based point cloud compression bitstream corresponding to the first object within a first set of tracks, wherein tracks within the first set of tracks belong to a first group associated with the first object, and wherein tracks within the first set of tracks are alternatives to one another; and

encapsulate the second video-based point cloud compression bitstream corresponding to the second object within a second set of tracks, wherein tracks within the second set of tracks belong to a second group associated with the second object, and wherein tracks within the second set of tracks are alternatives to one another.

8. The apparatus according to claim 6 , wherein the computer program code is configured to cause the apparatus to

include said indication in a syntax element defining an entity group.

9. The apparatus according to claim 6 , wherein the computer program code is configured to cause the apparatus to

include said indication in a syntax element defining user data.

10. The apparatus according claim 6 , wherein the computer program code is configured to cause the apparatus to

include said indication in a syntax element defining a track group.

11. A method comprising:

receiving a bitstream comprising a media file including or inferring to a container file comprising a first video-based point cloud compression bitstream and a second video-based point cloud compression bitstream, wherein said first and second video-based point cloud compression bitstreams are associated with a common group based on at least one logical context, and an indication about the common group between the first video-based point cloud compression bitstream and the second video-based point cloud compression bitstream;

parsing, from a dynamic adaptive streaming file, a first representation belonging to a first adaptation set associated with the first video-based point cloud compression bitstream corresponding to a first object, and a second representation belonging to a second adaptation set associated with the second video-based point cloud compression bitstream corresponding to a second object;

parsing, from the dynamic adaptive streaming file, at least one information element describing grouping information of the first representation belonging to the first adaptation set and the second representation belonging to the second adaptation set, wherein said information element is provided with at least one attribute indicating that said first and second video-based point cloud compression bitstreams, corresponding respectively to the first object and the second object, are user-switchable alternatives; and

selecting either the first representation or the second representation for rendering, based on user selection of the first object or the second object.

12. The method according to claim 11 , further comprising

decapsulating the first and the second video-based point cloud compression bitstream from a single-track or from a multi-track container.

13. The method according to claim 11 , further comprising

obtaining said indication from a syntax element defining an entity group.

14. The method according to claim 11 , further comprising

obtaining said indication from a syntax element defining user data.

15. The method according to claim 11 , further comprising

obtaining said indication from a syntax element defining a track group.

16. An apparatus comprising at least one processor and at least one non-transitory memory, said at least one memory stored with computer program code thereon, the at least one memory and the computer program code configured to, with the at least one processor, cause the apparatus at least to perform:

receive a bitstream comprising a media file including or inferring to a container file comprising a first video-based point cloud compression bitstream and a second video-based point cloud compression bitstream, wherein said first and second video-based point cloud compression bitstreams are associated with a common group based on at least one logical context, and an indication about the common group between the first video-based point cloud compression bitstream and the second video-based point cloud compression bitstream;

parse, from a dynamic adaptive streaming file, a first representation belonging to a first adaptation set associated with the first video-based point cloud compression bitstream corresponding to a first object and a second representation belonging to a second adaptation set associated with the second video-based point cloud compression bitstream corresponding to a second object;

parse, from the dynamic adaptive streaming file, at least one information element describing grouping information of the first representation belonging to the first adaptation set and the second representation belonging to the second adaptation set, wherein said information element is provided with at least one attribute indicating that said first and second video-based point cloud compression bitstreams, corresponding respectively to the first object and the second object, are user-switchable alternatives; and

select either the first representation or the second representation for rendering, based on user selection of the first object or the second object.

17. The apparatus according to claim 16 , wherein the computer program code is configured to cause the apparatus to

decapsulate the first video-based point cloud compression bitstream corresponding to the first object from a first set of tracks, wherein tracks within the first set of tracks belong to a first group associated with the first object, and wherein tracks within the first set of tracks are alternatives to one another; and

decapsulate the second video-based point cloud compression bitstream corresponding to the second object from a second set of tracks, wherein tracks within the second set of tracks belong to a second group associated with the second object, and wherein tracks within the second set of tracks are alternatives to one another.

18. The apparatus according to claim 16 , wherein the computer program code is configured to cause the apparatus to

obtain said indication from a syntax element defining an entity group.

19. The apparatus according to claim 16 , wherein the computer program code is configured to cause the apparatus to

obtain said indication from a syntax element defining user data.

20. The apparatus according to claim 16 , wherein the computer program code is configured to cause the apparatus to

obtain said indication from a syntax element defining a track group.

21. A non-transitory computer-readable medium comprising program instructions stored thereon which are configured to, when executed with at least one processor, cause the at least one processor to perform:

write, in a container file, a first video-based point cloud compression bitstream and a second video-based point cloud compression bitstream, wherein said first and second video-based point cloud compression bitstreams are associated with a common group based on at least one logical context;

write, in the container file, an indication about the common group between the first video-based point cloud compression bitstream and the second video-based point cloud compression bitstream;

generate a dynamic adaptive streaming file with a first representation belonging to a first adaptation set associated with the first video-based point cloud compression bitstream corresponding to a first object, and a second representation belonging to a second adaptation set associated with the second video-based point cloud compression bitstream corresponding to a second object; and

write, in the dynamic adaptive streaming file, at least one information element describing grouping information of the first representation belonging to the first adaptation set and the second representation belonging to the second adaptation set, wherein said information element is provided with at least one attribute indicating that said first and second video-based point cloud compression bitstreams, corresponding respectively to the first object and the second object, are user-switchable alternatives upon rendering.

22. A non-transitory computer-readable medium comprising program instructions stored thereon which are configured to, when executed with at least one processor, cause the at least one processor to perform:

receive a bitstream comprising a media file including or inferring to a container file comprising a first video-based point cloud compression bitstream and a second video-based point cloud compression bitstream, wherein said first and second video-based point cloud compression bitstreams are associated with a common group based on at least one logical context, and an indication about the common group between the first video-based point cloud compression bitstream and the second video-based point cloud compression bitstream;

parse, from a dynamic adaptive streaming file, a first representation belonging to a first adaptation set associated with the first video-based point cloud compression bitstream corresponding to a first object and a second representation belonging to a second adaptation set associated with the second video-based point cloud compression bitstream corresponding to a second object;

parse, from the dynamic adaptive streaming file, at least one information element describing grouping information of the first representation belonging to the first adaptation set and the second representation belonging to the second adaptation set, wherein said information element is provided with at least one attribute indicating that said first and second video-based point cloud compression bitstreams, corresponding respectively to the first object and the second object, are user-switchable alternatives; and

select either the first representation or the second representation for rendering, based on user selection of the first object or the second object.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 6, 2021
From: KAMMACHI SREEDHAR, KASHYAP; AKSU, EMRE; KONDRAD, LUKASZ
To: NOKIA TECHNOLOGIES OY
Reel/Frame 055839/0970 →
Continuity (2)
Provisional Application 63006259 · Apr 7, 2020
Related Publication 20210314626A1 · Oct 7, 2021