IP Library › Granted Patent US 12,231,611
Granted Patent B2
US 12,231,611 · App. 18/206,346 · Granted Feb 18, 2025

Viewpoint metadata for omnidirectional video

Inventors: Yong He (San Diego, CA); Yan Ye (San Diego, CA); Ahmed Hamza (Montreal, CA)
Assignee: InterDigital Madison Patent Holdings, SAS
H04N13/178H04N13/183
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,231,611
App. No.
18/206,346
Granted
Feb 18, 2025
Kind
B2
Abstract

Methods are described herein for signaling information regarding different viewpoints in a multi-viewpoint omnidirectional media presentation. Techniques disclosed include receiving, from a server, information identifying groups of viewpoints, including information identifying a first group of viewpoints and a second group of viewpoints. And further receiving, from the server, information identifying one or more omnidirectional videos captured from respective viewpoints belonging to the first group and information identifying one or more omnidirectional videos captured from respective viewpoints belonging to the second group. Based on the received information identifying groups of viewpoints, the disclosed techniques include transitioning from a first omnidirectional video of the one or more omnidirectional videos captured from respective viewpoints belonging to the first group to a second omnidirectional video of the one or more omnidirectional videos captured from respective viewpoints belonging to the second group.

Claims (43)

1. A method comprising:

receiving, from a server, information identifying groups of viewpoints, including information identifying a first group of viewpoints and a second group of other viewpoints, wherein the viewpoints, in the groups of viewpoints, represent different locations of cameras with which corresponding omnidirectional videos are captured;

receiving, from the server, information identifying one or more omnidirectional videos captured from respective viewpoints belonging to the first group and information identifying one or more omnidirectional videos captured from respective viewpoints belonging to the second group; and

transitioning, based on the received information identifying groups of viewpoints, from a first omnidirectional video of the one or more omnidirectional videos captured from respective viewpoints belonging to the first group to a second omnidirectional video of the one or more omnidirectional videos captured from respective viewpoints belonging to the second group.

2. The method of claim 1 , wherein the one or more omnidirectional videos captured from respective viewpoints belonging to the first group cover a first event and the one or more omnidirectional videos captured from respective viewpoints belonging to the second group cover a second event.

3. The method of claim 2 , wherein the first and the second events occur at the same time.

4. The method of claim 2 , wherein the first and the second events occur at different venues.

5. The method of claim 1 , wherein, for a group of viewpoints, the received information identifying the groups of viewpoints comprises:

an identity value, identifying the group of viewpoints.

6. The method of claim 1 , wherein, for a group of viewpoints, the received information identifying the groups of viewpoints comprises:

a location value, identifying the location of an event covered by one or more omnidirectional videos captured by respective viewpoints of the group.

7. The method of claim 6 , wherein, the location value includes a longitude value and a latitude value of a geolocation of the event.

8. The method of claim 1 , wherein, for a group of viewpoints, the received information identifying the groups of viewpoints comprises:

a number value, identifying the number of viewpoints in the group.

9. The method of claim 1 , wherein, for a group of viewpoints, the received information identifying the groups of viewpoints comprises:

an array of identity values, identifying respective viewpoints in the group.

10. The method of claim 1 , wherein, for a group of viewpoints, the received information identifying the groups of viewpoints comprises:

a string, identifying a name describing the group.

11. A system, comprising:

at least one processor; and

memory storing instructions that, when executed by the at least one processor, cause the system to:

receive, from a server, information identifying groups of viewpoints, including information identifying a first group of viewpoints and a second group of other viewpoints, wherein the viewpoints, in the groups of viewpoints, represent different locations of cameras with which corresponding omnidirectional videos are captured;

receive, from the server, information identifying one or more omnidirectional videos captured from respective viewpoints belonging to the first group and information identifying one or more omnidirectional videos captured from respective viewpoints belonging to the second group; and

transition, based on the received information identifying groups of viewpoints, from a first omnidirectional video of the one or more omnidirectional videos captured from respective viewpoints belonging to the first group to a second omnidirectional video of the one or more omnidirectional videos captured from respective viewpoints belonging to the second group.

12. The system of claim 11 , wherein the one or more omnidirectional videos captured from respective viewpoints belonging to the first group cover a first event and the one or more omnidirectional videos captured from respective viewpoints belonging to the second group cover a second event.

13. The system of claim 12 , wherein the first and the second events occur at the same time, occur at different venues, or occur at the same time at different venues.

14. The system of claim 11 , wherein, for a group of viewpoints, the received information identifying the groups of viewpoints comprises:

an identity value, identifying the group of viewpoints.

15. The system of claim 11 , wherein, for a group of viewpoints, the received information identifying the groups of viewpoints comprises:

a location value, identifying the location of an event covered by one or more omnidirectional videos captured by respective viewpoints of the group.

16. The system of claim 15 , wherein, the location value includes a longitude value and a latitude value of a geolocation of the event.

17. The system of claim 11 , wherein, for a group of viewpoints, the received information identifying the groups of viewpoints comprises:

a number value, identifying the number of viewpoints in the group.

18. The system of claim 11 , wherein, for a group of viewpoints, the received information identifying the groups of viewpoints comprises:

an array of identity values, identifying respective viewpoints in the group.

19. The system of claim 11 , wherein, for a group of viewpoints, the received information identifying the groups of viewpoints comprises:

a string, identifying a group name describing the group.

20. A non-transitory computer-readable medium comprising instructions executable by at least one processor to perform a method, the method comprising:

receiving, from a server, information identifying groups of viewpoints, including information identifying a first group of viewpoints and a second group of other viewpoints, wherein the viewpoints, in the groups of viewpoints, represent different locations of cameras with which corresponding omnidirectional videos are captured;

receiving, from the server, information identifying one or more omnidirectional videos captured from respective viewpoints belonging to the first group and information identifying one or more omnidirectional videos captured from respective viewpoints belonging to the second group; and

transitioning, based on the received information identifying groups of viewpoints, from a first omnidirectional video of the one or more omnidirectional videos captured from respective viewpoints belonging to the first group to a second omnidirectional video of the one or more omnidirectional videos captured from respective viewpoints belonging to the second group.

21. The method of claim 1 , wherein the viewpoints, in the groups of viewpoints, are grouped based on a geolocation of an event covered by the corresponding omnidirectional videos.

22. The system of claim 11 , wherein the viewpoints, in the groups of viewpoints, are grouped based on a geolocation of an event covered by the corresponding omnidirectional videos.

Continuity (4)
Continuation 17045104
Provisional Application 62675524 · May 23, 2018
Provisional Application 62653363 · Apr 5, 2018
Related Publication 20230319251A1 · Oct 5, 2023
References Cited (55)
US 9294757B1 · Lewis · 2016 [cited by examiner]
US 11093752B2 · Kim et al. · 2021 [cited by applicant]
US 11573632B2 · Canberk · 2023 [cited by examiner]
US 11710288B2 · Handa · 2023 [cited by examiner]
US 20010056396A1 · Goino · 2001 [cited by applicant]
US 20080049123A1 · Gloudemans et al. · 2008 [cited by applicant]
US 20080189371A1 · Shaffer · 2008 [cited by examiner]
US 20090128667A1 · Gloudemans · 2009 [cited by examiner]
US 20110161875A1 · Kankainen · 2011 [cited by applicant]
US 20110182366A1 · Froejdh et al. · 2011 [cited by applicant]
US 20110202575A1 · Froejdh et al. · 2011 [cited by applicant]
US 20110261050A1 · Smolic et al. · 2011 [cited by applicant]
US 20130135315A1 · Bares · 2013 [cited by examiner]
US 20140002439A1 · Lynch · 2014 [cited by applicant]
US 20140043340A1 · Sobhy et al. · 2014 [cited by applicant]
US 20140189772A1 · Yamagishi · 2014 [cited by examiner]
US 20170347026A1 · Hannuksela · 2017 [cited by applicant]
US 20180046363A1 · Miller · 2018 [cited by examiner]
US 20180063505A1 · Lee et al. · 2018 [cited by applicant]
US 20180143756A1 · Mildrew et al. · 2018 [cited by applicant]
US 20180150989A1 · Mitsui · 2018 [cited by examiner]
US 20180241988A1 · Zhou · 2018 [cited by examiner]
US 20180247463A1 · Rekimoto · 2018 [cited by applicant]
US 20190075269A1 · Nashida · 2019 [cited by applicant]
US 20190102940A1 · Nakao et al. · 2019 [cited by applicant]
US 20190104316A1 · Da et al. · 2019 [cited by applicant]
US 20190141359A1 · Taquet et al. · 2019 [cited by applicant]
US 20190246146A1 · Bustamante et al. · 2019 [cited by applicant]
US 20190287302A1 · Bhuruth · 2019 [cited by applicant]
US 20190306530A1 · Fan et al. · 2019 [cited by applicant]
US 20190313081A1 · Oh · 2019 [cited by applicant]
US 20190327425A1 · Kobayashi et al. · 2019 [cited by applicant]
US 20190387271A1 · Ojiro et al. · 2019 [cited by applicant]
US 20210029294A1 · Deshpande · 2021 [cited by examiner]
US 20210035535A1 · Kanda · 2021 [cited by examiner]
US 20210084346A1 · Tsukagoshi · 2021 [cited by examiner]
US 20210092466A1 · Suzuki · 2021 [cited by examiner]
CN 102246491A · 2011 [cited by applicant]
JP 2012505570A · 2012 [cited by applicant]
JP 2017184114A · 2017 [cited by applicant]
WO WO2017159063A1 · 2017 [cited by applicant]
WO WO2018027067A1 · 2018 [cited by applicant]
International Organization for Standardization, “Information technology—Coding of audio-Visual Objects—Part 12: ISO Base Media File Format; AMENDMENT 1: General Improvements Including Hint Tracks, Metadata Support And S… [cited by applicant]
International Search Report and Written Opinion of the International Searching Authority for PCT/US2019/025784 mailed Jun. 6, 2019, 12 pages. [cited by applicant]
Domański, M. et al., “Extended VSRS for 360 degree video,” ISO/IEC JTC1/SC29 WG11 Doc. MPEG M41990, Gwangju, Jan. 2018 (5 pages). [cited by applicant]
Rosenthal, Paul, et. al. “Image-Space Point Cloud Rendering”. Proceedings of Computer Graphics International, (2008) pp. 1-8. [cited by applicant]
Koenen et al., “Requirements MPEG-I phase 1b”, International Organisation for Standarisation (ISO), Coding of Moving Pictures and Audio, ISO/IEC JTC1/SC29/WG11, Document: N17331, Gwanghu, Korea, Jan. 2018, 8 pages. [cited by applicant]
“Revised text of ISO/IEC FDIS 23090-2 Omnidirectional Media Format”, 121.MPEG METING ;Jan. 22, 2018-Jan. 26, 2018; Gwangju; (Motion Picture Expert Group or ISO/IEC JTC1/SC29/WG11), No. N17399, Feb. 28, 2018, XP030024044. [cited by applicant]
International Standard, “Information technology—Coded representation of immersive media (MPEG-I)—Part 2: Omnidirectional Media Format”. ISO/IEC JTC 1/SC 29/WG11, ISO/IEC FDIS 23090-2, Feb. 7, 2018, 182 pages. [cited by applicant]
International Organization for Standardization, “Information Technology—Dynamic Adaptive Streaming Over HTTP (DASH), Part 1: Media Presentation Description and Segment Formats”. International Standard, ISO/IEC 23009-1, … [cited by applicant]
Fehn, Christoph. “Depth-Image-Based Rendering (DIBR), Compression, And Transmission For A New Approach on 3D-TV”. Stereoscopic Displays and Virtual Reality Systems XI, International Society for Optics and Photonics, vol… [cited by applicant]
Lee et al., “Comment on Timed Metadata Signaling for DASH in OMAF DIS”. LG Electronics, International Organization for Standardization, Coding of Moving Pictures and Audio, Motion Picture Expert Group (MPEG), ISO/IEC JT… [cited by applicant]
International Preliminary Report on Patentability for PCT/US2019/025784 completed on Jun. 26, 2020, 10 pages. [cited by applicant]
Fehn, Christoph: “Depth-image-based Rendering (DIBR), Compression and Transmission for a New Approach on 3D-TV”, Proceedings vol. 5291, Stereoscopic Displays and Virtual Reality Systems XI; (2004), 12 pages. [cited by applicant]
Written Opinion of the International Preliminary Examining Authority for PCT/US2019/025784 mailed Mar. 12, 2020, 7 pages. [cited by applicant]