IP Library › Granted Patent US 12,537,909
Granted Patent B2
US 12,537,909 · App. 17/872,564 · Granted Jan 27, 2026

Virtual field of view adjustment in live volumetric video

Inventors: Aaron K. Baughman (Cary, NC); Micah Forster (Round Rock, TX); Ashafaaq Minhas (Jersey City, NJ); Sarbajit K. Rakshit (Kolkata, IN)
Assignee: INTERNATIONAL BUSINESS MACHINES CORPORATION
H04N5/2723G06T11/60G06T19/006H04N13/117H04N13/243G06T2207/20221
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,537,909
App. No.
17/872,564
Granted
Jan 27, 2026
Kind
B2
Abstract

Video of a plurality of fields of view of a scene is captured, each field of view comprising data of the scene from a different vantage point. An excitement level is determined by analyzing a portion of the captured video. Using the excitement level, a time series of future excitement levels is forecast. Using the time series of future excitement levels, a virtual field of view path of the scene is forecast. An insert image is determined to be included in the virtual field of view path. Captured data from at least two video cameras in the plurality of video cameras is composited into a virtual field of view of the scene. A rendering of the insert image is inserted into the virtual field of view.

Claims (61)

1 . A computer-implemented method comprising:

capturing, using a plurality of video cameras, video of a plurality of fields of view of a scene, each field of view comprising data of the scene from a different vantage point, the capturing resulting in captured video of the plurality of fields of view;

determining, by analyzing a portion of the captured video, an excitement level;

forecasting, using the excitement level, a time series of future excitement levels;

forecasting, using the time series of future excitement levels, a virtual field of view path of the scene;

determining that an insert image is included in the virtual field of view path;

compositing, into a virtual field of view of the scene, captured data from at least two video cameras in the plurality of video cameras; and

inserting, into the virtual field of view, a rendering of the insert image.

2 . The computer-implemented method of claim 1 , further comprising:

determining, by analyzing a portion of audio data of a vicinity of the scene, an audio-based excitement level corresponding to the scene; and

incorporating, into the excitement level, the audio-based excitement level.

3 . The computer-implemented method of claim 1 , further comprising:

identifying, by analyzing a portion of the captured video, an object within the scene; and

determining an object excitement level corresponding to the object.

4 . The computer-implemented method of claim 3 , further comprising:

determining, by analyzing a frame of the captured video, a frame excitement level; and

incorporating, into the object excitement level, the frame excitement level.

5 . The computer-implemented method of claim 3 , further comprising:

determining, by analyzing event data of an event occurring in the scene, an event excitement level; and

incorporating, into the object excitement level, the event excitement level.

6 . The computer-implemented method of claim 1 , further comprising:

removing, from the virtual field of view, a second object with a transparency level above a threshold transparency level, the transparency level set according to a second excitement level corresponding to the second object.

7 . The computer-implemented method of claim 1 , further comprising:

rendering, in the virtual field of view according to a transparency level of a third object, the third object, the transparency level of the third object set according to a third excitement level corresponding to the third object.

8 . The computer-implemented method of claim 1 , wherein the rendering is generated with a specified maximum angular offset.

9 . The computer-implemented method of claim 1 , wherein the rendering is generated with an angular offset determined according to a payment amount.

10 . The computer-implemented method of claim 1 , wherein the rendering is generated with an angular offset determined according to an amount of computing resources required to render the insert image at a specified angular offset.

11 . A computer program product for virtual field of view adjustment in volumetric video, the computer program product comprising:

one or more computer readable storage media, and program instructions collectively stored on the one or more computer readable storage media, the stored program instructions comprising:

program instructions to capture, using a plurality of video cameras, video of a plurality of fields of view of a scene, each field of view comprising data of the scene from a different vantage point, the capturing resulting in captured video of the plurality of fields of view;

program instructions to determine, by analyzing a portion of the captured video, an excitement level;

program instructions to forecast, using the excitement level, a time series of future excitement levels;

program instructions to forecast, using the time series of future excitement levels, a virtual field of view path of the scene;

program instructions to determine that an insert image is included in the virtual field of view path;

program instructions to composite, into a virtual field of view of the scene, captured data from at least two video cameras in the plurality of video cameras; and

program instructions to insert, into the virtual field of view, a rendering of the insert image.

12 . The computer program product of claim 11 , the stored program instructions further comprising:

program instructions to determine, by analyzing a portion of audio data of a vicinity of the scene, an audio-based excitement level corresponding to the scene; and

program instructions to incorporate, into the excitement level, the audio-based excitement level.

13 . The computer program product of claim 11 , the stored program instructions further comprising:

program instructions to identify, by analyzing a portion of the captured video, an object within the scene; and

program instructions to determine an object excitement level corresponding to the object.

14 . The computer program product of claim 13 , the stored program instructions further comprising:

program instructions to determine, by analyzing a frame of the captured video, a frame excitement level; and

program instructions to incorporate, into the object excitement level, the frame excitement level.

15 . The computer program product of claim 13 , the stored program instructions further comprising:

program instructions to determine, by analyzing event data of an event occurring in the scene, an event excitement level; and

program instructions to incorporate, into the object excitement level, the event excitement level.

16 . The computer program product of claim 11 , the stored program instructions further comprising:

program instructions to remove, from the virtual field of view, a second object with a transparency level above a threshold transparency level, the transparency level set according to a second excitement level corresponding to the second object.

17 . The computer program product of claim 11 , wherein the stored program instructions are stored in the at least one of the one or more storage media of a local data processing system, and wherein the stored program instructions are transferred over a network from a remote data processing system.

18 . The computer program product of claim 11 , wherein the stored program instructions are stored in the at least one of the one or more storage media of a server data processing system, and wherein the stored program instructions are downloaded over a network to a remote data processing system for use in a computer readable storage device associated with the remote data processing system.

19 . The computer program product of claim 11 , wherein the computer program product is provided as a service in a cloud environment.

20 . A computer system comprising one or more processors, one or more computer-readable memories, and one or more computer-readable storage media, and program instructions stored on at least one of the one or more storage media for execution by at least one of the one or more processors via at least one of the one or more memories, the stored program instructions comprising:

program instructions to capture, using a plurality of video cameras, video of a plurality of fields of view of a scene, each field of view comprising data of the scene from a different vantage point, the capturing resulting in captured video of the plurality of fields of view;

program instructions to determine, by analyzing a portion of the captured video, an excitement level;

program instructions to forecast, using the excitement level, a time series of future excitement levels;

program instructions to forecast, using the time series of future excitement levels, a virtual field of view path of the scene;

program instructions to determine that an insert image is included in the virtual field of view path;

program instructions to composite, into a virtual field of view of the scene, captured data from at least two video cameras in the plurality of video cameras; and

program instructions to insert, into the virtual field of view, a rendering of the insert image.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jul 25, 2022
From: BAUGHMAN, AARON K.; FORSTER, MICAH; MINHAS, ASHAFAAQ; RAKSHIT, SARBAJIT K.
To: INTERNATIONAL BUSINESS MACHINES CORPORATION
Reel/Frame 060607/0635 →
Continuity (1)
Related Publication 20240031519A1 · Jan 25, 2024
References Cited (36)
US 9060210B2 · Packard · 2015 [cited by examiner]
US 10250932B2 · Osminer · 2019 [cited by examiner]
US 10341632B2 · Pang et al. · 2019 [cited by applicant]
US 10469873B2 · Pang et al. · 2019 [cited by applicant]
US 10567464B2 · Pang et al. · 2020 [cited by applicant]
US 11025985B2 · Stojancic · 2021 [cited by examiner]
US 11341543B2 · Govindgari · 2022 [cited by applicant]
US 20170026577A1 · You · 2017 [cited by examiner]
US 20170164015A1 · Abramov · 2017 [cited by examiner]
US 20170244948A1 · Pang et al. · 2017 [cited by applicant]
US 20170352191A1 · Zhou · 2017 [cited by examiner]
US 20180035134A1 · Pang et al. · 2018 [cited by applicant]
US 20180097867A1 · Pang et al. · 2018 [cited by applicant]
US 20190213423A1 · Haberstroh · 2019 [cited by examiner]
US 20200334833A1 · Gibbon · 2020 [cited by examiner]
US 20220066537A1 · Govindgari · 2022 [cited by applicant]
US 20220066550A1 · Govindgari · 2022 [cited by applicant]
US 20220067792A1 · Govindgari · 2022 [cited by applicant]
US 20220351519A1 · Kalirajan · 2022 [cited by examiner]
US 20220368964A1 · Uzaki · 2022 [cited by examiner]
US 20230063505A1 · Chastain · 2023 [cited by examiner]
US 20230328320A1 · Krishnamoorthi · 2023 [cited by examiner]
US 20240007716A1 · Panchaksharaiah · 2024 [cited by examiner]
JP 2020174971A · 2020 [cited by examiner]
Author: Merler et al.; Title: Automatic Curation of Sports Highlights using Multimodal Excitement Features; Publisher: IEEE, (Year: 2018). [cited by examiner]
Author: Merler et al.; Title: Automatic Curation of Golf Highlights using Multimodal Excitement Features; Publisher: Computer Vision Foundation (Year: 2017). [cited by examiner]
Canon Inc., Canon opens Volumetric Video Studio—Kawasaki, provides a new kind of visual experience for the entertainment industry, Sep. 1, 2020, https://global.canon/en/news/2020/20200901.html. [cited by applicant]
IBM, Canon, Inc. and IBM Launch Collaboration in Entertainment and the Arts in Japan, Jul. 15, 2021, https://newsroom.ibm.com/2021-07-15-Canon,-Inc-and-IBM-Launch-Collaboration-in-Entertainment-and-the-Arts-in-Japan. [cited by applicant]
Cohen, Welcome to the Netaverse, Where Brooklyn Nets Players Can Be Seen in a Whole New (3D) Light, Feb. 3, 2022, https://www.sporttechie.com/welcome-to-the-netaverse-where-brooklyn-nets-players-can-be-seen-in-a-whole-n… [cited by applicant]
Chen et al., Accelerated Stimulated Raman Projection Tomography by Sparse Reconstruction From Sparse-View Data, IEEE Transactions on Biomedical Engineering, vol. 67, No. 5, May 2020. [cited by applicant]
Kawamura et al., Real-Time Streaming of Sequential Volumetric Data for Augmented Reality Synchronized with Broadcast Video, 2019 IEEE 9th International Conference on Consumer Electronics (ICCE-Berlin), Sep. 8-11, 2019. [cited by applicant]
Sheikhi-Pour et al., Efficient 2D Video Coding of Volumetric Video Data, 2018 7th European Workshop on Visual Information Processing (EUVIP), Nov. 26-28, 2018. [cited by applicant]
Gul et al., Low-latency Cloud-based Volumetric Video Streaming Using Head Motion Prediction, NOSSDAV '20: Proceedings of the 30th ACM Workshop on Network and Operating Systems Support for Digital Audio and Video, Jun. 8… [cited by applicant]
Qian et al., Toward Practical Volumetric Video Streaming on Commodity Smartphones, HotMobile '19: Proceedings of the 20th International Workshop on Mobile Computing Systems and Applications, Feb. 27-28, 2019, pp. 135-14… [cited by applicant]
Son et al., Split Rendering for Mixed Reality: Interactive Volumetric Video in Action, SA '20: SIGGRAPH Asia 2020 XR, Dec. 4-13, 2020, No. 8, pp. 1-3. [cited by applicant]
ip.com, Contextualized Ad Generation via a Reinforced Con-GANs Framework, Mar. 14, 2022. [cited by applicant]