Automated spatial indexing of images to video
A spatial indexing system receives a video that is a sequence of frames depicting an environment, such as a floor of a construction site, and performs a spatial indexing process to automatically identify the spatial locations at which each of the images were captured. The spatial indexing system also generates an immersive model of the environment and provides a visualization interface that allows a user to view each of the images at its corresponding location within the model.
1 . A method comprising:
generating and displaying a three-dimensional rendering of an environment based at least in part on video captured by an image capture system as the image capture system moves through the environment;
generating a content annotation associated with a location within the environment based at least in part on metadata generated by the image capture system, the metadata representative of the capture of the video by the image capture system; and
modifying the displayed three-dimensional rendering of the environment to include the content annotation displayed at the location within the three-dimensional rendering of the environment and including within the displayed content annotation an identity of a user that created the content annotation and content generated by the user for inclusion within the content annotation.
2 . The method of claim 1 , wherein the content annotation comprises metadata describing one or more of the location at which the content annotation was created, a time at which the content annotation was created, the identity of the user that created the content annotation, and the content annotation system that creates the content annotation.
3 . The method of claim 2 , wherein the location is determined based on locations described by the metadata.
4 . The method of claim 1 , wherein the content annotation is created by the user that moves the image capture system through the environment.
5 . The method of claim 1 , wherein the content annotation comprises text.
6 . The method of claim 1 , wherein the content annotation comprises a comment associated with an image captured within the environment.
7 . The method of claim 1 , wherein the content annotation comprises one or more of images, timestamps, camera orientation information, and metadata associated with the content annotation.
8 . The method of claim 1 , wherein the three-dimensional rendering of the environment is aligned with a floorplan of the environment, and wherein the floorplan specifies positions of a plurality of physical features in the environment that are included within the three-dimensional rendering of the environment.
9 . The method of claim 1 , wherein the content annotation is associated with a feature of the environment, and wherein a displayed interface is modified to include the content annotation when the feature of the environment is shown within the three-dimensional rendering of the environment.
10 . The method of claim 1 , wherein the content annotation is generated from a plurality of users, and wherein the content annotation is displayed in a feed such that more than one content annotation is visible at once.
11 . A system comprising:
a processor; and
a non-transitory computer readable storage medium comprising computer program instructions that when executed by the processor, cause the processor to:
generating and displaying a three-dimensional rendering of an environment based at least in part on video captured by an image capture system as the image capture system moves through the environment;
generating a content annotation associated with a location within the environment based at least in part on metadata generated by the image capture system, the metadata representative of the capture of the video by the image capture system; and
modifying the displayed three-dimensional rendering of the environment to include the content annotation displayed at the location within the three-dimensional rendering of the environment and including within the displayed content annotation an identity of a user that created the content annotation and content generated by the user for inclusion within the content annotation.
12 . The system of claim 11 , wherein the content annotation comprises metadata describing one or more of the location at which the content annotation was created, a time at which the content annotation was created, the identity of the user that created the content annotation, and the content annotation system that creates the content annotation.
13 . The system of claim 12 , wherein the location is determined based on locations described by the metadata.
14 . The system of claim 11 , wherein the content annotation is created by the user that moves the image capture system through the environment.
15 . The system of claim 11 , wherein the content annotation comprises text.
16 . The system of claim 11 , wherein the content annotation comprises a comment associated with an image captured within the environment.
17 . The system of claim 11 , wherein the content annotation comprises one or more of images, timestamps, camera orientation information, and metadata associated with the content annotation.
18 . The system of claim 11 , wherein the three-dimensional rendering of the environment is aligned with a floorplan of the environment, and wherein the floorplan specifies positions of a plurality of physical features in the environment that are included within the three-dimensional rendering of the environment.
19 . The system of claim 11 , wherein the content annotation is associated with a feature of the environment, and wherein a displayed interface is modified to include the content annotation when the feature of the environment is shown within the three-dimensional rendering of the environment.
20 . The system of claim 11 , wherein the content annotation is generated from a plurality of users, and wherein the content annotation is displayed in a feed such that more than one content annotation is visible at once.