SYSTEMS AND METHODS FOR COHERENT MONITORING
Systems and methods are provided for intelligently monitoring environments, classifying objects within such environments, detecting events within such environments, receiving and propagating input concerning image information from multiple users in a collaborative environment, identifying and responding to situational abnormalities or situations of interest based on such detections and/or user inputs.
1 . A system for intelligently monitoring an environment, comprising:
one or more processors; and
a memory storing instructions that, when executed by the one or more processors, cause the system to:
obtain content representing an environment, the content comprising a plurality of video frames;
identify, based on the content, one or more discrete objects or events observed within the environment;
generate a media representation that augments the one or more discrete objects or events; and
augment a graphical representation of the environment with the generated media representation.
2 . The system of claim 1 , wherein the media representation comprises different frames corresponding to different perspectives of the one or more discrete objects or events captured by sensors at different orientations and positions.
3 . The system of claim 1 , wherein the generating of the media representation comprises automatically tagging the one or more discrete objects or events based on a confidence level of the one or more discrete objects or events matching one or more respective templates that define respective characteristics of the one or more discrete objects or events.
4 . The system of claim 3 , wherein the instructions that, when executed by the one or more processors, further cause the system to:
obtain a missed detection of the one or more discrete objects or events; and
based on the missed detection, adjust a criteria of detecting the one or more discrete objects or events based on features of the one or more discrete objects or events.
5 . The system of claim 1 , wherein instructions that, when executed by the one or more processors, further causes the system to:
present the graphical representation of the environment; and
simultaneously present a playback of the one or more discrete objects or events in a separate pane.
6 . The system of claim 1 , wherein the memory stored instructions that, when executed by the one or more processors, further causes the system to:
identify an operational status of the one or more discrete objects or events; and overlay an indication of the operational status of the one or more discrete objects or events.
7 . The system of claim 4 , wherein the adjusting of the criteria comprises adjusting the one or more respective templates.
8 . The system of claim 1 , wherein the detecting one or more events associated with the tracked object comprises:
detecting changes in the environment;
identifying candidate events based on the detected changes; and
comparing the candidate events with templates while accounting for a scaling, angle, or orientation difference between the object and corresponding objects in the templates.
9 . The system of claim 1 , wherein the instructions further cause the system to:
determine a view field of a sensor capturing the content; and
display an indication of the determined view field.
10 . The system of claim 1 , wherein the augmenting of the graphical representation of the environment with the generated media representation comprises presenting a snapshot of a particular frame corresponding to the one or more detected events or objects, wherein the snapshot corresponds to a flagged event.
11 . A method being implemented by a computing system including one or more physical processors and storage media storing machine-readable instructions, the method comprising:
obtaining content representing an environment, the content comprising a plurality of video frames;
identifying, based on the content, one or more discrete objects or events observed within the environment;
generating a media representation that augments the one or more discrete objects or events; and
augmenting a graphical representation of the environment with the generated media representation.
12 . The method of claim 11 , wherein the media representation comprises different frames corresponding to different perspectives of the one or more discrete objects or events captured by sensors at different orientations and positions.
13 . The method of claim 11 , wherein the generating of the media representation comprises automatically tagging the one or more discrete objects or events based on a confidence level of the one or more discrete objects or events matching one or more respective templates that define respective characteristics of the one or more discrete objects or events.
14 . The method of claim 13 , further comprising:
obtaining a missed detection of the one or more discrete objects or events; and
based on the missed detection, adjusting a criteria of detecting the one or more discrete objects or events based on features of the one or more discrete objects or events.
15 . The method of claim 11 , further comprising:
presenting the graphical representation of the environment; and
simultaneously presenting a playback of the one or more discrete objects or events in a separate pane.
16 . The method of claim 11 , further comprising:
identifying an operational status of the one or more discrete objects or events; and
overlaying an indication of the operational status of the one or more discrete objects or events.
17 . The method of claim 14 , wherein the adjusting of the criteria comprises adjusting the one or more respective templates.
18 . The method of claim 11 , wherein the detecting one or more events associated with the tracked object comprises:
detecting changes in the environment;
identifying candidate events based on the detected changes; and
comparing the candidate events with templates while accounting for a scaling, angle, or orientation difference between the object and corresponding objects in the templates.
19 . The method of claim 11 , further comprising:
determining a view field of a sensor capturing the content; and
displaying an indication of the determined view field.
20 . The method of claim 11 , wherein the augmenting of the graphical representation of the environment with the generated media representation comprises presenting a snapshot of a particular frame corresponding to the one or more detected events or objects, wherein the snapshot corresponds to a flagged event.