IP Library Granted Patent US 11,100,784
Granted Patent B2
US 11,100,784 · App. 16/808,543 · Granted Aug 24, 2021

Method and system for detecting and notifying actionable events during surveillance

Inventors: Subhas Chandra Mondal (Bangalore, IN); Vishal Kumar Pandey (Tempe, AZ)
Assignee: Wipro Limited
G08B25/006G08B29/183G16Y40/60H04N7/185
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,100,784
App. No.
16/808,543
Granted
Aug 24, 2021
Kind
B2
Abstract

The disclosure relates to method and system for detecting and notifying actionable events during surveillance. The method may include receiving initial multi-modal inputs from a geo-location during surveillance, determining an incident of interest based on an analysis of the initial multi-modal inputs, and collecting additional multi-modal inputs from at least one access device corresponding to at least one person in the geo-location upon determination of the incident of interest. The method may further include determining the actionable event based on an analysis of the initial and the additional multi-modal inputs, and providing a notification of the actionable event to one or more appropriate authorities.

Claims (66)

1. A method for detecting and notifying an actionable event during surveillance, the method comprising:

receiving initial multi-modal inputs from a geo-location during surveillance;

determining an incident of interest based on an analysis of the initial multi-modal inputs, wherein the determination of an incident of interest is based on:

generating a confidence score for each of a plurality of incidents of interest determined over a period of time, wherein the confidence score is based on a criticality of the actionable event;

creating a catalogue of the plurality of incidents of interest based on their respective confidence scores and actionable events, wherein the created catalogue is stored in form of a knowledge graph which is referred for decision making while evaluating an incident of interest from the plurality of incidents of interest; and

utilizing the catalogue for evaluating a new incident of interest;

collecting additional multi-modal inputs from at least one access device corresponding to at least one person in the geo-location upon determination of the incident of interest;

determining the actionable event based on an analysis of the initial and the additional multi-modal inputs; and

providing a notification of the actionable event to one or more appropriate authorities.

2. The method of claim 1 , wherein receiving the initial multi-modal inputs comprises receiving the initial multi-modal inputs from one or more surveillance devices, wherein the one or more surveillance devices comprise at least one of a closed-circuit television (CCTV) camera, an Internet Protocol (IP) camera, a microphone, an Internet-of-Things (IoT) sensor, a mobile device, a hand-held device, or a wearable device; and wherein the at least one access device comprises a mobile device, a hand-held device, or a wearable device.

3. The method of claim 1 , further comprising:

identifying a set of persons in the geo-location upon determination of the incident of interest by performing at least one of a facial recognition of a set of faces or a voice recognition of a set of voices in the initial multi-modal inputs against a plurality of persons in a population register;

determining a set of access devices corresponding to the set of persons from the population register; and

activating the at least one access device from the set of access devices for collecting additional multi-modal inputs.

4. The method of claim 1 , further comprising:

determining a plurality of access devices in the geo-location upon determination of the incident of interest based on inputs from one or more network operators;

identifying a plurality of persons corresponding to the plurality of access devices from a population register;

identifying a set of persons by performing at least one of a facial recognition of a set of faces or a voice recognition of a set of voices in the multi-modal inputs against the plurality of persons in the population register; and

activating the at least one access device corresponding to the at least one person from the set of persons for collecting additional multi-modal inputs.

5. The method of claim 1 , wherein the at least one access device is activated for collecting additional multi-modal inputs upon at least one of: a notification to the at least one person, and a permission from the at least one person.

6. The method of claim 1 , wherein determining the actionable event comprises correlating the initial and the additional multi-modal inputs so as to validate the incident of interest and gathering specific inputs with respect to the incident of interest.

7. The method of claim 1 , further comprising providing one or more recommendations to the one or more appropriate authorities based on the actionable event.

8. A system for detecting and notifying an actionable event during surveillance, the system comprising:

an edge server comprising a processor and a memory communicatively coupled to the processor, wherein the memory stores processor-executable instructions, which, on execution, causes the processor to:

receive initial multi-modal inputs from a geo-location during surveillance;

determine an incident of interest based on an analysis of the initial multi-modal inputs, wherein the determination of an incident of interest is based on:

generating a confidence score for each of a plurality of incidents of interest determined over a period of time, wherein the confidence score is based on a criticality of the actionable event;

creating a catalogue of the plurality of incidents of interest based on their respective confidence scores and actionable events, wherein the created catalogue is stored in form of a knowledge graph which is referred for decision making while evaluating an incident of interest from the plurality of incidents of interest; and

utilizing the catalogue for evaluating a new incident of interest;

collect additional multi-modal inputs from at least one access device corresponding to at least one person in the geo-location upon determination of the incident of interest;

determine the actionable event based on an analysis of the initial and the additional multi-modal inputs; and

provide a notification of the actionable event to one or more appropriate authorities.

9. The system of claim 8 , further comprising one or more surveillance devices for acquiring the initial multi-modal inputs, wherein the processor receive initial multi-modal inputs from the one or more surveillance devices, and wherein the one or more surveillance devices comprise at least one of a closed-circuit television (CCTV) camera, an Internet Protocol (IP) camera, a microphone, an Internet-of-Things (IoT) sensor, a mobile device, a hand-held device, or a wearable device.

10. The system of claim 8 , wherein the processor-executable instructions further cause the processor to:

identify a set of persons in the geo-location upon determination of the incident of interest by performing at least one of a facial recognition of a set of faces or a voice recognition of a set of voices in the initial multi-modal inputs against a plurality of persons in a population register;

determine a set of access devices corresponding to the set of persons from the population register; and

activate the at least one access device from the set of access devices for collecting additional multi-modal inputs.

11. The system of claim 8 , wherein the processor-executable instructions further cause the processor to:

determine a plurality of access devices in the geo-location upon determination of the incident of interest based on inputs from one or more network operators;

identify a plurality of persons corresponding to the plurality of access devices from a population register;

identify a set of persons by performing at least one of a facial recognition of a set of faces or a voice recognition of a set of voices in the multi-modal inputs against the plurality of persons in the population register; and

activate the at least one access device corresponding to the at least one person from the set of persons for collecting additional multi-modal inputs.

12. The system of claim 8 , and wherein the at least one access device comprises a mobile device, a hand-held device, or a wearable device, and wherein the at least one access device is activated for collecting additional multi-modal inputs upon at least one of: a notification to the at least one person, and a permission from the at least one person.

13. The system of claim 8 , wherein determining the actionable event comprises correlating the initial and the additional multi-modal inputs so as to validate the incident of interest and gathering specific inputs with respect to the incident of interest.

14. The system of claim 8 , wherein the processor-executable instructions further cause the processor to:

provide one or more recommendations to the one or more appropriate authorities based on the actionable event.

15. A non-transitory computer-readable medium storing computer-executable instructions for detecting and notifying an actionable event during surveillance that when executed by a processor, cause the processor to perform operations comprising:

receiving initial multi-modal inputs from a geo-location during surveillance;

determining an incident of interest based on an analysis of the initial multi-modal inputs, wherein the determination of an incident of interest is based on:

generating a confidence score for each of a plurality of incidents of interest determined over a period of time, wherein the confidence score is based on a criticality of the actionable event;

creating a catalogue of the plurality of incidents of interest based on their respective confidence scores and actionable events, wherein the created catalogue is stored in form of a knowledge graph which is referred for decision making while evaluating an incident of interest from the plurality of incidents of interest; and

utilizing the catalogue for evaluating a new incident of interest;

collecting additional multi-modal inputs from at least one access device corresponding to at least one person in the geo-location upon determination of the incident of interest;

determining the actionable event based on an analysis of the initial and the additional multi-modal inputs; and

providing a notification of the actionable event to one or more appropriate authorities.

16. The non-transitory computer-readable medium of claim 15 , further storing computer-executable instructions for:

identifying a set of persons in the geo-location upon determination of the incident of interest by performing at least one of a facial recognition of a set of faces or a voice recognition of a set of voices in the initial multi-modal inputs against a plurality of persons in a population register;

determining a set of access devices corresponding to the set of persons from the population register; and

activating the at least one access device from the set of access devices for collecting additional multi-modal inputs.

17. The non-transitory computer-readable medium of claim 15 , further storing computer-executable instructions for:

determining a plurality of access devices in the geo-location upon determination of the incident of interest based on inputs from one or more network operators;

identifying a plurality of persons corresponding to the plurality of access devices from a population register;

identifying a set of persons by performing at least one of a facial recognition of a set of faces or a voice recognition of a set of voices in the multi-modal inputs against the plurality of persons in the population register; and

activating the at least one access device corresponding to the at least one person from the set of persons for collecting additional multi-modal inputs.

18. The non-transitory computer-readable medium of claim 15 , further storing computer-executable instructions for:

providing one or more recommendations to the one or more appropriate authorities based on the actionable event.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 4, 2020
From: MONDAL, SUBHAS CHANDRA; PANDEY, VISHAL KUMAR
To: WIPRO LIMITED
Reel/Frame 052005/0868 →
Priority Claims (1)
IN 201941054420 · Dec 30, 2019 · national
Continuity (1)
Related Publication 20210201652A1 · Jul 1, 2021
Cited By (1)
US 12,361,803