IP Library Granted Patent US 12,452,365
Granted Patent B2
US 12,452,365 · App. 18/236,284 · Granted Oct 21, 2025

Method/system to identify, extract, and convey key words, phrases, and meanings in service of emergency communications (e.g., voice, text, and video)

Inventors: Adan K. Pope (Riverside, IL); Michel Brkovic (St-Eustache, CA)
Assignee: Intrado Life & Safety, Inc.
H04M3/42348G06V10/40G06V10/95G06V20/52G06V20/70G10L15/1815G10L15/22G10L15/30H04M3/5116H04N7/183H04N7/188H04N23/61H04W64/006G10L2015/088H04M2201/40H04M2201/41
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,452,365
App. No.
18/236,284
Granted
Oct 21, 2025
Kind
B2
Abstract

A method includes receiving visual content of a physical location, the visual content identified by a content timestamp, the physical location identified by a spatial identifier; receiving audio content of a call; performing voice recognition on the call to extract a first audio symbol; receiving a first timestamp of the call, the first timestamp indicating a time at which the call was initiated or a time extracted from the call by voice recognition; and determining a feature in the visual content, at least in part based on the first audio symbol, the feature defined by a person, object, or situation.

Claims (53)

1. A method, comprising:

receiving audio content of a call associated with a physical location;

performing voice recognition on the call to extract a first audio symbol and a spoken location;

determining a spatial identifier of the physical location, at least in part based on the spoken location;

receiving visual content of the physical location, at least in part based on the spatial identifier; and

determining a feature in the visual content, at least in part based on the first audio symbol, the feature defined by a person, object, or situation.

2. The method of claim 1 , further comprising:

triggering a camera to receive the visual content, at least in part based on the spatial identifier of the physical location.

3. The method of claim 2 , further comprising:

determining the spatial identifier, at least in part based on an initiating location of the call.

4. The method of claim 1 , further comprising:

transmitting the visual content and an identifier of the feature to an emergency system.

5. The method of claim 4 , further comprising:

receiving caller location information of the call;

comparing location information of the physical location to the caller location information to produce an identified location;

annotating the visual content with the identified location to produce an annotation; and

transmitting the annotation to the emergency system.

6. The method of claim 1 , further comprising:

determining a relationship between the first audio symbol and a second audio symbol, the second audio symbol extracted from the call by voice recognition; and

transmitting the relationship in an annotation.

7. An apparatus, comprising:

a network interface that receives audio content of a call associated with a physical location; and

a processing unit configured to perform voice recognition on the call to extract a first audio symbol and a spoken location and to determine a spatial identifier of the physical location, at least in part based on the spoken location, wherein

the network interface receives visual content of the physical location, at least in part based on the spatial identifier, and

the processing unit further is configured to determine a feature in the visual content, at least in part based on the first audio symbol, the feature defined by a person, object, or situation.

8. The apparatus of claim 7 , wherein the processing unit triggers a camera to receive the visual content, at least in part based on the spatial identifier of the physical location.

9. The apparatus of claim 8 , wherein the processing unit further is configured to determine the spatial identifier, at least in part based on an initiating location of the call.

10. The apparatus of claim 7 , wherein the network interface transmits the visual content and an identifier of the feature to an emergency system.

11. The apparatus of claim 10 , wherein

the network interface receives caller location information of the call,

the processing unit further is configured to compare location information of the physical location to the caller location information to produce an identified location,

the processing unit further is configured to annotate the visual content with the identified location to produce an annotation, and

the network interface transmits the annotation to the emergency system.

12. A non-transitory computer-readable medium including instructions that, when executed by a processing unit, perform operations comprising:

performing voice recognition on a call to extract a first audio symbol and a spoken location, wherein a network interface receives audio content of the call, the call associated with a physical location;

determining a spatial identifier of the physical location, at least in part based on the spoken location, wherein the network interface receives visual content of the physical location, at least in part based on the spatial identifier; and

determining a feature in the visual content, at least in part based on the first audio symbol, the feature defined by a person, object, or situation.

13. The medium of claim 12 , wherein the network interface triggers a camera to receive the visual content, at least in part based on the spatial identifier of the physical location.

14. The medium of claim 13 , the operations further comprising:

determining the spatial identifier, at least in part based on an initiating location of the call.

15. The medium of claim 12 , wherein the network interface transmits the visual content and an identifier of the feature to an emergency system.

16. The medium of claim 15 , the operations further comprising:

comparing location information of the physical location to caller location information to produce an identified location, wherein the network interface receives the caller location information of the call; and

annotating the visual content with the identified location to produce an annotation, wherein the network interface transmits the annotation to the emergency system.

17. The medium of claim 12 , the operations further comprising:

determining a relationship between the first audio symbol and a second audio symbol, the second audio symbol extracted from the call by voice recognition; and

transmitting the relationship in an annotation.

18. The method of claim 1 , further comprising:

extracting at least one frame of the visual content, at least in part based on the feature; and

transmitting the at least one frame of the visual content.

19. The apparatus of claim 7 , wherein the processing unit further is configured to extract at least one frame of the visual content, at least in part based on the feature, and the network interface transmits the at least one frame of the visual content.

20. The medium of claim 12 , the operations further comprising:

extracting at least one frame of the visual content, at least in part based on the feature, wherein the at least one frame of the visual content is transmitted by a computing device including the processing unit.

Assignments (1)
SECURITY INTEREST Recorded Oct 24, 2023
From: INTRADO LIFE & SAFETY, INC.
To: ROYAL BANK OF CANADA, AS COLLATERAL AGENT
Reel/Frame 065322/0333 →
Continuity (2)
Continuation 18236175 · Aug 21, 2023
Related Publication 20250069386A1 · Feb 27, 2025
References Cited (8)
US 9094615B2 · Aman · 2015 [cited by examiner]
US 10170114B2 · Printz · 2019 [cited by examiner]
US 11922930B1 · Leeds · 2024 [cited by examiner]
US 20110117878A1 · Barash et al. · 2011 [cited by applicant]
US 20140337733A1 · Rodriguez · 2014 [cited by examiner]
US 20200175961A1 · Thomson · 2020 [cited by examiner]
US 20240070251A1 · Maizels · 2024 [cited by examiner]
Non-Final Office Action in U.S. Appl. No. 18/236,284 dated Aug. 19, 2025, 7 pages. [cited by applicant]