IP Library Granted Patent US 11,636,674
Granted Patent B2
US 11,636,674 · App. 16/683,369 · Granted Apr 25, 2023

System and method for virtual assistant situation commentary

Inventors: Alfy Merican Ahmad Hambaly (Bayan Baru, MY); Hasrolnizam Mohd Mokhtar (Teluk Intan, MY)
Assignee: MOTOROLA SOLUTIONS, INC.
G06V20/20G06F3/165G10L15/22G10L15/24G10L25/78G10L25/93H04N7/183
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,636,674
App. No.
16/683,369
Granted
Apr 25, 2023
Kind
B2
Abstract

Techniques for virtual assistant situation commentary are provided. At least one image frame of a field of view (FOV) of a camera may be received, the at least one image frame intended to be sent to at least one participant of a talk group. A description associated with each element of a plurality of elements within the FOV of the camera may be generated. It may be determined that the at least one participant of the talk group is not currently visually engaged. Audio communication of a sender of the at least one image frame may be monitored to identify a reference to an element of the plurality of elements. The audio communication may be supplemented to include portions of the description of the element that were not included in the audio communication from the sender when it is determined that the at least one participant is not visually engaged.

Claims (56)

1. A method comprising:

receiving at least one image frame of a field of view (FOV) of a camera, the at least one image frame intended to be sent to at least one participant of a talk group;

generating a description associated with each element of a plurality of elements within the FOV of the camera;

determining that the at least one participant of the talk group is not currently visually engaged;

monitoring audio communication of a sender of the at least one image frame to identify a reference to an element of the plurality of elements; and

supplementing the audio communication to include portions of the description of the element that were not included in the audio communication from the sender when it is determined that the at least one participant of the talk group is not currently visually engaged.

2. The method of claim 1 wherein the at least one image frame of the FOV of the camera further comprises a streamed video.

3. The method of claim 1 wherein supplementing the audio further comprises:

detecting periods of silence within the audio communication from the sender; and

inserting the portions of the description of the element during the periods of silence.

4. The method of claim 3 further comprising:

detecting the sender has begun speaking again; and

discontinuing inserting the portions of the description of the element.

5. The method of claim 1 further comprising:

suppressing the supplemented audio communication when the at least one participant of the talk group is currently visually engaged.

6. The method of claim 1 wherein determining that the at least one participant of the talk group is not currently visually engaged further comprises at least one of:

determining the at least one participant of the talk group is using a device that is not equipped to display the at least one image frame; and

determining the at least one participant of the talk group is not currently viewing the at least one image frame.

7. A system comprising:

a processor; and

a memory coupled to the processor, the memory containing a set of instructions thereon that when executed by the processor cause the processor to:

receive at least one image frame of a field of view (FOV) of a camera, the at least one image frame intended to be sent to at least one participant of a talk group;

generate a description associated with each element of a plurality of elements within the FOV of the camera;

determine that the at least one participant of the talk group is not currently visually engaged;

monitor audio communication of a sender of the at least one image frame to identify a reference to an element of the plurality of elements; and

supplement the audio communication to include portions of the description of the element that were not included in the audio communication from the sender when it is determined that the at least one participant of the talk group is not currently visually engaged.

8. The system of claim 7 wherein the at least one image frame of the FOV of the camera further comprises a streamed video.

9. The system of claim 7 wherein the instructions to supplement the audio further comprises instructions to:

detect periods of silence within the audio communication from the sender; and

insert the portions of the description of the element during the periods of silence.

10. The system of claim 9 further comprising instructions to:

detect the sender has begun speaking again; and

discontinue inserting the portions of the description of the element.

11. The system of claim 7 further comprising instructions to:

suppress the supplemented audio communication when the at least one participant of the talk group is currently visually engaged.

12. The system of claim 7 wherein the instructions to determine that the at least one participant of the talk group is not currently visually engaged further comprises at least one of instructions to:

determine the at least one participant of the talk group is using a device that is not equipped to display the at least one image frame; and

determine the at least one participant of the talk group is not currently viewing the at least one image frame.

13. A non-transitory processor readable medium containing a set of instructions thereon that when executed by a processor cause the processor to:

receive at least one image frame of a field of view (FOV) of a camera, the at least one image frame intended to be sent to at least one participant of a talk group;

generate a description associated with each element of a plurality of elements within the FOV of the camera;

determine that the at least one participant of the talk group is not currently visually engaged;

monitor audio communication of a sender of the at least one image frame to identify a reference to an element of the plurality of elements; and

supplement the audio communication to include portions of the description of the element that were not included in the audio communication from the sender when it is determined that the at least one participant of the talk group is not currently visually engaged.

14. The medium of claim 13 wherein the at least one image frame of the FOV of the camera further comprises a streamed video.

15. The medium of claim 13 wherein the instructions to supplement the audio further comprises instructions to:

detect periods of silence within the audio communication from the sender; and

insert the portions of the description of the element during the periods of silence.

16. The medium of claim 15 further comprising instructions to:

detect the sender has begun speaking again; and

discontinue inserting the portions of the description of the element.

17. The medium of claim 13 further comprising instructions to:

suppress the supplemented audio communication when the at least one participant of the talk group is currently visually engaged.

18. The medium of claim 13 wherein the instructions to determine that the at least one participant of the talk group is not currently visually engaged further comprises at least one of instructions to:

determine the at least one participant of the talk group is using a device that is not equipped to display the at least one image frame; and

determine the at least one participant of the talk group is not currently viewing the at least one image frame.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 14, 2019
From: HAMBALY, ALFY MERICAN AHMAD; MOKHTAR, HASROLNIZAM MOHD
To: MOTOROLA SOLUTIONS INC.
Reel/Frame 051006/0236 →
Continuity (1)
Related Publication 20210150211A1 · May 20, 2021