IP Library Granted Patent US 11,294,474
Granted Patent B1
US 11,294,474 · App. 17/169,227 · Granted Apr 5, 2022

Controlling video data content using computer vision

Inventors: Aaron Michael Stewart (Raleigh, NC); Alden Rose (Durham, NC); Ellis Anderson (Greensboro, NC)
Assignee: Lenovo (Singapore) Pte. Ltd.
G06F3/017G06F3/167G06K9/00355G06K9/00389G06K9/00671
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,294,474
App. No.
17/169,227
Granted
Apr 5, 2022
Kind
B1
Abstract

A virtual collaboration system receives input video data including a participant. The system analyzes the input video data to identify a gesture or a movement made by the participant. The system selects an overlay image as a function of the gesture or the movement made by the participant, incorporates the overlay image into the input video data, thereby generating output video data that includes the overlay image, and transmits the output video data to one or more participant devices.

Claims (63)

1. A process comprising:

receiving, into a computer processor, input video data comprising a participant;

analyzing the input video data to identify a gesture or a movement made by the participant;

selecting an overlay image as a function of the gesture or the movement made by the participant;

incorporating the overlay image into the input video data, thereby generating output video data comprising the overlay image;

transmitting the output video data to one or more participant devices;

receiving an input from the participant indicating that the participant would like to record a personal gesture or movement;

capturing the personal gesture or movement made by the participant from the input video data; and

storing the personal gesture or movement in a database.

2. The process of claim 1 , comprising incorporating the overlay image into a location in the output video data as a function of a location of the gesture or the movement made by the participant in the input video data.

3. The process of claim 1 , comprising incorporating the overlay image into a location in the output video data as a function of a location in the output video data selected by the participant.

4. The process of claim 1 , comprising:

receiving a personal overlay image from the participant;

storing the personal overlay image in the database;

associating the personal overlay image from the participant with the personal gesture or movement received from the participant; and

upon detection of the personal gesture or movement, incorporating the personal overlay image into the output video data.

5. The process of claim 1 , comprising:

associating audio data with the gesture or movement made by the participant; and

transmitting the audio data to the one or more participant devices.

6. The process of claim 1 , comprising:

receiving a second gesture or second movement made by the participant; and

incorporating a second overlay into the output video data as a function of the second gesture or second movement.

7. The process of claim 1 , comprising removing the overlay image from the output video data within a preset time period.

8. The process of claim 1 , comprising:

determining that the input video data comprise a sign language; and

incorporating into the output video data a textual translation of the sign language.

9. The process of claim 1 , comprising selecting an appropriate overlay image as a function of a type of the virtual meeting or a type of participant associated with the virtual meeting.

10. The process of claim 1 , comprising selecting a version of the overlay image as a function of one or more of contrast, color, brightness, and visual complexity differences between the overlay image and a background of the output video data.

11. The process of claim 1 , comprising:

receiving the input video data from a plurality of participants at a substantially simultaneous time;

generating an order of the input video data from the plurality of participants; and

transmitting the output video data as a function of the order.

12. The process of claim 1 , wherein the overlay image comprises a substantially entire portion of the output video data.

13. The process of claim 1 , comprising:

determining that the output video data has been cropped;

determining that the cropped output video data has removed or partially removed the overlay image; and

relocating the overlay image in the output video data.

14. The process of claim 1 , comprising blurring the output video data as a function of the gesture or movement made by the participant.

15. The process of claim 1 , comprising selecting the overlay image as a function of a characteristic of the participant comprising one or more of skin color, gender, hair color, hair length, and eye color.

16. The process of claim 1 , wherein the input video data comprise a virtual meeting, a recorded camera feed, or a broadcast camera feed.

17. A non-transitory computer readable medium comprising instructions that when executed by a processor execute a process comprising:

receiving, into a computer processor, input video data comprising a participant;

analyzing the input video data to identify a gesture or a movement made by the participant;

selecting an overlay image as a function of the gesture or the movement made by the participant;

incorporating the overlay image into the input video data, thereby generating output video data comprising the overlay image;

transmitting the output video data to one or more participant devices;

receiving the input video data from a plurality of participants at a substantially simultaneous time;

generating an order of the input video data from the plurality of participants; and

transmitting the output video data as a function of the order.

18. A system comprising:

a computer processor; and

a computer storage device coupled to the computer processor;

wherein the computer processor and the computer storage device are operable for:

receiving, into a computer processor, input video data comprising a participant;

analyzing the input video data to identify a gesture or a movement made by the participant;

selecting an overlay image as a function of the gesture or the movement made by the participant;

incorporating the overlay image into the input video data, thereby generating output video data comprising the overlay image; and

transmitting the output video data to one or more participant devices;

wherein the overlay image comprises a substantially entire portion of the output video data.

19. The process of claim 1 , comprising:

analyzing a field of view of the input video data;

determining that the participant is no longer in the field of view of the input video data; and

transmitting an indication to the one or more participant devices regarding a status or availability of the participant.

Assignments (4)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Feb 20, 2025
From: LENOVO PC INTERNATIONAL LIMITED
To: LENOVO SWITZERLAND INTERNATIONAL GMBH
Reel/Frame 070269/0092 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Feb 19, 2025
From: LENOVO (SINGAPORE) PTE LTD.
To: LENOVO PC INTERNATIONAL LIMITED
Reel/Frame 070266/0821 →
CORRECTIVE ASSIGNMENT TO CORRECT THE ASSIGNEE'S NAME PREVIOUSLY RECORDED ON REEL 055168 FRAME 0969. ASSIGNOR(S) HEREBY CONFIRMS THE THE ASSIGNMENT. Recorded Feb 11, 2021
From: STEWART, AARON MICHAEL; ROSE, ALDEN; ANDERSON, ELLIS
To: LENOVO (SINGAPORE) PTE. LTD.
Reel/Frame 055280/0723 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Feb 5, 2021
From: STEWART, AARON; ROSE, ALDEN; ANDERSON, ELLIS
To: LENOVO (SINGAPORT) PTE. LTD.
Reel/Frame 055168/0969 →
Cited By (3)
US 12,536,837 US 12,603,794 US 12,619,317