IP Library › Granted Patent US 12,254,903
Granted Patent B2
US 12,254,903 · App. 18/311,358 · Granted Mar 18, 2025

Composite video generation

Inventors: Guillaume Oules (Bordeaux, FR); Anais Oules (Bordeaux, FR)
Assignee: GoPro, Inc.
G11B27/036G06T7/194G06T7/70G06V10/98G06V20/40G06V40/103G06V40/161H04N5/265G06T2207/10016G06T2207/20221G06T2207/30201
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,254,903
App. No.
18/311,358
Granted
Mar 18, 2025
Kind
B2
Abstract

Locations in which a person is depicted within video frames may be determined to identify portions of the video frames to be included in a composite video. A background image not including any depiction of the person may be generated, and the identified portions of the video frames may be inserted into the background image to generate the composite video.

Claims (43)

1. A system for generating composite videos, the system comprising:

one or more physical processors configured by machine-readable instructions to:

obtain video information defining a video having a progress length, the video including video frames, the video frames including depiction of a thing at different locations within the video frames;

determine multiple ones of the video frames of the video to be used to generate individual ones of composite video frames of a composite video;

select portions of the multiple ones of the video frames of the video to be included in the composite video based on the different locations of the depiction of the thing within the video frames; and

generate the composite video frames of the composite video based on the selected portions of the multiple ones of the video frames of the video, the composite video frames including multiple depictions of the thing;

wherein:

two or more of the selected portions of the multiple ones of the video frames of the video are combined into a single composite video frame based on different locations of a mask for the two or more of the selected portions of the multiple ones of the video frames;

the different locations of the mask include different placements and/or different sizes of the mask; and

the different placements and/or the different sizes of the mask are determined to prevent overlap of the mask for the single composite video frame.

2. The system of claim 1 , wherein the thing includes a living thing.

3. The system of claim 1 , wherein the thing includes a non-living thing.

4. The system of claim 1 , wherein the multiple ones of the video frames of the video to be used to generate the individual ones of composite video frames of the composite video are determined based on matching of audio associated with the video frames of the video.

5. The system of claim 1 , wherein the multiple ones of the video frames of the video to be used to generate the individual ones of composite video frames of the composite video are determined based on distance between the selected portions of the multiple ones of the video frames of the video.

6. A system for generating composite videos, the system comprising:

one or more physical processors configured by machine-readable instructions to:

obtain video information defining a video having a progress length, the video including video frames, the video frames including depiction of a thing at different locations within the video frames, wherein the thing includes a person;

determine multiple ones of the video frames of the video to be used to generate individual ones of composite video frames of a composite video, wherein the multiple ones of the video frames of the video to be used to generate the individual ones of composite video frames of the composite video are determined based on matching of poses of the person depicted within the video frames of the video;

select portions of the multiple ones of the video frames of the video to be included in the composite video based on the different locations of the depiction of the thing within the video frames; and

generate the composite video frames of the composite video based on the selected portions of the multiple ones of the video frames of the video, the composite video frames including multiple depictions of the thing.

7. The system of claim 6 , wherein two or more of the selected portions of the multiple ones of the video frames of the video are combined into a single composite video frame based on different locations of a mask for the two or more of the selected portions of the multiple ones of the video frames.

8. The system of claim 7 , wherein the different locations of the mask include different placements and/or different sizes of the mask.

9. The system of claim 6 , wherein the thing further includes a non-living thing.

10. A method for generating composite videos, the method performed by a computing system including one or more processors, the method comprising:

obtaining, by the computing system, video information defining a video having a progress length, the video including video frames, the video frames including depiction of a thing at different locations within the video frames;

determining, by the computing system, multiple ones of the video frames of the video to be used to generate individual ones of composite video frames of a composite video;

selecting, by the computing system, portions of the multiple ones of the video frames of the video to be included in the composite video based on the different locations of the depiction of the thing within the video frames; and

generating, by the computing system, the composite video frames of the composite video based on the selected portions of the multiple ones of the video frames of the video, the composite video frames including multiple depictions of the thing wherein:

two or more of the selected portions of the multiple ones of the video frames of the video are combined into a single composite video frame based on different locations of a mask for the two or more of the selected portions of the multiple ones of the video frames;

the different locations of the mask include different placements and/or different sizes of the mask; and

the different placements and/or the different sizes of the mask are determined to prevent overlap of the mask for the single composite video frame.

11. The method of claim 10 , wherein the thing includes a living thing.

12. The method of claim 10 , wherein the thing includes a non-living thing.

13. The method of claim 10 , wherein the multiple ones of the video frames of the video to be used to generate the individual ones of composite video frames of the composite video are determined based on matching of audio associated with the video frames of the video.

14. The method of claim 10 , wherein the multiple ones of the video frames of the video to be used to generate the individual ones of composite video frames of the composite video are determined based on distance between the selected portions of the multiple ones of the video frames of the video.

15. A method for generating composite videos, the method performed by a computing system including one or more processors, the method comprising:

obtaining, by the computing system, video information defining a video having a progress length, the video including video frames, the video frames including depiction of a thing at different locations within the video frames, wherein the thing includes a person;

determining, by the computing system, multiple ones of the video frames of the video to be used to generate individual ones of composite video frames of a composite video, wherein the multiple ones of the video frames of the video to be used to generate the individual ones of composite video frames of the composite video are determined based on matching of poses of the person depicted within the video frames of the video;

selecting, by the computing system, portions of the multiple ones of the video frames of the video to be included in the composite video based on the different locations of the depiction of the thing within the video frames; and

generating, by the computing system, the composite video frames of the composite video based on the selected portions of the multiple ones of the video frames of the video, the composite video frames including multiple depictions of the thing.

16. The method of claim 15 , wherein two or more of the selected portions of the multiple ones of the video frames of the video are combined into a single composite video frame based on different locations of a mask for the two or more of the selected portions of the multiple ones of the video frames.

17. The method of claim 16 , wherein the different locations of the mask include different placements and/or different sizes of the mask.

18. The method of claim 15 , wherein the thing further includes a non-living thing.

Assignments (3)
SECURITY INTEREST Recorded Aug 4, 2025
From: GOPRO, INC.
To: FARALLON CAPITAL MANAGEMENT, L.L.C., AS AGENT
Reel/Frame 072340/0676 →
SECURITY INTEREST Recorded Aug 4, 2025
From: GOPRO, INC.
To: WELLS FARGO BANK, NATIONAL ASSOCIATION, AS AGENT
Reel/Frame 072358/0001 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded May 5, 2023
From: OULES, GUILLAUME; OULES, ANAIS
To: GOPRO, INC.
Reel/Frame 063556/0243 →
Continuity (3)
Continuation 17847162 · Jun 22, 2022
Continuation 17244895 · Apr 29, 2021
Related Publication 20230274767A1 · Aug 31, 2023
References Cited (8)
US 11404088B1 · Oulès · 2022 [cited by examiner]
US 11688430B2 · Oules · 2023 [cited by examiner]
US 20100281375A1 · Pendergast · 2010 [cited by examiner]
US 20160110612A1 · Sabripour · 2016 [cited by examiner]
US 20170076154A1 · Schupp · 2017 [cited by examiner]
US 20200134313A1 · Endoh · 2020 [cited by examiner]
US 20200251146A1 · St. John Brislin · 2020 [cited by examiner]
US 20220351753A1 · Oules · 2022 [cited by applicant]