IP Library Granted Patent US 9,396,758
Granted Patent B2
US 9,396,758 · App. 14/170,621 · Granted Jul 19, 2016

Semi-automatic generation of multimedia content

Inventors: Ran Oz (Maccabim, IL); Dror Ginzberg (Nir Zvi, IL); Amotz Hoshen (Tel-Aviv, IL); Eitan Lavi (Tel-Aviv, IL); Igor Dvorkin (Be'er Sheva, IL); Ran Yakir (Hod Hasharon, IL)
Assignee: Wochit, Inc.
G11B27/031G06F17/30823H04N5/76H04N9/8205H04N9/8211H04N5/222H04N5/765
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 9,396,758
App. No.
14/170,621
Granted
Jul 19, 2016
Kind
B2
Abstract

A method for multimedia content generation includes receiving a textual input, and automatically retrieving from one or more media databases a plurality of media items that are relevant to the textual input. User input, which selects one or more of the automatically-retrieved media items and correlates one or more of the selected media items in time with the textual input, is received. A video clip, which includes an audio narration of the textual input and the selected media items scheduled in accordance with the user input, is constructed automatically.

Claims (31)

1. A method for multimedia content generation, comprising:

receiving a textual input;

before receiving an audio narration of the textual input, estimating occurrence times of respective components of the audio narration;

automatically retrieving from one or more media databases a plurality of media items that are relevant to the textual input;

receiving user input, which selects one or more of the automatically-retrieved media items and correlates one or more of the selected media items in time with the textual input; and

automatically constructing a video clip, which comprises the audio narration and the selected media items scheduled in accordance with the user input, by scheduling the selected media items in accordance with the estimated occurrence times.

2. The method according to claim 1 , wherein retrieving the media items comprises automatically deriving one or more search queries from the textual input, and querying the media databases with the search queries.

3. The method according to claim 1 , wherein retrieving the media items comprises assigning ranks to the retrieved media items in accordance with relevance to the textual input, filtering the media items based on the ranks, and presenting the filtered media items to a human moderator for producing the user input.

4. The method according to claim 1 , wherein receiving the user input comprises receiving from a human moderator an instruction to synchronize a selected media item with a component of the textual input, and wherein automatically constructing the video clip comprises synchronizing the selected media item with a narration of the component in the audio narration.

5. The method according to claim 1 , wherein receiving the user input is performed before receiving the audio narration.

6. The method according to claim 1 , wherein automatically constructing the video clip comprises defining multiple scheduling permutations of the selected media items, assigning respective scores to the scheduling permutations, and scheduling the selected media items in the video clip in accordance with a scheduling permutation having a best score.

7. A method for multimedia content generation, comprising:

receiving a textual input;

automatically retrieving from one or more media databases a plurality of media items that are relevant to the textual input;

receiving user input, which selects one or more of the automatically-retrieved media items and correlates one or more of the selected media items in time with the textual input; and

automatically constructing a video clip, which comprises an audio narration of the textual input and the selected media items scheduled in accordance with the user input, including dividing a timeline of the video clip into two or more segments and scheduling the selected media items separately in each of the segments,

wherein dividing the timeline comprises scheduling a video media asset whose audio is selected to appear as foreground audio in the video clip, configuring a first segment to end at a start time of the video media asset, and configuring a second segment to begin at an end time of the video media asset.

8. The method according to claim 1 , wherein automatically constructing the video clip comprises training a scheduling model using a supervised learning process, and scheduling the selected media items in accordance with the trained model.

9. Apparatus for multimedia content generation, comprising:

an interface for communicating over a communication network; and

a processor, which is configured to receive a textual input, to estimate, before receiving an audio narration of the textual input, occurrence times of respective components of the audio narration, to automatically retrieve, from one or more media databases over the communication network, a plurality of media items that are relevant to the textual input, to receive user input, which selects one or more of the automatically-retrieved media items and correlates one or more of the selected media items in time with the textual input, to receive an audio narration of the textual input, and to automatically construct a video clip, which comprises the audio narration and the selected media items scheduled in accordance with the user input, by scheduling the selected media items in accordance with the estimated occurrence times.

10. The apparatus according to claim 9 , wherein the processor is configured to automatically derive one or more search queries from the textual input, and to retrieve the media items by querying the media databases with the search queries.

11. The apparatus according to claim 9 , wherein the processor is configured to assign ranks to the retrieved media items in accordance with relevance to the textual input, to filter the media items based on the ranks, and to present the filtered media items to a human moderator for producing the user input.

12. The apparatus according to claim 9 , wherein the processor is configured to receive from a human moderator an instruction to synchronize a selected media item with a component of the textual input, and to automatically construct the video clip by synchronizing the selected media item with a narration of the component in the audio narration.

13. The apparatus according to claim 9 , wherein the processor is configured to receive the user input before receiving the audio narration.

14. The apparatus according to claim 9 , wherein the processor is configured to define multiple scheduling permutations of the selected media items, to assign respective scores to the scheduling permutations, and to schedule the selected media items in the video clip in accordance with a scheduling permutation having a best score.

15. The apparatus according to claim 9 , wherein the processor is configured to train a scheduling model using a supervised learning process, and to schedule the selected media items in accordance with the trained model.

16. Apparatus for multimedia content generation, comprising:

an interface for communicating over a communication network; and

a processor, which is configured to receive a textual input, to automatically retrieve, from one or more media databases over the communication network, a plurality of media items that are relevant to the textual input, to receive user input, which selects one or more of the automatically-retrieved media items and correlates one or more of the selected media items in time with the textual input, to receive an audio narration of the textual input, and to automatically construct a video clip, which comprises an audio narration of the textual input and the selected media items scheduled in accordance with the user input, including dividing a timeline of the video clip into two or more segments and scheduling the selected media items separately in each of the segments,

wherein the processor is configured to schedule a video media asset whose audio is selected to appear as foreground audio in the video clip, to configure a first segment to end at a start time of the video media asset, and to configure a second segment to begin at an end time of the video media asset.

Assignments (2)
SECURITY INTEREST Recorded Nov 3, 2017
From: WOCHIT INC.
To: SILICON VALLEY BANK
Reel/Frame 044026/0755 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Feb 2, 2014
From: OZ, RAN; GINZBERG, DROR; HOSHEN, AMOTZ; LAVI, EITAN; DVORKIN, IGOR; YAKIR, RAN
To: WOCHIT, INC.
Reel/Frame 032150/0504 →
Continuity (4)
Continuation In Part 13874496 · May 1, 2013
Provisional Application 61640748 · May 1, 2012
Provisional Application 61697833 · Sep 7, 2012
Related Publication 20140147095A1 · May 29, 2014