IP Library Granted Patent US 9,888,279
Granted Patent B2
US 9,888,279 · App. 14/483,507 · Granted Feb 6, 2018

Content based video content segmentation

Inventors: Faisal Ishtiaq (Chicago, IL); Benedito J. Fonseca, Jr. (Glen Ellyn, IL); Kevin L. Baum (Rolling Meadows, IL); Anthony J. Braskich (Palatine, IL); Stephen P. Emeott (Rolling Meadows, IL); Bhavan Gandhi (Vernon Hills, IL); Renxiang Li (Lake Zurich, IL); Alfonso Martinez Smith (Algonquin, IL); Michael L. Needham (Palatine, IL); Isselmou Ould Dellahy (Lake in the Hills, IL)
Assignee: ARRIS Enterprises LLC
H04N21/4316G06K9/00711G06K9/00765H04N21/2353H04N21/23418H04N21/4828H04N21/4884H04N21/812G06K2209/27
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 9,888,279
App. No.
14/483,507
Granted
Feb 6, 2018
Kind
B2
Abstract

A method receives video content and metadata associated with video content. The method then extracts features of the video content based on the metadata. Portions of the visual, audio, and textual features are fused into composite features that include multiple features from the visual, audio, and textual features. A set of video segments of the video content is identified based on the composite features of the video content. Also, the segments may be identified based on a user query.

Claims (65)

1. A method comprising:

in a video data analyzer of a first computing device, configuring an extraction, based on metadata associated with video content, of content features;

wherein the content features are selected from the group consisting of visual features of the video content, audio features of the video content, and textual features of the video content,

wherein one or more feature extractors corresponding to the content features are selected from the group consisting of a visual feature extractor for content features selected from visual features of the video content, an audio feature extractor for content features selected from audio features of the video content, and a text feature extractor for content features selected from textual features of the video content, and

wherein configuring the extraction comprises configuring the one or more selected feature extractors to extract the respective content features in accordance with one or more operating parameters that are used internally by the respective feature extractor, and that are tunable by the video data analyzer to alter an extraction behavior of the feature extractor based on the metadata;

creating a single data stream of fused information for rendering in a client computing device communicatively coupled to one or more distributed content servers, wherein the creating comprises:

fusing, in a plurality of fusion modules communicatively coupled to the one or more distributed content servers, portions of the content features into composite features that are generated from functions of the multiple features from the content features;

identifying, by one or more of the plurality of fusion modules, a plurality of video segments comprising one or more video segments of the video content based on the composite features; and

rendering the created single data stream, in a user interface of the client computing device, by rendering representations of the identified video segments.

2. The method of claim 1 , wherein some of the plurality of video segments are identified based on only one content feature.

3. The method of claim 1 , wherein identifying the plurality of video segments comprises combining non-contiguous segments from the video content into a segment.

4. The method of claim 1 , wherein the multiple features are based on at least two of the group consisting of visual features of the video content, audio features of the video content, and textual features of the video content.

5. The method of claim 1 , wherein:

the composite features include the multiple features from at least two of the visual feature extractor, the audio feature extractor, and the text feature extractor.

6. The method of claim 1 , wherein:

the extraction is performed by a plurality of extractors, and

the metadata is used to configure an extractor in the plurality of extractors to extract one of visual, audio, and textual features based on the metadata.

7. The method of claim 1 , wherein:

the identifying is performed by a plurality of fusion modules, and

the metadata is used to configure a fusion module in the plurality of fusion modules to fuse the multiple features into the composite features.

8. The method of claim 7 , wherein the fusion module determines a composite feature based on the metadata.

9. The method of claim 1 , further comprising classifying the plurality of video segments based on the metadata.

10. The method of claim 1 , wherein the metadata comprises program metadata received from an electronic program guide data source.

11. The method of claim 1 , further comprising:

displaying the plurality of video segments;

receiving a selection of one of the plurality of video segments; and

displaying the one of the plurality of video segments.

12. The method of claim 11 , further comprising adding supplemental content in association with the one of the plurality of video segments based on a feature associated with the one of the plurality of video segments.

13. The method of claim 12 , wherein the supplemental content is based on a type of user reaction to the one of the plurality of video segments.

14. An apparatus comprising:

a plurality of computer processors comprising a video data analyzer processor and one or more segment services processors;

at least one non-transitory computer readable storage memory coupled to each of the plurality of computer processors and comprising instructions that when executed by one or more of the computer processors cause the one or more of the computer processors to be configured for:

in the video data analyzer processor, configuring an extraction, based on metadata associated with video content, of content features;

wherein the content features are selected from the group consisting of visual features of the video content, audio features of the video content, and textual features of the video content,

wherein one or more feature extractors corresponding to the content features are selected from the group consisting of a visual feature extractor for content features selected from visual features of the video content, an audio feature extractor for content features selected from audio features of the video content, and a text feature extractor for content features selected from textual features of the video content, and

wherein configuring the extraction comprises configuring the one or more selected feature extractors to extract the respective content features in accordance with one or more operating parameters that are used internally by the respective feature extractor, and that are tunable by the video data analyzer to alter an extraction behavior of the feature extractor based on the metadata;

creating a single data stream of fused information for rendering in a client computing device communicatively coupled to one or more distributed content servers, wherein the creating comprises:

in a plurality of fusion modules in the segment services processors, fusing portions of the content features into composite features that include are generated from functions of the multiple features from the content features, wherein the segment services processors are communicatively coupled to the one or more distributed content servers;

identifying, by one or more of the plurality of fusion modules, a plurality of video segments comprising one or more video segments of the video content based on the composite features; and

rendering the created single data stream, in a user interface of the client computing device, by rendering representations of the identified video segments.

15. A method for creating a single data stream of fused information for rendering in a client computing device communicatively coupled to one or more distributed content servers, the method comprising:

receiving a search query comprising at least one word;

receiving a textual program index associated with each video program from a plurality of video programs stored on a content server;

identifying, by one or more of a plurality of fusion modules, matching video programs from the plurality of video programs based on the textual program index associated with each video program and the at least one word;

receiving a user selection of a matching video program from the matching video programs to identify a selected video program;

receiving a plurality of text records associated with the selected video program;

searching, by one or more of the plurality of fusion modules, the text records of the selected video program to identify matching text records based on the at least one word;

segmenting, by one or more of the plurality of fusion modules, at least one matching video program into a plurality of video segments that include at least one of the matching text records; and

rendering the created single data stream, in a user interface of the client computing device, by rendering representations of the identified video segments;

wherein the plurality of fusion modules is implemented by one or more segment services processors each comprising one or more computer processors, the one or more segment services processors communicatively coupled to the one or more distributed content servers.

16. The method of claim 15 , further comprising:

presenting a representation of the at least one matching video programs to a user; and

receiving a user selection that identifies a user-selected matching video program.

17. The method of claim 15 , further comprising:

presenting at least a portion of the generated at least one matching segment to a user.

18. The method of claim 15 , further comprising:

calculating a windowed text record rank from the plurality of text records associated with a matching video program;

identifying contiguous blocks of matching text records as a segment;

assigning a segment score to the segment based upon the text record rank of the segment; and

presenting at least one segment to a user based on the segment score associated with the at least one segment.

19. The method of claim 15 , wherein receiving the textual program index associated with each video program comprises:

creating a plurality of text records for each video program, wherein each text record comprises at least a start time and a representation of text for the text record;

creating the textual program index from the plurality of text records for each video program; and

storing the textual program index for each video program with an identifier for the video program.

20. The method of claim 15 , wherein the plurality of video segments are combined with previously generated segments based on extraction of features in the video program.

Assignments (16)
RELEASE OF SECURITY INTEREST AT REEL/FRAME 049905/0504 Recorded Dec 19, 2024
From: JPMORGAN CHASE BANK, N.A., AS COLLATERAL AGENT
To: ARRIS ENTERPRISES LLC (F/K/A ARRIS ENTERPRISES, INC.); ARRIS TECHNOLOGY, INC.; ARRIS SOLUTIONS, INC.; COMMSCOPE, INC. OF NORTH CAROLINA; COMMSCOPE TECHNOLOGIES LLC; RUCKUS WIRELESS, LLC (F/K/A RUCKUS WIRELESS, INC.)
Reel/Frame 071477/0255 →
PARTIAL TERMINATION AND RELEASE OF SECURITY INTEREST IN PATENTS RECORDED AT R/F 060752/0001 Recorded Apr 13, 2023
From: WILMINGTON TRUST
To: COMMSCOPE TECHNOLOGIES LLC; COMMSCOPE, INC. OF NORTH CAROLINA; ARRIS ENTERPRISES LLC
Reel/Frame 063322/0209 →
RELEASE OF SECURITY INTEREST Recorded Apr 7, 2023
From: WILMINGTON TRUST, NATIONAL ASSOCIATION
To: COMMSCOPE TECHNOLOGIES LLC; COMMSCOPE, INC. OF NORTH CAROLINA; ARRIS ENTERPRISES LLC
Reel/Frame 063270/0220 →
PARTIAL TERMINATION AND RELEASE OF SECURITY INTEREST IN PATENTS Recorded Jul 15, 2022
From: WILMINGTON TRUST, NATIONAL ASSOCIATION, AS COLLATERAL AGENT
To: COMMSCOPE TECHNOLOGIES LLC; ARRIS ENTERPRISES LLC; COMMSCOPE, INC. OF NORTH CAROLINA
Reel/Frame 060671/0324 →
PARTIAL RELEASE OF TERM LOAN SECURITY INTEREST Recorded Jul 13, 2022
From: JPMORGAN CHASE BANK, N.A.
To: COMMSCOPE TECHNOLOGIES LLC; ARRIS ENTERPRISES LLC; COMMSCOPE, INC. OF NORTH CAROLINA
Reel/Frame 060649/0286 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jul 13, 2022
From: ARRIS ENTERPRISES LLC
To: BISON PATENT LICENSING, LLC
Reel/Frame 060641/0130 →
PARTIAL RELEASE OF ABL SECURITY INTEREST Recorded Jul 13, 2022
From: JPMORGAN CHASE BANK, N.A.
To: COMMSCOPE TECHNOLOGIES LLC; ARRIS ENTERPRISES LLC; COMMSCOPE, INC. OF NORTH CAROLINA
Reel/Frame 060649/0305 →
SECURITY INTEREST Recorded Nov 19, 2021
From: ARRIS SOLUTIONS, INC.; ARRIS ENTERPRISES LLC; COMMSCOPE TECHNOLOGIES LLC; COMMSCOPE, INC. OF NORTH CAROLINA; RUCKUS WIRELESS, INC.
To: WILMINGTON TRUST
Reel/Frame 060752/0001 →
ABL SECURITY AGREEMENT Recorded Jul 3, 2019
From: COMMSCOPE, INC. OF NORTH CAROLINA; COMMSCOPE TECHNOLOGIES LLC; ARRIS ENTERPRISES LLC; ARRIS TECHNOLOGY, INC.; RUCKUS WIRELESS, INC.; ARRIS SOLUTIONS, INC.
To: JPMORGAN CHASE BANK, N.A.
Reel/Frame 049892/0396 →
TERM LOAN SECURITY AGREEMENT Recorded Jul 3, 2019
From: COMMSCOPE, INC. OF NORTH CAROLINA; COMMSCOPE TECHNOLOGIES LLC; ARRIS ENTERPRISES LLC; ARRIS TECHNOLOGY, INC.; RUCKUS WIRELESS, INC.; ARRIS SOLUTIONS, INC.
To: JPMORGAN CHASE BANK, N.A.
Reel/Frame 049905/0504 →
PATENT SECURITY AGREEMENT Recorded Jul 3, 2019
From: ARRIS ENTERPRISES LLC
To: WILMINGTON TRUST, NATIONAL ASSOCIATION, AS COLLATERAL AGENT
Reel/Frame 049820/0495 →
CHANGE OF NAME Recorded Jun 25, 2019
From: ARRIS ENTERPRISES, INC.
To: ARRIS ENTERPRISES LLC
Reel/Frame 049586/0470 →
TERMINATION AND RELEASE OF SECURITY INTEREST IN PATENTS Recorded Apr 8, 2019
From: BANK OF AMERICA, N.A., AS ADMINISTRATIVE AGENT
To: ARRIS GROUP, INC.; ARRIS ENTERPRISES, INC.; ARRIS INTERNATIONAL LIMITED; ARRIS TECHNOLOGY, INC.; ARCHIE U.S. MERGER LLC; ARCHIE U.S. HOLDINGS LLC; ARRIS GLOBAL SERVICES, INC.; ARRIS HOLDINGS CORP. OF ILLINOIS, INC.; ARRIS SOLUTIONS, INC.; BIG BAND NETWORKS, INC.; TEXSCAN CORPORATION; POWER GUARD, INC.; JERROLD DC RADIO, INC.; NEXTLEVEL SYSTEMS (PUERTO RICO), INC.; GIC INTERNATIONAL HOLDCO LLC; GIC INTERNATIONAL CAPITAL LLC
Reel/Frame 050721/0401 →
CHANGE OF NAME Recorded Mar 14, 2017
From: ARRIS ENTERPRISES INC
To: ARRIS ENTERPRISES LLC
Reel/Frame 041995/0031 →
SECURITY INTEREST Recorded Jun 26, 2015
From: ARRIS GROUP, INC.; ARRIS ENTERPRISES, INC.; ARRIS INTERNATIONAL LIMITED; ARRIS TECHNOLOGY, INC.; ARCHIE U.S. MERGER LLC; ARCHIE U.S. HOLDINGS LLC; ARRIS GLOBAL SERVICES, INC.; ARRIS HOLDINGS CORP. OF ILLINOIS, INC.; ARRIS SOLUTIONS, INC.; BIG BAND NETWORKS, INC.; TEXSCAN CORPORATION; POWER GUARD, INC.; JERROLD DC RADIO, INC.; NEXTLEVEL SYSTEMS (PUERTO RICO), INC.; GIC INTERNATIONAL HOLDCO LLC; GIC INTERNATIONAL CAPITAL LLC
To: BANK OF AMERICA, N.A., AS ADMINISTRATIVE AGENT
Reel/Frame 036020/0789 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 10, 2014
From: ISHTIAQ, FAISAL; FONSECA, BENEDITO J., JR.; BAUM, KEVIN L.; BRASKICH, ANTHONY J.; EMEOTT, STEPHEN P.; GANDHI, BHAVAN; LI, RENXIANG; SMITH, ALFONSO MARTINEZ; NEEDHAM, MICHAEL L.; OULD DELLAHY, ISSELMOU
To: ARRIS ENTERPRISES, INC.
Reel/Frame 034138/0976 →
Continuity (2)
Provisional Application 61877292 · Sep 13, 2013
Related Publication 20150082349A1 · Mar 19, 2015