IP Library Granted Patent US 9,635,219
Granted Patent B2
US 9,635,219 · App. 14/183,614 · Granted Apr 25, 2017

Supplementary media validation system

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 9,635,219
App. No.
14/183,614
Granted
Apr 25, 2017
Kind
B2
Abstract

A method for preparing supplementary media content for coordinated transmission with multimedia content includes accepting the multimedia content including an audio portion, accepting the supplementary media content associated with the multimedia content, and validating the supplementary media content according to the audio portion of the multimedia content.

Claims (44)

1. A computer-implemented method for preparing supplementary media content for coordinated transmission with multimedia content, the method comprising:

accepting, by a wordspotting engine, the multimedia content including an audio portion;

accepting, by the wordspotting engine, caption data included in the supplementary media content associated with the multimedia content;

validating the supplementary media content according to the audio portion of the multimedia content based on a phonetic process, wherein validating the supplementary media content comprises:

aligning, by the wordspotting engine, the caption data against the audio portion based on a phonetic process;

determining, for captions of the caption data, corresponding actual time intervals in the audio portion;

flagging captions which do not align within a predetermined tolerance level as misaligned;

determining a most likely first language of the caption data;

analyzing the audio content to determine a most likely second language for the audio portion of the multimedia content;

comparing the most likely first language and the most likely second language;

if the first and second language are the same language, generating an output indicating that there is no language mismatch; and

if the first and second language differ, generating an output indicating that there is a language mismatch.

2. The method of claim 1 wherein validating the supplementary media content further comprises comparing expected time intervals associated with the captions in the caption data to the actual time intervals in the audio portion determined by the wordspotting engine.

3. The method of claim 1 wherein the supplementary media content includes a plurality of segments of text and wherein validating the supplementary media content includes determining whether a text content of the plurality of segments of text corresponds to a speech content of the audio portion of the multimedia content.

4. The method of claim 3 wherein the plurality of segments of text includes text content in the first language, the audio portion of the multimedia content includes speech content in the second language, and determining whether a text content of the plurality of segments of text corresponds to a speech content of the audio portion of the multimedia content includes determining if the first language and the second language are the same language.

5. The method of claim 1 wherein the supplementary media content includes a plurality of segments of text and wherein validating the supplementary media content includes identifying a temporal misalignment between a text content of the plurality of segments of text and a speech content of the audio portion of the multimedia content.

6. The method of claim 5 further comprising determining a rate of change of temporal misalignment over a plurality of segments of text and the speech content of the audio portion of the multimedia content.

7. The method of claim 6 further comprising mitigating the temporal misalignment between the text content of the plurality of segments of text and the speech content of the audio portion of the multimedia content according to the determined rate of change of temporal misalignment.

8. The method of claim 1 wherein the supplementary media content includes an audio track and validating the supplementary media content includes determining whether at least some segments of the audio track which include speech content correspond to segments of the audio portion of the multimedia content which do not include speech content.

9. The method of claim 1 wherein the supplementary media content includes an audio track and validating the supplementary media content includes:

determining a total duration of speech activity in audio track;

determining a total duration of speech activity in the audio portion of the multimedia content; and

comparing the total duration of speech activity in the audio track to the total duration of speech activity in the audio portion of the multimedia content.

10. The method of claim 1 wherein the supplementary media content includes an audio track and validating the supplementary media content includes:

identifying time intervals including voice activity in the audio portion of the multimedia content;

identifying time intervals including voice activity in the audio track; and

determining whether any of the time intervals including voice activity in the audio portion of the multimedia content overlaps with the time intervals including voice activity in the audio track.

11. The method of claim 1 wherein the supplementary media content includes an audio track and validating the supplementary media content includes:

identifying a time offset between a content of the audio track and a content of the audio portion of the multimedia content; and

aligning the content of the audio track to the content of the audio portion of the multimedia content based on the identified time offset.

12. The method of claim 1 wherein the supplementary media content includes an audio track and validating the supplementary media content includes comparing the audio track to the audio portion of the multimedia content to identify a first plurality of segments in the audio portion that do exist in the audio track and a second plurality of segments in the audio portion that do not exist in the audio track.

13. The method of claim 1 wherein the supplementary media content includes a plurality of segments of text and wherein the plurality of segments of text is associated with an expected time alignment to the audio portion of the multimedia content and validating the supplementary media content includes:

determining an actual alignment of the plurality of segments of text to the audio portion of the multimedia content; and

comparing the actual alignment of the plurality of segments of text to the expected alignment of the plurality of segments of text.

14. The method of claim 13 wherein comparing the actual alignment of the plurality of segments of text to the expected alignment of the plurality of segments of text includes identifying one or more differences between the actual alignment of the plurality of segments of text and the expected alignment of the plurality of segments of text.

15. A non-transitory computer-readable medium comprising instructions stored thereon which when executed cause a data processing system to prepare supplementary media content for coordinated transmission with multimedia content by performing the operations of:

accepting the multimedia content including an audio portion;

accepting caption data included in the supplementary media content associated with the multimedia content;

validating the supplementary media content according to the audio portion of the multimedia content based on a phonetic process, wherein validating the supplementary media content comprises aligning, by a wordspotting engine, the caption data against the audio portion based on a phonetic process; determining for captions of the caption data corresponding actual time intervals in the audio portion and flagging captions which do not align within a predetermined tolerance level as misaligned;

determining a most likely first language of the caption data;

analyzing the audio content to determine a most likely second language for the audio portion of the multimedia content;

comparing the most likely first language and the most likely second language;

if the first and second language are the same language, generating an output indicating that there is no language mismatch; and

if the first and second language differ, generating an output indicating that there is a language mismatch.

Assignments (2)
PATENT SECURITY AGREEMENT Recorded Dec 6, 2016
From: NICE LTD.; NICE SYSTEMS INC.; AC2 SOLUTIONS, INC.; ACTIMIZE LIMITED; INCONTACT, INC.; NEXIDIA, INC.; NICE SYSTEMS TECHNOLOGIES, INC.
To: JPMORGAN CHASE BANK, N.A., AS ADMINISTRATIVE AGENT
Reel/Frame 040821/0818 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 24, 2015
From: GARLAND, JACOB B.; LANHAM, DREW
To: NEXIDIA, INC.
Reel/Frame 037134/0102 →