IP Library › Granted Patent US 8,311,832
Granted Patent B2
US 8,311,832 · App. 12/172,260 · Granted Nov 13, 2012

Hybrid-captioning system

Assignee: International Business Machines Corporation
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 8,311,832
App. No.
12/172,260
Granted
Nov 13, 2012
Kind
B2
Abstract

A hybrid-captioning system for editing captions for spoken utterances within video includes an editor-type caption-editing subsystem, a line-based caption-editing subsystem, and a mechanism. The editor-type subsystem is that in which captions are edited for spoken utterances within the video on a groups-of-line basis without respect to particular lines of the captions and without respect to temporal positioning of the captions in relation to the spoken utterances. The line-based subsystem is that in which captions are edited for spoken utterances within the video on a line-by-line basis with respect to particular lines of the captions and with respect to temporal positioning of the captions in relation to the spoken utterances. For each section of spoken utterances within the video, the mechanism is to select the editor-type or the line-based subsystem to provide captions for the section of spoken utterances in accordance with a predetermined criteria.

Claims (14)

1. A method performed by a hybrid-captioning system to edit captions for spoken utterances within video, the method comprising:

editing, in an editor-type caption-editing subsystem of the hybrid-captioning system, captions for spoken utterances within the video on a groups-of-lines basis without respect to particular lines of the captions and without respect to temporal positioning of the captions in relation to the spoken utterances;

editing, in a line-based caption-editing subsystem of the hybrid-captioning system, the captions for spoken utterances within the video on a line-by-line basis with respect to particular lines of the captions and with respect to temporal positioning of the captions in relation to the spoken utterances;

for each section of spoken utterances within the video, selecting, by a mechanism of the hybrid-captioning system, the editor-type caption-editing subsystem or the line-based caption-editing subsystem to provide the captions for the section of spoken utterances in accordance with a predetermined criteria,

wherein the mechanism, for each section of spoken utterances within the video, selects the editor-type caption-editing subsystem or the line-based caption-editing subsystem to provide the captions for the section of spoken utterances based on a certainty level of voice recognition as to the section of spoken utterances,

wherein, for each section of spoken utterances within the video, where the certainty level of voice recognition as to the section of spoken utterances is greater than a predetermined threshold, the mechanism selects the line-based caption-editing subsystem to provide the captions for the section of spoken utterances, and otherwise selects the editor-type caption-editing subsystem to provide the captions for the section of spoken utterances,

and wherein generation of the captions is independent of an input path selected from the group of input paths essentially consisting of: a microphone path, and a file path.

2. A non-transitory computer-readable storage medium storing a computer program that upon execution by a processor implements a hybrid-captioning system to edit captions for spoken utterances within video and causes the hybrid-captioning system to perform a method comprising:

editing, in an editor-type caption-editing subsystem of the hybrid-captioning system, captions for spoken utterances within the video on a groups-of-lines basis without respect to particular lines of the captions and without respect to temporal positioning of the captions in relation to the spoken utterances;

editing, in a line-based caption-editing subsystem of the hybrid-captioning system, the captions for spoken utterances within the video on a line-by-line basis with respect to particular lines of the captions and with respect to temporal positioning of the captions in relation to the spoken utterances;

for each section of spoken utterances within the video, selecting, by a mechanism of the hybrid-captioning system, the editor-type caption-editing subsystem or the line-based caption-editing subsystem to provide the captions for the section of spoken utterances in accordance with a predetermined criteria,

wherein the mechanism, for each section of spoken utterances within the video, selects the editor-type caption-editing subsystem or the line-based caption-editing subsystem to provide the captions for the section of spoken utterances based on a certainty level of voice recognition as to the section of spoken utterances,

wherein, for each section of spoken utterances within the video, where the certainty level of voice recognition as to the section of spoken utterances is greater than a predetermined threshold, the mechanism selects the line-based caption-editing subsystem to provide the captions for the section of spoken utterances, and otherwise selects the editor-type caption-editing subsystem to provide the captions for the section of spoken utterances,

and wherein generation of the captions is independent of an input path selected from the group of input paths essentially consisting of: a microphone path, and a file path.

Continuity (2)
Continuation 11294234 · Dec 4, 2005
Related Publication 20080270134A1 · Oct 30, 2008