IP Library Granted Patent US 12,549,821
Granted Patent B1
US 12,549,821 · App. 17/407,986 · Granted Feb 10, 2026

User placement of closed captioning

Inventors: Brant Candelore (Escondido, CA); Mahyar Nejat (San Diego, CA); Peter Shintani (San Diego, CA)
Assignee: SATURN LICENSING LLC
H04N21/4884H04N21/4307H04N21/4316H04N21/4394
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,549,821
App. No.
17/407,986
Granted
Feb 10, 2026
Kind
B1
Abstract

Placement of Closed Captioning (CC) in content by a content provider is overridden by means of a user interface (UI) that allows the user to place CC on screen on top of the video. The CC may be derived directly from the audio and synchronized with play of the audio and video.

Claims (45)

1 . An audio video display device assembly comprising:

an electronic device that comprises a microphone configured to receive audible commands; and

an audio video display device includes:

a display;

a processor that controls the display;

a wireless interface to connect to the electronic device that comprises a Bluetooth transceiver or a Near Feld Communication element to receive the audible commands; and

a user interface;

wherein the user interface is configured to enable a selection to present closed captioning (CC) rendered by a speech-to-text (STT) converter;

wherein the audio video display device is configured to simultaneously annotate the CC to correlate to a specific speaker;

wherein the audio video display device is configured to communicate with a first decoder for decoding audio received at the audio video display device at a first time in which the first decoder is implemented via a cloud system in which timing information of the CC is implemented by the cloud system; and

wherein the audio video display device is configured to decode the audio at a second time later than the first time to generate decoded audio.

2 . The audio video display device assembly of claim 1 , wherein the STT converter is implemented by the cloud system, and the audio video display device is configured to receive.

3 . The audio video display device assembly of claim 2 , wherein the audio video display device is configured to interact with a second decoder for communicating for decoding the audio at the second time.

4 . The audio video display device assembly of claim 3 , wherein, at least one of the decoders is implemented on a server configured to communicate with the audio video display device over a network.

5 . The audio video display device assembly of claim 4 , comprising at least one speaker for playing the decoded audio.

6 . The audio video display device assembly of claim 5 , wherein the audio video display device presents the CC in synchronization with playing the decoded audio from the second decoder on the at least one speaker using the timing information.

7 . The audio video display device assembly of claim 6 , wherein the first decoder comprises an AC3 decoder.

8 . The audio video display device assembly of claim 7 , wherein the STT converter comprises a field programmable gate array.

9 . The audio video display device assembly of claim 8 , wherein the STT converter comprises a processor in the audio video display device.

10 . The audio video display device assembly of claim 9 , wherein the display is configured to present video associated with the audio in a window comprising less than all of a video display area of the display, the window comprising only first and second corners of the video display area, the CC being presented in a CC region outside of the window.

11 . The audio video display device assembly of claim 9 , wherein the display is configured to present the user interface enabling selection to present the CC rendered by the STT converter and selection to present CC received from a content source in lieu of or in addition to the CC rendered by the STT converter.

12 . The audio video display device assembly of claim 1 , wherein the display is configured to display images at 4K resolution.

13 . The audio video display device assembly of claim 1 , wherein the audio video display device is configured to be touch-enabled.

14 . The audio video display device assembly of claim 1 , wherein the electronic device comprises multiple cameras, a geographic position receiver, and a biometric sensor.

15 . The audio video display device assembly of claim 1 , wherein the audio video display device comprises a position or location receiver.

16 . A method presenting of audio video content, comprising:

receiving audible commands by a microphone of an electronic device;

selecting a present closed captioning (CC) rendered by a speech-to-text (STT) converter from a user interface;

displaying the audio video content on an audio video display device;

annotating simultaneously the CC to correlate to a specific speaker;

wirelessly connecting the electronic device that comprises a Bluetooth transceiver or a Near Feld Communication element to receive the audible commands;

communicating, by the audio video display device, with a first decoder to decode audio received at the audio video display device at a first time, the first decoder being implemented via a cloud system in which timing information of the CC is implemented by the cloud system; and

decoding, by the audio video display device, the audio at a second time later than the first time to generate decoded audio.

17 . The audio video display device assembly of claim 1 , wherein the annotation of the CC is responsive to a change in the speaker.

18 . The audio video display device assembly of claim 1 , wherein the annotation of the CC is responsive to a voice fingerprint of the specific speaker.

19 . The method of claim 16 , further comprising executing a voice fingerprinting of the specific speaker.

20 . The audio video display device assembly of claim 1 , wherein the annotation of the CC includes an indicator of the specific speaker.

21 . The method of claim 16 , wherein the annotating of the CC includes an indicator of the specific speaker.

22 . An audio video display device comprising:

a processor that controls display; and

a user interface;

wherein the user interface is configured to enable a selection to present closed captioning (CC) rendered by a speech-to-text (STT) converter;

wherein the audio video display device is configured to simultaneously annotate the CC to correlate to a specific speaker;

wherein the audio video display device is configured to communicate with a first decoder for decoding audio received at the audio video display device at a first time in which the first decoder is implemented via a cloud system in which timing information of the CC is implemented by the cloud system; and

wherein the audio video display device is configured to decode the audio at a second time later than the first time to generate decoded audio.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jun 7, 2022
From: CANDELORE, BRANT; NEJAT, MAHYAR; SHINTANI, PETER
To: SONY CORPORATION
Reel/Frame 060123/0244 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jun 7, 2022
From: SONY CORPORATION
To: SATURN LICENSING LLC
Reel/Frame 060123/0443 →
Continuity (2)
Continuation 16537267 · Aug 9, 2019
Continuation 15646297 · Jul 11, 2017
References Cited (31)
US 5828416A · Ryan · 1998 [cited by examiner]
US 8713625B2 · Blanchard et al. · 2014 [cited by applicant]
US 9131280B2 · Xiong et al. · 2015 [cited by applicant]
US 9451207B2 · Mountain · 2016 [cited by applicant]
US 9568997B2 · Wilairat et al. · 2017 [cited by applicant]
US 10304458B1 · Woo · 2019 [cited by examiner]
US 20070253680A1 · Mizote · 2007 [cited by examiner]
US 20080293443A1 · Pettinato · 2008 [cited by examiner]
US 20080295040A1 · Crinon · 2008 [cited by applicant]
US 20090162036A1 · Fujii · 2009 [cited by examiner]
US 20090228948A1 · Guarin et al. · 2009 [cited by applicant]
US 20100014595A1 · Platzer · 2010 [cited by examiner]
US 20100111322A1 · Park · 2010 [cited by applicant]
US 20100303159A1 · Schultz et al. · 2010 [cited by applicant]
US 20110173537A1 · Hemphill · 2011 [cited by applicant]
US 20130027613A1 · Kim · 2013 [cited by examiner]
US 20130262123A1 · Boukadakis · 2013 [cited by examiner]
US 20140123195A1 · Han · 2014 [cited by examiner]
US 20140136195A1 · Abdossalami et al. · 2014 [cited by applicant]
US 20150172766A1 · Shin · 2015 [cited by examiner]
US 20150363389A1 · Zhang et al. · 2015 [cited by applicant]
US 20160066055A1 · Nir · 2016 [cited by examiner]
US 20160098849A1 · Shintani et al. · 2016 [cited by applicant]
US 20160100213A1 · Song et al. · 2016 [cited by applicant]
US 20160119438A1 · Abramson · 2016 [cited by examiner]
US 20180174600A1 · Chaudhuri · 2018 [cited by examiner]
US 20180220201A1 · Stoksik · 2018 [cited by examiner]
US 20180227424A1 · Dorsey et al. · 2018 [cited by applicant]
US 20180337964A1 · Caramma · 2018 [cited by applicant]
FR 2850821A1 · 2004 [cited by applicant]
WO WO2016039285A1 · 2016 [cited by examiner]