IP Library Granted Patent US 12,574,608
Granted Patent B2
US 12,574,608 · App. 18/144,907 · Granted Mar 10, 2026

User interface method and apparatus for video navigation using captions

Inventors: V. Michael Bove, Jr. (Wrentham, MA); Serhad Doken (Bryn Mawr, PA); Ning Xu (Irvine, CA); Tao Chen (Palo Alto, CA)
Assignee: Adeia Guides Inc.
H04N21/4884G06F3/0488G06F3/16H04N21/47217
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,574,608
App. No.
18/144,907
Granted
Mar 10, 2026
Kind
B2
Abstract

Systems and methods for navigating a video via interaction with text overlaid on the video are described. In one example, a method includes generating for display a video, and generating for display at least one line of text overlaid over the video. Then, in response to receiving a directional user interface input for at least a portion of the at least one line of text, the method includes modifying a play position of the video based on a direction of the directional user interface input for the at least a portion of the at least one line of text.

Claims (76)

1 . A method for media navigation comprising:

generating for display a video;

generating for display at least one line of text overlaid over the video;

in response to receiving a first directional swipe in a first direction and having a first speed, for at least a portion of the at least one line of text:

modifying a play position of the video in a first manner based on the first direction and the first speed of the first directional swipe for the at least a portion of the at least one line of text; and

modifying the display of the at least one line of text overlaid over the video based on the first direction and the first speed of the first directional swipe; and

in response to receiving a selection of a specific portion of text of the at least one line of text overlaid over the video:

identifying from the video a plurality of scenes related to the specific portion of text; and

providing a plurality of selectable options corresponding to the plurality of scenes wherein a selection of a respective selectable option causes a respective scene to be generated for display; and

in response to receiving a second directional swipe in a second direction orthogonal to the first direction and having a second speed, for at least a portion of the at least one line of text:

modifying the play position of the video in a second manner based on the second direction and the second speed of the second directional swipe for the at least a portion of the at least one line of text; and

modifying the display of the at least one line of text overlaid over the video based on the second direction and the second speed of the second directional swipe.

2 . The method of claim 1 , wherein modifying the play position of the video in the first manner comprises modifying the play position at a rate proportional to the first speed of the first directional swipe, and wherein modifying the play position of the video in the second manner comprises modifying the play position at a rate proportional to the second speed of the second directional swipe.

3 . The method of claim 1 , further comprising:

in response to receiving a directional swipe comprising a selection and upward drag of the at least a portion of the at least one line of text, modifying the play position of the video forward proportional to a speed of the upward drag; and

in response to receiving a directional swipe comprising a selection and downward drag of the at least a portion of the at least one line of text, modifying the play position of the video backward proportional to a speed of the downward drag.

4 . The method of claim 3 , wherein the directional swipe is received via a touch screen, and wherein the selection and upward drag comprises a tap and swipe.

5 . The method of claim 1 , further comprising:

generating for display a plurality of lines of text overlaid over the video;

generating for display a current line of text that corresponds to a current video segment in a middle position, wherein the current line of text is visually different from adjacent lines of text;

generating for display a first adjacent line of text that corresponds to an adjacent preceding video segment above the middle position; and

generating for display a second adjacent line of text that corresponds to an adjacent following video segment below the middle position.

6 . The method of claim 1 , further comprising:

in response to receiving a selection of a selected line of text of the at least one line of text overlaid over the video, modifying the play position of the video to a video segment corresponding to the selected line of text.

7 . The method of claim 1 , wherein the at least one line of text overlaid over the video comprises a first page of text generated for display in a page format and comprising a plurality of lines of text, the method further comprising:

in response to receiving the first directional swipe in the first direction and having the first speed, for at least a portion of the first page of text:

generating for display a second page of text comprising a second plurality of lines of text; and

modifying the play position of the video in the first manner to a first video segment corresponding to a first line of text of the second plurality of lines of text of the second page of text; and

in response to receiving the second directional swipe in the second direction orthogonal to the first direction and having the second speed, for at least a portion of the first page of text:

generating for display a third page of text comprising a third plurality of lines of text; and

modifying the play position of the video in the second manner to a second video segment corresponding to the first line of text of the third plurality of lines of text of the third page of text.

8 . The method of claim 1 , further comprising:

generating for display credits corresponding to the video; and

in response to receiving a user interface input selecting a selected portion of the credits:

generating for display an indication of one or more video segments that correspond to the selected portion of the credits;

loading the one or more video segments of the video that correspond to the selected portion of the credits; and

setting a current play position of the video to a first video segment of the one or more video segments that correspond to the selected portion of the credits.

9 . The method of claim 1 , further comprising:

generating for display song information for one or more songs played during the video; and

in response to receiving a user interface input selecting a selected portion of the song information:

generating for display an indication of one or more video segments that correspond to the selected portion of the song information; and

causing an audio output of audio corresponding to the selected portion of the song information.

10 . A system for media navigation comprising:

input/output circuitry configured to:

generate for display a video; and

generate for display at least one line of text overlaid over the video; and control circuitry configured to:

in response to receiving a first directional swipe in a first direction and having a first speed, for at least a portion of the at least one line of text:

modify a play position of the video in a first manner based on the first direction and the first speed of the first directional swipe for the at least a portion of the at least one line of text; and

modify the display of the at least one line of text overlaid over the video based on the first direction and the first speed of the first directional swipe; and

in response to receiving a selection of a specific portion of text of the at least one line of text overlaid over the video:

identify from the video a plurality of scenes related to the specific portion of text; and

provide a plurality of selectable options corresponding to the plurality of scenes wherein a selection of a respective selectable option causes a respective scene to be generated for display; and

in response to receiving a second directional swipe in a second direction orthogonal to the first direction and having a second speed, for at least a portion of the at least one line of text:

modify the play position of the video in a second manner based on the second direction and the second speed of the second directional swipe for the at least a portion of the at least one line of text; and

modify the display of the at least one line of text overlaid over the video based on the second direction and the second speed of the second directional swipe.

11 . The system of claim 10 , wherein the control circuitry is configured to modify the play position of the video in the first manner at a rate proportional to the first speed of the first directional swipe, and wherein the control circuitry is configured to modify the play position of the video in the second manner at a rate proportional to the second speed of the second directional swipe.

12 . The system of claim 10 , wherein the control circuitry is further configured to:

in response to receiving a directional swipe comprising a selection and upward drag of the at least a portion of the at least one line of text, modify the play position of the video forward proportional to a speed of the upward drag; and

in response to receiving a directional swipe comprising a selection and downward drag of the at least a portion of the at least one line of text, modify the play position of the video backward proportional to a speed of the downward drag.

13 . The system of claim 12 , wherein the directional swipe is received via a touch screen, and wherein the selection and upward drag comprises a tap and swipe.

14 . The system of claim 10 , wherein the input/output circuitry is further configured to:

generate for display a plurality of lines of text overlaid over the video;

generate for display a current line of text that corresponds to a current video segment in a middle position, wherein the current line of text is visually different from adjacent lines of text;

generate for display a first adjacent line of text that corresponds to an adjacent preceding video segment above the middle position; and

generate for display a second adjacent line of text that corresponds to an adjacent following video segment below the middle position.

15 . The system of claim 10 , wherein the control circuitry is further configured to:

in response to receiving a selection of a selected line of text of the at least one line of text overlaid over the video, modify the play position of the video to a video segment corresponding to the selected line of text.

16 . The system of claim 10 , wherein the at least one line of text overlaid over the video comprises a first page of text generated for display in a page format and comprising a plurality of lines of text, wherein the control circuitry is further configured to:

in response to receiving a directional user interface input for at least a portion of the first page of text:

generate for display a second page of text comprising a second plurality of lines of text; and

modify the play position of the video to a video segment corresponding to a first line of text of the second plurality of lines of text of the second page of text.

17 . The system of claim 10 , wherein the control circuitry is further configured to:

in response to receiving a user interface input selecting a selected portion of the at least one line of text overlaid over the video:

generate for display an indication of one or more video segments of the video corresponding to the selected portion of the at least one line of text;

load the one or more video segments of the video; and

generate for display, the one or more video segments of the video.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jul 10, 2023
From: BOVE JR., V. MICHAEL; DOKEN, SERHAD; XU, NING; CHEN, TAO
To: ADEIA GUIDES INC.
Reel/Frame 064195/0175 →
Continuity (1)
Related Publication 20240380948A1 · Nov 14, 2024
References Cited (41)
US 6061056A · Menard · 2000 [cited by examiner]
US 8479238B2 · Chen · 2013 [cited by examiner]
US 8610673B2 · Storrusten · 2013 [cited by examiner]
US 9681165B1 · Gupta · 2017 [cited by examiner]
US 10489496B1 · Sen · 2019 [cited by examiner]
US 12216700B2 · Taboriskiy · 2025 [cited by examiner]
US 20090164460A1 · Jung · 2009 [cited by examiner]
US 20100262938A1 · Woods · 2010 [cited by examiner]
US 20110282906A1 · Wong · 2011 [cited by examiner]
US 20140068692A1 · Archibong · 2014 [cited by examiner]
US 20140213332A1 · Otsuka · 2014 [cited by examiner]
US 20140282660A1 · Oztaskent · 2014 [cited by examiner]
US 20140365882A1 · Lemay · 2014 [cited by examiner]
US 20150016801A1 · Homma · 2015 [cited by examiner]
US 20150256763A1 · Niemi · 2015 [cited by examiner]
US 20150309686A1 · Morin · 2015 [cited by examiner]
US 20150382047A1 · Van Os · 2015 [cited by examiner]
US 20160094875A1 · Peterson · 2016 [cited by examiner]
US 20160295294A1 · Lan · 2016 [cited by examiner]
US 20170072301A1 · Abecassis · 2017 [cited by examiner]
US 20170091153A1 · Thimbleby · 2017 [cited by examiner]
US 20180025078A1 · Quennesson · 2018 [cited by examiner]
US 20180335922A1 · Nilo · 2018 [cited by examiner]
US 20180349494A1 · Zhao · 2018 [cited by examiner]
US 20190155955A1 · Castaneda · 2019 [cited by examiner]
US 20190179846A1 · Taboriskiy · 2019 [cited by examiner]
US 20190289359A1 · Sekar · 2019 [cited by examiner]
US 20190394419A1 · Zhang · 2019 [cited by examiner]
US 20200356593A1 · Azzinnari · 2020 [cited by examiner]
US 20200396497A1 · Liu · 2020 [cited by examiner]
US 20210006867A1 · Liu · 2021 [cited by examiner]
US 20210105538A1 · Ogawa · 2021 [cited by examiner]
US 20210150222A1 · Evans · 2021 [cited by examiner]
US 20210160571A1 · Menendez · 2021 [cited by examiner]
US 20210233427A1 · Pesta · 2021 [cited by examiner]
US 20210266641A1 · Selfors · 2021 [cited by examiner]
US 20220021950A1 · Wei · 2022 [cited by examiner]
US 20220321972A1 · Chandrashekar · 2022 [cited by examiner]
US 20230164296A1 · Chang · 2023 [cited by examiner]
US 20240007718A1 · Sheng · 2024 [cited by examiner]
CA 2572709A1 · 2007 [cited by applicant]