IP Library Granted Patent US 12701301
Granted Patent B2
US 12701301 · App. 18/875,644 · Granted Aug 4, 2026

Interaction method and apparatus, electronic device, and storage medium

Inventor: Xingge Li (Beijing, CN)
Assignee: BEIJING ZITIAO NETWORK TECHNOLOGY CO., LTD.
H04N21/4884H04N21/2387
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12701301
App. No.
18/875,644
Granted
Aug 4, 2026
Kind
B2
Abstract

The present disclosure relates to an interaction method, an apparatus, an electronic device and a storage medium. The method comprises: receiving a text display operation for a first media content, wherein the first media content includes a video content; in response to the text display operation, displaying in a preset region a list of text sentences of the first media content, wherein the list of text sentences contains at least two text sentences, and the at least two text sentences each have a corresponding audio sentence in target audio data of the first media content.

Claims (65)

1 . An interaction method, comprising:

receiving a text display operation for a first media content, wherein the first media content comprises a video content;

in response to the text display operation, displaying in a preset region a list of text sentences of the first media content, wherein the list of text sentences contains at least two text sentences, and the at least two text sentences each have a corresponding audio sentence in target audio data of the first media content; and

in response to a sentence selecting operation performed within the preset region, determining a first text sentence selected by the sentence selecting operation and executing at least one of:

in response to a copy operation for the first text sentence, copying the first text sentence,

in response to a sharing operation for the first text sentence, sharing the first text sentence,

in response to a play operation for the first text sentence, playing the first media content by taking a time node in the first media content corresponding to a starting point of the first text sentence as a playing starting point, or

in response to a media content generating operation for the first text sentence, generating a second media content containing the first text sentence.

2 . The method of claim 1 , wherein the at least two text sentences comprise a current text sentence and other text sentence apart from the current text sentence, and the current text sentence has a different display style from the other text sentence, and wherein the current text sentence is a text sentence currently being played.

3 . The method of claim 2 , wherein the current text sentence is displayed at a set position of the preset region, and the method further comprises:

in a case that the current text sentence changes, moving the changed current text sentence to the set position for display.

4 . The method of claim 2 , wherein after the displaying in a preset region a list of text sentences of the first media content, the method further comprises:

in response to a sentence switching operation performed within the preset region, switching a text sentence displayed in the preset region and displaying a position control, wherein the position control is provided to trigger movement of the current text sentence to a set position of the preset region for display.

5 . The method of claim 1 , wherein the generating a second media content containing the first text sentence comprises:

generating a target card containing the first text sentence; and

generating the second media content by using the target card as content and using music corresponding to at least one of the first text sentence or the target card as background music.

6 . The method of claim 1 , wherein the generating a second media content containing the first text sentence comprises: generating a second media content containing the first text sentence and posting the second media content; or wherein,

after the generating a second media content containing the first text sentence, the method further comprises: in response to a posting operation for the second media content, posting the second media content.

7 . The method of claim 1 , wherein after the determining a first text sentence selected by the sentence selecting operation, the method further comprises:

displaying the first text sentence in a selected state.

8 . The method of claim 1 , wherein a keyword to be searched contained in the at least two text sentences and other contents in the at least two text sentences apart from the keyword to be searched have different display styles, and the keyword to be searched is provided for triggering display of a search result matching with a triggered keyword to be searched.

9 . The method of claim 1 , wherein before the receiving a text display operation for a first media content, the method further comprises:

playing the first media content; and

wherein the displaying in a preset region a list of text sentences of the first media content comprises:

displaying in the preset region a list of text sentences of the first media content, and adjusting the first media content to be played outside the preset region.

10 . The method of claim 1 , wherein the text display operation comprises a media content jump operation for a third media content; and before the receiving a text display operation for a first media content, the method further comprises:

playing the third media content, wherein the third media content is a media content generated based on a second text sentence in the first media content; and

wherein the displaying in a preset region a list of text sentences of the first media content comprises:

playing the first media content outside the preset region and displaying a list of text sentences of the first media content in the preset region.

11 . The method of claim 10 , wherein before the receiving a text display operation for a first media content, the method further comprises:

in response to a control display operation for the third media content, displaying a jump control corresponding to the third media content, wherein the jump control is provided for triggering execution of the media content jump operation.

12 . The method of claim 10 , wherein the playing the first media content outside the preset region comprises:

playing the first media content outside the preset region by taking a starting point of the first media content as a playing starting point; or

playing the first media content outside the preset region by taking a time node in the first media content corresponding to a starting point of the second text sentence as a playing starting point.

13 . An electronic device, comprising:

at least one processor; and

a memory communicatively connected to the at least one processor; wherein,

the memory is stored with a computer program executable by the at least one processor, wherein the computer program is executed by the at least one processor to cause the at least one processor to perform operations comprising:

receiving a text display operation for a first media content, wherein the first media content comprises a video content;

in response to the text display operation, displaying in a preset region a list of text sentences of the first media content, wherein the list of text sentences contains at least two text sentences, and the at least two text sentences each have a corresponding audio sentence in target audio data of the first media content; and

in response to a sentence selecting operation performed within the preset region, determining a first text sentence selected by the sentence selecting operation and executing at least one of:

in response to a copy operation for the first text sentence, copying the first text sentence,

in response to a sharing operation for the first text sentence, sharing the first text sentence,

in response to a play operation for the first text sentence, playing the first media content by taking a time node in the first media content corresponding to a starting point of the first text sentence as a playing starting point, or

in response to a media content generating operation for the first text sentence, generating a second media content containing the first text sentence.

14 . A non-transitory computer-readable storage medium stored thereon with computer instructions, the computer instructions enabling a processor to perform operations comprising:

receiving a text display operation for a first media content, wherein the first media content comprises a video content;

in response to the text display operation, displaying in a preset region a list of text sentences of the first media content, wherein the list of text sentences contains at least two text sentences, and the at least two text sentences each have a corresponding audio sentence in target audio data of the first media content; and

in response to a sentence selecting operation performed within the preset region, determining a first text sentence selected by the sentence selecting operation and executing at least one of:

in response to a copy operation for the first text sentence, copying the first text sentence,

in response to a sharing operation for the first text sentence, sharing the first text sentence,

in response to a play operation for the first text sentence, playing the first media content by taking a time node in the first media content corresponding to a starting point of the first text sentence as a playing starting point, or

in response to a media content generating operation for the first text sentence, generating a second media content containing the first text sentence.

15 . The electronic device of claim 13 , wherein the at least two text sentences comprise a current text sentence and other text sentence apart from the current text sentence, and the current text sentence has a different display style from the other text sentence, and wherein the current text sentence is a text sentence currently being played.

16 . The electronic device of claim 15 , wherein the current text sentence is displayed at a set position of the preset region, and the operations further comprise:

in a case that the current text sentence changes, moving the changed current text sentence to the set position for display.

17 . The electronic device of claim 15 , wherein after the displaying in a preset region a list of text sentences of the first media content, the operations further comprise:

in response to a sentence switching operation performed within the preset region, switching a text sentence displayed in the preset region and displaying a position control, wherein the position control is provided to trigger movement of the current text sentence to a set position of the preset region for display.

18 . The electronic device of claim 13 , wherein the generating a second media content containing the first text sentence comprises:

generating a target card containing the first text sentence; and

generating the second media content by using the target card as content and using music corresponding to at least one of the first text sentence or the target card as background music.

19 . The non-transitory computer-readable storage medium of claim 14 , wherein the at least two text sentences comprise a current text sentence and other text sentence apart from the current text sentence, and the current text sentence has a different display style from the other text sentence, and wherein the current text sentence is a text sentence currently being played.

20 . The non-transitory computer-readable storage medium of claim 14 , the operations further comprising:

playing the first media content before receiving the text display operation for the first media content; and

adjusting to play the first media content outside the preset region while displaying the list of text sentences of the first media content in the preset region.