IP Library Granted Patent US 12713110
Granted Patent B2
US 12713110 · App. 18/727,313 · Granted Aug 18, 2026

Image processing method and apparatus, and device and storage medium

Inventors: Xingyi Wang (Beijing, CN); Daoyu Wang (Beijing, CN); Hui Sun (Beijing, CN)
Assignee: Beijing Zitiao Network Technology Co., Ltd.
H04N21/816H04N21/8106H04N21/8153
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12713110
App. No.
18/727,313
Granted
Aug 18, 2026
Kind
B2
Abstract

A method of image processing, an apparatus, a device and a storage medium are provided. The method includes: acquiring at least one target image from an image set corresponding to a target image work in response to receiving a first preset triggering operation on a preset display interface, wherein the target image work is displayed in a preset display form in the preset display interface, and the target image work includes at least one of an image work to-be-posted or a posted image work; and acquiring a target audio corresponding to the at least one target image and generating a target video with the at least one target image and the target audio, wherein the target video is set to be storable.

Claims (59)

1 . A method of image processing, comprising:

presenting at least two options corresponding to a first image work in a first display interface in response to receiving a second triggering operation on the first display interface, wherein the at least two options indicate processing manners for the first image work, and the at least two options comprise a video generation option;

acquiring at least one first image from an image set corresponding to the first image work in response to receiving a first triggering operation on a first display interface, wherein the first image work is displayed in a first display form in the first display interface, and the first image work comprises an image work to be posted or a posted image work; and

acquiring a fourth audio corresponding to the at least one first image and generating a first video using the at least one first image and the fourth audio, wherein the first video is configured to be storable.

2 . The method according to claim 1 , wherein the fourth audio comprises a third audio, the at least one first image comprises all images of the image set,

wherein the generating the first video with the at least one first image and the fourth audio comprises:

acquiring a first image show order and a single-image single-show duration; and

generating the first video using the at least one first image and the fourth audio based on the image show order and the single-image single-show duration, wherein a playing order of the at least one first image in the first video coincides with the image show order, and a playing duration of a single first image in the first video coincides with the single-image single-show duration.

3 . The method according to claim 2 , wherein the acquiring the first image show order and the single-image single-show duration comprises:

acquiring an image show order and a single-image single-show duration that are set by a user who creates the first image work as the first image show order and the single-image single-show duration.

4 . The method according to claim 2 , wherein the first display form comprises showing images in the image set in sequence in a process of playing the third audio;

wherein the acquiring the first image show order and the single-image single-show duration comprises:

acquiring an image show order and a single-image single-show duration corresponding to the first display form as the first image show order and the single-image single-show duration.

5 . The method according to claim 1 , wherein the third audio is an audio set by a user who creates the first image work or the third audio is an audio in the first image work.

6 . The method according to claim 1 , wherein the acquiring the at least one first image from the image set corresponding to the first image work in response to receiving the first triggering operation on the first display interface comprises:

acquiring the at least one first image from an image set corresponding to the first image work in response to receiving a selection operation for the video generation option at the first display interface.

7 . The method according to claim 6 , wherein the at least two options further comprise a current image acquisition option,

wherein the presenting the at least two options corresponding to the first image work in the first display interface in response to receiving the second triggering operation on the first display interface comprises:

maintaining presenting of a current image in the first display interface in response to receiving the second triggering operation on the first display interface, and showing the at least two options corresponding to the first image work in the first display interface.

8 . The method according to claim 1 , wherein the generating the first video with the at least one first image and the fourth audio comprises:

determining a first video generation manner, wherein the first video generation manner comprises at least one selected from a group consisting of an image playing order, a single image playing duration, a playing switching effect of neighbor images, an association of an image playing timing and a tempo of the fourth audio, and a video playing total duration; and

generating, based on the first video generation manner, the first video using the at least one first image and the fourth audio.

9 . The method according to claim 8 , wherein the determining the first video generation manner comprises:

acquiring an image source file corresponding to the at least one first image or an audio source file of the fourth audio, and a descriptive file corresponding to the first image work, and parsing the image source file or the audio source file, and the descriptive file, and loading a parsed result into a memory;

acquiring first attribute information of the at least one first image or second attribute information of the fourth audio based on the parsed result, determining a candidate video generation manner according to the first attribute information of the at least one first image or the second attribute information of the fourth audio; and

determining the first video generation manner according to a selection operation of a current user for the candidate video generation manner.

10 . The method according to claim 1 , wherein the acquiring the fourth audio corresponding to the at least one first image comprises:

acquiring a corresponding first image according to a selection operation of a current user for an image in the image set corresponding to the first image work.

11 . The method according to claim 1 , wherein the acquiring the fourth audio corresponding to the at least one first image comprises:

acquiring the fourth audio corresponding to the at least one first image according to an audio selection operation of a current user.

12 . An electronic device comprising:

at least one processor; and

a non-transitory memory with instructions thereon,

wherein the instructions upon execution by the processor, cause the processor to implement a method comprising:

presenting at least two options corresponding to a first image work in a first display interface in response to receiving a second triggering operation on the first display interface, wherein the at least two options indicate processing manners for the first image work, and the at least two options comprise a video generation option;

acquiring at least one first image from an image set corresponding to the first image work in response to receiving a first triggering operation on a first display interface, wherein the first image work is displayed in a first display form in the first display interface, and the first image work comprises an image work to be posted or a posted image work; and

acquiring a fourth audio corresponding to the at least one first image and generating a first video using the at least one first image and the fourth audio, wherein the first video is set to be storable.

13 . A non-transitory computer-readable storage medium, on which a computer program is stored, wherein when the computer program is executed by a processor, the processor implements a method comprising:

presenting at least two options corresponding to a first image work in a first display interface in response to receiving a second triggering operation on the first display interface, wherein the at least two options indicate processing manners for the first image work, and the at least two options comprise a video generation option;

acquiring at least one first image from an image set corresponding to the first image work in response to receiving a first preset triggering operation on a first display interface, wherein the first image work is displayed in a first display form in the first display interface, and the first image work comprises an image work to be posted or a posted image work; and

acquiring a fourth audio corresponding to the at least one first image and generating a first video using the at least one first image and the fourth audio, wherein the first video is set to be storable.

14 . The electronic device according to claim 12 , wherein the fourth audio comprises a third audio, the at least one first image comprises all images of the image set,

wherein the processor is further caused to:

acquire the first image show order and a single-image single-show duration; and

generate the first video using the at least one first image and the fourth audio based on the image show order and the single-image single-show duration, wherein a playing order of the at least one first image in the first video coincides with the image show order, and a playing duration of a single first image in the first video coincides with the single-image single-show duration.

15 . The electronic device according to claim 12 , wherein the third audio is an audio set by a user who creates the first image work or the third audio is an audio in the first image work.

16 . The electronic device according to claim 12 , wherein the processor is further caused to:

acquire at least one first image from an image set corresponding to the first image work in response to receiving a selection operation for the video generation option at the first display interface.

17 . The electronic device according to claim 12 , wherein the processor is further caused to:

determine a first video generation manner, wherein the first video generation manner comprises at least one selected from a group consisting of an image playing order, a single image playing duration, a playing switching effect of neighbor images, an association of an image playing timing and a tempo of the fourth audio, and a video playing total duration; and

generate, based on the first video generation manner, the first video using the at least one first image and the fourth audio.

18 . The electronic device according to claim 17 , wherein the processor is further caused to:

acquire an image source file corresponding to the at least one first image or an audio source file of the fourth audio, and a descriptive file corresponding to the first image work, and parsing the image source file or the audio source file, and the descriptive file, and loading a parsed result into a memory;

acquire first attribute information of the at least one first image or second attribute information of the fourth audio based on the parsed result, determining a candidate video generation manner according to the first attribute information of the at least one first image or the second attribute information of the fourth audio; and

determine the first video generation manner according to a selection operation of a current user for the candidate video generation manner.

19 . The electronic device according to claim 12 , wherein the acquiring the fourth audio corresponding to the at least one first image comprises:

acquiring a corresponding first image according to a selection operation of a current user for an image in the image set corresponding to the first image work.

20 . The electronic device according to claim 12 , wherein the acquiring the fourth audio corresponding to the first image comprises:

acquiring the fourth audio corresponding to the at least one first image according to an audio selection operation of a current user.