IP Library › Granted Patent US 12,425,529
Granted Patent B2
US 12,425,529 · App. 18/253,186 · Granted Sep 23, 2025

Video processing method and apparatus for triggering special effect, electronic device and storage medium

Inventors: Qinghua Zhou (Beijing, CN); Shiyin Wang (Beijing, CN)
Assignee: Beijing Bytedance Network Technology Co., Ltd.
H04N5/2621G06T5/50G06T5/70G06T7/12G06V10/25H04N5/272G06T2207/20221
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,425,529
App. No.
18/253,186
Granted
Sep 23, 2025
Kind
B2
Abstract

A video processing method and apparatus, an electronic device and a storage medium are provided in the present disclosure. In the video processing method provided in the present disclosure, a target frame image of a target video is fixed in response to a triggering instruction, and a target object in the target frame image is removed, then a padding operation is performed on a target area in the target frame image so as to generate and display a padding frame image. Thus, a special effect that the target object disappears from the target video can be realized, thereby providing visual special effects of more personalized and better visual perception for a user in a video application.

Claims (56)

1. A video processing method, comprising:

fixing a target frame image in a target video in response to a triggering instruction; and

removing a target object in the target frame image, and performing a padding operation on a target area in the target frame image so as to generate and display a padding frame image, wherein the target area comprises a vacant area after the target object is removed from the target frame image,

wherein the removing the target object in the target frame image and performing the padding operation on the target area in the target frame image comprises:

identifying various pixel points in the target frame image by using a preset object segmentation model so as to generate a target binary image with the same size as the target frame image; and

determining the target area of the target object in the target frame image according to the target binary image, wherein the target area in the target binary image comprises various pixel points with pixel values being target values.

2. The video processing method according to claim 1 , wherein after the generating and displaying the padding frame image, the method further comprises:

playing a first special effect sequence frame by using the padding frame image as a background, wherein the first special effect sequence frame is used to dynamically display special effect particles according to a preset path.

3. The video processing method according to claim 1 , wherein after the fixing the target frame image of the target video, the method further comprises:

displaying a preset second special effect, wherein the preset second special effect is used to enable the target frame image to present a visual blurring effect.

4. The video processing method according to claim 1 , wherein after the fixing the target frame image of the target video, the method further comprises:

displaying a preset third special effect, wherein the preset third special effect is used to enable the target frame image to present a visual shaking effect.

5. The video processing method according to claim 1 , wherein after generating and displaying the padding frame image, the method further comprises:

continuously performing a padding operation on the target area in a subsequent frame image of the target video, wherein the subsequent frame image is located after the padding frame image in the target video.

6. The video processing method according to claim 1 , further comprising:

determining that the target object is an object of a target type, wherein the target object or the object of the target type comprises at least one of a target person, a target animal, or a target building.

7. The video processing method according to claim 1 , wherein the triggering instruction comprises at least one of a target gesture instruction, a target voice instruction, a target expression instruction, a target limb instruction, or a target text instruction.

8. The video processing method according to claim 1 , wherein the performing the padding operation on the target area in the target frame image comprises:

fusing the target binary image with the target frame image to obtain a model inputting image;

inputting the model inputting image into an image patching model to generate a processing frame image; and

replacing the target area in the target frame image with a target area in the processing frame image to generate the padding frame image.

9. The video processing method according to claim 8 , wherein the image patching model is provided in a terminal device, and the terminal device processes the target video based on the image patching model.

10. The video processing method according to claim 8 , wherein the inputting the model inputting image into the image patching model to generate the processing frame image comprises:

inputting the model inputting image into a first image patching model to generate a first padding image;

performing pixel truncation on the first padding image by using a preset pixel threshold value so as to generate a second padding image;

inputting the second padding image into a second image patching model to generate a third padding image, wherein the second image patching model has higher patching precision than the first image patching model; and

performing pixel truncation on the third padding image by using the preset pixel threshold value so as to generate a fourth padding image, wherein the processing frame image comprises the fourth padding image.

11. An electronic device, comprising:

a processor; and

a memory, configured to store a computer program;

a display, configured to display a video after processing by the processor;

wherein the processor is configured to:

fix a target frame image in a target video in response to a triggering instruction;

remove a target object in the target frame image, and performing a padding operation on a target area in the target frame image so as to generate and display a padding frame image, wherein the target area comprises a vacant area after the target object is removed from the target frame image;

identify various pixel points in the target frame image by using a preset object segmentation model so as to generate a target binary image with the same size as the target frame image; and

determine the target area of the target object in the target frame image according to the target binary image, wherein the target area in the target binary image comprises various pixel points with pixel values being target values.

12. The electronic device according to claim 11 , wherein the processor is further configured to:

play a first special effect sequence frame by using the padding frame image as a background, wherein the first special effect sequence frame is used to dynamically display special effect particles according to a preset path.

13. The electronic device according to claim 11 , wherein the processor is further configured to:

display a preset second special effect, wherein the preset second special effect is used to enable the target frame image to present a visual blurring effect.

14. The electronic device according to claim 11 , wherein the processor is further configured to:

display a preset third special effect, wherein the preset third special effect is used to enable the target frame image to present a visual shaking effect.

15. The electronic device according to claim 11 , wherein the processor is further configured to:

continuously perform a padding operation on the target area in a subsequent frame image of the target video, wherein the subsequent frame image is located after the padding frame image in the target video.

16. A non-transitory computer readable storage medium, wherein the computer readable storage medium stores computer execution instructions, and when a processor executes the computer execution instructions, the following operations are executed:

fixing a target frame image in a target video in response to a triggering instruction; and

removing a target object in the target frame image, and performing a padding operation on a target area in the target frame image so as to generate and display a padding frame image, wherein the target area comprises a vacant area after the target object is removed from the target frame image,

wherein the removing the target object in the target frame image and performing the padding operation on the target area in the target frame image comprises:

identifying various pixel points in the target frame image by using a preset object segmentation model so as to generate a target binary image with the same size as the target frame image; and

determining the target area of the target object in the target frame image according to the target binary image, wherein the target area in the target binary image comprises various pixel points with pixel values being target values.

17. The non-transitory computer readable storage medium according to claim 16 , wherein after the generating and displaying the padding frame image, the following operation are further executed:

playing a first special effect sequence frame by using the padding frame image as a background, wherein the first special effect sequence frame is used to dynamically display special effect particles according to a preset path.

18. The non-transitory computer readable storage medium according to claim 16 , wherein after the fixing the target frame image of the target video, the following operation are further executed:

displaying a preset second special effect, wherein the preset second special effect is used to enable the target frame image to present a visual blurring effect.

19. The non-transitory computer readable storage medium according to claim 16 , wherein after the fixing the target frame image of the target video, the following operation are further executed:

displaying a preset third special effect, wherein the preset third special effect is used to enable the target frame image to present a visual shaking effect.

Assignments (3)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jul 11, 2023
From: ZHOU, QINGHUA
To: BEIJING ZITIAO NETWORK TECHNOLOGY CO., LTD.
Reel/Frame 064212/0310 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jul 11, 2023
From: WANG, SHIYIN
To: SHENZHEN JINRITOUTIAO TECHNOLOGY CO., LTD.
Reel/Frame 064212/0381 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jul 11, 2023
From: BEIJING ZITIAO NETWORK TECHNOLOGY CO., LTD.; SHENZHEN JINRITOUTIAO TECHNOLOGY CO., LTD.
To: BEIJING BYTEDANCE NETWORK TECHNOLOGY CO., LTD.
Reel/Frame 064212/0428 →
Priority Claims (1)
CN 202011280804.3 · Nov 16, 2020 · national
Continuity (1)
Related Publication 20230421716A1 · Dec 28, 2023
References Cited (54)
US 9514523B2 · Park et al. · 2016 [cited by applicant]
US 20100027961A1 · Gentile et al. · 2010 [cited by applicant]
US 20130346588A1 · Zhang et al. · 2013 [cited by applicant]
US 20140184858A1 · Yu · 2014 [cited by examiner]
US 20140219555A1 · Hsia et al. · 2014 [cited by applicant]
US 20150091900A1 · Yang et al. · 2015 [cited by applicant]
US 20160350598A1 · Yamaji · 2016 [cited by examiner]
US 20180075304A1 · Li · 2018 [cited by examiner]
US 20190132642A1 · Wang et al. · 2019 [cited by applicant]
US 20190196698A1 · Cohen · 2019 [cited by examiner]
US 20210012502A1 · Mulford · 2021 [cited by examiner]
US 20210248721A1 · Tian et al. · 2021 [cited by applicant]
US 20210407051A1 · Pardeshi · 2021 [cited by examiner]
CN 102164234A · 2011 [cited by applicant]
CN 104574311A · 2015 [cited by applicant]
CN 104680487A · 2015 [cited by applicant]
CN 108829893A · 2018 [cited by applicant]
CN 109215091A · 2019 [cited by applicant]
CN 109960453A · 2019 [cited by applicant]
CN 110225246A · 2019 [cited by applicant]
CN 110728639A · 2020 [cited by applicant]
CN 111161275A · 2020 [cited by applicant]
CN 111179159A · 2020 [cited by applicant]
CN 111260537A · 2020 [cited by applicant]
CN 111353071A · 2020 [cited by applicant]
CN 111416939A · 2020 [cited by applicant]
CN 111444921A · 2020 [cited by applicant]
CN 111556278A · 2020 [cited by applicant]
CN 111754528A · 2020 [cited by applicant]
CN 111832538A · 2020 [cited by applicant]
CN 112188058A · 2021 [cited by applicant]
CN 112199526A · 2021 [cited by applicant]
CN 112637517A · 2021 [cited by applicant]
EP 3945494A1 · 2022 [cited by applicant]
JP 2008005084A · 2008 [cited by applicant]
JP 2011096018A · 2011 [cited by applicant]
JP 2013077873A · 2013 [cited by applicant]
JP 2014096661A · 2014 [cited by applicant]
JP 2020129356A · 2020 [cited by applicant]
JP 7583165B2 · 2024 [cited by applicant]
WO 2020022055A1 · 2020 [cited by applicant]
WO 2020125739A1 · 2020 [cited by applicant]
China National Intellectual Property Administration, International Search Report and Written of Opinion Issued in Application No. PCT/CN2021/130708, Feb. 7, 2022, WIPO, 12 pages. [cited by applicant]
China National Intellectual Property Administration, Office Action and Search Report Issued in Application No. 202011280804.3, Jul. 7, 2022, 8 pages. [cited by applicant]
China National Intellectual Property Administration, Office Action and Search Report Issued in Application No. 202011280804.3, Apr. 15, 2022, 17 pages. [cited by applicant]
China National Intellectual Property Administration, Office Action and Search Report Issued in Application No. 202011280804.3, Feb. 8, 2022, 17 pages. [cited by applicant]
China National Intellectual Property Administration, Office Action and Search Report Issued in Application No. 202011280804.3, Sep. 21, 2022, 6 pages. [cited by applicant]
Ni, H. et al., “Large Damaged Area Image Inpainting Algorithm Based on Matching Model for Broken Structure Line,” Computer Science, vol. 43, No. 10, Oct. 2016, 6 pages. Submitted with English abstract. [cited by applicant]
“How To Improve the Communication Effect of Short Video Platform—Taking Tiktok Short Video as an Example,” News Dissemination 2018.5, 4 pages. Submitted with English abstract. [cited by applicant]
Japan Patent Office, Office Action Issued in Application No. 2023528594, Jun. 4, 2024, 8 pages. [cited by applicant]
Decision to Grant a Patent for Japanese Application No. 2023-528594, mailed Oct. 1, 2024, 5 pages. [cited by applicant]
European Patent Office, Extended European Search Report Issued in Application No. 21891255.8, Mar. 14, 2024, Germany, 7 pages. [cited by applicant]
European Patent Office, Communication pursuant to Article 94(3) EPC for European Application No. 21891255.8, mailed Nov. 20, 2024, 4 pages. [cited by applicant]
ISA China National Intellectual Property Administration, International Search Report for International Application No. PCT/CN2021/117199, mailed Dec. 17, 2021, 5 pages. [cited by applicant]