IP Library Granted Patent US 12,337,236
Granted Patent B2
US 12,337,236 · App. 18/763,530 · Granted Jun 24, 2025

Virtual-environment-based object construction method and apparatus, computer device, and computer-readable storage medium

Inventor: Chao Shen (Shenzhen, CN)
Assignee: Tencent Technology (Shenzhen) Company Limited
A63F13/52G06T7/55G06T7/74G06T7/90G06T19/20G06T2200/24G06T2207/10016G06T2207/10028G06T2207/30244G06T2210/12G06T2210/56G06T2219/2008G06T2219/2012G06T2219/2024
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,337,236
App. No.
18/763,530
Granted
Jun 24, 2025
Kind
B2
Abstract

A virtual-environment-based object construction method and apparatus, and a computer-readable storage medium are disclosed in this disclosure, relating to the field of virtual environment. Performed by a terminal comprising a camera, the method including: displaying an environment interface; receiving a capture operation being used for acquiring three-dimensional (3D) information of a to-be-acquired object; receiving a position input operation for determining a display position of the target object in the virtual environment; and displaying the target object at the display position in the virtual environment according to the capture operation and the position input operation.

Claims (71)

1. A virtual-environment-based object construction method, performed by a terminal, the method comprising:

displaying an environment interface, the environment interface comprising an image corresponding to a virtual environment;

obtaining a feature point cloud from a capture operation performed by a camera, the feature point cloud determining a style of a target object to be displayed;

receiving a position input operation indicating a display position of the target object in the virtual environment; and

displaying the target object at the display position in the virtual environment by filling a region corresponding to the feature point cloud with voxel blocks, wherein filling the region corresponding to the feature point cloud with the voxel blocks comprises:

performing three-dimensional (3D) slicing on a bounding box corresponding to the feature point cloud to obtain voxel regions; and

filling the voxel regions with the voxel blocks according to an inclusion relationship between the voxel regions and pixels in the feature point cloud by filling the voxel regions with the voxel blocks in response to a quantity of the pixels in the voxel regions being greater than a preset quantity.

2. The method according to claim 1 , further comprising:

receiving the capture operation to capture n frames of images around a to-be-acquired object, and capturing the n frames of images, n being a positive integer; and

generating the feature point cloud according to depth information corresponding to each of the n frames of images, or performing 3D reconstruction on the to-be-acquired object based on the n frames of images to obtain the feature point cloud.

3. The method according to claim 2 , wherein generating the feature point cloud according to the depth information corresponding to the each of the n frames of images comprises:

determining a relative position of the camera during capturing of the each of the n frames of images, the relative position being determined according to a relationship with a position in which the camera captures a key-frame image;

determining positions of pixels in a 3D space according to positions of the pixels in the each of the n frames of images, the depth information corresponding to the pixels, and the relative position of the camera; and

obtaining the feature point cloud according to the positions of the pixels in the 3D space.

4. The method according to claim 2 , wherein receiving the capture operation for capturing the n frames of images around the to-be-acquired object comprises one of:

receiving a video capture operation around the to-be-acquired object, a video stream captured by the video capture operation comprising the n frames of images; or

receiving a fixed-point capture operation around the to-be-acquired object, the fixed-point capture operation being used for capturing the n frames of images at designated positions around the to-be-acquired object.

5. The method according to claim 1 , further comprising:

receiving a 3D slicing operation to obtain a slicing mode corresponding to each dimension of 3D dimensions, the 3D slicing operation being used to perform the 3D slicing on the bounding box corresponding to the feature point cloud according to the slicing mode.

6. The method according to claim 5 , wherein receiving the 3D slicing operation comprises:

receiving slice quantities from a slice quantity input operation, the slice quantities corresponding to three dimensions of the feature point cloud; and performing the 3D slicing on the bounding box based on the slice quantities; or

receiving a sliding slicing operation indicating the slice quantities, and performing the 3D slicing on the bounding box according to the slice quantities, the slice quantities corresponding to three dimensions being used to determine a degree of refinement of the target object generated by the feature point cloud.

7. The method according to claim 1 , wherein filling the voxel regions with the voxel blocks comprises:

determining a weighted mean color of the pixels in the voxel regions to obtain a target color, and filling the voxel regions with the voxel blocks having a color closest to the target color.

8. The method according to claim 1 , wherein filling the voxel regions with the voxel blocks comprises:

determining a color with a highest proportion in distribution as a target color according to color distribution of the pixels in the voxel regions, and filling the voxel regions with the voxel blocks having a color closest to the target color.

9. A device for virtual-environment-based object construction, the device comprising:

a memory storing a plurality of computer instructions; and

a processor configured to execute the plurality of computer instructions, wherein upon execution of the plurality of computer instructions, the processor is configured to cause the device to:

display an environment interface, the environment interface comprising an image corresponding to a virtual environment;

obtain a feature point cloud from a capture operation performed by a camera, the feature point cloud determining a style of a target object to be displayed;

receive a position input operation indicating a display position of the target object in the virtual environment; and

display the target object at the display position in the virtual environment by filling a region corresponding to the feature point cloud with voxel blocks, wherein filling the region corresponding to the feature point cloud with the voxel blocks comprises:

performing three-dimensional (3D) slicing on a bounding box corresponding to the feature point cloud to obtain voxel regions; and

filling the voxel regions with the voxel blocks according to an inclusion relationship between the voxel regions and pixels in the feature point cloud by filling the voxel regions with the voxel blocks in response to a quantity of the pixels in the voxel regions being greater than a preset quantity.

10. The device according to claim 9 , wherein, when the processor, upon execution of the plurality of computer instructions, is further configured to cause the device to:

receive the capture operation to capture n frames of images around a to-be-acquired object, and capture the n frames of images, n being a positive integer; and

generate the feature point cloud according to depth information corresponding to each of the n frames of images, or perform 3D reconstruction on the to-be-acquired object based on the n frames of images to obtain the feature point cloud.

11. The device according to claim 10 , wherein in order to cause the device to generate the feature point cloud according to the depth information corresponding to the each of the n frames of images, the processor is configured to cause the device to:

determine a relative position of the camera during capturing of the each of the n frames of images, the relative position being determined according to a relationship with a position in which the camera captures a key-frame image;

determine positions of pixels in a 3D space according to positions of the pixels in the each of the n frames of images, the depth information corresponding to the pixels, and the relative position of the camera; and

obtain the feature point cloud according to the positions of the pixels in the 3D space.

12. The device according to claim 10 , wherein in order to cause the device to receive the capture operation to capture the n frames of images around the to-be-acquired object, the processor is configured to cause the device to:

receive a video capture operation around the to-be-acquired object, a video stream captured by the video capture operation comprising the n frames of images; or

receive a fixed-point capture operation around the to-be-acquired object, the fixed-point capture operation being used for capturing the n frames of images at designated positions around the to-be-acquired object.

13. The device according to claim 9 , wherein upon execution of the plurality of computer instructions, the processor is further configured to cause the device to:

receive a 3D slicing operation to obtain a slicing mode corresponding to each dimension of 3D dimensions, the 3D slicing operation being used to perform the 3D slicing on the bounding box corresponding to the feature point cloud according to the slicing mode.

14. The device according to claim 13 , wherein in order to receive the 3D slicing operation, the processor is configured to cause the device to:

receive slice quantities from a slice quantity input operation, the slice quantities corresponding to three dimensions of the feature point cloud; and perform the 3D slicing on the bounding box based on the slice quantities; or

receive a sliding slicing operation indicating the slice quantities, and perform the 3D slicing on the bounding box according to the slice quantities, the slice quantities corresponding to three dimensions being used to determine a degree of refinement of the target object generated by the feature point cloud.

15. The device according to claim 9 , wherein in order to cause the device to fill the voxel regions with the voxel blocks, the processor is configured to cause the device to:

determine a weighted mean color of the pixels in the voxel regions to obtain a target color, and fill the voxel regions with the voxel blocks having a color closest to the target color.

16. The device according to claim 9 , wherein in order to cause the device to fill the voxel regions with the voxel blocks, the processor is configured to cause the device to:

determine a color with a highest proportion in distribution as a target color according to color distribution of the pixels in the voxel regions, and fill the voxel regions with the voxel blocks having a color closest to the target color.

17. A non-transitory storage medium for storing computer readable instructions, the computer readable instructions, when executed by a processor in a device comprising a camera, causing the processor to:

display an environment interface, the environment interface comprising an image corresponding to a virtual environment;

obtain a feature point cloud from a capture operation performed by a camera, the feature point cloud determining a style of a target object to be displayed;

receive a position input operation indicating a display position of the target object in the virtual environment; and

display the target object at the display position in the virtual environment by filling a region corresponding to the feature point cloud with voxel blocks, wherein filling the region corresponding to the feature point cloud with the voxel blocks comprises:

performing three-dimensional (3D) slicing on a bounding box corresponding to the feature point cloud to obtain voxel regions; and

filling the voxel regions with the voxel blocks according to an inclusion relationship between the voxel regions and pixels in the feature point cloud by filling the voxel regions with the voxel blocks in response to a quantity of the pixels in the voxel regions being greater than a preset quantity.

18. The non-transitory storage medium according to claim 17 , wherein, when the computer readable instructions further cause the processor to:

receive the capture operation for capturing n frames of images around a to-be-acquired object, and capturing the n frames of images, n being a positive integer; and

generate the feature point cloud according to depth information corresponding to each of the n frames of images, or perform 3D reconstruction on the to-be-acquired object based on the n frames of images to obtain the feature point cloud.

19. The non-transitory storage medium according to claim 18 , wherein in order to cause the processor to generate the feature point cloud according to the depth information corresponding to the each of the n frames of images, the computer readable instructions cause the processor to:

determine a relative position of the camera during capturing of the each of the n frames of images, the relative position being determined according to a relationship with a position in which the camera captures a key-frame image;

determine positions of pixels in a 3D space according to positions of the pixels in the each of the n frames of images, the depth information corresponding to the pixels, and the relative position of the camera; and

obtain the feature point cloud according to the positions of the pixels in the 3D space.

20. The non-transitory storage medium according to claim 18 , wherein in order to cause the processor to receive the capture operation to capture the n frames of images around the to-be-acquired object, the computer readable instructions cause the processor to:

receive a video capture operation around the to-be-acquired object, a video stream captured by the video capture operation comprising the n frames of images; or

receive a fixed-point capture operation around the to-be-acquired object, the fixed-point capture operation being used for capturing the n frames of images at designated positions around the to-be-acquired object.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jul 3, 2024
From: SHEN, CHAO
To: TENCENT TECHNOLOGY (SHENZHEN) COMPANY LIMITED
Reel/Frame 067910/0205 →
Priority Claims (1)
CN 201910340401.4 · Apr 25, 2019 · national
Continuity (3)
Continuation 17345366 · Jun 11, 2021
Continuation PCTCN2020074910 · Feb 12, 2020
Related Publication 20240359095A1 · Oct 31, 2024
References Cited (23)
US 9779479B1 · Makinen et al. · 2017 [cited by applicant]
US 20140375629A1 · Samson · 2014 [cited by examiner]
US 20160247017A1 · Sareen et al. · 2016 [cited by applicant]
US 20160284123A1 · Hare · 2016 [cited by examiner]
US 20180264365A1 · Soederberg · 2018 [cited by examiner]
US 20190089942A1 · Hanamoto · 2019 [cited by examiner]
US 20190240580A1 · Walker · 2019 [cited by examiner]
CN 103971404A · 2014 [cited by applicant]
CN 108122275A · 2018 [cited by applicant]
CN 108136257A · 2018 [cited by applicant]
CN 108537876A · 2018 [cited by applicant]
CN 108550181A · 2018 [cited by applicant]
CN 108694741A · 2018 [cited by applicant]
CN 109410316A · 2019 [cited by applicant]
CN 109525831A · 2019 [cited by applicant]
CN 109641150A · 2019 [cited by applicant]
CN 110064200A · 2019 [cited by applicant]
WO WO2018148924A1 · 2018 [cited by applicant]
WO WO2018203029A1 · 2018 [cited by applicant]
Extended European Search Report for European Patent Application No. 20794587.4 dated Jun. 13, 2022 (8 pages). [cited by applicant]
International Search Report and Written Opinion with English tranlation for International Application No. PCT/CN2020/074910 mailed May 11, 2020 (13 pages). [cited by applicant]
Office Action with English Translation of Concise Explanation of Relevance for Chinese Patent Application No. 201910340401.4 dated Feb. 24, 2021 (15 pages). [cited by applicant]
Mo, E.; Minecraft My World, [My World 0.15 Surrounding] Pixel Image Generator, online video from 00:00 to 08:36 https://www.igiyi.com/w_19ruc9dgwh.html dated Jan. 20, 2017. [cited by applicant]