IP Library › Granted Patent US 12,316,844
Granted Patent B2
US 12,316,844 · App. 17/795,149 · Granted May 27, 2025

3D point cloud enhancement with multiple measurements

Inventors: Jiahao Pang (Plainsboro, NJ); Xue Zhang (Toronto, CA); Gene Cheung (Toronto, CA); Dong Tian (Boxborough, MA)
Assignee: InterDigital VC Holdings, Inc.
H04N19/124H04N19/172H04N19/463H04N19/597H04N19/85
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,316,844
App. No.
17/795,149
Granted
May 27, 2025
Kind
B2
Abstract

Systems and methods are described for refining first point cloud data using at least second point cloud data and one or more sets of quantizer shifts. An example point cloud decoding method includes obtaining data representing at least a first point cloud and a second point cloud; obtaining information identifying at least a first set of quantizer shifts associated with the first point cloud; and obtaining refined point cloud data based on at least the first point cloud, the first set of quantizer shifts, and the second point cloud. The obtaining of the refined point cloud data may include performing a subtraction based on at least the first set of quantizer shifts. Corresponding encoding systems and methods are also described.

Claims (38)

1. A point cloud decoding method comprising:

obtaining data representing at least a first point position in a first point cloud and a second point position in a second point cloud, the first point cloud representing a first observation of a scene and the second point cloud representing a second observation of the scene;

obtaining information identifying at least a first set of quantizer shifts associated with the first point cloud; and

obtaining refined point cloud data including a refined point position based on at least the first point position, the first set of quantizer shifts, and the second point position.

2. A point cloud decoder apparatus comprising a processor configured to perform at least:

obtaining data representing at least a first point position in a first point cloud and a second point position in a second point cloud, the first point cloud representing a first observation of a scene and the second point cloud representing a second observation of the scene;

obtaining information identifying at least a first set of quantizer shifts associated with the first point cloud; and

obtaining refined point cloud data including a refined point position based on at least the first point position, the first set of quantizer shifts, and the second point position.

3. The method of claim 1 , wherein obtaining the refined point cloud data comprises performing a subtraction based on at least the first set of quantizer shifts.

4. The method of claim 1 , further comprising:

obtaining information identifying a second set of quantizer shifts associated with the second point cloud, the second set of quantizer shifts being different from the first set of quantizer shifts;

wherein obtaining the refined point cloud data is further based on the second set of quantizer shifts.

5. The method of claim 1 wherein the first point cloud represents a left view of a scene and the second point cloud represents a right view of the scene.

6. The method of claim 1 , wherein the first point cloud and the second point cloud are frames associated with different times.

7. The method of claim 1 , wherein the first set of quantizer shifts comprises at least a first shift associated with the first point position in the first point cloud and a different second shift associated with a position of a different point in the first point cloud.

8. The method of claim 1 , wherein the data representing the first point cloud comprises a first parameter (y 1 ) representing the first point position having a first quantization range, the data representing the second point cloud comprises a second parameter (y 2 ) representing the second point position having a second quantization range, and wherein obtaining the refined point cloud data comprises obtaining a third parameter ({tilde over (x)}) representing the refined point position in the quantization range of both the first parameter (y 1 ) and the second parameter (y 2 ).

9. The method of claim 1 , wherein obtaining the refined point cloud data (x l ) based on the first point cloud (y l ) and the second point cloud (y r ) comprises selecting the refined first point cloud (x l ) to substantially maximize a product of factors comprising one or more of the following factors:

a conditional probability Pr(y l |x l ) of the first point cloud (y l ) given the refined first point cloud (x l ),

a conditional probability Pr(y r |g(x l )) of the second point cloud (y r ) given an estimate g(x l ) of a second refined point cloud (x r ), where the estimate g(x l ) is based on the first refined point cloud (x l ),

a prior probability Pr(x l ) of the first refined point cloud data (x l ), and

a prior probability Pr(g(x l )) of the estimate g(x l ) of the second refined point cloud (x r ).

10. A point cloud encoding method comprising:

obtaining data representing at least a first point position in a first point cloud and a second point position in a second point cloud, the first point cloud representing a first observation of a scene and the second point cloud representing a second observation of the scene;

quantizing the first point position in the first point cloud and the second point position in the second point cloud, wherein quantizing the first point position in the first point cloud data comprises adding at least a first set of quantizer shifts to the first point position such that quantization bins of the first point position and the second point position are not aligned; and

encoding in a bitstream the quantized first and second point clouds.

11. A point cloud encoding apparatus comprising a processor configured to perform at least:

obtaining data representing at least a first point position in a first point cloud and a second point position in a second point cloud, the first point cloud representing a first observation of a scene and the second point cloud representing a second observation of the scene;

quantizing the first point position in the first point cloud and the second point position in the second point cloud, wherein quantizing the first point position in the first point cloud data comprises adding at least a first set of quantizer shifts to the first point position such that quantization bins of the first point position and the second point position are not aligned; and

encoding in a bitstream the quantized first and second point clouds.

12. The method of claim 10 , further comprising encoding in the bitstream information indicating the first set of quantizer shifts.

13. The method of claim 10 , wherein quantizing the first point position and the second point position includes adding a second set of quantizer shifts to the second point cloud, further comprising encoding in the bitstream information indicating the second set of quantizer shifts.

14. The method of claim 10 , wherein the first point cloud represents a left view of a scene and the second point cloud represents a right view of the scene.

15. The decoder apparatus of claim 2 , wherein obtaining the refined point cloud data comprises performing a subtraction based on at least the first set of quantizer shifts.

16. The decoder apparatus of claim 2 , wherein the first point cloud represents a left view of a scene and the second point cloud represents a right view of the scene.

17. The decoder apparatus of claim 2 , wherein the first point cloud and the second point cloud are frames associated with different times.

18. The encoding method of claim 10 , wherein the first point cloud and the second point cloud are frames associated with different times.

19. The encoding apparatus of claim 11 , wherein the first point cloud represents a left view of a scene and the second point cloud represents a right view of the scene.

20. The encoding apparatus of claim 11 , wherein the first point cloud and the second point cloud are frames associated with different times.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 16, 2023
From: PCMS HOLDINGS, INC.
To: INTERDIGITAL VC HOLDINGS, INC.
Reel/Frame 062383/0528 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 9, 2022
From: PANG, JIAHAO; ZHANG, XUE; CHEUNG, GENE; TIAN, DONG
To: PCMS HOLDINGS, INC.
Reel/Frame 062042/0502 →
Continuity (2)
Provisional Application 62970956 · Feb 6, 2020
Related Publication 20230056576A1 · Feb 23, 2023
References Cited (44)
US 11549805B2 · Gan · 2023 [cited by examiner]
US 11889114B2 · Iguchi · 2024 [cited by examiner]
US 12159436B2 · Zhang · 2024 [cited by examiner]
US 12159437B2 · Yang · 2024 [cited by examiner]
US 20170347100A1 · Chou · 2017 [cited by applicant]
US 20190304139A1 · Joshi · 2019 [cited by examiner]
US 20190311502A1 · Mammou · 2019 [cited by applicant]
US 20200021844A1 · Yea · 2020 [cited by applicant]
US 20200021856A1 · Tourapis · 2020 [cited by applicant]
US 20200219290A1 · Tourapis · 2020 [cited by examiner]
US 20200234491A1 · Pöyhtäri · 2020 [cited by examiner]
US 20240114167A1 · Iguchi et al. · 2024 [cited by applicant]
CA 3070678A1 · 2019 [cited by applicant]
CN 108961320A · 2018 [cited by applicant]
CN 110691243A · 2020 [cited by applicant]
EP 3554074A1 · 2019 [cited by applicant]
JP 2024061790A · 2024 [cited by applicant]
WO 2019050931A1 · 2019 [cited by applicant]
WO 2019078696A1 · 2019 [cited by applicant]
WO 2019182811A1 · 2019 [cited by applicant]
WO 2019183113A1 · 2019 [cited by applicant]
WO 2020175176A1 · 2020 [cited by applicant]
G—PCC coding description v4; Jul. 2019. (Year: 2019). [cited by examiner]
Point cloud compression proposed by Sony; Oct. 2017. (Year: 2017). [cited by examiner]
Signaling of QP variations for Adaptive geometry quantization in Point cloud coding; 2019. (Year: 2019). [cited by examiner]
Kim, M. et al., “Development of Multiview Image Generation Simulator for Depth Map Quantization” In International Conference on Virtual, Augmented and Mixed Reality (pp. 58-64). Springer, Berlin, Heidelberg. Jul. 2013. [cited by applicant]
Wei, Ku-Chu . et al., “Quantization error reduction in depth maps.” In 2013 IEEE International Conference on Acoustics, Speech and Signal Processing, pp. 1543-1547. IEEE, 2013. [cited by applicant]
Gauthier, L. et al., “Understanding MPEG-I coding standardization in immersive VR/AR applications.” SMPTE motion imaging journal 128, No. 10 (2019): 33-39. [cited by applicant]
International Search Report and Written Opinion for PCT/US2021/016895 (10 pages). [cited by applicant]
International Preliminary Report on Patentability for PCT/2021/016895 issued Jul. 28, 2022 (7 pages). [cited by applicant]
Furukawa, Y. et al., “Accurate, Dense, and Robust Multiview Stereopsis.” IEEE transactions on pattern analysis and machine intelligence 32, No. 8, 2009 (8 pages). [cited by applicant]
Han, J. et al., “Enhanced computer vision with microsoft kinect sensor: A review.” IEEE transactions on cybernetics vol. 43, No. 5, 2013 pp. 1318-1334 (17 pages). [cited by applicant]
Wan, P. et al., “High Bit-Precision Image Acquisition and Reconstruction by Planned Sensor Distortion.” In 2014 IEEE International Conference on Image Processing (ICIP), pp. 1773-1777. IEEE, 2014 (5 pages). [cited by applicant]
Wan, P. et al., “Precision Enhancement of 3-D Surfaces From Compressed Multiview Depth Maps.” IEEE Signal Processing Letters 22, No. 10 (2015 (6 pages). [cited by applicant]
Schnabel, R. “Octree-based Point-Cloud Compression.” PBG@ SIGGRAPH 3, 2006 (11 pages). [cited by applicant]
Gumhold, S. et al., “Predictive point-cloud compression.” In ACM SIGGRAPH 2005 Sketches, p. 137, 2005 (1 page). [cited by applicant]
Loop, C. “Computing Rectifying Homographies For Stereo Vision.” In Proceedings. 1999 IEEE Computer Society Conference on Computer Vision and Pattern Recognition (Cat. No PR00149), vol. 1, IEEE, 1999 (14 pages). [cited by applicant]
Kang, Y.-S. et al., “An Efficient Image Rectification Method For Parallel Multi-Camera Arrangement.” IEEE transactions on Consumer Electronics vol. 57, No. 3, 2011 pp. 1041-1048 (8 pages). [cited by applicant]
Zhang, Zhengyou, “A Flexible New Technique for Camera Calibration”. Microsoft Corporation, MSR-TR-98-71, last updated on Aug. 13, 2008, (22 pages). [cited by applicant]
Zhang, Z., “A Flexible New Technique for Camera Calibration” in IEEE Transactions on Pattern Analysis & Machine Intelligence, vol. 22, No. 11, pp. 1330-1334, 2000 (5 pages). [cited by applicant]
Scharstein, D. et al., “High-accuracy stereo depth maps using structured light.” In 2003 IEEE Computer Society Conference on Computer Vision and Pattern Recognition, 2003. Proceedings., vol. 1, IEEE, 2003 (8 pages). [cited by applicant]
Jin, J. et al., “Region-aware 3-D warping for DIBR.” IEEE Transactions on Multimedia vol. 18, No. 6, Jun. 2016, pp. 953-966 (14 pages). [cited by applicant]
Cheung, G. “Graph spectral image processing.” Proceedings of the IEEE 106, No. 5, 2018 (22 pages). [cited by applicant]
Qu, Huamin, et al., “A Framework for Sample-Based Rendering with O-Buffers”. Proceedings of the 14th IEEE Visualization Conference (VIS' 2003), Oct. 19-24, 2003, Seattle, Washington, pp. 441-448 (8 pages). [cited by applicant]
Cited By (1)
US 12,712,935