IP Library Granted Patent US 12,437,448
Granted Patent B2
US 12,437,448 · App. 19/087,497 · Granted Oct 7, 2025

System and methods for multimodal series transformation for optimal compressibility with neural upsampling

Inventor: Brian Galvin (Silverdale, WA)
Assignee: ATOMBEAM TECHNOLOGIES INC.
G06T9/002G06T7/337G06T7/38G06T11/00G06T2207/10016G06T2207/10044G06T2207/20081G06T2207/20084H04N19/66H04N19/86
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,437,448
App. No.
19/087,497
Granted
Oct 7, 2025
Kind
B2
Abstract

Image series transformation for optimal compressibility is performed with neural upsampling and error resilience. A novel correlation network composed of convolutional layers for feature extraction that extract multi-dimensional features from the image and a channel-wise transformer with attention to capture complex inter-channel dependencies. An angle optimizer enhances compressibility of an image and an error resilience subsystem improves robustness against transmission errors and data loss. The error resilience subsystem applies forward error correction coding, data partitioning based on importance, and embeds error concealment hints. This hybrid approach addresses both local and global features, mitigates compression artifacts, improves image quality, and enhances data integrity during transmission. The correlation network incorporates error correction and concealment techniques during decoding. The model's outputs enable effective image reconstruction, achieving advanced compression while preserving information for accurate analysis.

Claims (41)

1. A computer system comprising:

a hardware memory, wherein the computer system is configured to execute software instructions stored on nontransitory machine-readable storage media that:

collects a plurality of multimodal data;

processes each modality through a corresponding specialized preprocessor;

aligns and registers the preprocessed multimodal data to a common spatial-temporal reference frame;

trains an angle optimizer using the aligned multimodal data to determine optimal slicing angles that maximize compressibility while preserving cross-modal relationships;

slices the multimodal data along an optimal angle, as determined by the angle optimizer;

reconstructs the sliced multimodal data into a plurality of reconstructed representations;

encodes the plurality of reconstructed multimodal data into a plurality of compressed representations;

applies error resilience techniques to the compressed representations; and

decodes the plurality of compressed representations into a plurality of decompressed representations.

2. The computer system of claim 1 , wherein the angle optimizer is a convolutional neural network.

3. The computer system of claim 1 , wherein the plurality of multimodal data includes optical data, thermal data, hyperspectral data, and LIDAR data.

4. The computer system of claim 1 , wherein applying error resilience techniques comprises:

forward error correction coding;

data partitioning based on importance; and

embedding error concealment hints.

5. The computer system of claim 4 , wherein decoding the plurality of compressed representations includes performing error correction and concealment based on the applied error resilience techniques.

6. The computer system of claim 4 , wherein the forward error correction coding uses Reed-Solomon codes or Low-Density Parity-Check codes.

7. The computer system of claim 4 , wherein the data partitioning separates the compressed representations into at least three partitions.

8. The computer system of claim 4 , wherein the error concealment hints include information about neighboring blocks or redundant feature data.

9. A computer-implemented method comprising the steps of:

collecting a plurality of multimodal data;

processing each modality through a corresponding specialized preprocessor;

aligning and registering the preprocessed multimodal data to a common spatial-temporal reference frame;

training an angle optimizer using the aligned multimodal data to determine optimal slicing angles that maximize compressibility while preserving cross-modal relationships;

slicing the multimodal data along an optimal angle, as determined by the angle optimizer;

reconstructing the sliced multimodal data into a plurality of reconstructed representations;

encoding the plurality of reconstructed multimodal data into a plurality of compressed representations;

applying error resilience techniques to the compressed representations; and

decoding the plurality of compressed representations into a plurality of decompressed representations.

10. The method of claim 9 , wherein the angle optimizer is a convolutional neural network.

11. The method of claim 9 , wherein the plurality of multimodal data includes optical data, thermal data, hyperspectral data, and LIDAR data.

12. The method of claim 9 , wherein applying error resilience techniques comprises:

forward error correction coding;

data partitioning based on importance; and

embedding error concealment hints.

13. The method of claim 12 , wherein decoding the plurality of compressed representations includes performing error correction and concealment based on the applied error resilience techniques.

14. The method of claim 12 , wherein the forward error correction coding uses Reed-Solomon codes or Low-Density Parity-Check codes.

15. The method of claim 12 , wherein the data partitioning separates the compressed representations into at least three partitions.

16. The method of claim 12 , wherein the error concealment hints include information about neighboring blocks or redundant feature data.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 25, 2025
From: GALVIN, BRIAN
To: ATOMBEAM TECHNOLOGIES INC.
Reel/Frame 070951/0994 →
Continuity (4)
Continuation In Part 18915030 · Oct 14, 2024
Continuation In Part 18668163 · May 18, 2024
Continuation In Part 18537728 · Dec 12, 2023
Related Publication 20250218053A1 · Jul 3, 2025
References Cited (21)
US 7629922B2 · Winstead et al. · 2009 [cited by applicant]
US 7876257B2 · Vetro et al. · 2011 [cited by applicant]
US 9060733B2 · Bruder et al. · 2015 [cited by applicant]
US 11656353B2 · Li et al. · 2023 [cited by applicant]
US 12015776B2 · Besenbruch · 2024 [cited by examiner]
US 12058333B1 · Li et al. · 2024 [cited by applicant]
US 20130301890A1 · Kaempfer · 2013 [cited by examiner]
US 20230090743A1 · Pinto et al. · 2023 [cited by applicant]
US 20230154055A1 · Besenbruch et al. · 2023 [cited by applicant]
US 20230236271A1 · Fessler et al. · 2023 [cited by applicant]
US 20230239500A1 · Yang et al. · 2023 [cited by applicant]
US 20230368438A1 · Shen et al. · 2023 [cited by applicant]
CN 106210742A · 2016 [cited by applicant]
CN 111182301A · 2020 [cited by applicant]
CN 111869206A · 2020 [cited by applicant]
CN 115706798A · 2023 [cited by applicant]
CN 117011619A · 2023 [cited by examiner]
CN 118902603A · 2024 [cited by examiner]
WO 2023017975A1 · 2023 [cited by applicant]
Lowe, David G., “Distinctive Image Features from Scale-Invariant Keypoints”, International Journal of Computer Vision, Jan. 5, 2004, pp. 1-28, Vancouver, B.C., Canada. [cited by applicant]
Kim, Youngseop et al; “Lossless Volumetric Medical Image Compression”, Proc. SPIE 3808, Applications of Digital Image Processing XXII, Oct. 18, 1999. [cited by applicant]