IP Library Granted Patent US 12,574,050
Granted Patent B2
US 12,574,050 · App. 19/026,329 · Granted Mar 10, 2026

Federated latent transformer deep learning core

Inventor: Brian Galvin (Silverdale, WA)
Assignee: AtomBeam Technologies Inc.
H03M7/3059G06N20/00H03M7/6005
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,574,050
App. No.
19/026,329
Filed
Jan 16, 2025
Granted
Mar 10, 2026
Kind
B2
Art Unit
2845
USPC
707/693
Abstract

A system and method for a federated deep learning platform utilizing homomorphically-compressed and encrypted data. The system comprises multiple client devices, each with a local dataset, and a central server hosting a deep learning core. Client devices convert local data into codewords, which are also homomorphically encrypted. The central server processes these encrypted codewords without decryption, preserving data privacy. The platform supports at least two architectural variants: a conventional Transformer trained on codewords, and a Latent Transformer operating on latent space vectors. Both variants eliminate the need for embedding and positional encoding layers. The system aggregates encrypted model updates from clients, enabling collaborative learning while maintaining data confidentiality. Additional features comprise differential privacy implementation and adaptive federated optimization techniques. This innovative approach allows for efficient, privacy-preserving distributed learning across diverse datasets, addressing key challenges in federated learning such as data heterogeneity, non-IID distributions, and communication efficiency.

Claims (37)

1 . A computer system comprising a hardware memory, wherein the computer system is configured to execute software instructions stored on nontransitory machine-readable storage media that, when executed by at least one processor, cause the computer system to:

orchestrate federated deep learning with a plurality of client devices by receiving a plurality of encrypted codewords from the plurality of client devices;

process the encrypted codewords using a deep learning core without decrypting the codewords; and

generate as output an encrypted codeword response to the input using the deep learning core;

wherein the deep learning core comprises a latent transformer core that was initially trained on encrypted codeword inputs to predict a plurality of probable future encrypted codewords that extend the input sequence.

2 . The system of claim 1 , wherein the deep learning core comprises a latent transformer architecture.

3 . The system of claim 2 , wherein each client device further comprises a variational autoencoder encoder that generates latent space vectors from the plurality of codewords.

4 . The system of claim 3 , wherein the latent transformer architecture processes the latent space vectors.

5 . The system of claim 1 , wherein the plurality of programming instructions further cause the computing device to:

aggregate encrypted model updates from the plurality of client devices;

update the deep learning core based on the aggregated encrypted model updates; and

facilitate federated learning by iteratively updating the deep learning core based on encrypted updates from the client devices.

6 . The system of claim 5 , wherein the plurality of programming instructions further cause the computing device to:

implement differential privacy by:

adding calibrated noise to the encrypted model updates before aggregation;

enforcing a privacy budget across multiple rounds of federated learning; and

dynamically adjusting the level of noise based on the privacy budget consumption;

thereby enhancing privacy guarantees for individual client datasets while maintaining model utility.

7 . A method for federated deep learning using homomorphically-compressed and encrypted data, comprising the steps of:

orchestrating federated deep learning with a plurality of client devices by receiving a plurality of encrypted codewords from the plurality of client devices;

processing the encrypted codewords using a deep learning core without decrypting the codewords; and

generating as output an encrypted codeword response to the input using the deep learning core;

wherein the deep learning core comprises a latent transformer core that was initially trained on encrypted codeword inputs to predict a plurality of probable future encrypted codewords that extend the input sequence.

8 . The method of claim 7 , wherein the deep learning core comprises a latent transformer architecture.

9 . The method of claim 8 , wherein each client device further comprises a variational autoencoder encoder that generates latent space vectors from the plurality of codewords.

10 . The method of claim 9 , wherein the latent transformer architecture processes the latent space vectors.

11 . The method of claim 10 , further comprising the step of generating output vectors from processed latent space vectors using a variational autoencoder decoder.

12 . The method of claim 7 , further comprising the steps of:

aggregating encrypted model updates from the plurality of client devices;

updating the deep learning core based on the aggregated encrypted model updates; and

facilitating federated learning by iteratively updating the deep learning core based on encrypted updates from the client devices.

13 . The method of claim 12 , further comprising the steps of:

implementing differential privacy by:

adding calibrated noise to the encrypted model updates before aggregation;

enforcing a privacy budget across multiple rounds of federated learning; and

dynamically adjusting the level of noise based on the privacy budget consumption;

thereby enhancing privacy guarantees for individual client datasets while maintaining model utility.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Feb 3, 2025
From: GALVIN, BRIAN
To: ATOMBEAM TECHNOLOGIES INC.
Reel/Frame 070095/0192 →
Continuity (36)
Continuation 18919394 · Oct 17, 2024
Continuation In Part 18898608 · Sep 26, 2024
Continuation In Part 18893984 · Sep 24, 2024
Continuation In Part 18890748 · Sep 19, 2024
Continuation In Part 18770652 · Jul 12, 2024
Continuation In Part 18770652 · Jul 12, 2024
Continuation In Part 18737906 · Jun 7, 2024
Continuation In Part 18736498 · Jun 6, 2024
Continuation In Part 18623018 · Mar 31, 2024
Continuation In Part 18503135 · Nov 6, 2023
Continuation 18305305 · Apr 21, 2023
Continuation In Part 18190044 · Mar 24, 2023
Continuation In Part 17875201 · Jul 27, 2022
Continuation In Part 17727913 · Apr 25, 2022
Continuation 17514913 · Oct 29, 2021
Continuation 17458747 · Aug 27, 2021
Continuation 17404699 · Aug 17, 2021
Continuation In Part 17404699 · Aug 17, 2021
Continuation In Part 17234007 · Apr 19, 2021
Continuation In Part 17180439 · Feb 19, 2021
Continuation In Part 16923039 · Jul 7, 2020
Continuation In Part 16923039 · Jul 7, 2020
Continuation In Part 16716098 · Dec 16, 2019
Continuation 16455655 · Jun 27, 2019
Continuation In Part 16455655 · Jun 27, 2019
Continuation In Part 16200466 · Nov 26, 2018
Continuation In Part 15975741 · May 9, 2018
Provisional Application 63651359 · May 23, 2024
Provisional Application 63485518 · Feb 16, 2023
Provisional Application 63388411 · Jul 12, 2022
Provisional Application 63232041 · Aug 11, 2021
Provisional Application 63140111 · Jan 21, 2021
Provisional Application 63027166 · May 19, 2020
Provisional Application 62926723 · Oct 28, 2019
Provisional Application 62578824 · Oct 30, 2017
Related Publication 20250167802A1 · May 22, 2025
References Cited (26)
US 8555400B2 · Shi et al. · 2013 [cited by applicant]
US 10091529B2 · Lee et al. · 2018 [cited by applicant]
US 11276001B1 · Sutherland · 2022 [cited by examiner]
US 11461690B2 · Szeto et al. · 2022 [cited by applicant]
US 20150154646A1 · Mishra · 2015 [cited by examiner]
US 20190026489A1 · Nerurkar · 2019 [cited by examiner]
US 20220156631A1 · Kanso et al. · 2022 [cited by applicant]
US 20230019128A1 · Zeghidour · 2023 [cited by examiner]
US 20230131694A1 · Saber et al. · 2023 [cited by applicant]
US 20230169709A1 · Im et al. · 2023 [cited by applicant]
US 20230222334A1 · Clement et al. · 2023 [cited by applicant]
US 20240185037A1 · Park et al. · 2024 [cited by applicant]
US 20240195438A1 · Isik et al. · 2024 [cited by applicant]
US 20240220829A1 · Hu et al. · 2024 [cited by applicant]
US 20240370464A1 · Lehmann et al. · 2024 [cited by applicant]
Gilard-Bachrach et al. (“Cryptonets: Applying Neural Networks to encrypted data with high throughput and Accuracy.” International Conference on Machine learning, New York, 1016) (Year: 2016). [cited by examiner]
Worrall et al. (Interpretable transforms with encoder-decoder networks. Proceedings of the IEEE International Conference on Computer Vision, 2017, pp. 5726-5735. (Year: 2017). [cited by examiner]
Balaneshin-Kordan, Saeid et al., “Deep Neural Architecture for Multi-Modal Retrieval based on Joint Embedding Space for Text and Images”, Association for Computing Machinery, Feb. 5-9, 2018, pp. 1-9, Marina Del Rey, CA,… [cited by applicant]
Khan, Abdul Rafae et al, “Coding Textual Inputs Boosts the Accuracy of Neural Networks”, 2020 Conference on Empirical Methods in Natural Language Processing, Nov. 16-20, 2020, pp. 1350-1360. [cited by applicant]
Messina, Nicola et al., “Towards Efficient Cross-Modal Visual Textual Retrieval using Transformer-Encoder Deep Features”, 2021 International Conference on Content-Based Multimedia Indexing, 2021, pp. 1-6, United States. [cited by applicant]
Pagnoni et al; “Byte Latent Transformer: Patches Scale Better Than Tokens”, FAIR at Meta, 2024. [cited by applicant]
Seo, Beomseok et al., “How Does A Transformer Learn Compression? An Attention Study on Huffman and LZ4”, Department of Electronic and Electrical Engineering, Dec. 12, 2023, pp. 1-10, vol. 11, Seoul, South Korea. [cited by applicant]
Tjandra et al; “Speech-to-Speech Translation Between Untranscribed Unknown Languages”, IEEE, Nara Institute of Science and Technology Japan, 2019. [cited by applicant]
Vaswani, Ashish et al., “Attention is All You Need”, 31st Conference on Neural Information Processing Systems, 2017, pp. 1-11, Long Beach, CA, USA. [cited by applicant]
Wang, Tianming et al., “T-CVAE: Transformer-Based Conditioned Variational Autoencoder for Story Completion”, Proceedings of the Twenty-Eigth Joint Conference on Artificial Intelligence, pp. 5233-5239. [cited by applicant]
Wieting, John et al., “A Bilingual Generative Transformer for Semantic Sentence Embedding”, Nov. 19, 2020, pp. 1-14. [cited by applicant]