IP Library Granted Patent US 12,436,826
Granted Patent B2
US 12,436,826 · App. 18/476,526 · Granted Oct 7, 2025

Method for generating and executing a data processing pipeline

Inventors: Sven Winkelmann (Nuremberg, DE); Michael Kelm (Erlangen, DE); Johann Pongratz (Lahntal, DE); Mikhail Limmer (Nuremberg, DE)
Assignee: Siemens Healthineers AG
G06F9/544G06F21/577
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,436,826
App. No.
18/476,526
Granted
Oct 7, 2025
Kind
B2
Abstract

One or more example embodiments of the present invention relates to a method for generating a data processing pipeline for a cloud computing system, wherein the data processing pipeline comprises at least one software module for execution in at least one execution environment, wherein the execution environment has allocated execution environment information data. The method includes receiving at least one software module selection of a user, receiving at least one input data selection of a user, receiving at least one user planning input, checking a compliance with at least one of data protection requirements or technical requirements of the data processing pipeline, which is to be generated.

Claims (69)

1. A method for generating a data processing pipeline for a cloud computing system, the data processing pipeline including at least one software module for execution in at least one execution environment, the execution environment having allocated execution environment information data, the method comprising:

receiving, from a user, at least one software module selection for the data processing pipeline, the software module selection including at least one software module from a source of software modules, each software module having allocated module requirement data;

receiving, from the user, at least one input data selection for the data processing pipeline, the input data selection including a selection of input data for the data processing pipeline the input data having allocated data contract data;

receiving, from the user, at least one user planning input, the at least one user planning input including a description of at least one of

an architecture of the data processing pipeline,

a data technical connection between the selected at least one software module of the data processing pipeline that is to be generated, or

a distribution of the selected at least on software module to execution environments;

determining test information by checking at least one of a compliance with data protection requirements of the data processing pipeline or a compliance with technical requirements of the data processing pipeline based on at least one of

the data contract data of the selected input data,

the module requirement data of the selected at least one software module,

the at least one user planning input, or

execution environment information data; and

providing the test information.

2. The method of claim 1 , further comprising:

generating the data processing pipeline based on the selected input data, the selected at least one software module and the at least one user planning input; and

providing the generated data processing pipeline.

3. The method of claim 1 , further comprising:

determining, based on the test information, at least one of a possible software module or a required software module for fulfilment, in response to the checking resulting in non-compliance with at least one of the data protection requirements or the technical requirements; and

providing a module proposal for integration into the data processing pipeline, the module proposal including at least one of the possible software module or the required software module.

4. The method of claim 1 , further comprising:

determining, based on the test information, at least one of a possible software module or a required software module for fulfilment, in response to the checking resulting in non-compliance with at least one of the data protection requirements or the technical requirements; and

generating the data processing pipeline based on the selected input data, the selected at least one software module, the at least one user planning input, and the determined at least one possible software module or required software module.

5. The method of claim 1 , wherein

each software module of the at least software module includes at least one input interface for acquiring module input data and at least one output interface for acquiring module output data,

at least one of the module input data or the module output data has allocated data contract data, and

the checking checks whether the data protection requirements are fulfilled based on a comparison of at least one of the data contract data of the module output data or the data contract data of the module input data with at least one of

module requirement data of further selected software modules,

the at least one user planning input, or

the execution environment information data.

6. The method of claim 3 , further comprising:

determining a software module having at least one of minimal data processing, minimal data minimization, minimal anonymization, or minimal pseudonymization of at least one of the input data or module input data as the at least one possible software module or required software module.

7. The method of claim 1 , wherein the checking the at least one of the compliance with data protection requirements of the data processing pipeline or the technical requirements of the data processing pipeline checks a fulfilment of technical requirements of the selected at least one software module for at least one of

an execution in the data processing pipeline,

an execution in combination with further software modules, or

an execution in the execution environment,

based on at least one of a comparison of the module requirement data of the respective software module with the execution environment information data, the at least one user planning input, or the module requirement data of the further software modules.

8. The method of claim 1 , wherein the checking the at least one of the compliance with data protection requirements of the data processing pipeline or the technical requirements of the data processing pipeline checks a fulfilment of the data protection requirements of the data processing pipeline based on a comparison of the data contract data of the selected input data with at least one of the module requirement data of the selected at least one software module, the at least one user planning input, or the execution environment information data.

9. The method of claim 2 , wherein the generating the data processing pipeline generates a hybrid data processing pipeline, wherein the hybrid data processing pipeline comprises at least one of a distribution, an allocation, or an execution of the selected at least one software module on at least two execution environments.

10. A data processing pipeline for execution on a cloud computing system, wherein the data processing pipeline has been obtained using the method of claim 1 .

11. A method for executing a data processing pipeline on a cloud computing system, wherein the cloud computing system has at least one execution environment, at least one module source, at least one input data source and a control module, the method comprising:

receiving a pipeline data set, wherein the pipeline data set includes a data processing pipeline that is generated according to the method of claim 1 ;

deploying, by the control module, software modules included in the data processing pipeline by at least one of distributing the software modules to the at least one execution environment or allocating the software modules to the at least one execution environment; and

executing the software modules by the at least one execution environment.

12. The method of claim 11 , further comprising:

generating data channels at least one of between the software modules of the data processing pipeline or between the execution environments by the control module, based on a comparison of the execution environment information data with at least one of the data contract data of the input data, module input data, or module output data.

13. The method of claim 12 , wherein the generating the data channels generates at least one of secured or encrypted data channels.

14. A non-transitory computer-readable medium including instructions which, when executed by at least one of a computer or a cloud computing system, cause the at least one computer or cloud computing system to perform the method of claim 1 .

15. A cloud computing system comprising at least one execution environment, an input data source, and a control module, wherein the cloud computing system is configured perform the method of claim 1 .

16. The method of claim 2 , further comprising:

determining at least one of a possible software module or a required software module for fulfilment based on the test information in an event of non-compliance with at least one of the data protection requirements or the technical requirements,

wherein the generating the data processing pipeline generates the data processing pipeline based on the selected input data, the selected at least one software module, the at least one user planning input, and at least one of the possible software module or the required software module.

17. The method of claim 16 , wherein

each software module of the at least one possible or required software module includes at least one input interface for acquiring module output data and at least one output interface acquiring module output data,

at least one of the module input data or the module output data has allocated data contract data, and

the checking checks whether the data protection requirements are fulfilled based on a comparison of at least one of the data contract data of the module output data or the data contract data of the module input data with at least one of

module requirement data of further selected software modules,

the at least one user planning input, or

the execution environment information data.

18. The method of claim 2 , wherein the checking the at least one of the compliance with data protection requirements of the data processing pipeline or the technical requirements of the data processing pipeline checks a fulfilment of technical requirements of the selected at least one software module for at least one of

an execution in the data processing pipeline,

an execution in combination with the further software modules, or

an execution in the execution environment,

based on at least one of a comparison of the module requirement data of the respective software module with the execution environment information data, the user planning input or the module requirement data of the further software modules.

19. The method of claim 18 , wherein the checking the at least one of the compliance with data protection requirements of the data processing pipeline or the technical requirements of the data processing pipeline checks a fulfilment of the data protection requirements of the data processing pipeline based on a comparison of the data contract data of the selected input data with at least one of the module requirement data of the selected at least one software module, the at least one user planning input, or the execution environment information data.

20. The method of claim 19 , wherein the generating the data processing pipeline generates a hybrid data processing pipeline, wherein the hybrid data processing pipeline comprises at least one of a distribution, an allocation, or an execution of the selected at least one software module on at least two execution environments.

21. The method of claim 1 , further comprising:

generating the data processing pipeline based on the selected input data, the selected at least one software module and the at least one user planning input in response to the test information indicating compliance with the data protection requirements and the technical requirements;

providing the data processing pipeline in response to generating the data processing pipeline; and

providing the test information in response to the test information indicating non-compliance with at least one of the data protection requirements or the technical requirements.

Assignments (4)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 20, 2025
From: SIEMENS HEALTHCARE DIAGNOSTICS PRODUCTS GMBH
To: SIEMENS HEALTHINEERS AG
Reel/Frame 072072/0157 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 15, 2025
From: PONGRATZ, JOHANN
To: SIEMENS HEALTHCARE DIAGNOSTICS PRODUCTS GMBH
Reel/Frame 072032/0607 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 15, 2025
From: WINKELMANN, SVEN; KELM, MICHAEL; LIMMER, MIKHAIL
To: SIEMENS HEALTHINEERS AG
Reel/Frame 072032/0685 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 20, 2023
From: SIEMENS HEALTHCARE GMBH
To: SIEMENS HEALTHINEERS AG
Reel/Frame 066267/0346 →
Priority Claims (1)
DE 10 2022 210 351.3 · Sep 29, 2022 · national
Continuity (1)
Related Publication 20240111613A1 · Apr 4, 2024
References Cited (10)
US 20230010019A1 · Muthuswamy · 2023 [cited by examiner]
Kubeflow “The Machine Learning Toolkit for Kubernetes”; 5] https://www.kubeflow.org/ [Online Sep. 21, 2022). [cited by applicant]
Orion Innovation “Orion Pseudonymization Tool”, https://www.orioninc.com/products/pseudonymization-tool/ [Online Sep. 21, 2022]. [cited by applicant]
“Power Automate” https://powerautomate.microsoft.com/de-de/ [Online Sep. 21, 2022]; and English translation thereof. [cited by applicant]
Caristix “De-Identify HL7 Sensitive Data and Protect PHI”, https://caristix.com/complete-and-specialized-hl7-fhir-solutions/cloak-desktop-live/ [Online Sep. 21, 2022). [cited by applicant]
Apache Airfllow “Airflow is a platform created by the community to programmatically author, schedule and monitor workflows”; https://airflow.apache.org/ [Online Sep. 21, 2022]. [cited by applicant]
“syngo.via VB60A”, DICOM Conformance Statement; https://cdn0.scrvt.com/39b415fb07de4d9656c7b516d8e2d907/6c2ec855974ae4f1/0b62fb04ba87/DCS_syngo_via_VB60A.pdf. [cited by applicant]
Apeer com “Automated Image Analysisscalable solution for reproducible results”; https://apeer.com/; Download Sep. 21, 2022. [cited by applicant]
Veil ai.:“solutions enabling better use of health data for research, development and innovation”, https://veil.ai/ Download Sep. 21, 2022. [cited by applicant]
Mabotuwana et al; An HL7 Data Pseudonymization Pipeline; International Conference on Healthcare Informatics; 978-1-4673-9548-9/15 $31.00 © 2015 IEEE; DOI 10.1109/ICHI.2015.43. [cited by applicant]