IP Library Granted Patent US 12,657,944
Granted Patent B2
US 12,657,944 · App. 18/298,400 · Granted Jun 16, 2026

OCR-based extraction of clinical data from DICOM SC images

Inventors: Poikavila Ullaskrishnan (Lebanon, NH); Ren-Yi Lo (Plainsboro, NJ)
Assignee: Siemens Healthineers AG
G06V30/153G16H30/40G16H50/70
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,657,944
App. No.
18/298,400
Granted
Jun 16, 2026
Kind
B2
Abstract

Techniques of facilitating processing of at least one DICOM SC image—e.g., using a PC or workstation in a hospital or an institution—to automatically extract clinical data therein are provided. Characters associated with the clinical data are extracted from the at least one DICOM SC image based on configuration information associated with the at least one DICOM SC image, which configuration information is obtained based on the at least one DICOM SC image.

Claims (51)

1 . A computer-implemented method comprising:

parsing, by at least one processor, at least one Digital Imaging and Communications in Medicine (DICOM) file comprising DICOM tags to identify at least one DICOM Secondary Capture (SC) image and to extract metadata associated with the at least one DICOM SC image;

determining, by the at least one processor, a SC image type based on at least one of a content pattern of pixel data of the at least one DICOM SC image or an image structure identified from the extracted metadata;

obtaining, by the at least one processor, configuration information associated with the determined SC image type, the configuration information comprising a grid-size defining sections within the pixel data of the at least one DICOM SC image, pre-defined variables associated with clinical data, and optical character recognition (OCR) page segmentation modes (PSMs) mapped to the sections;

segmenting, by the at least one processor, the pixel data of the at least one DICOM SC image into sub-images according to the grid-size;

selecting, by the at least one processor, one or more of the sub-images corresponding to the pre-defined variables associated with the clinical data and excluding remaining ones of the sub-images from further text extraction processing to condense the pixel data;

applying, by the at least one processor, optical character recognition to only the selected one or more sub-images using the OCR page segmentation modes;

generating, by the at least one processor, an output having a required structure of content comprising extracted textual content associated with the pre-defined variables; and

storing the output.

2 . The method of claim 1 , further comprising:

converting, by the at least one processor, the at least one DICOM SC image to any one of the following image formats: tag image file format, raw image format, bitmap image file format, or portable network graphic format.

3 . The method of claim 1 , further comprising:

trimming, by the at least one processor, margins of the at least one DICOM SC image.

4 . The method of claim 1 , wherein the configuration information comprises one or more keywords of the characters associated with the clinical data, and the segmenting is further based on the one or more keywords.

5 . The method of claim 4 , wherein the one or more keywords are determined by the applying optical character recognition to the at least one DICOM SC image.

6 . The method of claim 1 , wherein the configuration information comprises a template of the at least one DICOM SC image, and the segmenting is further based on the template.

7 . The method of claim 1 , further comprising:

determining, by the at least one processor, an arrangement of the characters in each of the selected one or more sub-images, wherein the arrangement comprises row-wise, column-wise, or tabular.

8 . The method of claim 7 , wherein the arrangement is row-wise, the method further comprising:

splitting, by the at least one processor, the selected one or more sub-images into rows; and

applying, by the at least one processor, the optical character recognition to each of the split rows to extract the characters therein.

9 . The method of claim 7 , wherein the arrangement is column-wise, the method further comprising:

splitting, by the at least one processor, the selected one or more sub-images into columns; and

applying, by the at least one processor, the optical character recognition to each of the split columns to extract the characters therein.

10 . The method of claim 7 , wherein the arrangement is tabular, the method further comprising:

splitting, by the at least one processor, the selected one or more sub-images into both rows and columns, respectively;

applying, by the at least one processor, the optical character recognition to each of the split rows and to each of the split columns to extract the characters therein, respectively; and

determining, by the at least one processor, a position within a table of each of the extracted characters based on positions in both the row-wisely and column-wisely extracted characters.

11 . The method of claim 1 , further comprising:

extracting, by the at least one processor, the DICOM tags from a header of the at least one DICOM SC image.

12 . The method of claim 11 , further comprising:

pairing, by the at least one processor, the extracted DICOM tags with the extracted characters associated with the clinical data; or

removing, by the at least one processor, patient health information from the extracted characters associated with the clinical data.

13 . A computing device comprising:

at least one processor; and

at least one memory;

wherein upon loading and executing program code from the at least one memory, the at least one processor is configured to:

parse at least one Digital Imaging and Communications in Medicine (DICOM) file comprising DICOM tags to identify at least one DICOM Secondary Capture (SC) image and to extract metadata associated with the at least one DICOM SC image;

determine a SC image type based on at least one of a content pattern of pixel data of the at least one DICOM SC image or an image structure identified from the extracted metadata:

obtain configuration information associated with the determined SC image type, the configuration information comprising a predefined spatial extraction profile including sub-region coordinate boundaries within the pixel data of the at least one DICOM SC image, variable identifiers associated with clinical data, and optical character recognition (OCR) segmentation parameters mapped to the predefined sub-region coordinate boundaries;

segment the pixel data of the at least one DICOM SC image into sub-images according to the sub-region coordinate boundaries;

identify a subset of the plurality of sub-images corresponding to the variable identifiers associated with the clinical data and excluding remaining ones of the sub-images from further text extraction processing;

apply optical character recognition to only the identified subset of sub-images using the OCR segmentation parameters;

generate structured extracted clinical data comprising extracted textual content associated with the variable identifiers; and

store the structured extracted clinical data in association with at least one DICOM metadata field.

14 . The computing device of claim 13 , wherein the at least one processor is further configured to trim margins of the at least one DICOM SC image.

15 . The computing device of claim 13 , wherein the at least one processor is further configured to determine an arrangement of the characters in each of the plurality of sub-images, wherein the arrangement comprises row-wise, column-wise, or tabular.

16 . The computing device of claim 13 , wherein the at least one processor is further configured to extract the DICOM tags from a header of the at least one DICOM SC image.

17 . The computing device of claim 16 , wherein the at least one processor is further configured to:

pair the extracted DICOM tags with the extracted characters associated with the clinical data; or

remove patient health information from the extracted characters associated with the clinical data.

Assignments (3)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 20, 2023
From: SIEMENS HEALTHCARE GMBH
To: SIEMENS HEALTHINEERS AG
Reel/Frame 066267/0346 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 21, 2023
From: SIEMENS MEDICAL SOLUTIONS USA, INC.
To: SIEMENS HEALTHCARE GMBH
Reel/Frame 063403/0691 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 13, 2023
From: ULLASKRISHNAN, POIKAVILA; LO, REN-YI
To: SIEMENS MEDICAL SOLUTIONS USA, INC.
Reel/Frame 063309/0241 →
Priority Claims (1)
EP 22169432 · Apr 22, 2022 · regional
Continuity (1)
Related Publication 20230343121A1 · Oct 26, 2023
References Cited (31)
US 10078725B2 · Kalafut et al. · 2018 [cited by applicant]
US 11183294B2 · Sargent et al. · 2021 [cited by applicant]
US 20040252871A1 · Tecotzky · 2004 [cited by examiner]
US 20080109250A1 · Walker et al. · 2008 [cited by applicant]
US 20130060579A1 · Yu et al. · 2013 [cited by applicant]
US 20150254401A1 · Sankhe · 2015 [cited by examiner]
US 20210174503A1 · Trautwein · 2021 [cited by examiner]
CN 103946885A · 2014 [cited by applicant]
CN 113972000A · 2022 [cited by applicant]
CN 114365181A · 2022 [cited by applicant]
Eapen et al., “DICODerma: A practical approach for metadata management of images in dermatology,” Feb. 17, 2021, arXiv: 2102.08673v1, https://doi.org/10.48550/arXiv.2102.08673 (Year: 2021). [cited by examiner]
Tsui et al., “Automatic Selective Removal of Embedded Patient Information From Image Content of DICOM Files,” American Journal of Roentgenology 2012 198:4, 769-772. (Year: 2012). [cited by examiner]
Graham et al., “DICOM demystified: A review of digital file formats and their use in radiological practice,” Clinical Radiology (2005), 60 1133-1140, Elsevier Ltd., doi:10.1016/j.crad.2005.07.003. (Year: 2005). [cited by examiner]
DICOM Tag Library. Retrieved Aug. 8, 2019. https://www.dicomlibrary.com/dicom/dicom-tags/ (Year: 2019). [cited by examiner]
Crane et al., “SIVIC: Open-Source, Standards-Based Software for DICOM MR Spectroscopy Workflows,” International Journal of Biomedical Imaging, vol. 2013, 2013, Article ID 169526, 12 pages, http://dx.doi.org/10.1155/2013… [cited by examiner]
Bridge et al., “Highdicom: A Python library for standardized encoding of image annotations and machine learning model outputs in pathology and radiology,” Sep. 9, 2021, arXiv:2106.07806v2 [eess.IV], http://doi.org/10.48… [cited by examiner]
Secondary Capture Image IOD. Retrieved Aug. 9, 2019. http://dicom.nema.org/medical/dicom/current/output/chtml/part03/sect_A.8.html (Year: 2019). [cited by examiner]
Secondary Capture Image IOD. Retrieved Aug. 9, 2019. http://dicom.nema.org/medical/dicom/current/output/chtml/part03/sect_A.8.html. [cited by applicant]
P. M. Manwatkar and Kavita R. Singh, “A Technical Review on Text recognition from Images,” 9th International Conference on Intelligent Systems and Control (ISCO), 2015, pp. 721-725. [cited by applicant]
DoseUtility Tool. Retrieved on Aug. 9, 2019. https://www.dclunie.com/pixelmed/software/webstart/DoseUtilityUsage.html. [cited by applicant]
Tsui, Gary and Chan, Tao, “Automatic Selective Removal of Embedded Patient Information from Image Content of DICOM Files”. American Journal of Roentgenology 2012 198:4, 769-772. [cited by applicant]
Optical Character Recognition (OCR) Nicomsoft.com. Retrieved Aug. 8, 2019. [cited by applicant]
Tesseract Wiki. Retrieved Aug. 8, 2019. https://github.com/tesseract-ocr/tesseract/blob/master/doc/tesseract.1.asc https://tesseract.patagames.com/help/html/T_Patagames_Ocr_Enums_PageSegMode.htm. [cited by applicant]
DICOM Tag Library. Retrieved Aug. 8, 2019. https://www.dicomlibrary.com/dicom/dicom-tags/. [cited by applicant]
Gupta, Maya R.; Jacobson, Nathaniel P.; Garcia, Eric K. (2007). “OCR binarisation and image pre-processing for searching historical documents” (PDF). Pattern Recognition. 40 (2): 389. doi:10.1016/j.patcog.2006.04.043. [cited by applicant]
Methods to Improve OCR Quality. Retrieved Aug. 8, 2019. https://github.com/tesseract-ocr/tesseract/wiki/ImproveQuality. [cited by applicant]
Holley, Rose (Apr. 2009). “How Good Can It Get? Analysing and Improving OCR Accuracy in Large Scale Historic Newspaper Digitisation Programs”. D-Lib Magazine. [cited by applicant]
Pydicom Library. Retrieved Aug. 8, 2019. https://pydicom.github.io/. [cited by applicant]
TIFF Image Format. Retrieved Aug. 8, 2019. https://www.loc.gov/preservation/digital/formats/fdd/fdd000072.shtml. [cited by applicant]
Extended European Search Report (EESR) mailed Sep. 22, 2022 in corresponding European Patent Application No. 22169432.6. [cited by applicant]
Kay, Anthony, “Tesseract: an Open-Source Optical Character Recognition Engine”. Linux Journal. (Jul. 2007). Retrieved Sep. 28, 2011. [cited by applicant]