Related Experiment Video
Updated: May 3, 2026

Quantitative Optical Microscopy: Measurement of Cellular Biophysical Features with a Standard Optical Microscope
Published on: April 7, 2014
A Method for Efficient De-identification of DICOM Metadata and Burned-in Pixel Text
Jacob A Macdonald1, Katelyn R Morgan2, Brandon Konkel2
1Department of Radiology, Duke University, Durham, NC, USA. jacob.macdonald@duke.edu.
Abstract:
De-identification of DICOM images is an essential component of medical image research. While many established methods exist for the safe removal of protected health information (PHI) in DICOM metadata, approaches for the removal of PHI "burned-in" to image pixel data are typically manual, and automated high-throughput approaches are not well validated. Emerging optical character recognition (OCR) models can potentially detect and remove PHI-bearing text from medical images but are very time-consuming to run on the high volume of images found in typical research studies. We present a data processing method that performs metadata de-identification for all images combined with a targeted approach to only apply OCR to images with a high likelihood of burned-in text. The method was validated on a dataset of 415,182 images across ten modalities representative of the de-identification requests submitted at our institution over a 20-year span. Of the 12,578 images in this dataset with burned-in text of any kind, only 10 passed undetected with the method. OCR was only required for 6050 images (1.5% of the dataset).
More Related Videos
09:21Human Brown Adipose Tissue Depots Automatically Segmented by Positron Emission Tomography/Computed Tomography and Registered Magnetic Resonance Images
Published on: February 18, 2015
10:39A Label-Free Segmentation Approach for Intravital Imaging of Mammary Tumor Microenvironment
Published on: May 24, 2022