Related Experiment Video
Updated: Jan 25, 2026

Deep Learning-Based Segmentation of Cryo-Electron Tomograms
Published on: November 11, 2022
Eye Tracking for Deep Learning Segmentation Using Convolutional Neural Networks
J N Stember1, H Celik2, E Krupinski3
1Department of Radiology, Columbia University Medical Center - NYPH, New York, NY, 10032, USA. joestember@gmail.com.
Eye tracking (ET) technology can generate accurate segmentation masks for medical images, comparable to manual annotation (HA). CNNs trained with ET masks perform similarly to those trained with HA masks, offering a promising solution for data annotation challenges.
Area of Science:
- Medical Imaging
- Deep Learning
- Computer Vision
Background:
- Convolutional Neural Networks (CNNs) show high accuracy in medical image semantic segmentation.
- A major limitation is the need for large-scale, precisely annotated imaging datasets for training.
- Eye tracking (ET) technology is explored as a novel method to address this data annotation bottleneck.
Purpose of the Study:
- To evaluate if segmentation masks generated by eye tracking (ET) are similar to hand annotations (HA).
- To determine if a CNN trained on ET masks is equivalent to one trained on HA masks.
- To assess the feasibility of ET for efficient medical image annotation.
Main Methods:
- Two steps were used: 1) Analysis of 19 public radiologic images for ET and HA mask generation. 2) Generation of ET and HA masks for 356 meningioma images.
- A U-net based CNN was trained on 306 image-mask pairs (ET and HA), with 50 images reserved for testing.
- Dice Similarity Coefficient (DSC) and Area Under the Curve (AUC) were used to compare mask and model performance.
Main Results:
- ET and HA masks showed high similarity, with average DSC of 0.86 for non-neurological images and 0.85 for meningioma images.
- CNNs trained with ET and HA masks achieved comparable performance on the test set (AUC 0.88 vs. 0.87).
- Trimmed DSCs between ET and HA predictions were statistically equivalent (p=0.015), indicating similar segmentation accuracy.
Conclusions:
- Eye tracking (ET) technology can generate segmentation masks suitable for deep learning semantic segmentation in medical imaging.
- ET offers a viable alternative to manual annotation, potentially accelerating the training of CNNs.
- Future research will focus on integrating ET into clinical workflows for faster and more natural mask generation.
Related Concept Videos
Convolution Properties II
The width property indicates that if the durations of input signals are T1 and T2, then the width of the output response equals the sum of both durations, irrespective of the shapes of the two functions. For instance, convolving two rectangular pulses with durations of 2 seconds and 1 second results in a function with a width of 3 seconds.
The area property asserts that the area under the...
Protein Networks
These interactions can be represented through maps depicting protein-protein interaction networks, represented as nodes and edges. Nodes are circles that are representative of a protein,...
Convolution Properties I
The commutative property reveals that the input and the impulse response of an LTI (Linear Time-Invariant) system can be interchanged without affecting the output:
Neural Regulation
Network Covalent Solids
To break or to melt a covalent network solid, covalent bonds must be broken. Because covalent bonds are relatively strong, covalent network solids are typically...
Convolution: Math, Graphics, and Discrete Signals
To simplify the convolution integral, it is assumed that both the input signal and impulse response are zero for negative time values. The graphical convolution process...

