Related Experiment Video
Updated: Sep 8, 2025

An R-Based Landscape Validation of a Competing Risk Model
Published on: September 16, 2022
Worst-Case Robustness Evaluation Methods for IMPT: A Critical Comparison
Chunbo Liu1,2, Chris J Beltran2, Jiajian Shen3
1Department of Radiation Oncology, The First Affiliated Hospital of Zhengzhou University, Zhengzhou, China.
Purpose:
Robustness evaluation is routinely used in clinics to ensure the intended dose delivery for intensity-modulated proton therapy (IMPT). Various methods have been proposed, but there is no consensus on which method should be adopted in clinical practice. This study examined various methods within the widely used worst-case approach to provide insights into IMPT plan evaluation.
Materials And Methods:
We evaluated the robustness of 20 clinical IMPT plans (10 prostate and 10 head and neck). Five robustness evaluation methods were assessed: error-bar dose distribution (ebDD), root-mean-square error dose (RMSED) distribution, voxel-wise worst-case, physical scenario worst-case, and dose-volume histogram (DVH) band. Correlations between these methods were analyzed. Each method was reviewed for their quantitative and qualitative capabilities to identify potential underdosing or overdosing.
Results:
Strong correlations were found between ebDD and RMSED, and between voxel-wise worst-case and physical scenario worst-case. The DVH band method provides a straightforward way to assess whether the worst DVH meets plan criteria and to illustrate dose variations but lacks spatial detail to pinpoint areas of potential underdosing or overdosing. The voxel-wise worst-case captures the worst dose distribution across all evaluation metrics, allowing spatial identification of areas of concern within a single distribution. The physical scenario worst-case also pinpoints specific areas of concern but requires individual assessment for each region of interest and evaluation metric, which can be cumbersome. A 3D visualization with ebDD and RMSED highlights regions of dose variation but does not necessarily indicate clinically meaningful impact.
Conclusion:
Different robustness evaluation methods offer different types of information. Our study provides valuable insights to help identify an effective and practical approach for clinical practice. Based on our findings, we propose a potential evaluation strategy: use the DVH band derived from physical uncertainty scenarios to assess whether the worst boundary values meet plan evaluation criteria, and, when concerns arise, apply the voxel-wise worst-case dose distribution to localize areas of potential risk.
More Related Videos
05:47Evidence-based Knowledge Synthesis and Hypothesis Validation: Navigating Biomedical Knowledge Bases via Explainable AI and Agentic Systems
Published on: June 13, 2025
07:28Psychophysically-anchored, Robust Thresholding in Studying Pain-related Lateralization of Oscillatory Prestimulus Activity
Published on: January 21, 2017
Related Concept Videos
Comparing the Survival Analysis of Two or More Groups
Multiple Comparison Tests
It would be easy to compare two samples using a significance alpha level of 0.05. In other words, there is only one sample pair to be compared. However, it would be difficult to identify a significantly different sample if the number...
Critical Region, Critical Values and Significance Level
In hypothesis testing, a sample statistic is converted to a test statistic using z, t, or chi-square distribution. A critical region is an area under the curve in probability distributions demarcated by the critical value. When the test statistic falls in this region, it suggests that the null hypothesis must be rejected. As this region contains all those values of the...
Bonferroni Test
The means of different samples are first paired in all possible combinations.
The null hypothesis of the...
Quantifying and Rejecting Outliers: The Grubbs Test
Detection of Gross Error: The Q Test