Related Experiment Video
Updated: Jul 2, 2026

12:18
A Machine Learning Approach to Design an Efficient Selective Screening of Mild Cognitive Impairment
Published on: January 11, 2020
Regression analysis of misclassified current status data with potentially unknown test accuracy
1Department of Statistics, University of South Carolina, Columbia, SC, USA.
Statistical Methods in Medical Research
|July 1, 2026
Summary
This study introduces a new regression analysis method for current status data with misclassified failure status. The approach uses monotone splines and an expectation-maximization algorithm, improving accuracy in epidemiological and medical studies.
Area of Science:
- Epidemiology
- Biostatistics
- Medical Statistics
Background:
- Current status data are common in cross-sectional studies, featuring censored failure times.
- Failure status determination can be error-prone, leading to misclassified current status data.
- Accurate analysis of misclassified data is crucial for reliable study outcomes.
Purpose of the Study:
- To develop a novel regression analysis approach for misclassified current status data.
- To address challenges posed by censored and misclassified failure times in observational studies.
- To propose an estimation method robust to unknown test accuracy.
Main Methods:
- Utilized a proportional odds model for regression analysis.
- Employed monotone splines to approximate the baseline odds function.
- Developed an expectation-maximization algorithm with data augmentation using latent variables.
- Extended the method to handle unknown test accuracy.
Main Results:
- The proposed method demonstrated excellent estimation performance in simulation studies.
- The approach effectively handles misclassified failure status in current status data.
- The method was successfully applied to uterine fibroid data.
Conclusions:
- The novel estimation approach provides a robust tool for analyzing misclassified current status data.
- The method enhances the reliability of regression analysis in epidemiological and medical research.
- The technique is valuable for studies where diagnostic tests may yield errors.
Related Concept Videos
Accuracy and Errors in Hypothesis Testing
Hypothesis testing is a fundamental statistical tool that begins with the assumption that the null hypothesis H0 is true. During this process, two types of errors can occur: Type I and Type II. A Type I error refers to the incorrect rejection of a true null hypothesis, while a Type II error involves the failure to reject a false null hypothesis.
In hypothesis testing, the probability of making a Type I error, denoted as α, is commonly set at 0.05. This significance level indicates a 5% chance...
In hypothesis testing, the probability of making a Type I error, denoted as α, is commonly set at 0.05. This significance level indicates a 5% chance...
Regression Analysis
Regression analysis is a statistical tool that describes a mathematical relationship between a dependent variable and one or more independent variables.
In regression analysis, a regression equation is determined based on the line of best fit– a line that best fits the data points plotted in a graph. This line is also called the regression line. The algebraic equation for the regression line is called the regression equation. It is represented as:
In regression analysis, a regression equation is determined based on the line of best fit– a line that best fits the data points plotted in a graph. This line is also called the regression line. The algebraic equation for the regression line is called the regression equation. It is represented as:
Detection of Gross Error: The Q Test
When one or more data points appear far from the rest of the data, there is a need to determine whether they are outliers and whether they should be eliminated from the data set to ensure an accurate representation of the measured value. In many cases, outliers arise from gross errors (or human errors) and do not accurately reflect the underlying phenomenon. In some cases, however, these apparent outliers reflect true phenomenological differences. In these cases, we can use statistical methods...
Regression Toward the Mean
Regression toward the mean (“RTM”) is a phenomenon in which extremely high or low values—for example, and individual’s blood pressure at a particular moment—appear closer to a group’s average upon remeasuring. Although this statistical peculiarity is the result of random error and chance, it has been problematic across various medical, scientific, financial and psychological applications. In particular, RTM, if not taken into account, can interfere when researchers try to extrapolate results...
Receiver Operating Characteristic Plot
A ROC (Receiver Operating Characteristic) plot is a graphical tool used to assess the performance of a binary classification model by illustrating the trade-off between sensitivity (true positive rate) and specificity (false positive rate). By plotting sensitivity against 1 - specificity across various threshold settings, the ROC curve shows how well the model distinguishes between classes, with a curve closer to the top-left corner indicating a more accurate model. The area under the ROC curve...
Errors In Hypothesis Tests
When performing a hypothesis test, there are four possible outcomes depending on the actual truth (or falseness) of the null hypothesis and the decision to reject or not.
