Related Experiment Video
Updated: Jan 1, 2026

Selecting Multiple Biomarker Subsets with Similarly Effective Binary Classification Performances
Published on: October 11, 2018
A Comparison of High Dimensional Variable Selection Methods with Missing Covariates in a Prostate Cancer Study
Abstract:
Prostate cancer is the most common cancer in American men. Dozens of specific genes have been shown to be correlated to prostate cancer, to benign and non-benign cancer cases, from a biology perspective. In this paper, we apply a penalized logistic regression model with different penalty functions to select genes that contribute to benign and non-benign cases, based on the data from a prostate cancer study. The tuning parameter is determined by cross validation. In order to take into account some specific genes that have been classified as prostate cancer genes through biology research but with missing values, multiple imputation is adopted to create complete data sets. We analyze the prostate cancer data by comparing the selection results with completely observed data only, and the results with imputed data. We also conduct a simulation study to validate our proposed method.
More Related Videos
Related Concept Videos
Cancer Survival Analysis
Comparing the Survival Analysis of Two or More Groups
Survival Tree
Building a Survival Tree
Constructing a...
Truncation in Survival Analysis
Left truncation occurs when individuals who experienced the event of interest before a certain time are not included in the study. This is often due to a "delayed entry" into the study where only those who survive until a certain entry point are...
Assumptions of Survival Analysis
Pharmacokinetic Models: Comparison and Selection Criterion
Physiological models take a detailed approach by considering specific molecular processes. They can predict drug distribution, metabolism, and elimination changes, providing a comprehensive understanding of how drugs interact with the body.

