Association Between Risk Factors and Major Cancers: Explainable Machine Learning Approach.

Xiayuan Huang1, Shushun Ren2, Xinyue Mao3

  • 1Department of Biostatistics, Yale University, New Haven, CT, United States.

JMIR Cancer
|May 2, 2025
PubMed
Summary

Explainable machine learning models identified key risk factors for major cancers, revealing common nontraditional factors like hyperlipidemia and diabetes. This aids in personalized cancer screening and prevention strategies.

Related Concept Videos

Cancer Prevention02:59

Cancer Prevention

Several factors can increase the risk of cancer in an individual. About 50% of cancer cases can be prevented by adopting a healthy lifestyle, regular exercise, eating healthy, and following a modest cancer prevention diet. Epidemiological studies have consistently shown that populations with vegetable and fruit-rich diets have reduced the incidence of cancer. On the other hand, populations who have a diet rich in animal fat, red meat, junk food, or high calories are predisposed to cancer.
Some...
6.0K
Cancer Survival Analysis01:21

Cancer Survival Analysis

Cancer survival analysis focuses on quantifying and interpreting the time from a key starting point, such as diagnosis or the initiation of treatment, to a specific endpoint, such as remission or death. This analysis provides critical insights into treatment effectiveness and factors that influence patient outcomes, helping to shape clinical decisions and guide prognostic evaluations. A cornerstone of oncology research, survival analysis tackles the challenges of skewed, non-normally...
301
Statistical Methods for Analyzing Epidemiological Data01:25

Statistical Methods for Analyzing Epidemiological Data

Epidemiological data primarily involves information on specific populations' occurrence, distribution, and determinants of health and diseases. This data is crucial for understanding disease patterns and impacts, aiding public health decision-making and disease prevention strategies. The analysis of epidemiological data employs various statistical methods to interpret health-related data effectively. Here are some commonly used methods:
239
Mouse Models of Cancer Study02:43

Mouse Models of Cancer Study

Mice have long served as models for studying human biology and pathology because of their phylogenetic and physiological similarity with humans. They are also easy to maintain and breed in the laboratory, and hence, many inbred strains are now available for research. Studies on mice have contributed immeasurably to our understanding of cancer biology.
The development of transgenic, knockout, and knock-in mice has led to an exponential increase in their use as model organisms in research,...
5.4K
Cause and Effect01:53

Cause and Effect

While variables are sometimes correlated because one does cause the other, it could also be that some other factor, a confounding variable, is actually causing the systematic movement in our variables of interest. For instance, as sales in ice cream increase, so does the overall rate of crime. Is it possible that indulging in your favorite flavor of ice cream could send you on a crime spree? Or, after committing crime do you think you might decide to treat yourself to a cone?
10.8K
Relative Risk01:12

Relative Risk

Relative risk (RR) is a statistical measure commonly used in epidemiology to compare the likelihood of a particular event occurring between two groups. This metric is important for evaluating the relationship between exposure to a specific risk factor and the probability of a particular outcome. It plays a crucial role in medical research, public health studies, and risk assessment. Relative risk quantifies how much more (or less) likely an event is to occur in an exposed group compared to an...
93