Related Experiment Video
Updated: Mar 27, 2026

Performing Data Mining And Integrative Analysis Of Biomarker in Breast Cancer Using Multiple Publicly Accessible Databases
Published on: May 17, 2019
Secondary analysis of large databases for hepatology research
Philip N Okafor1, Maria Chiejina2, Nicolo de Pretis3
1Division of Gastroenterology and Hepatology, Mayo Clinic, 200 First Street SW, Rochester, MN 55905, United States.
Abstract:
Secondary analysis of large datasets involves the utilization of existing data that has typically been collected for other purposes to advance scientific knowledge. This is an established methodology applied in health services research with the unique advantage of efficiently identifying relationships between predictor and outcome variables but which has been underutilized for hepatology research. Our review of 1431 abstracts published in the 2013 European Association for the Study of Liver (EASL) abstract book showed that less than 0.5% of published abstracts utilized secondary analysis of large database methodologies. This review paper describes existing large datasets that can be exploited for secondary analyses in liver disease research. It also suggests potential questions that could be addressed using these data warehouses and highlights the strengths and limitations of each dataset as described by authors that have previously used them. The overall goal is to bring these datasets to the attention of readers and ultimately encourage the consideration of secondary analysis of large database methodologies for the advancement of hepatology.

