Related Experiment Video
Updated: Jun 6, 2025

A Protocol for Using Gene Set Enrichment Analysis to Identify the Appropriate Animal Model for Translational Research
Published on: August 16, 2017
Use and Evaluation of GANs for Synthetic Data Generation in Pharmacogenetics
Dominic Aeschbacher1, Jessica Meisner1, Marko Miletic1
1Bern University of Applied Sciences, Switzerland.
Abstract:
Pharmacogenetics (PGx) explores the influence of genetic variability on drug efficacy and tolerability. Synthetic Data Generation (SDG) has emerged as a promising alternative to the labor-intensive process of collecting real-world PGx data, which is required for high-qualitative prediction models. This study investigates the performance of two Generative Adversarial Network (GAN) models, CTGAN and CTAB-GAN+, in generating synthetic PGx data. The benchmarking is based on utility metrics (Hellinger distance and Random Forest accuracy) and ϵ-identifiability. Results demonstrate that synthetic data generated by CTAB-GAN+ can surpass the original dataset in terms of utility. For instance, CTAB-GAN+ achieves higher Random Forest accuracy compared to the original data, indicating better predictive performance. These improvements suggest that synthetic data not only capture the essential patterns of the original data but also enhance model generalization and prediction capabilities, providing a more robust training ground for machine learning models. Consequently, SDG offers a promising solution to address data scarcity and imbalance in pharmacogenetic research.
Related Concept Videos
Analysis of Population Pharmacokinetic Data
Genome-wide Association Studies-GWAS
GWAS does not require the identification of the target gene involved in...
Synthetic Biology
Golden rice
Golden rice is a genetically modified...
What is Genetic Engineering?
In-vitro Mutagenesis

