Novel analytical methods to interpret large sequencing data from small sample sizes

Florence Lichou1, Sébastien Orazio2, Stéphanie Dulucq1

  • 1Laboratory of Mammary and Leukaemic Oncogenesis, Inserm U1218 ACTION, Bergonié Cancer Institute, University of Bordeaux, 146 rue Léo Saignat, bâtiment TP 4ème étage, case 50, 33076, Bordeaux, France.

Human Genomics
|September 1, 2019
PubMed
Abstract

Insights

New statistical methods analyze large genetic sequencing data from limited chronic myeloid leukemia patient samples. These methods identify genetic variants impacting imatinib treatment response, aiding personalized medicine development.

Area of Science:

  • Pharmacogenomics
  • Computational Biology
  • Oncology

Background:

  • Targeted therapies like imatinib have improved chronic myeloid leukemia (CML) treatment, but patient resistance due to genetic variations remains a challenge.
  • Pharmacogenetic studies are crucial for understanding treatment heterogeneity but are often limited by small sample sizes.
  • Classical statistical analyses are inadequate for large sequencing datasets derived from limited patient cohorts.

Purpose of the Study:

  • To introduce novel statistical methods for analyzing large-scale next-generation sequencing data from small patient sample sizes.
  • To identify genetic variants influencing drug response in CML patients treated with imatinib.
  • To overcome limitations of traditional statistical approaches in pharmacogenetic research.

Main Methods:

  • Next-generation sequencing was performed on 48 pharmacokinetic genes from 24 CML patients (sensitive and resistant to imatinib).
  • A graphical approach was employed to reduce 708 identified polymorphisms to a list of 115 candidate variants.
  • Analysis focused on gene-specific variant allele distribution to highlight potential drug-response-associated genes.

Main Results:

  • The novel methods successfully reduced a large set of genetic polymorphisms to a manageable list of candidates.
  • Key candidate genes, including UGT1A9, PTPN22, and ERCC5, were identified.
  • These highlighted genes have prior associations with drug transport, metabolism, and imatinib sensitivity.

Conclusions:

  • The developed statistical tests offer effective alternatives to inferential statistics for next-generation sequencing data from small sample sizes.
  • These approaches facilitate target reduction and identification of promising candidates for further pharmacogenetic studies.
  • The findings support the potential for personalized treatment strategies in CML based on genetic profiling.

Related Concept Videos

Sample Preparation for Analytical Characterization09:51

Sample Preparation for Analytical Characterization

Source: Laboratory of Dr. B. Jill Venton - University of Virginia
Sample preparation is the way in which a sample is treated to prepare for analysis. Careful sample preparation is critical in analytical chemistry to accurately generate either a standard or unknown sample for a chemical measurement. Errors in analytical chemistry methods are categorized as random or systematic. Random errors are errors due to change and are often due to noise in instrument. Systematic errors are due to...
88.2K
Bioequivalence Data: Statistical Interpretation01:16

Bioequivalence Data: Statistical Interpretation

Body:The statistical interpretation of bioequivalence data is a significant aspect of pharmaceutical research. Bioequivalence refers to the absence of any significant difference in the rate and extent to which the active ingredient in pharmaceutical products becomes available at the site of drug action when administered at the same molar dose under similar conditions. This helps determine if different drug products have similar absorption rates, ensuring their interchangeability.Statistical...
196
Optimization for Sequencing and Analysis of Degraded FFPE-RNA Samples07:30

Optimization for Sequencing and Analysis of Degraded FFPE-RNA Samples

This method describes the steps to improve the quality and quantity of sequence data that can be obtained from formalin-fixed paraffin-embedded (FFPE) RNA samples. We describe the methodology to more accurately assess the quality of FFPE-RNA samples, prepare sequencing libraries, and analyze the data from FFPE-RNA...
12.7K
A Method for Targeted 16S Sequencing of Human Milk Samples09:09

A Method for Targeted 16S Sequencing of Human Milk Samples

A semi-automated workflow is presented for targeted sequencing of 16S rRNA from human milk and other low-biomass sample...
10.3K
Data Analysis & Interpretation01:32

Data Analysis & Interpretation

Data analysis is crucial in marketing research to understand consumer behavior and guide business strategies. Two primary approaches—qualitative and quantitative data analysis—offer distinct advantages that help businesses refine their marketing efforts. Combining qualitative insights with quantitative evidence allows businesses to comprehensively understand the market and consumer behavior, leading to more effective and targeted marketing strategies.
Qualitative Data Analysis:...
990
Facilitating the Analysis of Immunological Data with Visual Analytic Techniques10:58

Facilitating the Analysis of Immunological Data with Visual Analytic Techniques

Visual analytics (VA) is a new approach of analyzing data interactively. In this video, we discuss the data overload problem brought on by high-throughput biological experiments, and propose VA as a solution to such problem. The video demonstrates analysis within and between immunological datasets using a VA tool called...
10.5K