Related Experiment Video
Updated: May 6, 2026

A Strategy for Sensitive, Large Scale Quantitative Metabolomics
Published on: May 27, 2014
Local neighbor Normalization: Reconciling accurate normalization and heterogeneity recovery in large-scale
Keyi Lu1, Yaru Liu1, Kian-Kai Cheng2
1Department of Electronic Science, National Institute for Data Science in Health and Medicine, Xiamen University, Xiamen, 361005, China.
Background:
Metabolomics studies often grapple with the dilution effect, where sample concentrations vary due to inconsistent handling or biological diversity, particularly in samples like urine, saliva, or cell extracts. This variation can mask true metabolic differences, complicating data interpretation. Traditional normalization methods, such as Constant Sum Normalization (CSN), Probabilistic Quotient Normalization (PQN), and Maximal Density Fold Change (MDFC), assume that all samples share a certain invariant statistic and overlook data heterogeneity, potentially erasing the dataset's heterogeneity essential for distinguishing biological subgroups.
Results:
To address this, we introduce Local Neighbor Normalization (LNN), a novel approach that corrects for dilution effects while preserving the intrinsic variability of metabolomics data. LNN identifies a neighbor set for each sample based on similarity metrics and normalizes each sample against a tailored reference spectrum derived from these neighbors. Through comprehensive evaluations on both simulated and real metabolomics datasets from NMR, GC-MS, and LC-MS platforms, LNN demonstrated superior performance over CSN, PQN, and MDFC. Specifically, it achieved better elimination of dilution effects, recovery of inter-sample heterogeneity and inter-metabolite correlations, as evidenced by metrics such as the D-statistic and correlation recovery rates. Notably, LNN excels in datasets with over 50 % differential metabolites, safeguarding local data structures critically for downstream analyses like biomarker discovery.
Significance And Novelty:
LNN constructs sample-specific reference spectra based on a local neighbor set. This approach ensures that normalization accounts for dilution effects without compromising local structure of the data, which is crucial for biological interpretation. Additionally, LNN demonstrates superior performance in recovering inter-sample heterogeneity and metabolite correlations, especially in datasets with high proportions of differential metabolites. This method's versatility, robustness against noise, and applicability across various metabolomics platforms make it a significant advancement in the field.
More Related Videos
08:27Large-Scale Multi-Omics Genome-Wide Association Studies Mo-GWAS: Guidelines for Sample Preparation and Normalization
Published on: July 27, 2021
11:02Identification and Quantification of Deranged Metabolites in Critically Ill Patients Using NMR-Based Metabolomics
Published on: November 29, 2024
Related Concept Videos
Conservative Site-specific Recombination and Phase Variation
The recognition sites for Cre recombinase called LoxP...
Sampling Plans
Random sampling is a method where each member of the population has an equal chance of being selected for the sample. It involves selecting individuals randomly, often using random number generators or lottery-type methods. For example, when analyzing the properties of a...
Contaminants and Errors
Another key consideration is determining the appropriate number of samples required to...
Deleterious Substances in Aggregate
Another type of impurity is clay and fine material that...
Distance Corrections
Methods of Medium Optimization