使用重新抽样策略评估变量重要性稳定性,以提高代谢学中的模型解释性和可靠性
1School of Pharmaceutical Sciences, University of Geneva, Geneva, Switzerland; Institute of Pharmaceutical Sciences of Western Switzerland, University of Geneva, Geneva, Switzerland.
Analytica chimica acta
|March 9, 2026
概括
这项研究引入了一种新的方法来评估代谢学数据中的变量重要性,提高识别关键生物信号的可靠性. 该方法提高了复杂数据集中发现的稳定性和可解释性.
科学领域:
- 代谢学 代谢学 代谢学
- 生物信息学是一种生物信息学.
- 统计建模 统计建模
背景情况:
- 多变量分析对于代谢学知识的发现至关重要.
- 矩阵分解方法有助于发现复杂数据集中的趋势和变量关系.
- 代谢学数据的高维度挑战了变量重要性评估的可靠性.
研究的目的:
- 开发一种可靠的方法来评估从部分最小平方 (PLS) 回归模型中预测变量重要性 (VIP) 的稳定性.
- 提供一种可靠的工具,用于识别信息变量,并删除代谢学数据中的非信息信号.
主要方法:
- 一种新的方法,结合了引导重新抽样和排列来评估VIP稳定性.
- 使用稳定性指数和诊断图进行可靠的评估.
- 构建从真实和置变量重要性值的经验分布.
主要成果:
- 拟议的方法有效地评估了代谢学中有意义的变量的可靠性.
- 在多种不同的实验配置中,证明了在去除无信息信号方面的潜力.
- 通过提供更稳定的信息变量的子集,提高可解释性,优于既定方法.
结论:
- 该方法在计算上高效,通用,不需要数据分布假设.
- 促进更一致和可重复的代谢学研究.
- 旨在通过改进数据分析来推进对代谢模式的理解.
相关概念视频
Mechanistic Models: Compartment Models in Individual and Population Analysis
314
Mechanistic models are utilized in individual analysis using single-source data, but imperfections arise due to data collection errors, preventing perfect prediction of observed data. The mathematical equation involves known values (Xi), observed concentrations (Ci), measurement errors (εi), model parameters (ϕj), and the related function (ƒi) for i number of values. Different least-squares metrics quantify differences between predicted and observed values. The ordinary least...
314
Variability: Analysis
598
Measures of variability are statistical metrics that reveal the dispersion pattern within a dataset. They are pivotal in biostatistics, providing insights into the heterogeneity within health and biological data. Variability signifies the degree to which data points diverge from one another, helping researchers understand the potential range of values and associated uncertainty within the data.
The range is a simple measure of variability, indicating the difference between the highest and...
The range is a simple measure of variability, indicating the difference between the highest and...
598
Biostatistics: Overview
993
Biostatistics plays a crucial role in understanding and analyzing data in healthcare and biology. Biostatisticians conduct experiments, gather evidence, and draw meaningful conclusions using statistical methods and techniques. Different variables form the foundation of biostatistical analysis, allowing researchers to understand and interpret data effectively. These variables are classified into different types, each serving a specific purpose in statistical analysis.
Discrete variables are...
Discrete variables are...
993
Statistical Analysis: Overview
16.8K
When we take repeated measurements on the same or replicated samples, we will observe inconsistencies in the magnitude. These inconsistencies are called errors. To categorize and characterize these results and their errors, the researcher can use statistical analysis to determine the quality of the measurements and/or suitability of the methods.
One of the most commonly used statistical quantifiers is the mean, which is the ratio between the sum of the numerical values of all results and the...
One of the most commonly used statistical quantifiers is the mean, which is the ratio between the sum of the numerical values of all results and the...
16.8K
Data Validation
3.3K
Method validation is a crucial process in analytical chemistry designed to confirm that a given method consistently produces reliable and high-quality results. This process is essential when a method is applied to different sample matrices or when procedural modifications are made, ensuring that the results meet acceptable standards across various applications.
Key parameters for method validation include:
Key parameters for method validation include:
3.3K


