复杂的动态系统与内部噪声的高维分析中的相关,隐藏和丧的信息
Chiara Lionello1, Matteo Becchi1, Simone Martino1
1Department of Applied Science and Technology, Politecnico di Torino, Torino 10129, Italy.
Journal of chemical theory and computation
|July 2, 2025
概括
对分子动态的高维分析通常是不必要的. 局部分子环境数据的单一维度 (原子位置的平滑重叠,SOAP) 有效地区分水和冰相,挑战了关于维度的假设.
科学领域:
- 计算化学是一种计算化学.
- 分子动力学分子动力学
- 统计力学就是统计力学.
背景情况:
- 从轨迹数据中理解复杂的分子系统是具有挑战性的.
- 高维分析通常被认为是保留信息所必需的.
- 在此类分析中,高维度的真正好处仍然不清楚.
研究的目的:
- 研究分子系统中高维分析的必要性和益处.
- 挑战传统的假设,即更多的维度总是更好.
- 探索复杂分子数据的维度减小策略.
主要方法:
- 对共存的液态水和冰的原子分子动力学轨迹的分析.
- 使用局部分子环境的高维描述符,特别是原子位置的光滑重叠 (SOAP).
- 分析了2.56 × 10^6 576维SOAP光谱的大数据集.
主要成果:
- 一个单一的SOAP数据维度,占<0.001%的差异,成功地分类了水和冰相和接口层.
- 超出这个单一有效维度的增加维度被证明是无效和有害的.
- "丧的信息"现象和噪音积累被观察到具有更高的维度.
结论:
- 高维分析对于阐明复杂的分子系统并不是本质上优越的.
- 低维的方法可能更有效,特别是在有噪音的情况下.
- 仔细选择相关尺寸对于准确的分子数据分析至关重要.
相关概念视频
¹H NMR: Interpreting Distorted and Overlapping Signals
1.1K
Spin systems where the difference in chemical shifts of the coupled nuclei is greater than ten times J are called first-order spin systems. These nuclei are weakly coupled, and their chemical shifts and coupling constant can generally be estimated from the well-separated signals in the spectrum.
As Δν decreases and the signals move closer, the doublets appear increasingly distorted. The intensities of the inner lines increase at the cost of those of the outer lines as the signals are...
As Δν decreases and the signals move closer, the doublets appear increasingly distorted. The intensities of the inner lines increase at the cost of those of the outer lines as the signals are...
1.1K
Propagation of Uncertainty from Systematic Error
894
The atomic mass of an element varies due to the relative ratio of its isotopes. A sample's relative proportion of oxygen isotopes influences its average atomic mass. For instance, if we were to measure the atomic mass of oxygen from a sample, the mass would be a weighted average of the isotopic masses of oxygen in that sample. Since a single sample is not likely to perfectly reflect the true atomic mass of oxygen for all the molecules of oxygen on Earth, the mass we obtain from this...
894
Mechanistic Models: Compartment Models in Algorithms for Numerical Problem Solving
103
Mechanistic models play a crucial role in algorithms for numerical problem-solving, particularly in nonlinear mixed effects modeling (NMEM). These models aim to minimize specific objective functions by evaluating various parameter estimates, leading to the development of systematic algorithms. In some cases, linearization techniques approximate the model using linear equations.
In individual population analyses, different algorithms are employed, such as Cauchy's method, which uses a...
In individual population analyses, different algorithms are employed, such as Cauchy's method, which uses a...
103
Random Error
1.7K
Random or indeterminate errors originate from various uncontrollable variables, such as variations in environmental conditions, instrument imperfections, or the inherent variability of the phenomena being measured. Usually, these errors cannot be predicted, estimated, or characterized because their direction and magnitude often vary in magnitude and direction even during consecutive measurements. As a result, they are difficult to eliminate. However, the aggregate effect of these errors can be...
1.7K
Noncompartmental Analysis: Statistical Moment Theory
186
Noncompartmental analyses leverage statistical moment theory to examine time-related changes in macroscopic events, encapsulating the collective outcomes stemming from the constituent elements in play. Statistical moment theory is a mathematical approach used to describe the time course of drug concentration in the body without assuming a specific compartmental model. SMT provides insights into drug absorption, distribution, metabolism, and elimination by treating drug concentration versus time...
186
Uncertainty: Overview
995
In analytical chemistry, we often perform repetitive measurements to detect and minimize inaccuracies caused by both determinate and indeterminate errors. Despite the cares we take, the presence of random errors means that repeated measurements almost never have exactly the same magnitude. The collective difference between these measurements - observed values - and the estimated or expected value is called uncertainty. Uncertainty is conventionally written after the estimated or expected value.
995


