Related Experiment Video
Updated: Oct 13, 2025

Evidence-based Knowledge Synthesis and Hypothesis Validation: Navigating Biomedical Knowledge Bases via Explainable AI and Agentic Systems
Published on: June 13, 2025
NIR robustness model of variable selection investigation of critical quality attributes coupled with different
Na Zhao1, Lijuan Ma1, Kaiyi Wang1
1Beijing University of Chinese Medicine, Beijing 100102, China; Pharmaceutical Engineering and New Drug Development of TCM of Ministry of Education, Beijing 100102, China.
Abstract:
variable selection is critical to select characteristic variables of critical quality attributes to improve model performance and interpret the identified variables in multivariate calibration. However, classical variable selection methods were developed and optimized by the prediction error. It is rare for the robustness evaluation of variable selection methods. In this study, the robustness of four different variable selection methods was investigated by adding different types of simulate noises to validation set and calibration and validation sets, respectively. The reproducibility as well as root mean squared error of prediction (RMSEP) were used together as common measure in assessing the robustness of different variable selection methods. The robustness of four variable selection methods method was investigated using two near infrared (NIR) datasets including open-source dataset of corn and Chinese herbal medicine (CHM) dataset. The result illustrated that variable importance in projection (VIP) was substantially more robust to additive noise, with smaller RMSEP value and high reproducibility. This provides a novel strategy for the reliability evaluation of variable selection methods in NIR model of critical quality attributes.
Related Concept Videos
Random and Systematic Errors
Variability: Analysis
The range is a simple measure of variability, indicating the difference between the highest and...
Expected Frequencies in Goodness-of-Fit Tests
Reliability and Validity
Propagation of Uncertainty from Systematic Error
Accuracy and Errors in Hypothesis Testing
In hypothesis testing, the probability of making a Type I error, denoted as α, is commonly set at 0.05. This significance level indicates a 5%...

