Related Experiment Video
Updated: Apr 11, 2026

Development of a Virtual Reality Assessment of Everyday Living Skills
Published on: April 23, 2014
Examining the Predictive Validity of NIH Peer Review Scores
Mark D Lindner1, Richard K Nakamura1
1Center for Scientific Review, National Institutes of Health, 6701 Rockledge Dr., Bethesda, Maryland, United States of America.
National Institutes of Health (NIH) peer review scores do not accurately predict research success. Using bibliometric indices alone to assess impact is invalid and exacerbates existing problems.
Area of Science:
- Biomedical Research Funding
- Grant Review Processes
- Scientific Evaluation Metrics
Background:
- Empirical demonstration of the predictive validity of National Institutes of Health (NIH) peer review remains limited.
- A common assumption suggests correlating peer review scores with bibliometric indices could validate predictive accuracy.
- This study investigates the feasibility and validity of using bibliometric data to assess NIH peer review's predictive power.
Purpose of the Study:
- To empirically test the predictive validity of NIH peer review.
- To determine if correlating peer review scores with bibliometric indices is a valid method for assessing predictive validity.
- To analyze the relationship between grant application scores and subsequent publication impact.
Main Methods:
- Utilized a large dataset of NIH grant applications and funded projects.
- Examined the correlation between peer review percentile scores and bibliometric indices (e.g., citation counts) of resulting publications.
- Analyzed the distribution of scores for funded versus unfunded applications and considered post-review negotiations.
Main Results:
- Significant restriction of range was observed in applications selected for funding.
- Funded applications with lower peer review scores were not random or representative.
- Negotiations between NIH institutes and applicants altered projects, meaning initial scores did not reflect final funded work.
- Citation metrics alone are insufficient and inappropriate for measuring scientific impact.
Conclusions:
- Retrospective correlation between NIH peer review scores and bibliometric indices is not a valid test of predictive validity.
- Bibliometric indices alone are inadequate measures of scientific impact and may worsen existing research issues.
- The current NIH peer review and evaluation system requires re-evaluation for true predictive accuracy.
Related Concept Videos
Reliability and Validity
Spearman's Rank Correlation Test
Spearman's test calculates correlation by...
Sensitivity, Specificity, and Predicted Value
Sensitivity is the...
Testing a Claim about Population Proportion
There are two methods of testing a claim about a population proportion: (1) Using the sample proportion from the data where a binomial distribution is approximated to the normal distribution and (2) Using the binomial probabilities calculated from the data.
The first method uses normal distribution as an approximation to the binomial distribution. The requirements are as follows: sample size is large...
Statistical Significance
Accuracy and Errors in Hypothesis Testing
In hypothesis testing, the probability of making a Type I error, denoted as α, is commonly set at 0.05. This significance level indicates a 5%...