Related Experiment Video
Updated: Aug 20, 2026

Association Between Sleep Quality and Cognitive Symptoms in Patients with Major Depressive Disorder
Published on: April 26, 2024
The Hamilton Depression Rating Scale: has the gold standard become a lead weight?
R Michael Bagby1, Andrew G Ryder, Deborah R Schuller
1Centre for Addiction and Mental Health, 250 College St., Toronto, Ont., Canada M5T 1R8. michael_bagby@camh.net
Insights
The Hamilton Depression Rating Scale shows psychometric and conceptual flaws, questioning its use. It is time to adopt a new gold standard for depression assessment.
Area of Science:
- Psychiatry
- Psychometrics
- Clinical Psychology
Background:
- The Hamilton Depression Rating Scale (HDRS) has been a long-standing measure for depression severity.
- Increasing criticism questions the HDRS's psychometric properties and clinical utility.
- A comprehensive review of studies published since 1979 was conducted.
Purpose of the Study:
- To evaluate the psychometric properties of the HDRS.
- To determine if the HDRS remains a justified measure for assessing depression treatment outcomes.
- To assess the validity, reliability, and item characteristics of the HDRS.
Main Methods:
- A systematic literature search of MEDLINE was performed for studies published after 1979.
- Seventy studies examining the HDRS's psychometric properties were identified and selected.
- Studies were categorized based on reliability, item-response characteristics, and validity.
Main Results:
- The HDRS demonstrates adequate internal reliability, but many items lack contribution and exhibit poor interrater/retest reliability.
- Response formats for several items are suboptimal, and content validity is poor.
- Convergent and discriminant validity are adequate, but the factor structure lacks consistent replication across samples.
Conclusions:
- The HDRS exhibits significant psychometric and conceptual limitations.
- The identified flaws are substantial, making revision efforts inadvisable.
- A new, superior instrument is needed to serve as the gold standard for depression assessment.
Objective:
The Hamilton Depression Rating Scale has been the gold standard for the assessment of depression for more than 40 years. Criticism of the instrument has been increasing. The authors review studies published since the last major review of this instrument in 1979 that explicitly examine the psychometric properties of the Hamilton depression scale. The authors' goal is to determine whether continued use of the Hamilton depression scale as a measure of treatment outcome is justified.
Method:
MEDLINE was searched for studies published since 1979 that examine psychometric properties of the Hamilton depression scale. Seventy studies were identified and selected, and then grouped into three categories on the basis of the major psychometric properties examined-reliability, item-response characteristics, and validity.
Results:
The Hamilton depression scale's internal reliability is adequate, but many scale items are poor contributors to the measurement of depression severity; others have poor interrater and retest reliability. For many items, the format for response options is not optimal. Content validity is poor; convergent validity and discriminant validity are adequate. The factor structure of the Hamilton depression scale is multidimensional but with poor replication across samples.
Conclusions:
Evidence suggests that the Hamilton depression scale is psychometrically and conceptually flawed. The breadth and severity of the problems militate against efforts to revise the current instrument. After more than 40 years, it is time to embrace a new gold standard for assessment of depression.