Related Experiment Video
Updated: Jan 6, 2026

10:32
Development of a Virtual Reality Assessment of Everyday Living Skills
Published on: April 23, 2014
19.0K
What is test accuracy? Comparing unitary accuracy metrics for cognitive screening instruments
1Cognitive Function Clinic, Walton Center for Neurology & Neurosurgery, Lower Lane, Fazakerley, Liverpool, L9 7LJ, UK.
Neurodegenerative Disease Management
|October 4, 2019
Summary
This study evaluated four accuracy metrics for cognitive screening tools. The Matthews Correlation Coefficient (MCC) showed theoretical advantages over other measures for assessing diagnostic accuracy in dementia and mild cognitive impairment.
Area of Science:
- Neuroscience
- Gerontology
- Psychometrics
Background:
- Cognitive screening instruments are crucial for early detection of dementia and mild cognitive impairment.
- Accurate assessment of these instruments relies on appropriate statistical metrics.
- Existing metrics may have limitations in reflecting true diagnostic performance.
Purpose of the Study:
- To compare four common accuracy metrics: correct classification accuracy, area under the receiver operating characteristic curve (AUC), F-measure (F1 score), and Matthews Correlation Coefficient (MCC).
- To evaluate the performance of these metrics in assessing widely used cognitive screening instruments.
- To determine the most suitable metric for accuracy studies of cognitive tests.
Main Methods:
- Extracted raw data from accuracy studies of six cognitive screening instruments: Mini-Mental State Examination, Montreal Cognitive Assessment, Mini-Addenbrooke's Cognitive Examination, Six-item Cognitive Impairment Test, informant AD8, and Free-Cog.
- Calculated correct classification accuracy, AUC, F1 score, and MCC for each instrument.
- Analyzed the ordering of instruments based on each metric for diagnosing dementia and mild cognitive impairment.
Main Results:
- All four metrics produced a similar ranking of the cognitive screening instruments.
- Area under the ROC curve (AUC) provided the most optimistic accuracy values, while MCC yielded the most pessimistic.
- Correct classification accuracy and F1 score fell between AUC and MCC in terms of reported accuracy.
Conclusions:
- Each accuracy metric has inherent limitations and cannot be recommended as a sole definitive measure.
- The Matthews Correlation Coefficient (MCC) offers theoretical advantages and may be a more robust choice for future studies.
- Further research is needed to establish standardized outcome measures for cognitive screening test accuracy.
Related Concept Videos
Measures of Intelligence
8.2K
Psychologists measure intelligence by using standardized tests that produce a score known as the intelligence quotient or IQ. To understand IQ tests, it's important to recognize the key principles behind their construction: validity, reliability, and standardization.
Validity refers to how well a test measures what it claims to measure. An intelligence test should accurately assess intelligence rather than another characteristic, like anxiety. Criterion validity is one way to evaluate this;...
Validity refers to how well a test measures what it claims to measure. An intelligence test should accurately assess intelligence rather than another characteristic, like anxiety. Criterion validity is one way to evaluate this;...
8.2K
Self-Report Tests of Personality
729
Self-report inventories are objective personality assessments that use multiple-choice items or numbered scales, typically ranging from 1 (strongly disagree) to 5 (strongly agree). They are often called Likert scales after Rensis Likert. These inventories are widely used due to their ease of administration and cost-effectiveness. One of the most prominent examples is the Minnesota Multiphasic Personality Inventory (MMPI), initially developed in the 1940s to assess abnormal personality traits.
729
Wechsler's Contribution to Measures of Intelligence
2.0K
David Wechsler, a psychologist who worked with World War I veterans, developed a significant IQ test in 1939 called the Wechsler-Bellevue Intelligence Scale. This test was innovative because it combined several subtests that measured both verbal and nonverbal skills, reflecting Wechsler's belief that intelligence is a global capacity involving purposeful action, rational thinking, and effective interaction with the environment. This test later evolved into the Wechsler Adult Intelligence...
2.0K
Multiple Comparison Tests
4.4K
Multiple comparison test, abbreviated as MCT, is a post hoc analysis generally performed after comparing multiple samples with one or more tests. An MCT will help identify a significantly different sample among multiple samples or a factor among multiple factors.
It would be easy to compare two samples using a significance alpha level of 0.05. In other words, there is only one sample pair to be compared. However, it would be difficult to identify a significantly different sample if the number...
It would be easy to compare two samples using a significance alpha level of 0.05. In other words, there is only one sample pair to be compared. However, it would be difficult to identify a significantly different sample if the number...
4.4K
Binet's Contribution to Measures of Intelligence
1.6K
Alfred Binet, along with his student Théophile Simon, was tasked by the French Ministry of Education in 1904 to create a method for identifying students who struggled to learn through conventional classroom instruction. This initiative aimed to address overcrowding by placing such students in specialized schools. Binet and Simon developed an intelligence test comprising 30 tasks, ranging from simple commands, like touching one's nose or ear, to more complex tasks, such as drawing...
1.6K

