Related Experiment Video
Updated: Jun 7, 2026

06:11
Nest Building Behavior as an Early Indicator of Behavioral Deficits in Mice
Published on: October 19, 2019
Agreement between two different scoring procedures for goal attainment scaling is low
Thamar J H Bovend'Eerdt1, Helen Dawes, Hooshang Izadi
1Department of Movement Science, Maastricht University, Maastricht, The Netherlands. thamar.bovendeerdt@maastrichtuniversity.nl
Journal of Rehabilitation Medicine
|November 3, 2010
Summary
Therapist and independent assessor agreement on patient goal attainment was poor. Improving goal attainment scaling reproducibility is crucial for its use in blinded randomized trials.
Area of Science:
- Neurology
- Rehabilitation Medicine
- Clinical Psychology
Background:
- Goal attainment scaling (GAS) is a method to measure patient progress.
- Its reliability as an outcome measure in clinical trials requires validation.
Purpose of the Study:
- To assess the agreement between treating therapists and independent assessors in scoring patient goal attainment.
- To evaluate the reliability of GAS in neurological patient populations.
Main Methods:
- Therapists set 2-4 goals for neurological patients in a randomized trial.
- Goal attainment was scored by treating therapists and an independent assessor after six weeks.
- A semi-structured interview and direct assessment were used for scoring.
Main Results:
- Analysis included 112 goals from 29 neurological patients.
- The intraclass correlation coefficient (ICC) indicated poor agreement (0.478).
- Limits of agreement showed substantial variability, with no systematic bias detected.
Conclusions:
- There is low agreement between therapist and independent scoring of goal attainment.
- The reproducibility of goal attainment scaling needs improvement for use in blinded randomized controlled trials.
- Further research is needed to enhance the reliability of GAS as an outcome measure.
Related Concept Videos
Standard Deviation
The most commonly used measure of variation is the standard deviation. It is a numerical value measuring how far data values are from their mean. The standard deviation value is small when the data are concentrated close to the mean, exhibiting slight variation or spread. The standard deviation value is never negative, it is either positive or zero. The standard deviation is larger when the data values are more spread out from the mean, which means the data values are exhibiting more...
Ratio Level of Measurement
The way a set of data is measured is called its level of measurement. Correct statistical procedures depend on a researcher being familiar with levels of measurement. For analysis, data are classified into four levels of measurement—nominal, ordinal, interval, and ratio.
A set of data measured using the ratio scale takes care of the ratio problem and provides complete information. Ratio scale data are like interval scale data, except they have a zero point and ratios can be calculated. For...
A set of data measured using the ratio scale takes care of the ratio problem and provides complete information. Ratio scale data are like interval scale data, except they have a zero point and ratios can be calculated. For...
One-Way ANOVA: Equal Sample Sizes
One-Way ANOVA can be performed on three or more samples with equal or unequal sample sizes. When one-way ANOVA is performed on two datasets with samples of equal sizes, it can be easily observed that the computed F statistic is highly sensitive to the sample mean.
Different sample means can result in different values for the variance estimate: variance between samples. This is because the variance between samples is calculated as the product of the sample size and the variance between the...
Different sample means can result in different values for the variance estimate: variance between samples. This is because the variance between samples is calculated as the product of the sample size and the variance between the...
Weighted Mean
While taking the arithmetic, geometric, or harmonic mean of a sample data set, equal importance is assigned to all the data points. However, all the values may not always be equally important in some data sets. An intrinsic bias might make it more important to give more weightage to specific values over others.
For example, consider the number of goals scored in the matches of a tournament. While computing the average number of goals scored in the tournament, it may be more important to...
For example, consider the number of goals scored in the matches of a tournament. While computing the average number of goals scored in the tournament, it may be more important to...
Kendall's Coefficient of Concordance
Kendall's Coefficient of Concordance (W), also known as Kendall's W, is a non-parametric statistical measure used to assess the agreement or concordance between multiple raters or judges when they rank a set of items. It is often used when you have ordinal data (ranks) and you want to see if there is consistency or consensus among the raters. It is widely applied in research areas such as psychology, medicine, and social sciences, where multiple judges are asked to rank or rate subjects or...
z Scores and Unusual Values
The z score is one of the three measures of relative standing. It describes the location of a value in a dataset relative to the mean. z scores are obtained after the standardization of the values in a dataset. The z score for the mean is 0.
This score indicates how far a value is from the mean in terms of standard deviation. For example, if a data value has a z score of +1, the researcher can infer that the particular data value is one standard deviation above the mean. If another data value...
This score indicates how far a value is from the mean in terms of standard deviation. For example, if a data value has a z score of +1, the researcher can infer that the particular data value is one standard deviation above the mean. If another data value...
