Functional Symmetry Observation Scale, Version 2: discriminant validity and recommendations for the development of
Mary Rahlin1, Steven Miller2, Bernadette Sarmiento3
1Board-Certified Clinical Specialist in Pediatric Physical Therapy, Department of Physical Therapy, Rosalind Franklin University of Medicine and Science, North Chicago, IL, USA.
Physiotherapy Theory and Practice
|December 3, 2025
Summary
The Functional Symmetry Observation Scale-V2 (FSOS-V2) shows scoring bias affecting infant symmetry assessment. Revisions are recommended to improve its reliability and validity for clinical use in congenital muscular torticollis.
Area of Science:
- Pediatric Physical Therapy
- Infant Motor Development
- Clinical Assessment Tools
Background:
- The Functional Symmetry Observation Scale-V2 (FSOS-V2) is a video-based tool for assessing infant movement symmetry in congenital muscular torticollis.
- While possessing good validity and intrarater reliability, the FSOS-V2 exhibits poor to moderate interrater reliability.
- Previous data suggested potential scoring bias related to infant age and motor development, impacting the scale's discriminant validity.
Purpose of the Study:
- To evaluate the discriminant validity of the FSOS-V2.
- To conduct individual item analyses using existing reliability data.
- To develop recommendations for modifying the FSOS-V2 into a shorter, more reliable form.
Main Methods:
- Secondary analysis of FSOS-V2 reliability data from 50 infants with congenital muscular torticollis (CMT).
- Correlation analyses between FSOS-V2 scores, infant age, and gross motor milestones.
- Item response theory (IRT) modeling for individual item analysis and scale modification recommendations.
Main Results:
- One rater's scores significantly correlated with infant age and motor milestones, indicating potential scoring bias.
- IRT modeling provided specific recommendations for item removal and scale modification.
- Evidence of scoring bias impacting discriminant validity and interrater reliability was confirmed.
Conclusions:
- Scoring bias in the FSOS-V2 compromises its discriminant validity and interrater reliability.
- Revising the scale into a shorter form and refining scoring guidelines are necessary.
- These modifications are crucial for accurate and consistent assessment of symmetry in infants with CMT.
More Related Videos
Related Concept Videos
Self-Report Tests of Personality
740
Self-report inventories are objective personality assessments that use multiple-choice items or numbered scales, typically ranging from 1 (strongly disagree) to 5 (strongly agree). They are often called Likert scales after Rensis Likert. These inventories are widely used due to their ease of administration and cost-effectiveness. One of the most prominent examples is the Minnesota Multiphasic Personality Inventory (MMPI), initially developed in the 1940s to assess abnormal personality traits.
740
Strategies of Self-Presentation III: Self-Monitoring
183
Self-monitoring is a central construct in understanding individual differences in self-presentation strategies across social contexts. It refers to how individuals observe, regulate, and control their expressive behavior and self-presentation following situational cues. Self-monitoring reflects a person's sensitivity to social appropriateness and willingness to adapt behavior to fit varying interpersonal demands.High vs. Low Self-Monitoring IndividualsIndividuals high in self-monitoring are...
183
Reliability and Validity
13.7K
Reliability and validity are two important considerations that must be made with any type of data collection. Reliability refers to the ability to consistently produce a given result. In the context of psychological research, this would mean that any instruments or tools used to collect data do so in consistent, reproducible ways.
13.7K
Ordinal Level of Measurement
31.9K
The way a set of data is measured is called its level of measurement. Correct statistical procedures depend on a researcher being familiar with levels of measurement. For analysis, data are classified into four levels of measurement—nominal, ordinal, interval, and ratio.
Data measured using an ordinal scale are similar to nominal scale data, but there is one major difference. The ordinal scale data can be ordered. An example of ordinal scale data is a list of the top five national parks...
Data measured using an ordinal scale are similar to nominal scale data, but there is one major difference. The ordinal scale data can be ordered. An example of ordinal scale data is a list of the top five national parks...
31.9K
Routh-Hurwitz Criterion II
906
In the application of the Routh-Hurwitz criterion, two specific scenarios can arise that complicate stability analysis.
The first scenario occurs when a singular zero appears in the first column of the Routh table. This situation creates a division by zero issues. To resolve this, a small positive or negative number, denoted as epsilon (∈), is substituted for the zero. The stability analysis proceeds by assuming a sign for ∈. If ∈ is positive, any sign change in the first...
The first scenario occurs when a singular zero appears in the first column of the Routh table. This situation creates a division by zero issues. To resolve this, a small positive or negative number, denoted as epsilon (∈), is substituted for the zero. The stability analysis proceeds by assuming a sign for ∈. If ∈ is positive, any sign change in the first...
906


