适度测量:在早上测量有效性受到损害吗?
Georgios Sideridis1,2, Fathima Jaffari3
1Boston Children's Hospital and Harvard Medical School, Boston, MA, United States.
Frontiers in psychology
|September 7, 2023
概括
在一般能力测试 (GAT) 上的早晨测试损害了其有效性和可靠性. 晚上测试表明,在沙特阿拉伯的学生中,心理测量特性优越,表现水平更高.
科学领域:
- 教育测量教育的测量
- 心理测量 心理测量 心理测量
- 认知心理学 认知心理学
背景情况:
- 一般能力测试 (GAT) 是沙特阿拉伯评估能力和成绩的关键国家工具.
- 了解测试条件的影响,例如一天中的时间,对于确保准确可靠的评估至关重要.
研究的目的:
- 根据每天的时间来评估通用能力测试 (GAT) 的可靠性和有效性.
- 在早上和晚上测试期间调查测量不变性和差异物品功能 (DIF).
主要方法:
- 在722名学生参加GAT的早上和晚上课程中,采用了个人预后设计.
- 参与者与关键的人口统计和测试中心变量相匹配.
- 分析了心理测量属性,包括内部一致性,维度和测量不变性.
主要成果:
- 发现了显著的不合适,表明由于一天的时间而导致的差异性项目功能 (DIF).
- 与晚上相比,内部一致性可靠性在早上较低.
- 早晨测试与所有领域学生表现的显著下降有关.
结论:
- 一天中的时间显著影响了通用能力测试 (GAT) 的心理测量属性和有效性.
- 晨间测试管理可能会损害能力和成绩测量的准确性.
- 公共测试政策应该考虑这些发现,以确保有效和可靠的评估.
相关概念视频
Reliability and Validity
12.8K
Reliability and validity are two important considerations that must be made with any type of data collection. Reliability refers to the ability to consistently produce a given result. In the context of psychological research, this would mean that any instruments or tools used to collect data do so in consistent, reproducible ways.
12.8K
Measures of Intelligence
7.6K
Psychologists measure intelligence by using standardized tests that produce a score known as the intelligence quotient or IQ. To understand IQ tests, it's important to recognize the key principles behind their construction: validity, reliability, and standardization.
Validity refers to how well a test measures what it claims to measure. An intelligence test should accurately assess intelligence rather than another characteristic, like anxiety. Criterion validity is one way to evaluate this;...
Validity refers to how well a test measures what it claims to measure. An intelligence test should accurately assess intelligence rather than another characteristic, like anxiety. Criterion validity is one way to evaluate this;...
7.6K
Binet's Contribution to Measures of Intelligence
1.3K
Alfred Binet, along with his student Théophile Simon, was tasked by the French Ministry of Education in 1904 to create a method for identifying students who struggled to learn through conventional classroom instruction. This initiative aimed to address overcrowding by placing such students in specialized schools. Binet and Simon developed an intelligence test comprising 30 tasks, ranging from simple commands, like touching one's nose or ear, to more complex tasks, such as drawing...
1.3K
Correlations
33.0K
Correlation means that there is a relationship between two or more variables (such as ice cream consumption and crime), but this relationship does not necessarily imply cause and effect. When two variables are correlated, it simply means that as one variable changes, so does the other. We can measure correlation by calculating a statistic known as a correlation coefficient. A correlation coefficient is a number from -1 to +1 that indicates the strength and direction of the relationship between...
33.0K
Self-Report Tests of Personality
380
Self-report inventories are objective personality assessments that use multiple-choice items or numbered scales, typically ranging from 1 (strongly disagree) to 5 (strongly agree). They are often called Likert scales after Rensis Likert. These inventories are widely used due to their ease of administration and cost-effectiveness. One of the most prominent examples is the Minnesota Multiphasic Personality Inventory (MMPI), initially developed in the 1940s to assess abnormal personality traits.
380
Wechsler's Contribution to Measures of Intelligence
1.5K
David Wechsler, a psychologist who worked with World War I veterans, developed a significant IQ test in 1939 called the Wechsler-Bellevue Intelligence Scale. This test was innovative because it combined several subtests that measured both verbal and nonverbal skills, reflecting Wechsler's belief that intelligence is a global capacity involving purposeful action, rational thinking, and effective interaction with the environment. This test later evolved into the Wechsler Adult Intelligence...
1.5K


