在多式大型语言模型中评估跨种族情绪识别能力,使用阅读眼睛中的思维测试

Elad Refoua1, Zohar Elyoseph2,3, David Piterman4

  • 1Department of Psychology, Bar-Ilan University, Ramat-Gan, Israel. eladrefoua@gmail.com.

Scientific reports
|February 20, 2026
PubMed
概括

聊天GPT-4o显示了先进的情感识别,在眼部情感测试中超过了不同种族群体的人类准确性. 这突显了社会认知任务的多式大型语言模型 (MLLMs) 的重大进展.