Competency of Large Language Models in Evaluating Appropriate Responses to Suicidal Ideation: Comparative Study

Ryan K McBain1,2,3, Jonathan H Cantor4, Li Ang Zhang4

  • 1RAND, Arlington, VA, United States.

Summary

Large language models (LLMs) show an upward bias in rating responses for suicidal ideation, but two models match or exceed mental health professional performance. This study assessed LLM competency in evaluating suicide risk responses.