Study of comparative performance of general-purpose LLM-based systems in predicting IVF outcomes

Can Dinç1, Ömer Faruk Öz2, Saltuk Buğra Arıkan2

  • 1Department of Gynecology and Obstetrics, Akdeniz University, Antalya, Turkey. candinc@akdeniz.edu.tr.

Summary

General-purpose AI models like ChatGPT, DeepSeek, and Gemini show limited accuracy in predicting in vitro fertilization (IVF) outcomes. Their current performance is insufficient for clinical decision support in IVF, requiring further validation and research.