Large language models show promising performance for some systematic review tasks but call for cautious

Florian Laignelot1, Guillaume L Martin1, Mohamad Ossman1

  • 1Sorbonne Université, INSERM, Institut Pierre Louis d'Epidémiologie et de Santé Publique, UMR-S 1136, AP-HP, Hôpital Pitié-Salpêtrière, Département de Santé Publique, Paris, France.

PubMed
Summary

Large language models (LLMs) show promise for automating systematic review tasks like screening, with newer models performing better. Careful implementation is key for integrating LLMs into systematic reviews.