Related Experiment Video
Updated: Jul 9, 2025

Author Spotlight: Exploring the Lifespan Dynamics of Healthy Human Hematopoiesis
Published on: December 8, 2023
Evaluating the performance of large language models in haematopoietic stem cell transplantation decision-making
Ivan Civettini1,2, Arianna Zappaterra1,2,3, Bianca Maria Granelli1,2
1Department of Medicine and Surgery, University of Milano-Bicocca, Monza, Italy.
Abstract:
In a first-of-its-kind study, we assessed the capabilities of large language models (LLMs) in making complex decisions in haematopoietic stem cell transplantation. The evaluation was conducted not only for Generative Pre-trained Transformer 4 (GPT-4) but also conducted on other artificial intelligence models: PaLm 2 and Llama-2. Using detailed haematological histories that include both clinical, molecular and donor data, we conducted a triple-blind survey to compare LLMs to haematology residents. We found that residents significantly outperformed LLMs (p = 0.02), particularly in transplant eligibility assessment (p = 0.01). Our triple-blind methodology aimed to mitigate potential biases in evaluating LLMs and revealed both their promise and limitations in deciphering complex haematological clinical scenarios.
Related Concept Videos
Lineage Commitment
Regulation of Hematopoietic Stem Cells
Bone Marrow Sampling and Transplants
The transplant begins with high doses of chemotherapy and radiation treatment, which aim to destroy...

