Benchmarking large language models GPT-4o, llama 3.1, and qwen 2.5 for cancer genetic variant classification

Kuan-Hsun Lin1,2, Tzu-Hang Kao3, Lei-Chi Wang3,4

  • 1Department of Information Management, Taipei Veterans General Hospital, Taipei, Taiwan, ROC.

PubMed
Summary

Large language models (LLMs) show promise in classifying cancer genetic variants for precision oncology. GPT-4o achieved the highest accuracy, but further optimization is needed for clinical use.