Comparing the performance of ChatGPT, DeepSeek, and Gemini in systematic and umbrella review tasks over time

Abstract