An Empirical Evaluation of the GPT-4 Multimodal Language Model on Visualization Literacy Tasks.

Summary

Large Language Models (LLMs) like GPT-4 show promise for visualization research, accurately identifying trends and design principles. However, they struggle with data retrieval, color distinction, and can exhibit inconsistencies.