Summary

Vision-language foundation (VLF) models show promise for ocular disease screening. Context-aware VLF models improved diabetic retinopathy grading and generalized to other ocular conditions, demonstrating enhanced robustness.