EyeVQA: Benchmarking Ophthalmic Vision-Language Models from Recognition to Spatial Grounding
Vision-language models (VLMs) have shown increasing potential for medical image understanding, yet their capabilities in ophthalmic imaging remain insufficiently characterized. Existing ophthalmic datasets are typically designed for individual diseases or specialized tasks, making it difficult to systematically evaluat...