Aug 2026· Cell Reports Medicine· Vol 7, pp. 102969· 0 citations· 45 references
Medicine
TL;DR
AgentEYE is an auditable multimodal agent that routes ocular images to specialized fundus and B-scan tools, retrieves guideline/web evidence, and synthesizes evidence-grounded reports that supports evidence grounding and citation auditability.
Abstract
Summary Multimodal ophthalmic diagnosis requires integrating fundus photography, B-scan ultrasonography, and medical evidence, yet most artificial intelligence (AI) systems remain single-task or weakly grounded. AgentEYE is an auditable multimodal agent that routes ocular images to specialized fundus and B-scan tools, retrieves guideline/web evidence, and synthesizes evidence-grounded reports. In a 302-case internal benchmark, AgentEYE shows higher diagnostic correctness and completeness than large language model (LLM)-only baselines and an ablation without specialized imaging tools; performance remains similar to the no-retrieval ablation, indicating that retrieval mainly supports evidence grounding and citation auditability. Blinded evaluation of 200 cases by three ophthalmologists confirms improved diagnostic correctness, completeness, safety, and citation grounding versus an LLM-only self-citation baseline. External analyses show distribution-dependent performance. These findings support AgentEYE as a traceable decision-support prototype requiring prospective multicenter validation.
Medical visual agents can use tools to inspect images and retrieve external knowledge, but indiscriminate tool use may introduce noisy or misleading evidence. Reliable diagnosis therefore requires not only acquiring additional observations, but also verifying whether tool actions are necessary and whether the resulting...
Sheng-Zhi Wang, Jun Yang, Kai Wu et al.· 0 citations
Recent advances in ophthalmic foundation models have accelerated the application of artificial intelligence in ophthalmology, improving disease detection, progression assessment, and treatment evaluation. These advances are supported by the increasing availability of multimodal ophthalmic imaging data, including optica...
Li-Ping Ren, Jun-Yang Huang, Z. Du et al.· IEEE journal of biomedical a...· 0 citations
An engineering-oriented deployment framework is proposed, integrating modality-driven model selection, structured preprocessing pipelines, multi-level clinical validation, computational feasibility assessment, and explainability, together with a clinical deployment readiness model spanning validation maturity, data div...
Enoch Jacob Dodo, Amos Takai Yayock, Gregory Onwodi et al.· Journal of Science Research...· 0 citations
Multimodal clinical decision-making requires reliable reasoning over heterogeneous evidence from electronic health records, medical images, and physiological signals. Existing models typically map these inputs directly to diagnoses without explicitly assessing evidence sufficiency, tool-use requirements, or diagnostic...
This narrative review summarizes the methodological evolution of ophthalmic AI, including traditional machine learning, task‐specific deep learning, self‐supervised learning, foundation models, multimodal AI, and generative AI, and examines their applications across major ophthalmic diseases.
Yu-Xi Liu, Han-Ruo Liu· Eye & ENT Research· 0 citations
Thyroid ultrasound diagnosis requires coordinated lesion localization, measurement, risk stratification and reporting, yet most AI systems address these tasks in isolation and provide limited support for clinical review. We present ThyroidXAgent, a clinician-interactive agentic AI system that coordinates specialized di...
Haifan Gong, Shiyu Chen, Bodong Wang et al.· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.