Skip to content

6 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Preprint Aug 2026

TRIAGE: Risk-Controlled Pseudo-Label Admission for Annotation-Efficient Semi-Supervised Retinal OCT Classification

The advanced retinal disease diagnosing imaging modality, optical coherence tomography (OCT), encounters a lack of automation because of the high expenses for annotations performed by specialists. The use of SSL solves the problem of insufficient annotations using unlabeled B-scans; however, most of the current techniq...

Md. Ashraful Hossen Akash, Shyla Afroge, A. al Mamun et al. · 0 citations
Review Open access Aug 2026

Cross-modal bias in medical vision-language models: a pipeline-aware framework for mechanisms, evaluation, and mitigation

Medical vision-language models encode images and clinical text in a shared representation. Across radiology and ophthalmology, their diagnostic performance now approaches that of specialist clinicians. The mechanism behind that performance is also the source of a problem that has gone largely unexamined. These models a...

Rafid Mehda, Ramisa Anjum Oishi, Tamzid Tanvi Alam et al. · 0 citations
#generative ai Open access Sep 2026

Class balanced diabetic retinopathy image synthesis using a latent diffusion framework

This study establishes the Diabetic Retinopathy Latent Diffusion Synthesizer (DR-LDS) as a highly resource-efficient solution to the medical data bottleneck by leveraging a fine-tuned Variational Autoencoder for domain-adapted latent space compression, alongside an optimized U-Net.

Touhid Alam, Tze Hui Liew, M. Morol et al. · 0 citations
#computer vision Preprint Sep 2026

IchthyoNoma: Nomenclature and Context Sensitivity of Zero-Shot Biological Vision--Language Models for Bangladeshi Freshwater Fish Recognition

Zero-shot vision-language models (VLMs) are increasingly used as training-free species recognizers, but reported accuracy can reflect more than visual species knowledge. We audit CLIP, BioCLIP, BioCLIP2, and a multilingual Jina CLIP v2 control on seven freshwater-fish categories from two Bangladeshi sources (10,321 ima...

Nazim-E-Alam, Tarek Rahman, M. Morol · 0 citations
#large language models Review Open access Sep 2026

Multimodal medical diagnosis: a mini review of LLM–vision fusion models in low-resource healthcare settings

Recent advances in large language models (LLMs) and vision transformers have enabled multimodal systems that integrate clinical text with medical imaging for diagnostic decision-making. While these systems show promising results on benchmark datasets in well-resourced research settings, their applicability in low-resou...

Kahakashan Ashraf, Md.Hamid Hosen, N. Farah et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.