Jul 2026
Benchmarking and AI-assisted human-like evaluation of retrieval-augmented generation for Arabic and English documents
An important practical reproducible framework for multilingual RAG benchmarking and insights for optimizing performance on resource-constrained devices are contributed and support the latent language hypothesis by suggesting an internal model bias toward high-resource languages.
B. J. Mohd, Khalil M. Ahmad Yousef, Salah G. Abu Ghalyon
· Language Resources and Evalu... · 0 citations