From Translation to Retrieval: Evaluating LLM-Based Information Retrieval for Hausa and Fongbe
A large LLM-capability gap between the two languages is confirmed, and data augmentation experiments across three encoder models show that LLM-generated text consistently hurts downstream NER tasks while producing mixed effects on POS tagging, motivating careful language-specific IR evaluation.