Research & Publications
I am actively involved in research, focusing on Natural Language Processing, Machine Learning, and Speech Processing for low-resource languages. Here are my selected publications and papers (you can also view my complete research profile on Google Scholar):
The WAXAL ASR Benchmark: Fine-Tuned Edge Models Across 19 African Languages
2026VT Olufemi, O Babatunde, R Njema, B Gbotemi, WL Yen, J Uzodinma, S Ajayi, O Williams, K Moshood, IE Anyaele, AT Arefaine, C Hunzwi, WD Daniel, EI Namuganga, C Kadima, AB Bahizire, O Ranaivoson, E Aaron, ND Ladislaus, I Muhammed, JE Simenya, M Koome, MT Endaylalu, PI Adeyemo, HP Birindwa, UA Eze-Mbey, Y Oduro-Yeboah, T Aremu, P Adjovi, MK Ngueajio, P Mitra
arXiv preprint arXiv:2606.02375
Ehugbo-QA: Diagnosing the Alignment Gap in Dialectal Cross-Lingual Information Retrieval
2026U Eze-Mbey, VT Olufemi, AB Bahizire, P Mitra, MK Ngueajio
Proceedings of the 49th International ACM SIGIR Conference on Research and Development in Information Retrieval (Resources Track)
A Survey of Text and Speech Resources for Hausa and Fongbe: Availability, Quality, and Gaps for NLP Development
2026MP Adjovi, VT Olufemi, R Eiselen, P Mitra
Proceedings of the 2026 IEEE Swiss Conference on Data Science and AI (SDS)
Evaluating Yoruba Text-to-Speech Systems for Accessible Computer-Based Testing in Visually Impaired Learners
2026KY Moshood, VT Olufemi, OB Babatunde, E Bolarinwa, W Oluwademilade
Proceedings of the 7th Workshop on African Natural Language Processing
Closing the Gap in Low-Resource ASR: Leveraging Multilingual Models for Code-Switched Yoruba-English Speech
2026E Bolarinwa, O Babatunde, V Olufemi, K Moshood, O Williams
Deep Learning Indaba
Challenging Multimodal LLMs with African Standardized Exams: A Document VQA Evaluation
2025VT Olufemi, OB Babatunde, E Bolarinwa, KY Moshood
Proceedings of the Sixth Workshop on African Natural Language Processing
Beyond monolingual limits: Fine-tuning monolingual asr for yoruba-english code-switching
2025OB Babatunde, VT Olufemi, E Bolarinwa, KY Moshood, CC Emezue
Proceedings of the 7th Workshop on Computational Approaches to Linguistic...
When Endangered Voices Speak: Building the First Ehugbo Dialect Audio Dataset Through Grassroots Collaboration
2025UA EZE-MBEY, VT Olufemi, UC Eze-Mbey
Women in Machine Learning Workshop @ NeurIPS 2025
Automatic Speech Recognition for Nigerian-Accented English
2023OB Babatunde, E Akeweje, S Ibejih, VT Olufemi, SO Folorunso
Deep Learning Indaba 2023 Conference