Research & Publications

I am actively involved in research, focusing on Natural Language Processing, Machine Learning, and Speech Processing for low-resource languages. Here are my selected publications and papers (you can also view my complete research profile on Google Scholar):

The WAXAL ASR Benchmark: Fine-Tuned Edge Models Across 19 African Languages

2026

VT Olufemi, O Babatunde, R Njema, B Gbotemi, WL Yen, J Uzodinma, S Ajayi, O Williams, K Moshood, IE Anyaele, AT Arefaine, C Hunzwi, WD Daniel, EI Namuganga, C Kadima, AB Bahizire, O Ranaivoson, E Aaron, ND Ladislaus, I Muhammed, JE Simenya, M Koome, MT Endaylalu, PI Adeyemo, HP Birindwa, UA Eze-Mbey, Y Oduro-Yeboah, T Aremu, P Adjovi, MK Ngueajio, P Mitra

arXiv preprint arXiv:2606.02375

Ehugbo-QA: Diagnosing the Alignment Gap in Dialectal Cross-Lingual Information Retrieval

2026

U Eze-Mbey, VT Olufemi, AB Bahizire, P Mitra, MK Ngueajio

Proceedings of the 49th International ACM SIGIR Conference on Research and Development in Information Retrieval (Resources Track)

A Survey of Text and Speech Resources for Hausa and Fongbe: Availability, Quality, and Gaps for NLP Development

2026

MP Adjovi, VT Olufemi, R Eiselen, P Mitra

Proceedings of the 2026 IEEE Swiss Conference on Data Science and AI (SDS)

Evaluating Yoruba Text-to-Speech Systems for Accessible Computer-Based Testing in Visually Impaired Learners

2026

KY Moshood, VT Olufemi, OB Babatunde, E Bolarinwa, W Oluwademilade

Proceedings of the 7th Workshop on African Natural Language Processing

Closing the Gap in Low-Resource ASR: Leveraging Multilingual Models for Code-Switched Yoruba-English Speech

2026

E Bolarinwa, O Babatunde, V Olufemi, K Moshood, O Williams

Deep Learning Indaba

Challenging Multimodal LLMs with African Standardized Exams: A Document VQA Evaluation

2025

VT Olufemi, OB Babatunde, E Bolarinwa, KY Moshood

Proceedings of the Sixth Workshop on African Natural Language Processing

Beyond monolingual limits: Fine-tuning monolingual asr for yoruba-english code-switching

2025

OB Babatunde, VT Olufemi, E Bolarinwa, KY Moshood, CC Emezue

Proceedings of the 7th Workshop on Computational Approaches to Linguistic...

When Endangered Voices Speak: Building the First Ehugbo Dialect Audio Dataset Through Grassroots Collaboration

2025

UA EZE-MBEY, VT Olufemi, UC Eze-Mbey

Women in Machine Learning Workshop @ NeurIPS 2025

Automatic Speech Recognition for Nigerian-Accented English

2023

OB Babatunde, E Akeweje, S Ibejih, VT Olufemi, SO Folorunso

Deep Learning Indaba 2023 Conference