Contributing toacademic knowledge
Advancing the fields of AI, NLP, cybersecurity, and African language technology through rigorous research and open collaboration.
TukaBench: A Culturally Grounded Jailbreak Benchmark for African Languages
Safety evaluation of Large Language Models (LLMs) remains heavily English-centric, leaving Low-Resource Languages (LRLs), particularly African ones, critically underexplored. We introduce TukaBench, a jailbreak benchmark for seven African languages that extends JailbreakBench beyond direct translation, isolating the effect of language, cultural grounding, and prompt evasiveness on model safety.
Authors:
Victor Akinode, Senyu Li, Wassim Hamidouche, Waqas Zamir, Inbal Becker-Reshef, David Ifeoluwa Adelani
MasakhaNER: Named Entity Recognition for African Languages
This addressed the under-representation of the African continent in NLP research by bringing together different stakeholders to create the first large, publicly available, high-quality dataset for named entity recognition (NER) in ten African languages.
Authors:
Victor Akinode
Prediction of Urban House Rental Prices in Lagos - Nigeria: A Machine Learning Approach
This work aims to predict house rental prices in Lagos, Nigeria, using machine learning by examining the relationship between the rental price and features such as the number of bedrooms, bathrooms, toilets, location and house status(newly built, furnished, and/or serviced).
Authors:
Victor Akinode
BibleTTS
BibleTTS is a large, high-quality, open speech dataset for ten languages spoken in Sub-Saharan Africa. The corpus contains up to 86 hours of aligned, studio quality 48kHz single speaker recordings per language, enabling the development of high-quality text-to-speech models.
Authors:
Victor Akinode
VocalTweets: Investigating Social Media Offensive Language Among Nigerian Musicians
In this study, we introduce VocalTweets, a code-switched and multilingual dataset comprising tweets from 12 prominent Nigerian musicians, labeled with a binary classification method as Normal or Offensive.
Authors:
Victor Akinode
Development of Text-to-Speech Synthesis for Yoruba Language using Deep Learning
The study employs the BibleTTS corpus, consisting of audio recordings and text transcripts, to develop a TTS synthesis method that accommodates diverse rhythms and pitches..
Authors:
Victor Akinode
Exploring ChatGPT’s Bias: Can AI Pass West Africa’s WAEC Exams the Way It Passed US.-Based Exams?
In this article, I put ChatGPT to the test with a dataset of 1000 questions and answers, 200 randomly selected across five subjects (Government, Civic Education, Agricultural Science, Economics, Commerce and Geography),
Authors:
Victor Akinode