Skip to content
Research Article Open access CC BY 4.0

Developing Statistical Machine Translation System for English and Nigerian Languages

Ignatius Ikechukwu Ayogu, Adebayo Olusola Adetunmbi, Bolanle Adefowoke Ojokoh

Asian Journal of Research in Computer Science · pp. 1–8 · Published 23 Oct 2018

10.9734/ajrcos/2018/v1i424761

Abstract

The global demand for translation and translation tools currently surpasses the capacity of available solutions. Besides, there is no one-solution-fits-all, off-the-shelf solution for all languages. Thus, the need and urgency to increase the scale of research for the development of translation tools and devices continue to grow, especially for languages suffering under the pressure of globalisation. This paper discusses our experiments on translation systems between English and two Nigerian languages: Igbo and Yorùbá. The study is setup to build parallel corpora, train and experiment English-to-Igbo, (), English-to-Yorùbá, () and Igbo-to-Yorùbá, () phrase-based statistical machine translation systems. The systems were trained on parallel corpora that were created for each language pair using text from the religious domain in the course of this research. A BLEU score of 30.04, 29.01 and 18.72 respectively was recorded for the English-to-Igbo, English-to-Yorùbá and Igbo-to-Yorùbá MT systems. An error analysis of the systems’ outputs was conducted using a linguistically motivated MT error analysis approach and it showed that errors occurred mostly at the lexical, grammatical and semantic levels. While the study reveals the potentials of our corpora, it also shows that the size of the corpora is yet an issue that requires further attention. Thus an important target in the immediate future is to increase the quantity and quality of the data.  

Machine translation Igbo language Yoruba language parallel corpora SMT

Cited by 9

Bridging Gaps in Natural Language Processing for Yorùbá: A Systematic Review of a Decade of Progress and Prospects

Toheeb A. Jimoh, Tabea De Wille, Nikola S. Nikolov · Natural Language Processing Journal · 2025

Bilingual Neural Machine Translation From English To Yoruba Using A Transformer Model

Adeboje Olawale Timothy, Adetunmbi Olusola Adebayo, A. J. Gabriel · International Journal of Innovative Science and Research Technology · 2024

A Rule-based Approach to English-Okun Prepositional Phrase Machine Translation

A. Esan, A. Sobowale, T. Adebiyi · Dutse Journal of Pure and Applied Sciences · 2024

Case Study on Data Collection of Kreol Morisien, a Low-Resourced Creole Language

David J. Bastien, Vijay Prakash Chumroo, Johan Patrice Bastien · 2022 IST-Africa Conference (IST-Africa) · 2022

Developing an Open-Source Corpus of Yoruba Speech

Alexander Gutkin, Isin Demirsahin, Oddur Kjartansson · Interspeech · 2020

Development of a Recurrent Neural Network Model for English to Yorùbá Machine Translation

A. Esan, J. Oladosu, C. Oyeleye · International Journal of Advanced Computer Science and Applications · 2020

A Transformer-Based Yoruba to English Machine Translation (TYEMT) System with Rouge Score

Oluwatoki, Tolani Grace, Adetunmbi, Olusola Adebayo, Boyinbode, Olutayo Kehinde · International Journal of Innovative Science and Research Technology (IJISRT) · 2024

Machine Translation Systems for Nigerian Indigenous Languages: A Statistical Overview

Tolani Grace Oluwatoki, Adebayo Olusola Adetunmbi, Olutayo Kehinde Boyinbode · 2025

Article metrics

Real usage data collected on this platform.

0

Page views

0

PDF downloads

0

Outbound clicks

9

Citations

Views by country

Approximate, from request IP at view time — not citizenship or institution. Countries with fewer than 5 views are grouped as "Other".

No views recorded yet.

Traffic sources

Referring site, by host.

No traffic recorded yet.

Views and downloads exclude known bots/crawlers. Citations combines this platform's own DOI-resolved index with each external source's own reported total — see Cited by above for individually listed citing works. Last refreshed 0 seconds ago.