Machine Translation for Low-Resource English-Mizo Pair Encountering Tonal Words

Authors

  • Vanlalmuansangi Khenglawt Mizoram University
  • Sahinur Rahman Laskar National Institute of Technology
  • Partha Pakray National Institute of Technology
  • Riyanka Manna Gandhi Institute of Technology and Management
  • Ajoy Kumar Khan Mizoram University

DOI:

https://doi.org/10.13053/cys-26-3-4358

Keywords:

English-Mizo, machine translation, low-resource, tonal

Abstract

Machine translation is one of the most powerful natural language processing applications for preserving and upgrading low-resource language. Mizo language is considered as low-resource since there is limited availability of resources. Therefore, it is a challenging task for English-Mizo language pair translation. Moreover, Mizo is a tonal language, where a word can express different meanings depending on a variety of tones. There are four variations of tones, namely high, low, rising, and falling. A tone marker is used to represent each of the tones, which is added to the vowels to indicate tone variation. Addressing tonal words in machine translation for such a low-resource pair is another challenging issue. In this paper, the English-Mizo corpus is developed where parallel sentences having tonal words are incorporated. The different machine translation models are explored based on statistical machine translation and neural machine translation for the baseline systems. Furthermore, the proposed approach attempts to augment the train data by expanding parallel data having tonal words and achieves state-of-the-art results for both forward

Author Biographies

Vanlalmuansangi Khenglawt, Mizoram University

Department of Computer Engineering

Sahinur Rahman Laskar, National Institute of Technology

Department of Computer Science and Engineering

Partha Pakray, National Institute of Technology

Department of Computer Science and Engineering

Riyanka Manna, Gandhi Institute of Technology and Management

Department of Computer Science and Engineering

Ajoy Kumar Khan, Mizoram University

Department of Computer Engineering

Downloads

Published

2022-08-31

Issue

Section

Articles