SciELO - Scientific Electronic Library Online

 
vol.26 número3Semi-Automatic Alignment of Multilingual Parts of Speech TagsetsTropical Cyclone Simulations with WRF Using High Performance Computing índice de autoresíndice de materiabúsqueda de artículos
Home Pagelista alfabética de revistas  

Servicios Personalizados

Revista

Articulo

Indicadores

Links relacionados

  • No hay artículos similaresSimilares en SciELO

Compartir


Computación y Sistemas

versión On-line ISSN 2007-9737versión impresa ISSN 1405-5546

Resumen

KHENGLAWT, Vanlalmuansangi et al. Machine Translation for Low-Resource English-Mizo Pair Encountering Tonal Words. Comp. y Sist. [online]. 2022, vol.26, n.3, pp.1377-1398.  Epub 02-Dic-2022. ISSN 2007-9737.  https://doi.org/10.13053/cys-26-3-4358.

Machine translation is one of the most powerful natural language processing applications for preserving and upgrading low-resource language. Mizo language is considered as low-resource since there is limited availability of resources. Therefore, it is a challenging task for English-Mizo language pair translation. Moreover, Mizo is a tonal language, where a word can express different meanings depending on a variety of tones. There are four variations of tones, namely high, low, rising, and falling. A tone marker is used to represent each of the tones, which is added to the vowels to indicate tone variation. Addressing tonal words in machine translation for such a low-resource pair is another challenging issue. In this paper, the English-Mizo corpus is developed where parallel sentences having tonal words are incorporated. The different machine translation models are explored based on statistical machine translation and neural machine translation for the baseline systems. Furthermore, the proposed approach attempts to augment the train data by expanding parallel data having tonal words and achieves state-of-the-art results for both forward and backward translations encountering tonal words.

Palabras llave : English-Mizo; machine translation; low-resource; tonal.

        · texto en Inglés     · Inglés ( pdf )