SciELO - Scientific Electronic Library Online

 
vol.26 número3Development of a Normalized Hadith Narrator Encyclopedia with TEIMetaphor Interpretation Using Word Embeddings índice de autoresíndice de materiabúsqueda de artículos
Home Pagelista alfabética de revistas  

Servicios Personalizados

Revista

Articulo

Indicadores

Links relacionados

  • No hay artículos similaresSimilares en SciELO

Compartir


Computación y Sistemas

versión On-line ISSN 2007-9737versión impresa ISSN 1405-5546

Resumen

DOMOTOR, Andrea; KAKONYI, Tibor  y  YANG, Zijian Győző. What’s Your Style?Automatic Genre Identification with Neural Network. Comp. y Sist. [online]. 2022, vol.26, n.3, pp.1293-1299.  Epub 02-Dic-2022. ISSN 2007-9737.  https://doi.org/10.13053/cys-26-3-4350.

Genre identification is an important task in natural language processing that can be useful for many practical and research purposes. The challenge of this task is that genre is not a homogeneous and unequivocal property of the texts and it is often hard to separate from the topic. In this paper we compare the performance of two different automatic genre identification methods. We classified six text types: literary, academic, legal, press, spoken and personal. In one part of our research we did experiments with traditional machine learning methods using linguistic, n-gram and error features. In the other part we tested the same task with a word embedding based neural network. In this part we did experiments with different training data (words only, POS-tags only, words and POS-tags etc.). Our results revealed that neural network is a suitable method for this task while traditional machine learning showed significantly lower performance. We gained high (around 70%) accuracy with our word embedding based method. The results of the different text categories seemed to depend on the stylistic properties of the studied genres.

Palabras llave : Genre identification; text classification; machine learning; neural networks; word embedding; stylistics.

        · texto en Inglés     · Inglés ( pdf )