paper

Language discrimination and clustering via a neural network approach

arXiv:1507.04116

Abstract

We classify twenty-one Indo-European languages starting from written text. We use neural networks in order to define a distance among different languages, construct a dendrogram and analyze the ultrametric structure that emerges. Four or five subgroups of languages are identified, according to the "cut" of the dendrogram, drawn with an entropic criterion. The results and the method are discussed.

10 pages, 12 figures

References in corpus (1)

Language discrimination and clustering via a neural network approach · wovepaper