Measuring diachronic language distance using perplexity. Application to English, Portuguese and Spanish.

The objective of this work is to set a corpus-driven methodology to quantify automatically diachronic language distance
between chronological periods of several languages. We apply a perplexity-based measure to written text representing
different historical periods of three languages: European English, European Portuguese and European Spanish. For this
purpose, we have built historical corpora for each period, which have been compiled from different open corpus sources


RSS - Artikulua-rako harpidetza egin