Affiliation:
1. Department of Comparative Language Science and Center for the Interdisciplinary Study of Language Evolution (ISLE) , 27217 University of Zurich , Zurich , Switzerland
Abstract
Abstract
Both areal and phylogenetic affiliation have been discussed as driving factors of the distribution of word order in the languages of the world. However, disentangling the interaction of these two factors is challenging. Here we take Indo-European as a test case. Word order in this family is largely homogeneous both within areas and within branches, which makes it difficult to assess which factor was more important in shaping the present-day distribution. To break out of this impasse we turn to corpus data and explicit statistical modeling. Building on a parallel corpus of movie subtitles, we investigate word order on the sentence level under stable pragmatic conditions. We measure the similarity of word order variation between pairs of languages with an information-theoretic distance metric. Using cluster analysis and variation partitioning methods these distance metrics show that phylogenetic distance predicts more variation than geographical distance, but the most important predictor is the shared fraction where phylogeny and area overlap. We conclude that word order has evolved along both dimensions and cannot be reduced to a single one.