Word Formation Is Aware of Morpheme Family Size
Please always quote using this URN: urn:nbn:de:bvb:20-opus-112848
- Words are built from smaller meaning bearing parts, called morphemes. As one word can contain multiple morphemes, one morpheme can be present in different words. The number of distinct words a morpheme can be found in is its family size. Here we used Birth-Death-Innovation Models (BDIMs) to analyze the distribution of morpheme family sizes in English and German vocabulary over the last 200 years. Rather than just fitting to a probability distribution, these mechanistic models allow for the direct interpretation of identified parameters. DespiteWords are built from smaller meaning bearing parts, called morphemes. As one word can contain multiple morphemes, one morpheme can be present in different words. The number of distinct words a morpheme can be found in is its family size. Here we used Birth-Death-Innovation Models (BDIMs) to analyze the distribution of morpheme family sizes in English and German vocabulary over the last 200 years. Rather than just fitting to a probability distribution, these mechanistic models allow for the direct interpretation of identified parameters. Despite the complexity of language change, we indeed found that a specific variant of this pure stochastic model, the second order linear balanced BDIM, significantly fitted the observed distributions. In this model, birth and death rates are increased for smaller morpheme families. This finding indicates an influence of morpheme family sizes on vocabulary changes. This could be an effect of word formation, perception or both. On a more general level, we give an example on how mechanistic models can enable the identification of statistical trends in language change usually hidden by cultural influences.…
Author: | Daniela Barbara Keller, Jörg Schultz |
---|---|
URN: | urn:nbn:de:bvb:20-opus-112848 |
Document Type: | Journal article |
Faculties: | Fakultät für Biologie / Theodor-Boveri-Institut für Biowissenschaften |
Language: | English |
Year of Completion: | 2014 |
Source: | PLoS ONE 9(4): e93978. doi:10.1371/journal.pone.0093978 |
DOI: | https://doi.org/10.1371/journal.pone.0093978 |
Dewey Decimal Classification: | 5 Naturwissenschaften und Mathematik / 57 Biowissenschaften; Biologie / 570 Biowissenschaften; Biologie |
Tag: | birth rates; chi square tests; culture; death rates; language; linguistic morphology; psycholinguistics; vocabulary |
Release Date: | 2015/05/15 |
Collections: | Open-Access-Publikationsfonds / Förderzeitraum 2014 |
Licence (German): | CC BY: Creative-Commons-Lizenz: Namensnennung |