Quantifying Linguistic Divergence: Methodological Evolutions and Case Studies in Lexicostatistics
DOI:
https://doi.org/10.61173/wrm65k84Keywords:
Lexicostatistics, Historical linguistics, Statistics, Central limit theorem, Swadesh listAbstract
Lexicostatistics acts as a strong interdisciplinary tool, bringing together the application of statistics and mathematics to language evolution. This paper critically assesses the traditional Swadesh method, highlighting its major steps in the lexicostatistical procedure. It is apparent that, although each step in the traditional method is subject to methodological pitfalls, which greatly impede its accuracy, recent improvements in the method have greatly enhanced its accuracy in glottochronology. From the assessment of the application of lexicostatistics in history, it is apparent that most studies focus on Indo- European languages. This is in line with most linguistic theories, although some aspects in the studies are unique and divergent from conventional theories. This paper, therefore, concludes that lexicostatistics is an emerging field with immense potential for growth, especially through the application of modern technologies like computer intelligence. By extending its scope to other languages, lexicostatistics is likely to give greater and accurate insights into language evolution. This work should be a good guide in the area of Lexicostatistics.
References
[1] Swadesh M. Towards greater accuracy in lexicostatistic dating. International Journal of American Linguistics, 1955.
[2] Piwowarczyk D. Computational approaches to linguistic chronology and subgrouping//Olander T, ed. The Indo-European Language Family. Cambridge: Cambridge University Press, 2022: 33-51.
[3] Serva M. Evolution of the lexicon: A probabilistic point of view. Journal of Statistical Mechanics: Theory and Experiment, 2025(11): 113404.
[4] Zhang M, Gong T. How many is enough?—Statistical principles for lexicostatistics. Frontiers in Psychology, 2016, 7: 1916.
[5] Dyen I, Kruskal J B, Black P. An Indoeuropean classification: A lexicostatistical experiment. Transactions of the American Philosophical Society, 1992.
[6] Kassian A S, Zhivlov M, Starostin G, et al. Rapid radiation of the inner Indo-European languages: An advanced approach to Indo-European lexicostatistics. Linguistics, 2021, 59(4): 949- 979.
[7] Pasquini M, Serva M. Stability of meanings versus rate of replacement of words: An experimental test. Journal of Quantitative Linguistics, 2019.
[8] Ii H I H, Zulfitri Z, Amin T S. Lexicostatistics study of Mandailing and Angkola languages. Jurnal Educatio FKIP UNMA, 2021, 7(1): 265-275.
[9] Peter the Great Museum of Anthropology and Ethnography of the Russian Academy of Sciences, Kassian A. The Dene- Caucasian macrofamily: Lexicostatistical classification and homeland. Etnografia, 2023(3).
[10] Hymes D H. Lexicostatistics so far. Current Anthropology, 1960.
Downloads
Published
Issue
Section
License
Copyright (c) 2026 by the authors.

This work is licensed under a Creative Commons Attribution 4.0 International License.
