Glottochronology
Lexicostatistical Dating Method · Also known as: Lexicostatistics, Glottochronological Dating
Glottochronology, or lexicostatistics, is a quantitative method in historical linguistics that estimates the time of divergence between related languages based on the proportion of shared cognates in their basic vocabularies. Developed by Morris Swadesh in 1950, the method assumes that core vocabulary items change at a relatively constant rate over time, allowing linguists to calculate a 'time depth'—how long ago two languages shared a common ancestor. Though controversial due to its restrictive assumptions, glottochronology provides rough temporal estimates when archaeological or written records are unavailable.
Read the full method
Sign in with a free account to read this section.
Method map
The neighbourhood of related methods — select a node to explore.
When to use it
Use glottochronology to estimate the relative time depth of language divergence when archaeological or documentary evidence is absent or unreliable. It is most useful for establishing rough chronologies of prehistoric language families, for testing competing hypotheses about divergence dates, and for identifying periods of rapid versus gradual change. However, expect significant margins of error and consider it a supplementary tool to be checked against independent evidence like archaeology or written records.
Strengths & limitations
- Provides quantitative time estimates for language divergence when other evidence is sparse, useful for prehistoric periods.
- Simple in concept and implementation, making it accessible for educational purposes and quick comparative studies.
- Can generate testable hypotheses that can be cross-checked against archaeological or paleontological timelines.
- Has demonstrated rough accuracy in some cases (e.g., dating of Romance languages, some Austronesian splits).
- Assumes a constant rate of vocabulary replacement, which is violated in real languages due to borrowing, social disruption, and differential innovation rates.
- Works poorly for very distant time depths (beyond 6000-7000 years) where the amount of shared vocabulary becomes too small for reliable estimation.
- Sensitive to which core vocabulary list is used; different lists can yield different results, and the choice of 'core' vocabulary is somewhat arbitrary.
- Does not account for variable change rates in different populations or for the effects of language contact and creolization.
Frequently asked
Why do glottochronological dates often not match historical records?
Several factors can skew results: borrowing inflates the cognate count (suggesting recent divergence when divergence is actually older), uneven rates of change, and loss of entire semantic domains due to cultural shift. Additionally, the assumed retention rate (0.81-0.86 per millennium) is an average; real languages vary. Always compare glottochronological estimates to archaeology, written records, and comparative linguistic evidence.
What is the Swadesh list, and why is it important?
The Swadesh list is a standardized set of 100 or 200 basic vocabulary items (like 'I,' 'you,' 'water,' 'fire,' 'earth') chosen because they are thought to be relatively resistant to replacement and borrowing. The idea is that core vocabulary changes slower than specialized vocabulary, so using a standard list makes comparisons across language families more uniform. However, which items count as 'core' varies by language, and some items are more stable in some cultures than others.
Can glottochronology handle three or more languages at once?
Yes, but the calculations become more complex. You can compute pairwise cognate percentages for all language pairs, then use phylogenetic methods (like neighbor-joining or maximum parsimony) to construct a tree showing the probable order of divergences. Modern computational approaches treat this as a Bayesian phylogenetic problem and account for uncertainty in the data.
What confidence interval should I use for a glottochronological date?
The method itself does not provide a formal confidence interval; you must estimate it from the quality of data and external checks. A rough rule of thumb is ±10-20% of the calculated date for recent divergences (< 2000 years) and wider margins (±30-50%) for deeper time. Always support glottochronological dates with other evidence before publishing them as established fact.
Sources
- Swadesh, M. (1950). Salish internal relationships. International Journal of American Linguistics, 16(3), 157-167. DOI: 10.1086/464084 ↗
- Swadesh, M. (1955). Towards greater accuracy in lexicostatistic dating. International Journal of American Linguistics, 21(2), 121-137. DOI: 10.1086/464321 ↗
- Embleton, S. M. (1986). Statistics in Historical Linguistics. Bochum: Brockmeyer. link ↗
How to cite this page
ScholarGate. (2026, June 3). Lexicostatistical Dating Method. ScholarGate. https://scholargate.app/en/linguistics/glottochronology
Which method?
Set this method beside its closest kin and read them side by side — the library lays the books on the table; the choice is yours.
- Comparative MethodLinguistics↔ compare
- Internal ReconstructionLinguistics↔ compare