Process / pipelineLinguisticsQuantitative historical linguisticsPipeline

Glottochronology (Lexical Dating)

Also known as: Glottochronology, Lexicostatistic Dating, Linguistic Dating

OriginatorMorris SwadeshYear1952Sources2Related methods7

Glottochronology is Morris Swadesh's method for estimating the time depth at which two related languages separated, derived from the proportion of basic-vocabulary cognates they still share. Building directly on lexicostatistics, it adds a crucial extra assumption — a 'glottoclock' — that basic vocabulary is lost at an approximately constant rate over time, analogous to radioactive decay. Plugging the observed cognate proportion into a logarithmic decay formula yields an estimated separation date in years. The method is historically important but has been heavily criticized, and most historical linguists today treat its dates with great caution.

Key highlights

  • Offers a simple, fully explicit formula that turns observable cognate counts into a concrete time estimate, making the reasoning transparent and reproducible.
  • Historically catalysed the entire enterprise of quantitative and computational dating in linguistics.
  • Provides a first, order-of-magnitude orientation to time depth when no other dating evidence is available.
  • Frames language change as a measurable process, motivating the better-calibrated rate models that followed.

Intuition

This section is available to Pro members. Upgrade to Pro

How it works

This section is available to Pro members. Upgrade to Pro

When to use it

Glottochronology is appropriate, if at all, only as a rough exploratory heuristic when you have reliable cognate counts for clearly related languages and want a crude, explicitly caveated sense of relative time depth. It can be pedagogically useful for illustrating how cognate decay relates to divergence. It is not suitable for producing dates to be cited as established facts, for families with significant borrowing, or for any case where the constant-rate assumption is doubtful — which is most cases. Where absolute dating matters, modern Bayesian phylogenetic methods, which estimate rates from the data rather than assuming a single fixed clock, are strongly preferred, and reconstruction itself should rest on the comparative method.

Strengths & limitations

Strengths
  • Offers a simple, fully explicit formula that turns observable cognate counts into a concrete time estimate, making the reasoning transparent and reproducible.
  • Historically catalysed the entire enterprise of quantitative and computational dating in linguistics.
  • Provides a first, order-of-magnitude orientation to time depth when no other dating evidence is available.
  • Frames language change as a measurable process, motivating the better-calibrated rate models that followed.
Limitations
  • Its foundational assumption — a constant, universal rate of basic-vocabulary replacement — is empirically false; rates vary widely across languages and items.
  • Borrowing and language contact corrupt the cognate count, biasing dates in unpredictable directions.
  • The clock was calibrated on very few language pairs with known histories, so the rate r is poorly grounded.
  • Error margins are large and rarely propagated, yet point dates are often reported as if precise.

Common pitfalls

This section is available to Pro members. Upgrade to Pro

Applications

This section is available to Pro members. Upgrade to Pro

Frequently asked

What is the glottochronology formula and what do its terms mean?

The canonical Swadesh formula is t = log(C) / (2 log r). Here C is the proportion of basic-vocabulary cognates two related languages still share (between 0 and 1), r is the assumed retention rate — the fraction of core vocabulary kept per millennium, conventionally about 0.86 for the 100-word list — and t is the estimated time since the two languages separated, expressed in millennia. The 2 appears because both daughters have been independently losing vocabulary since the split.

Why do most historical linguists distrust glottochronology?

Its central premise is that basic vocabulary is replaced at a single constant rate everywhere and always. Empirically this is not true: replacement rates differ across languages, individual words, and historical situations, and language contact injects loanwords that distort cognate counts. The clock was also calibrated on only a handful of language pairs with documented histories, and the reported dates carry large but usually unstated uncertainty. For these reasons the method is regarded as unreliable for absolute dating.

Has glottochronology been replaced by anything better?

Yes. Modern computational phylogenetic methods, especially Bayesian inference as in the Indo-European studies by Gray and Atkinson and by Bouckaert and colleagues, infer divergence dates while estimating rates from the data and allowing them to vary across the tree, rather than assuming one fixed Swadesh clock. These methods also quantify uncertainty explicitly. They are not free of controversy, but they address the core methodological flaws that discredited classical glottochronology.

Sources

  1. 1.
    Swadesh, M. (1955). Towards greater accuracy in lexicostatistic dating. International Journal of American Linguistics, 21(2), 121–137.
  2. 2.
    Campbell, L. (2013). Historical Linguistics: An Introduction (3rd ed.). Edinburgh University Press.
    ISBN 9780748675593

You have read it. What now?

Cite this page

ScholarGate. (2026, June 22). Glottochronology (Lexical Dating). ScholarGate. https://scholargate.app/linguistics/glottochronology-lexical-dating