Skip to contentScholarGate
LibraryBookshelfDeskReview StudioAssistant
Sign in
On this page
IntuitionHow it worksWhen to use itStrengths & limitationsCommon pitfallsApplicationsFrequently asked🔒 Read the full methodSourcesRelated methods
Cite this pageSpotted an issue on this page? Report or suggest a fix →
Home›Decision-making›Bradley-Terry Model
Regression modelRanking models

Bradley-Terry Model

Bradley-Terry Model for Paired Comparisons · Also known as: BT Model, Bradley-Terry-Luce Model, Paired Comparison Model, İkili Karşılaştırma Modeli

The Bradley-Terry model is a probabilistic model for paired comparisons that assigns a latent strength parameter to each item and predicts the probability that one item beats another in a head-to-head contest. Introduced by Ralph A. Bradley and Milton E. Terry in 1952, it provides a principled statistical framework for ranking items from pairwise preference data, including incomplete comparison designs where not every pair is directly observed.

ScholarGate
  1. Regression model
  2. v1
  3. 1 Sources
  4. PUBLISHED
Cite this page →
Tools & resources
Download slides
Learn & explore

Read the full method

Members only

Sign in with a free account to read this section.

Sign in

Method map

The neighbourhood of related methods — select a node to explore.

Bradley-Terry Model
Elo RatingLogistic RegressionPlackett-Luce ModelRank AggregationThurstone ScalingTrueSkillUnfolding Model

When to use it

Use the Bradley-Terry model when your data consist of pairwise comparisons or head-to-head contest outcomes and your goal is to rank items or estimate relative preference strengths. It is appropriate for tournaments, consumer preference studies, sports analytics, and crowdsourced ranking tasks. Key assumptions include transitivity of preferences, independence of comparisons, and no ties (or tie-handling extensions). Avoid it when comparisons are highly intransitive, when context effects invalidate independence, or when ordinal rankings without a probabilistic model suffice.

Strengths & limitations

Strengths
  • Provides a principled probabilistic foundation for ranking from pairwise data, yielding interpretable strength estimates with standard errors.
  • Handles incomplete comparison designs gracefully, requiring only that the comparison graph be connected rather than fully balanced.
  • Scales to large item sets through efficient maximum likelihood algorithms and has well-developed extensions for ties, home advantage, and covariates.
  • Produces a unique global ranking consistent with the law of comparative judgment, avoiding the cycles and contradictions of naive win-rate rankings.
Limitations
  • Assumes strict transitivity of preferences, which may not hold in real-world scenarios involving intransitive cycles or context-dependent choices.
  • Does not accommodate ties natively in its basic formulation; extensions such as the Davidson model are required for tied outcomes.
  • Requires a connected comparison graph for identifiability; isolated items or disconnected groups cannot be jointly ranked.
  • Maximum likelihood estimates are undefined when some items win or lose all comparisons, necessitating regularization or Bayesian priors.

Frequently asked

How does the Bradley-Terry model differ from the Elo rating system?

Both models are grounded in the same paired comparison probability formula, but they differ in estimation approach. Elo updates ratings incrementally after each match using a fixed K-factor, making it well suited for online or sequential settings. Bradley-Terry uses batch maximum likelihood estimation over all observed outcomes simultaneously, yielding statistically efficient estimates with uncertainty quantification but requiring all data to be available at once.

Can the Bradley-Terry model handle more than two items in a single comparison?

The basic Bradley-Terry model is defined for pairwise comparisons only. Its generalization to rankings of three or more items simultaneously is the Plackett-Luce model, which decomposes a full ranking into a sequence of paired choices. If your data consist of complete or partial orderings rather than binary win-loss outcomes, the Plackett-Luce model is the appropriate extension.

What sample size is needed for reliable Bradley-Terry estimates?

Reliability depends on the number of comparisons per pair rather than on total item count alone. As a rule of thumb, at least five to ten comparisons per pair produce reasonably stable estimates. With fewer comparisons the likelihood surface becomes flat, yielding wide confidence intervals. Bayesian regularization with weakly informative priors on the strength parameters can improve stability when comparison data are sparse.

Sources

  1. Bradley, R. A., & Terry, M. E. (1952). Rank analysis of incomplete block designs: I. The method of paired comparisons. Biometrika, 39(3/4), 324–345. DOI: 10.2307/2334029 ↗

How to cite this page

ScholarGate. (2026, June 2). Bradley-Terry Model for Paired Comparisons. ScholarGate. https://scholargate.app/en/decision-making/bradley-terry-model

Related methods

Elo RatingLogistic RegressionPlackett-Luce Model

Which method?

Set this method beside its closest kin and read them side by side — the library lays the books on the table; the choice is yours.

  • Elo RatingDecision-making↔ compare
  • Logistic RegressionResearch Statistics↔ compare
  • Plackett-Luce ModelDecision-making↔ compare
Compare side by side →

Referenced by

Elo RatingPlackett-Luce ModelRank AggregationThurstone ScalingTrueSkillUnfolding Model

Similar methods

Plackett-Luce ModelElo RatingThurstone ScalingRank AggregationTrueSkillOrdered LogitPaired Comparison MethodBayesian Logistic Regression

Related reference concepts

Learning to RankItem Response TheoryProbabilistic Retrieval ModelsBayesian Inference FoundationsMaximum Likelihood EstimationBayesian Model Comparison and Selection

Spotted an issue on this page? Report or suggest a fix →

ScholarGate — Bradley-Terry Model (Bradley-Terry Model for Paired Comparisons). Retrieved 2026-07-21 from https://scholargate.app/en/decision-making/bradley-terry-model · Dataset: https://doi.org/10.5281/zenodo.20539026
Quick facts
Originator
Ralph Bradley & Milton Terry
Year
1952
Type
Probabilistic paired comparison model
Subfamily
Ranking models
Estimation
Maximum likelihood estimation
Output
Latent strength parameters and win probabilities
Related methods
Elo RatingLogistic RegressionPlackett-Luce Model
ScholarGate

A content-first reference library for research methods — what each one is, how it works, and where it comes from.

Open data (CC-BY)

Explore

  • Library
  • Search the library…
  • Browse by field
  • Fields
  • Journey
  • Compare
  • Which method?

Reference

  • Subjects
  • Atlas
  • Glossary
  • Methodology
  • Philosophy

Your tools

  • Bookshelf
  • Desk
  • Chat

Company

  • About
  • Pricing
  • Contact
  • Suggest a method

Entries are compiled from published sources for reference. Verifying the accuracy and suitability of any information for your own use remains your responsibility.

© 2026 ScholarGate · A research-method reference library
  • Privacy
  • Cookies
  • Terms
  • Delete account