Skip to contentScholarGate
LibraryBookshelfDeskReview StudioAssistant
Sign in
On this page
IntuitionHow it worksWhen to use itStrengths & limitationsCommon pitfallsApplicationsFrequently asked🔒 Read the full methodSourcesRelated methods
Cite this pageSpotted an issue on this page? Report or suggest a fix →
Home›Statistics›Correspondence Analysis
Latent structureDimensionality reduction

Correspondence Analysis

Also known as: CA, Simple Correspondence Analysis, Reciprocal Averaging, Karşılıklı Uyum Analizi

Correspondence Analysis (CA) is an exploratory multivariate technique for visualizing the association structure of a two-way contingency table. Developed systematically by Jean-Paul Benzécri in France during the 1960s–1970s and brought to an English-language audience by Michael Greenacre in 1984, CA decomposes the chi-square statistic of a cross-tabulation to produce a low-dimensional joint display — called a biplot — in which rows and columns are represented as points whose proximities reflect their associations.

ScholarGate
  1. Latent structure
  2. v1
  3. 1 Sources
  4. PUBLISHED
Cite this page →
Tools & resources
Download slides
Learn & explore

Read the full method

Members only

Sign in with a free account to read this section.

Sign in

Method map

The neighbourhood of related methods — select a node to explore.

Correspondence Analysis
BiplotMultiple Correspondence…Bayesian Multiple Corres…Multidimensional ScalingPerceptual and Preferenc…Robust Correspondence An…Robust Multiple Correspo…Thurstone ScalingUnfolding Model

When to use it

Use Correspondence Analysis when you have a two-way contingency table of count data and wish to explore associations between categorical row and column variables. The method assumes the data are non-negative counts (or frequencies) and that an exploratory, rather than confirmatory, goal is primary. CA is appropriate when categories are nominal or ordinal; for ordinal variables consider non-linear CA variants. It is not suitable for continuous or interval-level data — use Principal Component Analysis instead. When more than two categorical variables are involved, Multiple Correspondence Analysis is the appropriate extension.

Strengths & limitations

Strengths
  • Produces an intuitive joint visual display (biplot) of row and column categories in the same low-dimensional space.
  • Scales naturally to large contingency tables with many categories without requiring distributional assumptions.
  • Total inertia provides a chi-square-based global measure of association, giving statistical grounding to the geometric solution.
  • Handles asymmetric data structures well; both symmetric and asymmetric biplots can be derived from the same decomposition.
Limitations
  • Interpretation relies on proximity between points in a reduced space, which can be misleading if the retained dimensions capture only a small fraction of total inertia.
  • CA is purely exploratory; it does not produce p-values or confidence intervals for specific category pairs without additional resampling procedures.
  • Results are sensitive to rare categories or cells with very small counts, which can disproportionately pull the solution.
  • The method applies only to two-way tables; richer multi-table designs require Multiple Correspondence Analysis or other extensions.

Frequently asked

How is Correspondence Analysis different from Principal Component Analysis?

PCA operates on continuous variables using Euclidean distance and a covariance or correlation matrix. CA operates on count data in a contingency table using chi-square distance between row or column profiles. The two methods share an SVD backbone but differ fundamentally in the data type they accept and the distance metric they optimize.

How many dimensions should I retain in a CA solution?

Retain enough dimensions so that the cumulative inertia (percentage of total chi-square explained) is acceptably high — commonly 70–80% or more. A scree plot of singular values helps identify the elbow beyond which additional dimensions contribute little. The maximum number of meaningful dimensions is min(I-1, J-1) where I and J are the numbers of rows and columns.

Can I apply CA to a table that contains zeros?

Structural zeros (cells that cannot logically contain observations) are problematic because CA relies on log-ratio geometry implicitly and zero cells produce undefined profile coordinates. Sampling zeros (rare but possible events) are generally tolerable if sparse; a common remedy is to merge rare categories or add a small constant, though the latter alters the chi-square structure and should be reported transparently.

Sources

  1. Greenacre, M. J. (1984). Theory and Applications of Correspondence Analysis. Academic Press. ISBN: 978-0-12-299050-2

How to cite this page

ScholarGate. (2026, June 2). Correspondence Analysis. ScholarGate. https://scholargate.app/en/statistics/correspondence-analysis

Related methods

BiplotMultiple Correspondence Analysis

Which method?

Set this method beside its closest kin and read them side by side — the library lays the books on the table; the choice is yours.

  • BiplotStatistics↔ compare
  • Multiple Correspondence AnalysisStatistics↔ compare
Compare side by side →

Referenced by

Bayesian Multiple Correspondence AnalysisBiplotMultidimensional ScalingMultiple Correspondence AnalysisPerceptual and Preference MappingRobust Correspondence AnalysisRobust Multiple Correspondence AnalysisThurstone ScalingUnfolding Model

Similar methods

Multiple Correspondence AnalysisRobust Correspondence AnalysisBayesian Multiple Correspondence AnalysisRobust Multiple Correspondence AnalysisBiplotCross-tabulation analysisMultidimensional ScalingCanonical Correlation Analysis

Related reference concepts

Canonical Correlation AnalysisCategorical Data AnalysisDimension ReductionMultidimensional ScalingChi-Squared and Fisher Exact TestsPrincipal Component Analysis

Spotted an issue on this page? Report or suggest a fix →

ScholarGate — Correspondence Analysis (Correspondence Analysis). Retrieved 2026-07-21 from https://scholargate.app/en/statistics/correspondence-analysis · Dataset: https://doi.org/10.5281/zenodo.20539026
Quick facts
Originator
Jean-Paul Benzécri; Michael Greenacre
Year
1984
Type
Exploratory multivariate technique for categorical data
Subfamily
Dimensionality reduction
Input
Two-way contingency table (cross-tabulation)
Output
Low-dimensional biplot of row and column profiles
Related methods
BiplotMultiple Correspondence Analysis
ScholarGate

A content-first reference library for research methods — what each one is, how it works, and where it comes from.

Open data (CC-BY)

Explore

  • Library
  • Search the library…
  • Browse by field
  • Fields
  • Journey
  • Compare
  • Which method?

Reference

  • Subjects
  • Atlas
  • Glossary
  • Methodology
  • Philosophy

Your tools

  • Bookshelf
  • Desk
  • Chat

Company

  • About
  • Pricing
  • Contact
  • Suggest a method

Entries are compiled from published sources for reference. Verifying the accuracy and suitability of any information for your own use remains your responsibility.

© 2026 ScholarGate · A research-method reference library
  • Privacy
  • Cookies
  • Terms
  • Delete account