Skip to contentScholarGate
LibraryBookshelfDeskReview StudioAssistant
Sign in
On this page
IntuitionHow it worksWhen to use itStrengths & limitationsCommon pitfallsApplicationsFrequently asked🔒 Read the full methodSourcesRelated methods
Cite this pageSpotted an issue on this page? Report or suggest a fix →
Home›Model Evaluation›Goodness-of-Fit Testing
MCDMStatistical testing

Goodness-of-Fit Testing

Goodness-of-Fit Testing Framework · Also known as: goodness of fit test, GOF test, model fit assessment

Goodness-of-fit (GOF) testing is a framework for assessing whether observed data are consistent with a hypothesized probability distribution or model. Originating from Karl Pearson's chi-square test (1900), GOF tests quantify the discrepancy between data and model predictions, yielding p-values to judge whether observed deviations are statistically significant or due to random chance.

ScholarGate
  1. MCDM
  2. v1
  3. 3 Sources
  4. PUBLISHED
Cite this page →
Tools & resources
Download slides
Learn & explore

Read the full method

Members only

Sign in with a free account to read this section.

Sign in

Method map

The neighbourhood of related methods — select a node to explore.

Goodness-of-Fit
Akaike Information Crite…Bayesian Information Cri…Mean Squared ErrorR-squared

When to use it

Use GOF tests to assess whether a fitted model is statistically adequate. Common applications: testing if data are normally distributed before applying parametric tests, or checking if a regression model's residuals are appropriately distributed. GOF tests are particularly useful for model criticism and diagnostic checking. However, large samples cause tests to reject even trivial deviations; visual inspection (Q-Q plots, histograms) is essential alongside formal tests.

Strengths & limitations

Strengths
  • Hypothesis test framework: provides formal statistical assessment
  • Objective decision rule: reject/accept based on p-value threshold
  • Multiple test variants: chi-square, Kolmogorov-Smirnov, Anderson-Darling, Shapiro-Wilk, etc.
  • Foundation for model criticism and diagnostic procedures
Limitations
  • Power depends on sample size; large samples reject even minor deviations, small samples fail to detect real misfit
  • Assumes a completely specified null model; incorrect specification leads to incorrect p-values
  • P-value interpretation: p > 0.05 does not prove the model is correct, only that data are consistent with it
  • Does not measure magnitude of deviation; two very different models can have identical p-values

Frequently asked

What does p > 0.05 really mean in a GOF test?

It means: if the model is true, you would observe data this extreme or more extreme only 5% of the time by random chance. It does NOT mean the model is correct; it only means data are consistent with the model. Many models might fit equally well.

Why do large samples reject models that seem to fit well?

Test power increases with sample size. Large samples detect tiny deviations; a model may be practically adequate but statistically significantly different. Always combine GOF tests with visual inspection (Q-Q plots, residual plots).

Which GOF test should I use?

Depends on context: chi-square for discrete/categorical data, Kolmogorov-Smirnov or Anderson-Darling for continuous, Shapiro-Wilk for normality. Check your specific question and data type; consult software documentation.

Can I use multiple GOF tests on the same data?

Avoid multiple hypothesis testing on the same data; you inflate false positive risk. If testing multiple hypotheses, apply correction methods (Bonferroni, Holm). Better to choose one pre-specified test.

Does GOF testing replace visual inspection?

No. Combine both: GOF tests provide formal hypothesis testing, but residual plots, Q-Q plots, and histograms reveal the nature and magnitude of deviations. Together they give complete diagnostic insight.

Sources

  1. Pearson, K. (1900). On the criterion that a given system of deviations from the probable in the case of a correlated system of variables is such that it can be reasonably supposed to have arisen from random sampling. Philosophical Magazine, 50(302), 157-175. DOI: 10.1080/14786440009463897 ↗
  2. Cramér, H. (1928). On the composition of elementary errors. Skandinavisk Aktuarietidskrift, 11, 141-180. link ↗
  3. Kolmogorov, A. N. (1933). Sulla determinazione empirica di una legge di distribuzione. Giornale dell'Istituto Italiano degli Attuari, 4, 83-91. link ↗

How to cite this page

ScholarGate. (2026, June 3). Goodness-of-Fit Testing Framework. ScholarGate. https://scholargate.app/en/model-evaluation/goodness-of-fit

Related methods

Akaike Information CriterionBayesian Information CriterionMean Squared ErrorR-squared

Which method?

Set this method beside its closest kin and read them side by side — the library lays the books on the table; the choice is yours.

  • Akaike Information CriterionModel Evaluation↔ compare
  • Bayesian Information CriterionModel Evaluation↔ compare
  • Mean Squared ErrorModel Evaluation↔ compare
  • R-squaredModel Evaluation↔ compare
Compare side by side →

Similar methods

Anderson-Darling TestChi-square goodness-of-fit testKolmogorov-Smirnov TestLilliefors TestShapiro-Wilk testP-Value and Statistical SignificanceNull Hypothesis TestingChi-square test

Related reference concepts

Goodness of FitStatistical Hypothesis TestingChi-Squared and Fisher Exact TestsHypothesis Testing FrameworkLikelihood-Ratio TestsData Distribution and Normality

Spotted an issue on this page? Report or suggest a fix →

ScholarGate — Goodness-of-Fit (Goodness-of-Fit Testing Framework). Retrieved 2026-07-21 from https://scholargate.app/en/model-evaluation/goodness-of-fit · Dataset: https://doi.org/10.5281/zenodo.20539026
Quick facts
Originator
Karl Pearson
Subfamily
Statistical testing
Year
1900
Type
Hypothesis testing framework for model adequacy
Related methods
Akaike Information CriterionBayesian Information CriterionMean Squared ErrorR-squared
ScholarGate

A content-first reference library for research methods — what each one is, how it works, and where it comes from.

Open data (CC-BY)

Explore

  • Library
  • Search the library…
  • Browse by field
  • Fields
  • Journey
  • Compare
  • Which method?

Reference

  • Subjects
  • Atlas
  • Glossary
  • Methodology
  • Philosophy

Your tools

  • Bookshelf
  • Desk
  • Chat

Company

  • About
  • Pricing
  • Contact
  • Suggest a method

Entries are compiled from published sources for reference. Verifying the accuracy and suitability of any information for your own use remains your responsibility.

© 2026 ScholarGate · A research-method reference library
  • Privacy
  • Cookies
  • Terms
  • Delete account