Summative Evaluation
Also known as: Outcome Judgement Evaluation, Accountability Evaluation
Summative evaluation is evaluation conducted to render an overall judgement of a program, policy or product — its merit, worth, effectiveness or impact — typically after it has been implemented or has matured. Named by Michael Scriven in his 1967 essay 'The Methodology of Evaluation' as the counterpart to formative evaluation, its purpose is to inform consequential decisions: whether to continue, expand, replicate, defund or certify an intervention. It addresses the bottom-line question 'did it work, and was it worth it?' for audiences such as funders, policymakers and the public.
Key highlights
- Provides a defensible, bottom-line judgement to inform high-stakes decisions and accountability.
- Renders findings against explicit, justified criteria of merit and worth.
- Supports comparison across programs and over time when consistent criteria are used.
- Can incorporate rigorous impact designs to substantiate causal claims of effectiveness.
Intuition
This section is available to Pro members. Upgrade to Pro
How it works
This section is available to Pro members. Upgrade to Pro
When to use it
Use summative evaluation when a consequential judgement is required about a completed or mature program — decisions on continuation, expansion, replication, funding or certification. It is appropriate when the intervention has stabilised enough to be fairly judged and when stakeholders need an authoritative conclusion rather than improvement feedback. It assumes meaningful outcome data can be obtained and that defensible criteria of success exist. It is the wrong choice during early development, when the program is still being shaped (use formative evaluation) or when it is a genuinely emergent innovation (use developmental evaluation). For credible causal claims about impact, a summative evaluation should be built on a sound impact-evaluation design such as a counterfactual or quasi-experimental approach.
Strengths & limitations
- Provides a defensible, bottom-line judgement to inform high-stakes decisions and accountability.
- Renders findings against explicit, justified criteria of merit and worth.
- Supports comparison across programs and over time when consistent criteria are used.
- Can incorporate rigorous impact designs to substantiate causal claims of effectiveness.
- Arrives too late to improve the program it judges, offering little to current implementers.
- A single overall verdict can obscure variation in who the program worked for and why.
- Credible impact claims often require costly experimental or quasi-experimental designs that may be infeasible.
- The judgement is only as sound as the explicitness and legitimacy of the value criteria chosen.
Common pitfalls
This section is available to Pro members. Upgrade to Pro
Applications
This section is available to Pro members. Upgrade to Pro
Frequently asked
How does summative evaluation differ from formative evaluation?
Formative evaluation improves a program during development and reports timely feedback to implementers; summative evaluation judges a completed or mature program and reports a verdict to decision-makers. Scriven introduced both terms in 1967. The difference is purpose, not method: the same data can serve either role, but summative work emphasises independence, defensible criteria and conclusions about merit and worth rather than suggestions for improvement.
Is summative evaluation the same as impact evaluation?
Not quite. Impact evaluation specifically estimates the causal effect of an intervention, usually against a counterfactual. Summative evaluation is broader: it judges overall merit and worth, which may include impact but also cost, quality, sustainability and value against alternatives. A rigorous summative evaluation of effectiveness will often incorporate an impact-evaluation design, but summative judgement encompasses more than the impact estimate alone.
Why does Scriven insist evaluators take a stance on values?
Scriven argued that evaluation is inherently about determining merit and worth, which cannot be done without value criteria. Pretending to be value-free, he held, simply hides the values being used. A credible summative judgement therefore requires making the standards of merit explicit and justifying them, so that the verdict rests on defensible, transparent criteria rather than the evaluator's concealed preferences. This insistence shaped the field's understanding of evaluation as a value-laden but disciplined activity.
Sources
- 1.Scriven, M. (1967). The methodology of evaluation. In R. W. Tyler, R. M. Gagné, & M. Scriven (Eds.), Perspectives of Curriculum Evaluation (pp. 39–83). Chicago: Rand McNally.ISBN 9780528616600
You have read it. What now?
Cite this page
ScholarGate. (2026, June 22). Summative Evaluation. ScholarGate. https://scholargate.app/public-policy/summative-evaluation