Skip to contentScholarGate
LibraryBookshelfDeskReview StudioAssistant
Sign in
On this page
IntuitionHow it worksWhen to use itStrengths & limitationsCommon pitfallsApplicationsFrequently asked🔒 Read the full methodSourcesRelated methods
Cite this pageSpotted an issue on this page? Report or suggest a fix →
Home›Qualitative›Content Analysis — Systematic Coding of Text and Media
Process / pipeline

Content Analysis — Systematic Coding of Text and Media

Content Analysis · Also known as: İçerik Analizi, systematic content coding, quantitative content analysis

Content analysis is a systematic research technique for reducing text, visual, or media material into coded categories so that patterns can be counted, compared, and interpreted. Formalised by Klaus Krippendorff in his widely cited methodology textbook (latest edition 2018), the method sits at the boundary of qualitative and quantitative inquiry: it imposes structured, replicable coding on inherently meaning-laden material.

ScholarGate
  1. Process / pipeline
  2. v1
  3. 1 Sources
  4. PUBLISHED
Cite this page →
Tools & resources
Download slides
Learn & explore

Read the full method

Members only

Sign in with a free account to read this section.

Sign in

Method map

The neighbourhood of related methods — select a node to explore.

Content Analysis
Discourse AnalysisGrounded TheorySentiment AnalysisText ClassificationThematic AnalysisAxial CodingCase StudyComparative Content anal…Comparative Typological…Constant Comparative Met…

+72 more

When to use it

Content analysis is appropriate when research questions concern the presence, frequency, or pattern of specific features in text or media, and when systematic, replicable coding across a corpus is required. It suits descriptive and exploratory purposes across health, social science, education, communication, and business research. The minimum viable corpus is approximately ten documents; below that, reliable theme production is not possible and thematic analysis of a smaller, richer dataset is preferable. The method does not require normality assumptions and works on purely textual variables.

Strengths & limitations

Strengths
  • Provides a transparent, auditable, and replicable process for analysing large bodies of text or media.
  • Intercoder reliability statistics give a quantitative quality check that is absent from purely interpretive methods.
  • Bridges qualitative richness and quantitative summation — categories can be counted, compared, and entered into further statistical analyses.
  • Applicable across domains and content types: news articles, interview transcripts, social media posts, policy documents, images, and broadcasts.
Limitations
  • Developing a reliable codebook is labour-intensive; it typically requires multiple pilot rounds and revisions.
  • Reliability above 0.70 Kappa is required — with poorly defined categories or complex content, achieving this threshold is difficult.
  • The method captures what is explicitly present in content but is less suited to latent meaning, irony, or context that lies outside the text.
  • Corpus must contain at least ten documents to sustain meaningful category frequencies; very small corpora should use thematic analysis instead.

Frequently asked

What is the difference between content analysis and thematic analysis?

Content analysis emphasises systematic, rule-based coding that can be checked for intercoder reliability and produces quantifiable frequencies. Thematic analysis is more interpretive, seeks latent patterns across a smaller dataset, and does not require multiple coders or a formal reliability check. When you have a large corpus and need replicable counts, content analysis is the right choice; when you have a smaller, richer dataset and want to explore meaning in depth, thematic analysis is preferable.

How do I know my coding is reliable enough?

Compute Cohen's Kappa or Krippendorff's Alpha between at least two independent coders after each pilot round. A Kappa above 0.70 is the conventional minimum for acceptability. If reliability falls below this threshold, revise the codebook — typically by adding decision rules and exemplars for the categories where coders disagreed — and recode before accepting results.

Can content analysis be done by a single researcher?

A single coder can apply the codebook, but without a second independent coder there is no way to compute intercoder reliability. Single-coder studies are therefore less credible in peer review. Where a second human coder is unavailable, some researchers use automated coding (rule-based or machine-learning) as a comparison point and report agreement with the automated system.

How large does my corpus need to be?

The minimum for meaningful content analysis is approximately ten documents. Below that, frequency-based findings are unstable and a thematic analysis of the available material is more appropriate. There is no hard upper limit — large corpora can be sampled — but every document in the sample must be coded according to the codebook.

Sources

  1. Krippendorff, K. (2018). Content Analysis: An Introduction to Its Methodology (4th ed.). Sage. ISBN: 978-1506395661

How to cite this page

ScholarGate. (2026, June 1). Content Analysis. ScholarGate. https://scholargate.app/en/qualitative/content-analysis

Related methods

Discourse AnalysisGrounded TheorySentiment AnalysisText ClassificationThematic Analysis

Which method?

Set this method beside its closest kin and read them side by side — the library lays the books on the table; the choice is yours.

  • Discourse AnalysisQualitative Research↔ compare
  • Grounded TheoryQualitative Research↔ compare
  • Sentiment AnalysisText mining↔ compare
  • Text ClassificationText mining↔ compare
  • Thematic AnalysisQualitative Research↔ compare
Compare side by side →

Referenced by

Axial CodingCase StudyComparative Content analysisComparative Typological AnalysisConstant Comparative MethodContent Analysis of TreatiesConversation AnalysisCritical Discourse AnalysisCultivation AnalysisCurriculum AnalysisDelphi MethodDelphi TechniqueDigital Case StudyDigital Content analysisDigital Critical Discourse AnalysisDigital Document AnalysisDigital EthnographyDigital Hermeneutic AnalysisDigital Reflexive Thematic AnalysisDigital Semiotic AnalysisDigital Visual AnalysisDiscourse Completion TaskDocument CollectionDocument-based Curriculum AnalysisDocument-based Program EvaluationFeminist Content AnalysisField-based Content AnalysisField-based Document AnalysisField-based Metaphor AnalysisField-based Semiotic AnalysisFoucauldian Discourse AnalysisFramework AnalysisFraming AnalysisHistorical Archival ResearchIn Vivo CodingIntercoder ReliabilityKrippendorff's AlphaLearning Analytics MethodLegal Content AnalysisLIWC Text AnalysisLongitudinal Content AnalysisLongitudinal Discourse AnalysisLongitudinal Historical Archival ResearchLongitudinal Metaphor AnalysisLongitudinal Visual AnalysisLongitudinal Web ScrapingManifest Content AnalysisMedia Richness AnalysisMetaphor AnalysisMixed Methods ResearchMulti-source Document CollectionMultiple Case-Based Metaphor AnalysisMultiple Case-Based Semiotic AnalysisMultiple Case-Based Visual AnalysisNarrative AnalysisNetnographyNominal Group TechniqueOnline Document CollectionOpen CodingOpportunity to Learn AnalysisParticipatory Content AnalysisPostcolonial AnalysisQ-Sort in CommunicationReflexive Thematic AnalysisSelective CodingSemantic Network AnalysisSemiotic AnalysisSentiment Analysis in CommunicationSequential Mixed Methods MatrixTextual CriticismTriangulated Document CollectionTypological AnalysisVisual analysisVisual elicitation content analysisVisual Elicitation Document AnalysisVisual Elicitation NetnographyWeb Scraping

Similar methods

Quantitative Content AnalysisQualitative Content AnalysisManifest Content AnalysisIntercoder ReliabilityInterpretive content analysisComparative Content analysisComparative Quantitative Content AnalysisDigital Content analysis

Related reference concepts

Qualitative Research MethodsMeasurement Validity and ReliabilityContent ValidityText ClusteringCategorical Data AnalysisUser Research Methods

Spotted an issue on this page? Report or suggest a fix →

ScholarGate — Content Analysis (Content Analysis). Retrieved 2026-07-21 from https://scholargate.app/en/qualitative/content-analysis · Dataset: https://doi.org/10.5281/zenodo.20539026
Quick facts
Originator
Klaus Krippendorff (systematic formulation); roots in early 20th-century communications research
Year
Systematised through Krippendorff's methodology work; 4th edition 2018
Type
Qualitative / mixed-method research technique
DataType
Text, visual, or media content
Output
Coded categories with frequencies or themes
ReliabilityThreshold
Cohen's Kappa > 0.70
MinimumSample
10
Difficulty
2
Related methods
Discourse AnalysisGrounded TheorySentiment AnalysisText ClassificationThematic Analysis
ScholarGate

A content-first reference library for research methods — what each one is, how it works, and where it comes from.

Open data (CC-BY)

Explore

  • Library
  • Search the library…
  • Browse by field
  • Fields
  • Journey
  • Compare
  • Which method?

Reference

  • Subjects
  • Atlas
  • Glossary
  • Methodology
  • Philosophy

Your tools

  • Bookshelf
  • Desk
  • Chat

Company

  • About
  • Pricing
  • Contact
  • Suggest a method

Entries are compiled from published sources for reference. Verifying the accuracy and suitability of any information for your own use remains your responsibility.

© 2026 ScholarGate · A research-method reference library
  • Privacy
  • Cookies
  • Terms
  • Delete account