In Vivo Coding — Qualitative Coding from Participants' Own Words
In Vivo Coding · Also known as: verbatim coding, literal coding, first-cycle in vivo coding, indigenous coding
In vivo coding is a qualitative first-cycle coding strategy in which the researcher uses the participants' own words or short phrases verbatim as code labels, rather than imposing researcher-generated or theoretical language. The technique preserves the voice, meaning, and conceptual priorities of participants, making it especially valuable in grounded theory, phenomenology, and any study where honouring the emic (insider) perspective is central to analytic integrity.
Read the full method
Sign in with a free account to read this section.
Method map
The neighbourhood of related methods — select a node to explore.
When to use it
In vivo coding is appropriate when the research aim is to understand phenomena from the participants' own conceptual frame — that is, when emic fidelity matters more than imposing a predetermined theoretical structure on the data. It is particularly well suited to grounded theory studies where theoretical sensitivity to participants' categories is a core methodological commitment, and to phenomenological studies where the goal is to let the participants' lived-experience language guide the analysis. It is most effective early in a study or at the first-cycle coding stage, before the researcher moves to more abstract categories. It is less appropriate when the research question requires direct comparison against an established theoretical framework, when data are primarily numerical or visual rather than textual, or when the primary coding purpose is deductive hypothesis testing rather than inductive discovery.
Strengths & limitations
- Preserves the emic (insider) voice and conceptual priorities of participants, reducing the risk of analyst-imposed distortion.
- Produces vivid, evocative codes that are easy to trace back to the original data, supporting audit trails and member checking.
- Particularly powerful in grounded theory and phenomenology where staying close to participants' language is a methodological requirement.
- Low barrier to entry — requires no specialised software or prior coding framework, making it accessible for novice qualitative researchers.
- Often surfaces unexpected vocabulary or metaphors that open new analytic directions the researcher had not anticipated.
- Codes derived from one participant's idiomatic expression may not travel meaningfully across a diverse sample, creating fragmented code lists.
- Highly context-dependent phrasing can be difficult to aggregate into coherent themes without losing the nuance that made the code valuable.
- Does not constitute a complete analytic method on its own — it is a first-cycle strategy that must be followed by higher-order analysis to produce findings.
- Language-specific: in multilingual or cross-cultural studies, direct verbatim coding raises translation decisions that can compromise the 'verbatim' principle.
Frequently asked
Is in vivo coding the same as grounded theory coding?
In vivo coding originated within the grounded theory tradition and remains central to constructivist grounded theory (Charmaz), but it is now widely used as a first-cycle strategy in many qualitative traditions, including phenomenology, thematic analysis, and content analysis. Grounded theory involves additional coding stages (focused coding, theoretical coding, memoing) that extend well beyond in vivo coding alone.
Can I use in vivo coding alongside other coding types?
Yes, and this is the norm. In vivo coding is almost always combined with other first-cycle strategies — for example, descriptive coding or process coding — and then followed by second-cycle approaches such as focused or pattern coding. The in vivo codes serve as an evidential anchor for the more abstract categories that emerge in later analytic cycles.
How do I handle multilingual data?
The verbatim principle is most cleanly applied in the original language of the interview. If you translate before coding you risk losing the specific connotations that made a phrase analytically meaningful. Best practice in multilingual studies is to code in the original language where feasible, translate the surrounding context for reporting, and document the translation decisions transparently.
What software works best for in vivo coding?
Any CAQDAS package — NVivo, ATLAS.ti, MAXQDA, Dedoose — supports in vivo coding by allowing the selected passage text to be used directly as the code name. The software handles storage and retrieval; the intellectual act of deciding which participant phrases constitute meaningful units remains with the researcher.
How many in vivo codes is too many?
There is no fixed ceiling, but a code list with hundreds of unrelated verbatim phrases is a sign that selectivity has broken down. A practical heuristic: if you cannot begin to cluster your codes into 10–20 tentative categories after completing first-cycle coding, the list is probably too granular. Revisit your data and apply a more selective criterion for what counts as a meaningful unit.
Sources
- Saldaña, J. (2021). The Coding Manual for Qualitative Researchers (4th ed.). Sage. ISBN: 978-1529731743
- Charmaz, K. (2006). Constructing Grounded Theory: A Practical Guide Through Qualitative Analysis. Sage. ISBN: 978-0761973539
How to cite this page
ScholarGate. (2026, June 3). In Vivo Coding. ScholarGate. https://scholargate.app/en/qualitative/in-vivo-coding
Which method?
Set this method beside its closest kin and read them side by side — the library lays the books on the table; the choice is yours.
- Content AnalysisQualitative↔ compare
- EthnographyQualitative↔ compare
- Grounded TheoryQualitative Research↔ compare
- Narrative AnalysisQualitative↔ compare
- PhenomenologyQualitative↔ compare
- Thematic AnalysisQualitative Research↔ compare