Visual Elicitation Metaphor Analysis — VEMA / ZMET
Visual Elicitation Metaphor Analysis (VEMA) · Also known as: VEMA, ZMET, Zaltman Metaphor Elicitation Technique, visual metaphor elicitation
Visual Elicitation Metaphor Analysis (VEMA) is a qualitative technique in which participants select or create images that represent their thoughts, feelings, or experiences about a topic, and then articulate the metaphors embedded in those images during a guided interview. Originally formalised as the Zaltman Metaphor Elicitation Technique (ZMET) by Gerald Zaltman in 1995, the approach rests on the premise that most human thought is nonverbal and structured through metaphor, making images a more direct gateway to deep mental models than verbal questioning alone.
Read the full method
Sign in with a free account to read this section.
Method map
The neighbourhood of related methods — select a node to explore.
When to use it
Visual Elicitation Metaphor Analysis is most appropriate when the research goal is to surface subconscious, emotionally laden, or hard-to-articulate beliefs — for example, brand perception, patient experiences, organisational culture, or identity. It is particularly valuable when verbal interviewing alone is likely to yield surface-level or socially desirable responses. The method suits exploratory studies in marketing, psychology, health sciences, education, and organisational research. It requires participants who can engage with images and invest time before and during the interview (typically 60–90 minutes per participant). It is not suitable when the research question calls for frequency or statistical inference, when participants cannot engage with projective tasks (e.g., very young children without adaptation, or populations with severe cognitive impairment), or when resources do not allow for individual in-depth interviews.
Strengths & limitations
- Accesses subconscious and emotionally complex constructs that direct verbal questioning rarely reaches.
- Images serve as a shared projective medium that transcends language barriers and literacy constraints.
- The metaphor-mapping output provides a visual, communicable model of collective mental structures.
- Highly participant-led — the images and the narratives they generate belong to the participant, reducing interviewer bias.
- Applicable across diverse contexts: consumer research, health, education, organisational studies.
- Time-intensive for both participants (pre-session image collection) and researchers (individual in-depth interviews plus metaphor coding).
- Requires skilled interviewers trained in projective and metaphor-elicitation probing; poorly conducted interviews yield thin data.
- Sample sizes are typically small (8–20 participants), limiting breadth of perspective.
- Metaphor interpretation carries inherent subjectivity; inter-rater reliability checks are recommended but not always reported.
- The ZMET variant is a patented commercial process; adaptations in academic research must depart from the proprietary protocol.
Frequently asked
How many participants are typical for a visual elicitation metaphor analysis study?
Most published studies use 8–20 participants. Because each session is long (60–90 minutes) and the data are rich, smaller samples can yield substantial findings. The stopping criterion is thematic saturation — when new participants' metaphors no longer add new constructs to the consensus map. Fewer than 8 participants risks missing important metaphor clusters; beyond 25, the consensus-mapping process becomes unwieldy without team-based analysis.
What kinds of images should participants bring?
Participants are instructed to collect any images — photographs, magazine clippings, drawings, digital images — that feel personally connected to their thoughts or feelings about the topic. No thematic guidance is given. Typically 8–12 images per participant is practical; too few limit the range of constructs, while too many make the interview unfocused.
Is ZMET the same as visual elicitation metaphor analysis?
ZMET (Zaltman Metaphor Elicitation Technique) is the original, patented commercial instantiation of this approach, with a specific proprietary protocol including seven universal deep metaphors. Since the ZMET patent expired in 2015, academic researchers have adapted the core idea under various labels. VEMA or 'visual elicitation metaphor analysis' refers to this broader family of adaptations that share the same logic — visual prompts, metaphor elicitation, construct mapping — without being bound to the ZMET protocol.
How is this different from simple photo elicitation?
Standard photo elicitation uses images primarily to prompt stories, memories, or opinions — the image is a conversation starter. Visual elicitation metaphor analysis goes further: images are treated as projective devices for surfacing metaphorical constructs, and the analysis explicitly maps the deep conceptual structures (metaphors) that participants use to make sense of the topic, rather than simply coding the content or themes of what they say.
How do I ensure rigour in metaphor interpretation?
Common rigour strategies include: (1) using a second coder to independently code metaphors and reporting inter-rater reliability (e.g., Cohen's kappa); (2) member-checking — returning the consensus map to participants and asking whether it reflects their experience; (3) an audit trail documenting coding decisions; and (4) reflexive memos about how the researcher's own metaphor frameworks may have influenced interpretation.
Sources
How to cite this page
ScholarGate. (2026, June 3). Visual Elicitation Metaphor Analysis (VEMA). ScholarGate. https://scholargate.app/en/qualitative/visual-elicitation-metaphor-analysis
Which method?
Set this method beside its closest kin and read them side by side — the library lays the books on the table; the choice is yours.
- Metaphor AnalysisQualitative↔ compare
- Narrative AnalysisQualitative↔ compare
- PhenomenologyQualitative↔ compare
- Thematic AnalysisQualitative Research↔ compare