Process / pipelineLinguisticsExperimental syntax / second-language acquisitionPipeline

Grammaticality Judgment Task

Also known as: Grammaticality Judgement Task, GJT, Sentence Grammaticality Judgment

OriginatorNoam Chomsky (generative-linguistics tradition)Year1965Sources3Related methods7

The grammaticality judgment task asks speakers to decide whether a sentence is grammatical — well-formed according to the rules of their language — and treats that decision as evidence about the mental grammar that produces it. Rooted in Noam Chomsky's generative program, where the native speaker's intuition is the primary data of linguistics, the task ranges from a single linguist consulting their own intuitions to large controlled experiments with binary, scaled, or forced-choice responses. It is a workhorse of syntactic theory and of second-language acquisition research, where it probes what learners know about a target language beyond what they can produce.

Key highlights

  • Yields negative evidence — what is impossible — that production and corpus data cannot provide.
  • Accesses structures too rare or complex to appear in naturalistic data, giving direct purchase on grammatical theory.
  • Is fast, inexpensive, and adaptable, from a single linguist's intuition to large multi-participant experiments.
  • In second-language research, exposes knowledge of the target grammar that learners possess but do not reliably produce.

Intuition

This section is available to Pro members. Upgrade to Pro

How it works

This section is available to Pro members. Upgrade to Pro

When to use it

Use a grammaticality judgment task when you need evidence about what a grammar permits and forbids, especially for structures that are rare, complex, or never produced spontaneously, where corpus or production data are silent. It is central to theoretical syntax, to testing predictions of competing grammatical analyses, and to second-language research that probes learners' implicit and explicit knowledge of a target structure. It is less appropriate when the question concerns frequency or processing rather than well-formedness, when introspection is unreliable for the population (e.g., young children or low-proficiency learners), or when gradient acceptability is expected — in which case a quantified acceptability judgment task is the better instrument.

Strengths & limitations

Strengths
  • Yields negative evidence — what is impossible — that production and corpus data cannot provide.
  • Accesses structures too rare or complex to appear in naturalistic data, giving direct purchase on grammatical theory.
  • Is fast, inexpensive, and adaptable, from a single linguist's intuition to large multi-participant experiments.
  • In second-language research, exposes knowledge of the target grammar that learners possess but do not reliably produce.
Limitations
  • Conflates grammatical competence with many performance factors — memory load, plausibility, processing difficulty — that also color judgments.
  • Single-linguist introspection is unreplicated and vulnerable to theoretical bias toward the analyst's own predictions.
  • Binary verdicts force a sharp line on phenomena that are often gradient, discarding information about degrees of well-formedness.
  • Learner and non-linguist judgments can be unreliable, sensitive to instructions, response mode, and metalinguistic awareness.

Common pitfalls

This section is available to Pro members. Upgrade to Pro

Applications

This section is available to Pro members. Upgrade to Pro

Frequently asked

What is the difference between grammaticality and acceptability?

Grammaticality is a property of the abstract grammar — whether a sentence is generated by the speaker's internalized rule system (competence). Acceptability is what a person actually reports or feels about a sentence, which reflects the grammar plus performance factors such as memory load, plausibility, and parsing difficulty. A sentence can be grammatical yet hard to accept (a deeply center-embedded sentence) or ungrammatical yet easy to process. Strictly, informants give acceptability judgments, and the analyst infers grammaticality from them, which is why modern practice often reframes the task as an acceptability judgment task.

Can a single linguist's intuition count as data?

In the traditional generative practice, yes — the analyst's own clear intuitions about sharp contrasts (such as classic island violations) have proved robust and are often confirmed when later tested formally. But single-linguist judgments are unreplicated and open to theoretical bias, and for subtle or contested contrasts the field increasingly demands judgments from multiple naive informants under controlled conditions, as the experimental-syntax program advocates.

How should grammaticality judgments be elicited from second-language learners?

Carefully, because learner judgments are sensitive to task conditions. Gass and others show that reliability depends on whether items are presented in writing or aurally, whether time pressure is applied (timed tasks tap more implicit knowledge, untimed tasks allow explicit rule application), how the instructions frame 'grammatical', and whether learners can correct or only classify. Mixing grammatical and ungrammatical items with fillers, randomizing order, and reporting these conditions are essential for interpretable learner data.

Sources

  1. 1.
    Chomsky, N. (1965). Aspects of the Theory of Syntax. MIT Press.
    ISBN 9780262530071
  2. 2.
    Schütze, C. T. (2016). The Empirical Base of Linguistics: Grammaticality Judgments and Linguistic Methodology. Language Science Press.
  3. 3.
    Gass, S. (1994). The reliability of second-language grammaticality judgments. In E. Tarone, S. Gass, & A. Cohen (Eds.), Research Methodology in Second-Language Acquisition (pp. 303–322). Lawrence Erlbaum.
    ISBN 9780805814569

You have read it. What now?

Cite this page

ScholarGate. (2026, June 22). Grammaticality Judgment Task. ScholarGate. https://scholargate.app/linguistics/grammaticality-judgment-task