ScholarGate
עוזר

השוואת שיטות

סקרו את השיטות שבחרתם זו לצד זו; שורות שבהן יש הבדל מודגשות.

Collostructional Analysis×N-gram Analysis×
תחוםבלשנותבלשנות
משפחהProcess / pipelineProcess / pipeline
שנת המקור20031999
הוגה השיטהAnatol Stefanowitsch & Stefan Th. GriesCorpus linguists (Douglas Biber; lexical bundles tradition)
סוגStatistical association analysis of lexemes and grammatical constructionsFrequency analysis of contiguous word sequences
מקור מכונןStefanowitsch, A., & Gries, S. T. (2003). Collostructions: Investigating the interaction of words and constructions. International Journal of Corpus Linguistics, 8(2), 209–243. DOI ↗Biber, D., Johansson, S., Leech, G., Conrad, S., & Finegan, E. (1999). Longman Grammar of Spoken and Written English. Longman. ISBN: 9780582237254
כינוייםCollexeme Analysis, Distinctive Collexeme Analysis, Co-varying Collexeme AnalysisLexical Bundle Analysis, Cluster Analysis (corpus linguistics), Contiguous Sequence Analysis
קשורות44
תקצירCollostructional analysis is a family of corpus-based methods, introduced by Anatol Stefanowitsch and Stefan Th. Gries in 2003, that quantify the mutual attraction or repulsion between specific words (lexemes) and the grammatical constructions they occur in. Rooted in construction grammar, it treats a construction — such as the ditransitive "V NP NP" or the "into-causative" — as a meaningful unit and asks which words are statistically drawn to it or kept from it. The core technique, simple collexeme analysis, cross-tabulates how often a lexeme appears in the construction against how often each appears elsewhere, and measures the strength of association, conventionally with a Fisher–Yates exact test. Two extensions handle near-synonymous constructions (distinctive collexeme analysis) and the joint behavior of two slots within one construction (co-varying collexeme analysis), making the method a rigorous quantitative window onto the lexis–grammar interface.N-gram analysis is a corpus-linguistic technique that extracts and ranks every contiguous sequence of n words (or characters) in a corpus, exposing the recurrent multi-word units — two-word bigrams, three-word trigrams, and longer 'lexical bundles' — that make up a register or text type. By counting how often each sequence recurs, it reveals the prefabricated, formulaic backbone of language that single-word frequency lists cannot capture.
ScholarGateמערך נתונים
  1. v1
  2. 2 מקורות
  3. PUBLISHED
  1. v1
  2. 3 מקורות
  3. PUBLISHED

מעבר לחיפוש הורדת מצגת

ScholarGateהשוואת שיטות: Collostructional Analysis · N-gram Analysis. אוחזר בתאריך 2026-06-24 מתוך https://scholargate.app/he/compare