ScholarGate
Trợ lý

So sánh phương pháp

Xem các phương pháp đã chọn cạnh nhau; những hàng khác biệt được làm nổi bật.

Phân tích từ ngữ cố định×Phân tích phụ thuộc×Độ đa dạng từ vựng×Phân tích tần suất văn bản×
Lĩnh vựcKhai phá văn bảnKhai phá văn bảnKhai phá văn bảnKhai phá văn bản
HọProcess / pipelineProcess / pipelineProcess / pipelineProcess / pipeline
Năm ra đời19901949
Người khởi xướngChurch & HanksGeorge K. Zipf (frequency-distribution foundation)
LoạiStatistical text-mining techniqueNLP syntactic-analysis taskText quantification / lexical richness measurementDescriptive text-mining analysis
Công trình gốcChurch, K.W. & Hanks, P. (1990). Word Association Norms, Mutual Information, and Lexicography. Computational Linguistics, 16(1), 22-29. link ↗Nivre, J. (2005). Dependency Grammar and Dependency Parsing. MSI Report. link ↗McCarthy, P. M. & Jarvis, S. (2010). MTLD, vocd-D, and HD-D: A validation study of sophisticated approaches to lexical diversity assessment. Behavior Research Methods, 42(2), 381-392. DOI ↗Zipf, G. K. (1949). Human Behavior and the Principle of Least Effort. Addison-Wesley. link ↗
Tên gọi khácword association, collocation extraction, Birliktelik Analizi (Collocation Analysis)syntactic dependency analysis, dependency tree parsing, Bağımlılık Ayrıştırma (Dependency Parsing)lexical richness, vocabulary richness, Sözcüksel Çeşitlilik Analiziword frequency analysis, n-gram frequency analysis, Metin Frekans Analizi
Liên quan3334
Tóm tắtCollocation analysis is a statistical text-mining technique that identifies word pairs or expressions that frequently occur together, using association measures rather than chance co-occurrence. Introduced in the lexicography work of Church and Hanks (1990), it is used for terminology extraction and language analysis, surfacing the multi-word units that carry meaning in a corpus.Dependency parsing is a natural-language-processing task that reveals the syntactic dependency relations between the words of a sentence as a tree structure. Surveyed in the dependency-grammar tradition by Nivre (2005) and made fast and accurate with neural networks by Chen and Manning (2014), it is commonly used as a prerequisite step for information extraction and relation detection.Lexical diversity analysis quantifies how varied the vocabulary of a text is — how rich an author's word choice is — using measures such as the type-token ratio (TTR), MTLD, vocd-D, and Yule's K. The MTLD and vocd-D measures were validated by McCarthy and Jarvis (2010), building on earlier work by Tweedie and Baayen (1998) on the stability of lexical-richness measures.Text frequency analysis is a descriptive text-mining method that counts how often words, n-grams, and phrases occur in a corpus to reveal content patterns and dominant themes. It rests on the frequency-distribution insight formalised by George K. Zipf (1949), that a few terms occur very often while most are rare, and it is one of the most basic and widely used entry points into quantitative text analysis.
ScholarGateBộ dữ liệu
  1. v1
  2. 2 Nguồn tài liệu
  3. PUBLISHED
  1. v1
  2. 2 Nguồn tài liệu
  3. PUBLISHED
  1. v1
  2. 2 Nguồn tài liệu
  3. PUBLISHED
  1. v1
  2. 2 Nguồn tài liệu
  3. PUBLISHED

Đến trang tìm kiếm Tải xuống bản trình chiếu

ScholarGateSo sánh phương pháp: Collocation Analysis · Dependency Parsing · Lexical Diversity · Text Frequency Analysis. Truy cập ngày 2026-06-18 từ https://scholargate.app/vi/compare