So sánh phương pháp
Xem các phương pháp đã chọn cạnh nhau; những hàng khác biệt được làm nổi bật.
| Phân tích diễn ngôn× | Phân loại văn bản× | |
|---|---|---|
| Lĩnh vực | Khai phá văn bản | Khai phá văn bản |
| Họ | Process / pipeline | Process / pipeline |
| Năm ra đời≠ | 1988 (RST); 2008 (PDTB 2.0) | — |
| Người khởi xướng≠ | Mann & Thompson (RST); Prasad et al. (PDTB) | — |
| Loại≠ | NLP discourse-structure analysis task | Supervised NLP classification task |
| Công trình gốc≠ | Mann, W. C. & Thompson, S. A. (1988). Rhetorical Structure Theory: Toward a functional theory of text organization. Text, 8(3), 243-281. DOI ↗ | Joachims, T. (1998). Text Categorization with Support Vector Machines: Learning with Many Relevant Features. ECML 1998. Lecture Notes in Computer Science, vol 1398. Springer. DOI ↗ |
| Tên gọi khác | rhetorical structure analysis, RST parsing, PDTB parsing, Söylem Ayrıştırma (Discourse Parsing) | text categorization, document classification, topic classification, metin sınıflandırma |
| Liên quan≠ | 3 | 4 |
| Tóm tắt≠ | Discourse parsing is a natural-language-processing task that models the rhetorical relations between sentences and paragraphs of a text — relations such as cause, contrast, and elaboration — and represents them as a tree structure. It works within established frameworks, principally Rhetorical Structure Theory (RST), introduced by Mann and Thompson in 1988, and the Penn Discourse TreeBank (PDTB), released by Prasad and colleagues in 2008. | Text classification, also called text categorization, is a supervised natural-language-processing task that automatically assigns documents to predefined categories. Building on the support-vector-machine approach to text categorization established by Joachims (1998) and consolidated in the text-mining literature by Aggarwal and Zhai (2012), it powers tasks such as spam detection and topic classification by learning from labelled examples. |
| ScholarGateBộ dữ liệu ↗ |
|
|