방법 비교
선택한 방법을 나란히 검토하세요. 서로 다른 행은 강조 표시됩니다.
| 담화 분석× | 텍스트 분류× | |
|---|---|---|
| 분야 | 텍스트 마이닝 | 텍스트 마이닝 |
| 계열 | Process / pipeline | Process / pipeline |
| 기원 연도≠ | 1988 (RST); 2008 (PDTB 2.0) | — |
| 창시자≠ | Mann & Thompson (RST); Prasad et al. (PDTB) | — |
| 유형≠ | NLP discourse-structure analysis task | Supervised NLP classification task |
| 원전≠ | Mann, W. C. & Thompson, S. A. (1988). Rhetorical Structure Theory: Toward a functional theory of text organization. Text, 8(3), 243-281. DOI ↗ | Joachims, T. (1998). Text Categorization with Support Vector Machines: Learning with Many Relevant Features. ECML 1998. Lecture Notes in Computer Science, vol 1398. Springer. DOI ↗ |
| 별칭 | rhetorical structure analysis, RST parsing, PDTB parsing, Söylem Ayrıştırma (Discourse Parsing) | text categorization, document classification, topic classification, metin sınıflandırma |
| 관련≠ | 3 | 4 |
| 요약≠ | Discourse parsing is a natural-language-processing task that models the rhetorical relations between sentences and paragraphs of a text — relations such as cause, contrast, and elaboration — and represents them as a tree structure. It works within established frameworks, principally Rhetorical Structure Theory (RST), introduced by Mann and Thompson in 1988, and the Penn Discourse TreeBank (PDTB), released by Prasad and colleagues in 2008. | Text classification, also called text categorization, is a supervised natural-language-processing task that automatically assigns documents to predefined categories. Building on the support-vector-machine approach to text categorization established by Joachims (1998) and consolidated in the text-mining literature by Aggarwal and Zhai (2012), it powers tasks such as spam detection and topic classification by learning from labelled examples. |
| ScholarGate데이터셋 ↗ |
|
|