So sánh phương pháp
Xem các phương pháp đã chọn cạnh nhau; những hàng khác biệt được làm nổi bật.
| Mạng bộ nhớ dài-ngắn hạn (LSTM)× | Phân loại dựa trên BERT× | |
|---|---|---|
| Lĩnh vực | Học sâu | Học sâu |
| Họ | Machine learning | Machine learning |
| Năm ra đời≠ | 1997 | 2019 |
| Người khởi xướng≠ | Hochreiter, S. & Schmidhuber, J. | Devlin, J., Chang, M.-W., Lee, K., & Toutanova, K. (Google AI Language) |
| Loại≠ | Recurrent neural network with gated memory cells | Pre-trained language model with fine-tuning |
| Công trình gốc≠ | Hochreiter, S. & Schmidhuber, J. (1997). Long short-term memory. Neural Computation, 9(8), 1735–1780. DOI ↗ | Devlin, J., Chang, M.-W., Lee, K., & Toutanova, K. (2019). BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding. In Proceedings of NAACL-HLT 2019 (pp. 4171–4186). Association for Computational Linguistics. DOI ↗ |
| Tên gọi khác | LSTM, LSTM network, LSTM-RNN, long short-term memory RNN | BERT classifier, BERT fine-tuning for classification, BERT text classification, BERT-CLS |
| Liên quan | 4 | 4 |
| Tóm tắt≠ | Long Short-Term Memory (LSTM) is a gated recurrent neural network architecture introduced by Hochreiter and Schmidhuber in 1997. It was designed to learn dependencies across long sequences by using dedicated memory cells and three learned gates — forget, input, and output — that control what information is retained, updated, or passed forward at each time step. | BERT-based Classification fine-tunes Google's Bidirectional Encoder Representations from Transformers model on a labelled text dataset, replacing the generic pre-trained head with a task-specific classification layer. It exploits deep bidirectional context from hundreds of millions of pre-trained parameters to deliver state-of-the-art accuracy on short- and medium-length text classification tasks with relatively modest amounts of labelled data. |
| ScholarGateBộ dữ liệu ↗ |
|
|