ScholarGate
助手

方法对比

并排查看您选择的方法;存在差异的行会高亮显示。

文本回归×BERT 嵌入×
领域文本挖掘文本挖掘
方法族Process / pipelineProcess / pipeline
起源年份2019
提出者Devlin, Chang, Lee & Toutanova (Google AI)
类型Supervised regression on text featuresContextual transformer text-representation method
开创性文献Gentzkow, M., Kelly, B. & Taddy, M. (2019). Text as Data. Journal of Economic Literature, 57(3), 535-574. DOI ↗Devlin, J., Chang, M.-W., Lee, K. & Toutanova, K. (2019). BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding. NAACL-HLT, 4171-4186. DOI ↗
别名text-as-data regression, predicting numeric outcomes from text, Metin Tabanlı Regresyoncontextual embeddings, transformer embeddings, BERT Tabanlı Metin Gömülmeleri
相关44
摘要Text-based regression predicts a continuous target variable using features extracted from text — TF-IDF scores, embeddings, or n-grams — as the independent variables. Building on the text-as-data programme consolidated by Gentzkow, Kelly and Taddy (2019), it lets a numeric outcome such as a price, a rating, or a sentiment score be estimated directly from documents, and is widely used in social-science, economics, and finance applications.BERT-based text embeddings, introduced by Devlin and colleagues at Google AI in 2019, turn text into context-sensitive dense vectors using a bidirectional Transformer encoder. Because the meaning of a word shifts with its context, BERT produces richer representations than static methods such as Word2Vec or topic models like LDA.
ScholarGate数据集
  1. v1
  2. 2 来源
  3. PUBLISHED
  1. v1
  2. 2 来源
  3. PUBLISHED

前往搜索 下载幻灯片

ScholarGate方法对比: Text Regression · BERT Embeddings. 于 2026-06-17 检索自 https://scholargate.app/zh/compare