Fetching the paper…
Reading the bibliography…
Pre-trained language models such as BERT have achieved great success in a broad range of natural language processing tasks.
Clinicalbert: Modeling clinical notes and predicting hospital readmission
Huang, K.; Altosaar, J.; and Ranganath, R. 2019 · 1904
Earlier work this paper cites.
Ernie: Enhanced representation through knowledge integration
Sun, Y.; Wang, S.; Li, Y.; Feng, S.; Chen, X.; Zhang, H.; Tian, X.; Zhu, D.; Tian, H.; and Wu, H. 2019 · 1904
Earlier work this paper cites.
ERNIE: Enhanced language representation with informative entities
Zhang, Z.; Han, X.; Liu, Z.; Jiang, X.; Sun, M.; and Liu, Q. 2019 · 1905
Earlier work this paper cites.
Roberta: A robustly optimized bert pretraining approach
Liu, Y.; Ott, M.; Goyal, N.; Du, J.; Joshi, M.; Chen, D.; Levy, O.; Lewis, M.; Zettlemoyer, L.; and Stoyanov, V. 2019 · 1907
Earlier work this paper cites.
Albert: A lite bert for self-supervised learning of language representations
Lan, Z.; Chen, M.; Goodman, S.; Gimpel, K.; Sharma, P.; and Soricut, R. 2019 · 1909
Earlier work this paper cites.
KEPLER: A unified model for knowledge embedding and pre-trained language representation
Wang, X.; Gao, T.; Zhu, Z.; Liu, Z.; Li, J.; and Tang, J. 2019 · 1911
Earlier work this paper cites.
K-adapter: Infusing knowledge into pre-trained models with adapters
Wang, R.; Tang, D.; Duan, N.; Wei, Z.; Huang, X.; Cao, C.; Jiang, D.; Zhou, M.; et al. 2020 · 2002
Earlier work this paper cites.
Don’t Stop Pretraining: Adapt Language Models to Domains and Tasks
Gururangan, S.; Marasović, A.; Swayamdipta, S.; Lo, K.; Beltagy, I.; Downey, D.; and Smith, N. A. 2020 · 2004
Earlier work this paper cites.
The Effect of Natural Distribution Shift on Question Answering Models
Miller, J.; Krauth, K.; Recht, B.; and Schmidt, L. 2020 · 2004
Earlier work this paper cites.
Language models are few-shot learners
Brown, T. B.; Mann, B.; Ryder, N.; Subbiah, M.; Kaplan, J.; Dhariwal, P.; Neelakantan, A.; Shyam, P.; Sastry, G.; Askell, A.; et al. 2020 · 2005
Earlier work this paper cites.
Distributed representations of words to guide bootstrapped entity classifiers
Gupta, S.; and Manning, C. D. 2015 · 2015
Cited alongside, same era.
Inferring networks of substitutable and complementary products
McAuley, J.; Pandey, R.; and Leskovec, J. 2015 · 2015
Cited alongside, same era.
Semeval-2016 task 5: Aspect based sentiment analysis
Pontiki, M.; Galanis, D.; Papageorgiou, H.; Androutsopoulos, I.; Manandhar, S.; Al-Smadi, M.; Al-Ayyoub, M.; Zhao, Y.; Qin, B.; De Clercq, O.; et al. 2016 · 2016
Cited alongside, same era.
Deep Contextualized Word Representations
Peters, M.; Neumann, M.; Iyyer, M.; Gardner, M.; Clark, C.; Lee, K.; and Zettlemoyer, L. 2018 · 2018
Cited alongside, same era.
Improving language understanding by generative pre-training
Radford, A.; Narasimhan, K.; Salimans, T.; and Sutskever, I. 2018 · 2018
Cited alongside, same era.
Automated phrase mining from massive text corpora
Shang, J.; Liu, J.; Jiang, M.; Ren, X.; Voss, C. R.; and Han, J. 2018 · 2018
Justifying recommendations using distantly-labeled reviews and fine-grained aspects
Ni, J.; Li, J.; and McAuley, J. 2019 · 2019
Later among the works it cites.
Knowledge Enhanced Contextual Word Representations
Peters, M. E.; Neumann, M.; Logan, R.; Schwartz, R.; Joshi, V.; Singh, S.; and Smith, N. A. 2019 · 2019
Later among the works it cites.
BERT Post-Training for Review Reading Comprehension and Aspect-based Sentiment Analysis
Xu, H.; Liu, B.; Shu, L.; and Philip, S. Y. 2019 · 2019
Later among the works it cites.
Xlnet: Generalized autoregressive pretraining for language understanding
Yang, Z.; Dai, Z.; Yang, Y.; Carbonell, J.; Salakhutdinov, R. R.; and Le, Q. V. 2019 · 2019
Later among the works it cites.
Spanbert: Improving pre-training by representing and predicting spans
Joshi, M.; Chen, D.; Liu, Y.; Weld, D. S.; Zettlemoyer, L.; and Levy, O. 2020 · 2020
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
SciBERT: A Pretrained Language Model for Scientific Text
Beltagy, I.; Lo, K.; and Cohan, A. 2019 · 2019
Cited alongside, same era.
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Devlin, J.; Chang, M.-W.; Lee, K.; and Toutanova, K. 2019 · 2019
Cited alongside, same era.
PubMedQA: A Dataset for Biomedical Research Question Answering
Jin, Q.; Dhingra, B.; Liu, Z.; Cohen, W.; and Lu, X. 2019 · 2019
Cited alongside, same era.
Domain Adaptation with BERT-based Domain Classification and Data Selection
Ma, X.; Xu, P.; Wang, Z.; Nallapati, R.; and Xiang, B. 2019 · 2019
Cited alongside, same era.
Lee, J.; Yoon, W.; Kim, S.; Kim, D.; Kim, S.; So, C. H.; and Kang, J. 2020 · 2020
Closest in time.
K-BERT: Enabling Language Representation with Knowledge Graph
Liu, W.; Zhou, P.; Zhao, Z.; Wang, Z.; Ju, Q.; Deng, H.; and Wang, P. 2020 · 2020
Closest in time.
Adapt or Get Left Behind: Domain Adaptation through BERT Language Model Finetuning for Aspect-Target Sentiment Classification
Rietzler, A.; Stabinger, S.; Opitz, P.; and Engl, S. 2020 · 2020
Closest in time.
ERNIE 2.0: A Continual Pre-Training Framework for Language Understanding
Sun, Y.; Wang, S.; Li, Y.-K.; Feng, S.; Tian, H.; Wu, H.; and Wang, H. 2020 · 2020
Closest in time.