Fetching the paper…
Reading the bibliography…
Active learning strives to reduce annotation costs by choosing the most critical examples to label.
Assessing BERT’s syntactic abilities
Yoav Goldberg. 2019 · 1901
Earlier work this paper cites.
RoBERTa: A robustly optimized bert pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019 · 1907
Earlier work this paper cites.
On warm-starting neural network training
Jordan T. Ash and Ryan P. Adams. 2019 · 1910
Earlier work this paper cites.
A mathematical theory of communication
Claude Elwood Shannon. 1948 · 1948
Earlier work this paper cites.
Silhouettes: A graphical aid to the interpretation and validation of cluster analysis
Peter J. Rousseeuw. 1987 · 1987
Earlier work this paper cites.
A sequential algorithm for training text classifiers
David D. Lewis and William A. Gale. 1994 · 1994
Earlier work this paper cites.
When is “nearest neighbor” meaningful?
Kevin Beyer, Jonathan Goldstein, Raghu Ramakrishnan, and Uri Shaft. 1999 · 1999
Earlier work this paper cites.
Fine-tuning pretrained language models: Weight initializations, data orders, and early stopping
Jesse Dodge, Gabriel Ilharco, Roy Schwartz, Ali Farhadi, Hannaneh Hajishirzi, and Noah Smith. 2020 · 2002
Earlier work this paper cites.
Representative sampling for text classification using support vector machines
Zhao Xu, Kai Yu, Volker Tresp, Xiaowei Xu, and Jizhi Wang. 2003 · 2003
Earlier work this paper cites.
Multi-criteria-based active learning for named entity recognition
Dan Shen, Jie Zhang, Jian Su, Guodong Zhou, and Chew-Lim Tan. 2004 · 2004
Earlier work this paper cites.
TREC-COVID: Constructing a pandemic information retrieval test collection
Ellen Voorhees, Tasmeer Alam, Steven Bedrick, Dina Demner-Fushman, William R. Hersh, Kyle Lo, Kirk Roberts, Ian Soboroff, and Lucy Lu Wang. 2020 · 2005
Earlier work this paper cites.
k-means++: The advantages of careful seeding
David Arthur and Sergei Vassilvitskii. 2006 · 2006
Earlier work this paper cites.
Active learning for word sense disambiguation with methods for addressing the class imbalance problem
Jingbo Zhu and Eduard Hovy. 2007 · 2007
Earlier work this paper cites.
Visualizing data using t-SNE
Laurens van der Maaten and Geoffrey Hinton. 2008 · 2008
Earlier work this paper cites.
Multiple-instance active learning
Burr Settles, Mark Craven, and Soumya Ray. 2008 · 2008
Earlier work this paper cites.
Jingbo Zhu, Huizhen Wang, Tianshun Yao, and Benjamin K. Tsou. 2008 · 2008
Earlier work this paper cites.
Active learning literature survey
Burr Settles. 2009 · 2009
Earlier work this paper cites.
Active learning with clustering
Zalán Bodó, Zsolt Minier, and Lehel Csató. 2011 · 2010
Earlier work this paper cites.
Off to a good start: Using clustering to select the initial training set in active learning
Rong Hu, Brian Mac Namee, and Sarah Jane Delany. 2010 · 2010
Earlier work this paper cites.
Domain adaptation meets active learning
Piyush Rai, Avishek Saha, Hal Daumé III, and Suresh Venkatasubramanian. 2010 · 2010
Cited alongside, same era.
Two faces of active learning
Sanjoy Dasgupta. 2011 · 2011
Cited alongside, same era.
Learning word vectors for sentiment analysis
Andrew L. Maas, Raymond E. Daly, Peter T. Pham, Dan Huang, Andrew Y. Ng, and Christopher Potts. 2011 · 2011
Cited alongside, same era.
Closing the loop: Fast, interactive semi-supervised annotation with queries on features and instances
Burr Settles. 2011 · 2011
Cited alongside, same era.
Active learning for imbalanced sentiment classification
Shoushan Li, Shengfeng Ju, Guodong Zhou, and Xiaojun Li. 2012 · 2012
Cited alongside, same era.
Recursive deep models for semantic compositionality over a sentiment treebank
Richard Socher, Alex Perelygin, Jean Wu, Jason Chuang, Christopher D. Manning, Andrew Y. Ng, and Christopher Potts. 2013 · 2013
Active learning for convolutional neural networks: A core-set approach
Ozan Sener and Silvio Savarese. 2018 · 2018
Later among the works it cites.
Deep active learning for named entity recognition
Yanyao Shen, Hyokun Yun, Zachary C. Lipton, Yakov Kronrod, and Animashree Anandkumar. 2018 · 2018
Later among the works it cites.
Deep bayesian active learning for natural language processing: Results of a large-scale empirical study
Aditya Siddhant and Zachary C. Lipton. 2018 · 2018
Later among the works it cites.
SciBERT: A pretrained language model for scientific text
Iz Beltagy, Kyle Lo, and Arman Cohan. 2019 · 2019
Later among the works it cites.
Commonsense knowledge mining from pretrained models
Joe Davison, Joshua Feldman, and Alexander M. Rush. 2019 · 2019
Later among the works it cites.
BERT: Pre-training of deep bidirectional transformers for language understanding
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
The role of hubness in clustering high-dimensional data
Nenad Tomasev, Milos Radovanovic, Dunja Mladenic, and Mirjana Ivanovic. 2013 · 2013
Cited alongside, same era.
A new active labeling method for deep learning
Dan Wang and Yi Shang. 2014 · 2014
Cited alongside, same era.
Active transfer learning under model shift
Xuezhi Wang, Tzu-Kuo Huang, and Jeff Schneider. 2014 · 2014
Cited alongside, same era.
Early gains matter: A case for preferring generative over discriminativecrowdsourcing models
Paul Felt, Eric Ringger, Kevin Seppi, Kevin Black, and Robbie Haertel. 2015 · 2015
Cited alongside, same era.
Active learning by learning
Wei-Ning Hsu and Hsuan-Tien Lin. 2015 · 2015
Cited alongside, same era.
Character-level convolutional networks for text classification
Xiang Zhang, Junbo Zhao, and Yann LeCun. 2015 · 2015
Cited alongside, same era.
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Later among the works it cites.
BatchBALD: Efficient and diverse batch acquisition for deep Bayesian active learning
Andreas Kirsch, Joost van Amersfoort, and Yarin Gal. 2019 · 2019
Later among the works it cites.
Decoupled weight decay regularization
Ilya Loshchilov and Frank Hutter. 2019 · 2019
Later among the works it cites.
Practical obstacles to deploying active learning
David Lowell, Zachary C. Lipton, and Byron C. Wallace. 2019 · 2019
Later among the works it cites.
Language models as knowledge bases?
Fabio Petroni, Tim Rocktäschel, Sebastian Riedel, Patrick Lewis, Anton Bakhtin, Yuxiang Wu, and Alexander Miller. 2019 · 2019
Later among the works it cites.
Energy and policy considerations for deep learning in NLP
Emma Strubell, Ananya Ganesh, and Andrew McCallum. 2019 · 2019
Later among the works it cites.
What do you learn from context? Probing for sentence structure in contextualized word representations
Ian Tenney, Patrick Xia, Berlin Chen, Alex Wang, Adam Poliak, R. Thomas McCoy, Najoung Kim, Benjamin Van Durme, Samuel R. Bowman, Dipanjan Das, et al. 2019 · 2019
Later among the works it cites.
XLNet: Generalized autoregressive pretraining for language understanding
Zhilin Yang, Zihang Dai, Yiming Yang, Jaime Carbonell, Ruslan Salakhutdinov, and Quoc V. Le. 2019 · 2019
Later among the works it cites.
Deep batch active learning by diverse, uncertain gradient lower bounds
Jordan T. Ash, Chicheng Zhang, Akshay Krishnamurthy, John Langford, and Alekh Agarwal. 2020 · 2020
Closest in time.
What BERT is not: Lessons from a new suite of psycholinguistic diagnostics for language models
Allyson Ettinger. 2020 · 2020
Closest in time.
Pretrained transformers improve out-of-distribution robustness
Dan Hendrycks, Xiaoyuan Liu, Eric Wallace, Adam Dziedzic, Rishabh Krishnan, and Dawn Song. 2020 · 2020
Closest in time.
Masked language model scoring
Julian Salazar, Davis Liang, Toan Q. Nguyen, and Katrin Kirchhoff. 2020 · 2020
Closest in time.
Interactive refinement of cross-lingual word embeddings
Michelle Yuan, Mozhi Zhang, Benjamin Van Durme, Leah Findlater, and Jordan Boyd-Graber. 2020 · 2020
Closest in time.