Fetching the paper…
Reading the bibliography…
Many pairwise classification tasks, such as paraphrase detection and open-domain question answering, naturally have extreme label imbalance (e.g., $99.99\%$ of examples are negatives).
R. Nogueira and K. Cho. 2019 · 1901
Earlier work this paper cites.
Learning and evaluating general linguistic intelligence
D. Yogatama, C. de M. d’Autume, J. Connor, T. Kocisky, M. Chrzanowski, L. Kong, A. Lazaridou, W. Ling, L. Yu, C. Dyer, et al. 2019 · 1901
Earlier work this paper cites.
XLNet: Generalized autoregressive pretraining for language understanding
Z. Yang, Z. Dai, Y. Yang, J. Carbonell, R. Salakhutdinov, and Q. V. Le. 2019 · 1906
Earlier work this paper cites.
RoBERTa: A robustly optimized BERT pretraining approach
Y. Liu, M. Ott, N. Goyal, J. Du, M. Joshi, D. Chen, O. Levy, M. Lewis, L. Zettlemoyer, and V. Stoyanov. 2019 · 1907
Earlier work this paper cites.
Overview of the first TREC text retrieval conference
D. K. Harman. 1992 · 1992
Earlier work this paper cites.
A sequential algorithm for training text classifiers
David D Lewis and William A Gale. 1994 · 1994
Earlier work this paper cites.
Evaluating and optimizing autonomous text classification systems
D. D. Lewis. 1995 · 1995
Earlier work this paper cites.
Editorial: Special issue on learning from imbalanced data sets
N. V. Chawla, N. Japkowicz, and A. R. Kolcz. 2004 · 2004
Earlier work this paper cites.
Rcv1: A new benchmark collection for text categorization research
D. D. Lewis, Y. Yang, T. G. Rose, and F. Li. 2004 · 2004
Earlier work this paper cites.
Margin based active learning
Maria-Florina Balcan, Andrei Broder, and Tong Zhang. 2007 · 2007
Earlier work this paper cites.
Learning on the border: active learning in imbalanced data classification
Seyda Ertekin, Jian Huang, Leon Bottou, and Lee Giles. 2007 · 2007
Earlier work this paper cites.
Introduction to information retrieval , volume 1
C. Manning, P. Raghavan, and H. Schütze. 2008 · 2008
Earlier work this paper cites.
Undersampling approach for imbalanced training sets and induction from multi-label text-categorization domains
S. Dendamrongvit and M. Kubat. 2009 · 2009
Earlier work this paper cites.
Active learning literature survey
Burr Settles. 2009 · 2009
Earlier work this paper cites.
On strategies for imbalanced text classification using svm: A comparative study
A. Sun, E. Lim, and ying Liu. 2009 · 2009
Earlier work this paper cites.
Why label when you can search? alternatives to active learning for applying human resources to build classification models under extreme class imbalance
J. Attenberg and F. Provost. 2010 · 2010
Earlier work this paper cites.
Overview of the TAC 2011 knowledge base population track
H. Ji, R. Grishman, and H. Trang Dang. 2011 · 2011
Earlier work this paper cites.
Active and passive learning of linear separators under log-concave distributions
Maria-Florina Balcan and Phil Long. 2013 · 2013
Cited alongside, same era.
A large annotated corpus for learning natural language inference
S. Bowman, G. Angeli, C. Potts, and C. D. Manning. 2015 · 2015
Cited alongside, same era.
Learning hybrid representations to retrieve semantically equivalent questions
C. dos Santos, L. Barbosa, D. Bogdanova, and B. Zadrozny. 2015 · 2015
Cited alongside, same era.
Batch normalization: Accelerating deep network training by reducing internal covariate shift
S. Ioffe and C. Szegedy. 2015 · 2015
Cited alongside, same era.
Cqadupstack: Gold or silver?
D. Hoogeveen, Karin M. Verspoor, and Timothy Baldwin. 2016 · 2016
Cited alongside, same era.
A corpus and cloze evaluation for deeper understanding of commonsense stories
N. Mostafazadeh, N. Chambers, X. He, D. Parikh, D. Batra, L. Vanderwende, P. Kohli, and J. Allen. 2016 · 2016
Stress test evaluation for natural language inference
A. Naik, A. Ravichander, N. Sadeh, C. Rose, and G. Neubig. 2018 · 2018
Later among the works it cites.
Hypothesis only baselines in natural language inference
A. Poliak, J. Naradowsky, A. Haldar, R. Rudinger, and B. V. Durme. 2018 · 2018
Later among the works it cites.
WikiQA: A challenge dataset for open-domain question answering
Y. Yang, W. Yih, and C. Meek. 2015 · 2018
Later among the works it cites.
A benchmark and comparison of active learning for logistic regression
Yazhou Yang and Marco Loog. 2018 · 2018
Later among the works it cites.
BERT: Pre-training of deep bidirectional transformers for language understanding
J. Devlin, M. Chang, K. Lee, and K. Toutanova. 2019 · 2019
Later among the works it cites.
Posing fair generalization tasks for natural language inference
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Reading Wikipedia to answer open-domain questions
D. Chen, A. Fisch, J. Weston, and A. Bordes. 2017 · 2017
Cited alongside, same era.
First quora dataset release: Question pairs
Shankar Iyer, Nikhil Dandekar, and Kornél Csernai. 2017 · 2017
Cited alongside, same era.
Billion-scale similarity search with gpus
Jeff Johnson, Matthijs Douze, and Hervé Jégou. 2017 · 2017
Cited alongside, same era.
Zero-shot relation extraction via reading comprehension
O. Levy, M. Seo, E. Choi, and L. Zettlemoyer. 2017 · 2017
Cited alongside, same era.
The effect of different writing tasks on linguistic style: A case study of the ROC story cloze task
R. Schwartz, M. Sap, Y. Konstas, L. Zilles, Y. Choi, and N. A. Smith. 2017 · 2017
Cited alongside, same era.
Inter-weighted alignment network for sentence pair modeling
G. Shen, Y. Yang, and Z. Deng. 2017 · 2017
Cited alongside, same era.
A. Geiger, I. Cases, L. Karttunen, and C. Potts. 2019 · 2019
Later among the works it cites.
Learning dense representations for entity retrieval
D. Gillick, S. Kulkarni, L. Lansing, A. Presta, J. Baldridge, E. Ie, and D. Garcia-Olano. 2019 · 2019
Later among the works it cites.
Latent retrieval for weakly supervised open domain question answering
K. Lee, M. Chang, and K. Toutanova. 2019 · 2019
Later among the works it cites.
Right for the wrong reasons: Diagnosing syntactic heuristics in natural language inference
R. T. McCoy, E. Pavlick, and T. Linzen. 2019 · 2019
Later among the works it cites.
Sentence-BERT: Sentence embeddings using siamese BERT-networks
N. Reimers and I. Gurevych. 2019 · 2019
Later among the works it cites.
MultiQA: An empirical investigation of generalization and transfer in reading comprehension
A. Talmor and J. Berant. 2019 · 2019
Later among the works it cites.
Glue: A multi-task benchmark and analysis platform for natural language understanding
A. Wang, A. Singh, J. Michael, F. Hill, O. Levy, and S. R. Bowman. 2019 · 2019
Later among the works it cites.
A compare-aggregate model with latent clustering for answer selection
S. Yoon, F. Dernoncourt, D. S. Kim, T. Bui, and K. Jung. 2019 · 2019
Later among the works it cites.
Selection bias explorations and debias methods for natural language sentence matching datasets
G. Zhang, B. Bai, J. Liang, K. Bai, S. Chang, M. Yu, C. Zhu, and T. Zhao. 2019 · 2019
Later among the works it cites.
TANDA: Transfer and adapt pre-trained transformer models for answer sentence selection
S. Garg, T. Vu, and A. Moschitti. 2020 · 2020
Closest in time.
ALBERT: A lite BERT for self-supervised learning of language representations
Z. Lan, M. Chen, S. Goodman, K. Gimpel, P. Sharma, and R. Soricut. 2020 · 2020
Closest in time.