On the realization of compositionality in neural networks
Original
Joris Baan, Jana Leible, Mitja Nikolaus, David Rau, Dennis Ulmer, Tim Baumgärtner, Dieuwke Hupkes, and Elia Bruni. 2019 · 1906
Earlier work this paper cites.
WINOGRANDE: an adversarial winograd schema challenge at scale
Original
Keisuke Sakaguchi, Ronan Le Bras, Chandra Bhagavatula, and Yejin Choi. 2019 · 1907
Earlier work this paper cites.
Towards debiasing fact verification models
Original
Tal Schuster, Darsh J Shah, Yun Jie Serene Yeo, Daniel Filizzola, Enrico Santus, and Regina Barzilay. 2019 · 1908
Earlier work this paper cites.
Simple but effective techniques to reduce biases
Original
Rabeeh Karimi Mahabadi and James Henderson. 2019 · 1909
Earlier work this paper cites.
Adversarial nli: A new benchmark for natural language understanding
Original
Yixin Nie, Adina Williams, Emily Dinan, Mohit Bansal, Jason Weston, and Douwe Kiela. 2019 · 1910
Earlier work this paper cites.
Huggingface’s transformers: State-of-the-art natural language processing
Original
Thomas Wolf, Lysandre Debut, Victor Sanh, Julien Chaumond, Clement Delangue, Anthony Moi, Pierric Cistac, Tim Rault, R’emi Louf, Morgan Funtowicz, and Jamie Brew. 2019 · 1910
Earlier work this paper cites.
Semi-supervised sequence modeling with cross-view training
Kevin Clark, Minh-Thang Luong, Christopher D. Manning, and Quoc V. Le. 2018 · 1925
Earlier work this paper cites.
Universal grammar
Richard Montague. 1970 · 1970
Earlier work this paper cites.
Adversarial filters of dataset biases
Original
Ronan Le Bras, Swabha Swayamdipta, Chandra Bhagavatula, Rowan Zellers, Matthew E Peters, Ashish Sabharwal, and Yejin Choi. 2020 · 2002
Earlier work this paper cites.
Smote: synthetic minority over-sampling technique
Nitesh V Chawla, Kevin W Bowyer, Lawrence O Hall, and W Philip Kegelmeyer. 2002 · 2002
Earlier work this paper cites.
An investigation of why overparameterization exacerbates spurious correlations
Original
Shiori Sagawa, Aditi Raghunathan, Pang Wei Koh, and Percy Liang. 2020b · 2005
Earlier work this paper cites.
Covariate shift adaptation by importance weighted cross validation
Masashi Sugiyama, Matthias Krauledat, and Klaus-Robert MÞller. 2007 · 2007
Earlier work this paper cites.
An empirical study on robustness to spurious correlations using pre-trained language models
Original
Lifu Tu, Garima Lalwani, Spandana Gella, and He He. 2020 · 2007
Earlier work this paper cites.
Self-Paced Learning for Latent Variable Models
M Pawan Kumar, Benjamin Packer, and Daphne Koller. 2010 · 2010
Earlier work this paper cites.
Evaluating performance of biomedical image retrieval systems - an overview of the medical image retrieval task at imageclef 2004-2013
Jayashree Kalpathy-Cramer, Alba Garcia Seco de Herrera, Dina Demner-Fushman, Sameer K. Antani, Steven Bedrick, and Henning Müller. 2015 · 2013
Earlier work this paper cites.
Glove: Global vectors for word representation
Jeffrey Pennington, Richard Socher, and Christopher Manning. 2014 · 2014
Earlier work this paper cites.
Stochastic optimization with importance sampling for regularized loss minimization
Peilin Zhao and Tong Zhang · 2015
Earlier work this paper cites.
Learning what data to learn
Original
Yang Fan, Fei Tian, Tao Qin, Jiang Bian, and Tie-Yan Liu. 2017 · 2017
Earlier work this paper cites.