Fetching the paper…
Reading the bibliography…
A `state of the art' model A surpasses humans in a benchmark B, but fails on similar benchmarks C, D, and E.
Guidelines for drinking-water quality
W. H. Organization · 1993
Earlier work this paper cites.
Indoor air quality and health
A. P. Jones · 1999
Earlier work this paper cites.
Understanding power quality problems
M. H. Bollen · 2000
Earlier work this paper cites.
Food quality and safety: consumer perception and demand
K. G. Grunert · 2005
Earlier work this paper cites.
Dqi: Measuring data quality in nlp
S. Mishra, A. Arunkumar, B. Sachdeva, C. Bryan, and C. Baral · 2005
Earlier work this paper cites.
Our evaluation metric needs an update to encourage generalization
S. Mishra, A. Arunkumar, C. Bryan, and C. Baral · 2007
Earlier work this paper cites.
Unbiased look at dataset bias
A. Torralba and A. A. Efros · 2011
Earlier work this paper cites.
A large annotated corpus for learning natural language inference
S. R. Bowman, G. Angeli, C. Potts, and C. D. Manning · 2015
Earlier work this paper cites.
A corpus and evaluation framework for deeper understanding of commonsense stories
N. Mostafazadeh, N. Chambers, X. He, D. Parikh, D. Batra, L. Vanderwende, P. Kohli, and J. Allen · 2016
Earlier work this paper cites.
The effect of different writing tasks on linguistic style: A case study of the roc story cloze task
R. Schwartz, M. Sap, I. Konstas, L. Zilles, Y. Choi, and N. A. Smith · 2017
Earlier work this paper cites.
A broad-coverage challenge corpus for sentence understanding through inference
A. Williams, N. Nangia, and S. R. Bowman · 2017
Cited alongside, same era.
Annotation artifacts in natural language inference data
S. Gururangan, S. Swayamdipta, O. Levy, R. Schwartz, S. R. Bowman, and N. A. Smith · 2018
Cited alongside, same era.
How much reading does reading comprehension require? a critical investigation of popular benchmarks
D. Kaushik and Z. C. Lipton · 2018
Cited alongside, same era.
Resound: Towards action recognition without representation bias
Y. Li, Y. Li, and N. Vasconcelos · 2018
Cited alongside, same era.
Hypothesis only baselines in natural language inference
A. Poliak, J. Naradowsky, A. Haldar, R. Rudinger, and B. Van Durme · 2018
Unlearn dataset bias in natural language inference by fitting the residual
H. He, S. Zha, and H. Wang · 2019
Later among the works it cites.
Learning the difference that makes a difference with counterfactually-augmented data
D. Kaushik, E. Hovy, and Z. C. Lipton · 2019
Later among the works it cites.
Repair: Removing representation bias by dataset resampling
Y. Li and N. Vasconcelos · 2019
Later among the works it cites.
Inoculation by fine-tuning: A method for analyzing challenge datasets
N. F. Liu, R. Schwartz, and N. A. Smith · 2019
Later among the works it cites.
simple but effective techniques to reduce biases
R. K. Mahabadi and J. Henderson · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Know what you don’t know: Unanswerable questions for squad
P. Rajpurkar, R. Jia, and P. Liang · 2018
Cited alongside, same era.
T. Wang, J.-Y. Zhu, A. Torralba, and A. A. Efros · 2018
Cited alongside, same era.
Swag: A large-scale adversarial dataset for grounded commonsense inference
R. Zellers, Y. Bisk, R. Schwartz, and Y. Choi · 2018
Cited alongside, same era.
Don’t take the easy way out: Ensemble based methods for avoiding known dataset biases
C. Clark, M. Yatskar, and L. Zettlemoyer · 2019
Cited alongside, same era.
Data shapley: Equitable valuation of data for machine learning
A. Ghorbani and J. Zou · 2019
Cited alongside, same era.
Adversarial nli: A new benchmark for natural language understanding
Y. Nie, A. Williams, E. Dinan, M. Bansal, J. Weston, and D. Kiela · 2019
Later among the works it cites.
Winogrande: An adversarial winograd schema challenge at scale
K. Sakaguchi, R. L. Bras, C. Bhagavatula, and Y. Choi · 2019
Later among the works it cites.
Evaluating nlp models via contrast sets
M. Gardner, Y. Artzi, V. Basmova, J. Berant, B. Bogin, S. Chen, P. Dasigi, D. Dua, Y. Elazar, A. Gottumukkala, et al · 2020
Closest in time.
Adversarial filters of dataset biases
R. Le Bras, S. Swayamdipta, C. Bhagavatula, R. Zellers, M. E. Peters, A. Sabharwal, and Y. Choi · 2020
Closest in time.