Fetching the paper…
Reading the bibliography…
Large pretrained language models are critical components of modern NLP pipelines.
Martin Arjovsky, Léon Bottou, Ishaan Gulrajani, and David Lopez-Paz. 2019 · 1907
Earlier work this paper cites.
Roberta: A robustly optimized BERT pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019b · 1907
Earlier work this paper cites.
Distilbert, a distilled version of BERT: smaller, faster, cheaper and lighter
Victor Sanh, Lysandre Debut, Julien Chaumond, and Thomas Wolf. 2019 · 1910
Earlier work this paper cites.
Causality for machine learning
Bernhard Schölkopf. 2019 · 1911
Earlier work this paper cites.
Complete identification methods for the causal hierarchy
Ilya Shpitser and Judea Pearl. 2008 · 1979
Earlier work this paper cites.
Generalizing from several related classification tasks to a new unlabeled sample
Gilles Blanchard, Gyemin Lee, and Clayton Scott. 2011 · 2011
Earlier work this paper cites.
Learning causal semantic representation for out-of-distribution prediction
Chang Liu, Xinwei Sun, Jindong Wang, Tao Li, Tao Qin, Wei Chen, and Tie-Yan Liu. 2020 · 2011
Earlier work this paper cites.
Robust solutions of optimization problems affected by uncertain probabilities
Aharon Ben-Tal, Dick den Hertog, Anja De Waegenaere, Bertrand Melenberg, and Gijs Rennen. 2013 · 2013
Earlier work this paper cites.
Domain generalization via invariant feature representation
Krikamol Muandet, David Balduzzi, and Bernhard Schölkopf. 2013 · 2013
Earlier work this paper cites.
Pointer sentinel mixture models
Stephen Merity, Caiming Xiong, James Bradbury, and Richard Socher. 2016 · 2016
Earlier work this paper cites.
Causal inference by using invariant prediction: identification and confidence intervals
Jonas Peters, Peter Bühlmann, and Nicolai Meinshausen. 2016 · 2016
Earlier work this paper cites.
Gender as a variable in natural-language processing: Ethical considerations
Brian Larson. 2017 · 2017
Earlier work this paper cites.
Elements of Causal Inference: Foundations and Learning Algorithms
Jonas Martin Peters, Dominik Janzing, and Bernard Schölkopf. 2017 · 2017
Earlier work this paper cites.
Men also like shopping: Reducing gender bias amplification using corpus-level constraints
Jieyu Zhao, Tianlu Wang, Mark Yatskar, Vicente Ordonez, and Kai-Wei Chang. 2017 · 2017
Earlier work this paper cites.
Metareg: Towards domain generalization using meta-regularization
Yogesh Balaji, Swami Sankaranarayanan, and Rama Chellappa. 2018 · 2018
Cited alongside, same era.
Knowledge distillation by on-the-fly native ensemble
xu lan, Xiatian Zhu, and Shaogang Gong. 2018 · 2018
Cited alongside, same era.
Learning to generalize: Meta-learning for domain generalization
Da Li, Yongxin Yang, Yi-Zhe Song, and Timothy M. Hospedales. 2018a · 2018
Cited alongside, same era.
Gender bias in neural natural language processing
Kaiji Lu, Piotr Mardziel, Fangjing Wu, Preetam Amancharla, and Anupam Datta. 2018 · 2018
Cited alongside, same era.
Theoretical impediments to machine learning with seven sparks from the causal revolution
Judea Pearl. 2018 · 2018
Cited alongside, same era.
Invariant risk minimization games
Kartik Ahuja, Karthikeyan Shanmugam, Kush R. Varshney, and Amit Dhurandhar. 2020 · 2020
Later among the works it cites.
Language models are few-shot learners
Tom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, Sandhini Agarwal, Ariel Herbert-Voss, Gretchen Krueger, Tom Henighan, Rewon Child, Aditya Ramesh, Daniel M. Ziegler, Jeffrey Wu, Clemens Winter, Christopher Hesse, Mark Chen, Eric Sigler, Mateusz Litwin, Scott Gray, Benjamin Chess, Jack Clark, Christopher Berner, Sam McCandlish, Alec Radford, Ilya Sutskever, and Dario Amodei. 2020 · 2020
Later among the works it cites.
The pile: An 800gb dataset of diverse text for language modeling
Leo Gao, Stella Biderman, Sid Black, Laurence Golding, Travis Hoppe, Charles Foster, Jason Phang, Horace He, Anish Thite, Noa Nabeshima, et al. 2020 · 2020
Later among the works it cites.
An empirical study on robustness to spurious correlations using pre-trained language models
Lifu Tu, Garima Lalwani, Spandana Gella, and He He. 2020 · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Learning gender-neutral word embeddings
Jieyu Zhao, Yichao Zhou, Zeyu Li, Wei Wang, and Kai-Wei Chang. 2018 · 2018
Cited alongside, same era.
Adversarial Invariant Feature Learning with Accuracy Constraint for Domain Generalization
Kei Akuzawa, Yusuke Iwasawa, and Yutaka Matsuo. 2019 · 2019
Cited alongside, same era.
Identifying and reducing gender bias in word-level language models
Shikha Bordia and Samuel R. Bowman. 2019 · 2019
Cited alongside, same era.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Cited alongside, same era.
Domain generalization via model-agnostic learning of semantic features
Qi Dou, Daniel Coelho de Castro, Konstantinos Kamnitsas, and Ben Glocker. 2019 · 2019
Cited alongside, same era.
Benchmarking neural network robustness to common corruptions and perturbations
Dan Hendrycks and Thomas G. Dietterich. 2019 · 2019
Cited alongside, same era.
Probing neural network comprehension of natural language arguments
Timothy Niven and Hung-Yu Kao. 2019 · 2019
Cited alongside, same era.
Domain Adversarial Fine-Tuning as an Effective Regularizer
Giorgos Vernikos, Katerina Margatina, Alexandra Chronopoulou, and Ion Androutsopoulos. 2020 · 2020
Later among the works it cites.
Transformers: State-of-the-art natural language processing
Thomas Wolf, Lysandre Debut, Victor Sanh, Julien Chaumond, Clement Delangue, Anthony Moi, Pierric Cistac, Tim Rault, Rémi Louf, Morgan Funtowicz, Joe Davison, Sam Shleifer, Patrick von Platen, Clara Ma, Yacine Jernite, Julien Plu, Canwen Xu, Teven Le Scao, Sylvain Gugger, Mariama Drame, Quentin Lhoest, and Alexander M. Rush. 2020 · 2020
Later among the works it cites.
Domain generalization via entropy regularization
Shanshan Zhao, Mingming Gong, Tongliang Liu, Huan Fu, and Dacheng Tao. 2020 · 2020
Later among the works it cites.
Learning to generate novel domains for domain generalization
Kaiyang Zhou, Yongxin Yang, Timothy M. Hospedales, and Tao Xiang. 2020 · 2020
Later among the works it cites.
Out-of-distribution prediction with invariant risk minimization: The limitation and an effective fix
Ruocheng Guo, Pengchuan Zhang, Hao Liu, and Emre Kiciman. 2021 · 2021
Closest in time.
Evaluating the robustness of neural language models to input perturbations
Milad Moradi and Matthias Samwald. 2021 · 2021
Closest in time.
StereoSet: Measuring stereotypical bias in pretrained language models
Moin Nadeem, Anna Bethke, and Siva Reddy. 2021 · 2021
Closest in time.
Better than average: Paired evaluation of NLP systems
Maxime Peyrard, Wei Zhao, Steffen Eger, and Robert West. 2021 · 2021
Closest in time.
Toward causal representation learning
B. Schölkopf, F. Locatello, S. Bauer, N. R. Ke, N. Kalchbrenner, A. Goyal, and Y. Bengio. 2021 · 2021
Closest in time.
Domain generalization with mixstyle
Kaiyang Zhou, Yongxin Yang, Yu Qiao, and Tao Xiang. 2021 · 2021
Closest in time.