Fetching the paper…
Reading the bibliography…
Present language understanding methods have demonstrated extraordinary ability of recognizing patterns in texts via machine learning.
Approximation capabilities of multilayer feedforward networks
Kurt Hornik. 1991 · 1991
Earlier work this paper cites.
Information theory and statistics
Solomon Kullback. 1997 · 1997
Earlier work this paper cites.
Syntactic structures
Noam Chomsky. 2002 · 2002
Earlier work this paper cites.
A generalized parallelogram law
Alan Nash. 2003 · 2003
Earlier work this paper cites.
Learning what makes a difference from counterfactual examples and gradient supervision
Damien Teney, Ehsan Abbasnedjad, and Anton van den Hengel. 2020 · 2004
Earlier work this paper cites.
Introduction to information retrieval , volume 39
Hinrich Schütze, Christopher D Manning, and Prabhakar Raghavan. 2008 · 2008
Earlier work this paper cites.
Convex optimization & Euclidean distance geometry
Jon Dattorro. 2010 · 2010
Earlier work this paper cites.
Counterfactual variable control for robust and interpretable question answering
Sicheng Yu, Yulei Niu, Shuohang Wang, Jing Jiang, and Qianru Sun. 2020 · 2010
Earlier work this paper cites.
Techniques and applications for sentiment analysis
Ronen Feldman. 2013 · 2013
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba. 2014 · 2014
Earlier work this paper cites.
A large annotated corpus for learning natural language inference
Samuel R. Bowman, Gabor Angeli, Christopher Potts, and Christopher D. Manning. 2015 · 2015
Earlier work this paper cites.
Gaussian error linear units (gelus)
Dan Hendrycks and Kevin Gimpel. 2016 · 2016
Earlier work this paper cites.
Thinking, fast and slow
Kahneman Daniel. 2017 · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Earlier work this paper cites.
Learning independent causal mechanisms
Giambattista Parascandolo, Niki Kilbertus, Mateo Rojas-Carulla, and Bernhard Schölkopf. 2018 · 2018
Earlier work this paper cites.
Zero-shot learning of classifiers from natural language quantification
Shashank Srivastava, Igor Labutov, and Tom Mitchell. 2018 · 2018
Cited alongside, same era.
A broad-coverage challenge corpus for sentence understanding through inference
Adina Williams, Nikita Nangia, and Samuel Bowman. 2018 · 2018
Cited alongside, same era.
Hierarchical graph representation learning with differentiable pooling
Rex Ying, Jiaxuan You, Christopher Morris, Xiang Ren, William L Hamilton, and Jure Leskovec. 2018 · 2018
Cited alongside, same era.
Confidence predictions affect performance confidence and neural preparation in perceptual decision making
Annika Boldt, Anne-Marike Schiffer, Florian Waszak, and Nick Yeung. 2019 · 2019
Cited alongside, same era.
Addressing failure prediction by learning model confidence
Charles Corbière, Nicolas Thome, Avner Bar-Hen, Matthieu Cord, and Patrick Pérez. 2019 · 2019
Cited alongside, same era.
A unified mrc framework for named entity recognition
Xiaoya Li, Jingrong Feng, Yuxian Meng, Qinghong Han, Fei Wu, and Jiwei Li. 2020 · 2020
Later among the works it cites.
Learning to contrast the counterfactual samples for robust visual question answering
Zujie Liang, Weitao Jiang, Haifeng Hu, and Jiaying Zhu. 2020 · 2020
Later among the works it cites.
Long-tailed classification by keeping the good and removing the bad momentum causal effect
Kaihua Tang, Jianqiang Huang, and Hanwang Zhang. 2020 · 2020
Later among the works it cites.
An empirical study on robustness to spurious correlations using pre-trained language models
Lifu Tu, Garima Lalwani, Spandana Gella, and He He. 2020 · 2020
Later among the works it cites.
Mind the trade-off: Debiasing nlu models without degrading the in-distribution performance
Prasetya Ajie Utama, Nafise Sadat Moosavi, and Iryna Gurevych. 2020 · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
BERT: pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Cited alongside, same era.
Learning the difference that makes a difference with counterfactually-augmented data
Divyansh Kaushik, Eduard Hovy, and Zachary Lipton. 2019 · 2019
Cited alongside, same era.
Entity-relation extraction as multi-turn question answering
Xiaoya Li, Fan Yin, Zijun Sun, Xiayu Li, Arianna Yuan, Duo Chai, Mingxin Zhou, and Jiwei Li. 2019 · 2019
Cited alongside, same era.
Roberta: A robustly optimized bert pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019 · 2019
Cited alongside, same era.
The seven tools of causal inference, with reflections on machine learning
Judea Pearl. 2019 · 2019
Cited alongside, same era.
Counterfactual samples synthesizing for robust visual question answering
Long Chen, Xin Yan, Jun Xiao, Hanwang Zhang, Shiliang Pu, and Yueting Zhuang. 2020 · 2020
Cited alongside, same era.
Counterfactual vision-and-language navigation via adversarial path sampler
Tsu-Jui Fu, Xin Eric Wang, Matthew F Peterson, Scott T Grafton, Miguel P Eckstein, and William Yang Wang. 2020 · 2020
Cited alongside, same era.
Corefqa: Coreference resolution as query-based span prediction
Wei Wu, Fei Wang, Arianna Yuan, Fei Wu, and Jiwei Li. 2020 · 2020
Later among the works it cites.
Generating plausible counterfactual explanations for deep transformers in financial text classification
Linyi Yang, Eoin Kenny, Tin Lok James Ng, Yi Yang, Barry Smyth, and Ruihai Dong. 2020 · 2020
Later among the works it cites.
Counterfactual generator: A weakly-supervised method for named entity recognition
Xiangji Zeng, Yunliang Li, Yuchen Zhai, and Yin Zhang. 2020 · 2020
Later among the works it cites.
Counterfactual off-policy training for neural dialogue generation
Qingfu Zhu, Weinan Zhang, Ting Liu, and William Yang Wang. 2020 · 2020
Later among the works it cites.
Should graph convolution trust neighbors? a simple causal inference method
Fuli Feng, Weiran Huang, Xin Xin, Xiangnan He, and Tat-Seng Chua. 2021 · 2021
Closest in time.
Counterfactual vqa: A cause-effect look at language bias
Yulei Niu, Kaihua Tang, Hanwang Zhang, Zhiwu Lu, Xian-Sheng Hua, and Ji-Rong Wen. 2021 · 2021
Closest in time.
Counterfactual generative networks
Axel Sauer and Andreas Geiger. 2021 · 2021
Closest in time.
” click” is not equal to” like”: Counterfactual recommendation for mitigating clickbait issue
Wenjie Wang, Fuli Feng, Xiangnan He, Hanwang Zhang, and Tat-Seng Chua. 2021 · 2021
Closest in time.
Counterfactual zero-shot and open-set visual recognition
Zhongqi Yue, Tan Wang, Hanwang Zhang, Qianru Sun, and Xian-Sheng Hua. 2021 · 2021
Closest in time.