Fetching the paper…
Reading the bibliography…
Pre-trained language models have been successful on text classification tasks, but are prone to learning spurious correlations from biased datasets, and are thus vulnerable when making inferences in a new domain.
Recognizing contextual polarity in phrase-level sentiment analysis
T. Wilson, J. Wiebe, and P. Hoffmann · 2005
Earlier work this paper cites.
Luke Zettlemoyer and M. Collins · 2005
Earlier work this paper cites.
Axiomatic characterizations of probabilistic and cardinal-probabilistic interaction indices
K. Fujimoto, I. Kojadinovic, and J. Marichal · 2006
Earlier work this paper cites.
Active learning literature survey
Burr Settles · 2009
Earlier work this paper cites.
Generalized expectation criteria for semi-supervised learning with weakly labeled data
Gideon S. Mann and Andrew McCallum · 2010
Earlier work this paper cites.
A survey on transfer learning
Sinno Jialin Pan and Qiang Yang · 2010
Earlier work this paper cites.
Human model evaluation in interactive supervised learning
R. Fiebrink, P. Cook, and D. Trueman · 2011
Earlier work this paper cites.
Domain adaptation for large-scale sentiment classification: A deep learning approach
Xavier Glorot, Antoine Bordes, and Yoshua Bengio · 2011
Earlier work this paper cites.
Unbiased look at dataset bias
A. Torralba and Alexei A. Efros · 2011
Earlier work this paper cites.
A short introduction to probabilistic soft logic
Angelika Kimmig, Stephen H. Bach, Matthias Broecheler, Bert Huang, and Lise Getoor · 2012
Earlier work this paper cites.
The harm in hate speech
R. Delgado · 2013
Earlier work this paper cites.
Recursive deep models for semantic compositionality over a sentiment treebank
R. Socher, Alex Perelygin, J. Wu, Jason Chuang, Christopher D. Manning, A. Ng, and Christopher Potts · 2013
Earlier work this paper cites.
Decaf: A deep convolutional activation feature for generic visual recognition
J. Donahue, Y. Jia, Oriol Vinyals, Judy Hoffman, Ning Zhang, Eric Tzeng, and Trevor Darrell · 2014
Earlier work this paper cites.
Unsupervised domain adaptation by backpropagation
Yaroslav Ganin and V. Lempitsky · 2015
Earlier work this paper cites.
Distilling the knowledge in a neural network
Geoffrey E. Hinton, Oriol Vinyals, and J. Dean · 2015
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P. Kingma and Jimmy Ba · 2015
Earlier work this paper cites.
Principles of explanatory debugging to personalize interactive machine learning
Todd Kulesza, Margaret Burnett, Weng-Keen Wong, and Simone Stumpf · 2015
Earlier work this paper cites.
Transfer learning for speech and language processing
D. Wang and T. Zheng · 2015
Earlier work this paper cites.
Ups and downs: Modeling the visual evolution of fashion trends with one-class collaborative filtering
Ruining He and Julian McAuley · 2016
Cited alongside, same era.
Return of frustratingly easy domain adaptation
Baochen Sun, Jiashi Feng, and Kate Saenko · 2016
Cited alongside, same era.
Using millions of emoji occurrences to learn any-domain representations for detecting sentiment, emotion and sarcasm
Bjarke Felbo, A. Mislove, Anders Søgaard, I. Rahwan, and S. Lehmann · 2017
Cited alongside, same era.
Lime: Low-light image enhancement via illumination map estimation
Xiaojie Guo, Y. Li, and Haibin Ling · 2017
Cited alongside, same era.
Representation stability as a regularizer for improved text analytics transfer learning
M. Riemer, Elham Khabiri, and Richard Goodwin · 2017
Cited alongside, same era.
Bert: Pre-training of deep bidirectional transformers for language understanding
J. Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova · 2019
Later among the works it cites.
Learning explainable models using attribution priors
G. Erion, J. Janizek, Pascal Sturmfels, Scott M. Lundberg, and Su-In Lee · 2019
Later among the works it cites.
A survey of methods for explaining black box models
Riccardo Guidotti, A. Monreale, F. Turini, D. Pedreschi, and F. Giannotti · 2019
Later among the works it cites.
Incorporating priors with feature attribution on text classification
Frederick Liu and B. Avci · 2019
Later among the works it cites.
Roberta: A robustly optimized bert pretraining approach
Y. Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, M. Lewis, Luke Zettlemoyer, and Veselin Stoyanov · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Right for the right reasons: Training differentiable models by constraining their explanations
A. Ross, M. Hughes, and Finale Doshi-Velez · 2017
Cited alongside, same era.
Axiomatic attribution for deep networks
M. Sundararajan, Ankur Taly, and Qiqi Yan · 2017
Cited alongside, same era.
Hate me, hate me not: Hate speech detection on facebook
F. D. Vigna, A. Cimino, Felice Dell’Orletta, M. Petrocchi, and M. Tesconi · 2017
Cited alongside, same era.
Neural domain adaptation for biomedical question answering
Georg Wiese, Dirk Weissenborn, and Mariana Neves · 2017
Cited alongside, same era.
Central moment discrepancy (cmd) for domain-invariant representation learning
W. Zellinger, Thomas Grubinger, E. Lughofer, T. Natschläger, and Susanne Saminger-Platz · 2017
Cited alongside, same era.
Visualizing deep neural network decisions: Prediction difference analysis
Luisa M. Zintgraf, T. Cohen, Tameem Adel, and M. Welling · 2017
Cited alongside, same era.
Hate speech dataset from a white supremacy forum
Ona de Gibert, Naiara Perez, Aitor García-Pablos, and Montse Cuadros · 2018
Cited alongside, same era.
Moment matching for multi-source domain adaptation
Xingchao Peng, Qinxun Bai, Xide Xia, Zijun Huang, Kate Saenko, and Bo Wang · 2019
Later among the works it cites.
Explanatory interactive machine learning
Stefano Teso and Kristian Kersting · 2019
Later among the works it cites.
Xisen Jin, Junyi Du, Zhongyu Wei, X. Xue, and Xiang Ren · 2020
Later among the works it cites.
Contextualizing hate speech classifiers with post-hoc explanation
Brendan Kennedy, Xisen Jin, Aida Mostafazadeh Davani, Morteza Dehghani, and Xiang Ren · 2020
Later among the works it cites.
Albert: A lite bert for self-supervised learning of language representations
Zhenzhong Lan, Mingda Chen, Sebastian Goodman, Kevin Gimpel, Piyush Sharma, and Radu Soricut · 2020
Later among the works it cites.
Find: human-in-the-loop debugging deep text classifiers
Piyawat Lertvittayakumjorn, Lucia Specia, and Francesca Toni · 2020
Later among the works it cites.
Model adaptation: Unsupervised domain adaptation without source data
Rui Li, Qianfen Jiao, Wenming Cao, H. Wong, and Si Wu · 2020
Later among the works it cites.
Do we really need to access the source data? source hypothesis transfer for unsupervised domain adaptation
Jian Liang, D. Hu, and Jiashi Feng · 2020
Later among the works it cites.
Regularizing black-box models for improved interpretability (hill 2019 version)
Gregory Plumb, Maruan Al-Shedivat, Eric Xing, and Ameet Talwalkar · 2020
Later among the works it cites.
Interpretations are useful: penalizing explanations to align neural networks with prior knowledge
Laura Rieger, Chandan Singh, William Murdoch, and Bin Yu · 2020
Later among the works it cites.
Learning from explanations with neural execution tree
Ziqi Wang, Yujia Qin, Wenxuan Zhou, Jun Yan, Qinyuan Ye, Leonardo Neves, Z. Liu, and X. Ren · 2020
Later among the works it cites.
Transformers: State-of-the-art natural language processing
Thomas Wolf, Lysandre Debut, Victor Sanh, Julien Chaumond, Clement Delangue, Anthony Moi, Pierric Cistac, Tim Rault, Remi Louf, Morgan Funtowicz, Joe Davison, Sam Shleifer, Patrick von Platen, Clara Ma, Yacine Jernite, Julien Plu, Canwen Xu, Teven Le Scao, Sylvain Gugger, Mariama Drame, Quentin Lhoest, and Alexander Rush · 2020
Later among the works it cites.
Teaching machine comprehension with compositional explanations
Qinyuan Ye, Xiao Huang, Elizabeth Boschee, and Xiang Ren · 2020
Later among the works it cites.