Fetching the paper…
Reading the bibliography…
Understanding predictions made by deep neural networks is notoriously difficult, but also crucial to their dissemination.
Language model pre-training for hierarchical document representations
Chang, Ming-Wei, Kristina Toutanova, Kenton Lee, and Jacob Devlin. 2019 · 1901
Earlier work this paper cites.
Using text embeddings for causal inference
Veitch, Victor, Dhanya Sridhar, and David M. Blei. 2019 · 1905
Earlier work this paper cites.
Explaining classifiers with causal concept effect (cace)
Goyal, Yash, Uri Shalit, and Been Kim. 2019 · 1907
Earlier work this paper cites.
Roberta: A robustly optimized BERT pretraining approach
Liu, Yinhan, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019 · 1907
Earlier work this paper cites.
Huggingface’s transformers: State-of-the-art natural language processing
Wolf, Thomas, Lysandre Debut, Victor Sanh, Julien Chaumond, Clement Delangue, Anthony Moi, Pierric Cistac, Tim Rault, Rémi Louf, Morgan Funtowicz, and Jamie Brew. 2019 · 1910
Earlier work this paper cites.
Causal diagrams for empirical research
Pearl, Judea. 1995 · 1995
Earlier work this paper cites.
Linguistic inquiry and word count: Liwc 2001
Pennebaker, James W, Martha E Francis, and Roger J Booth. 2001 · 2001
Earlier work this paper cites.
Thumbs up? sentiment classification using machine learning techniques
Pang, Bo, Lillian Lee, and Shivakumar Vaithyanathan. 2002 · 2002
Earlier work this paper cites.
Latent dirichlet allocation
Blei, David M., Andrew Y. Ng, and Michael I. Jordan. 2003 · 2003
Earlier work this paper cites.
Making things happen: A theory of causal explanation
Woodward, James. 2005 · 2005
Earlier work this paper cites.
Biographies, bollywood, boom-boxes and blenders: Domain adaptation for sentiment classification
Blitzer, John, Mark Dredze, and Fernando Pereira. 2007 · 2007
Earlier work this paper cites.
Mostly harmless econometrics: An empiricist’s companion
Angrist, Joshua D and Jörn-Steffen Pischke. 2008 · 2008
Earlier work this paper cites.
Software framework for topic modelling with large corpora
Řehůřek, Radim and Petr Sojka. 2010 · 2010
Earlier work this paper cites.
Learning word vectors for sentiment analysis
Maas, Andrew L., Raymond E. Daly, Peter T. Pham, Dan Huang, Andrew Y. Ng, and Christopher Potts. 2011 · 2011
Earlier work this paper cites.
Counterfactual reasoning and learning systems: the example of computational advertising
Bottou, Léon, Jonas Peters, Joaquin Quiñonero Candela, Denis Xavier Charles, Max Chickering, Elon Portugaly, Dipankar Ray, Patrice Y. Simard, and Ed Snelson. 2013 · 2013
Earlier work this paper cites.
Machine reading tea leaves: Automatically evaluating topic coherence and topic model quality
Lau, Jey Han, David Newman, and Timothy Baldwin. 2014 · 2014
Earlier work this paper cites.
Structural topic models for open-ended survey responses
Roberts, Margaret E, Brandon M Stewart, Dustin Tingley, Christopher Lucas, Jetson Leder-Luis, Shana Kushner Gadarian, Bethany Albertson, and David G Rand. 2014 · 2014
Earlier work this paper cites.
The effect of wording on message propagation: Topic- and author-controlled natural experiments on twitter
Tan, Chenhao, Lillian Lee, and Bo Pang. 2014 · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization
Kingma, Diederik P. and Jimmy Ba. 2015 · 2015
Earlier work this paper cites.
Aligning books and movies: Towards story-like visual explanations by watching movies and reading books
Zhu, Yukun, Ryan Kiros, Richard S. Zemel, Ruslan Salakhutdinov, Raquel Urtasun, Antonio Torralba, and Sanja Fidler. 2015 · 2015
Earlier work this paper cites.
Discovery of treatments from text corpora
Fong, Christian and Justin Grimmer. 2016 · 2016
Earlier work this paper cites.
Domain-adversarial training of neural networks
Ganin, Yaroslav, Evgeniya Ustinova, Hana Ajakan, Pascal Germain, Hugo Larochelle, François Laviolette, Mario Marchand, and Victor S. Lempitsky. 2016 · 2016
Earlier work this paper cites.
Learning representations for counterfactual inference
Johansson, Fredrik D., Uri Shalit, and David A. Sontag. 2016 · 2016
Earlier work this paper cites.
Examples are not enough, learn to criticize! criticism for interpretability
Kim, Been, Oluwasanmi Koyejo, and Rajiv Khanna. 2016a · 2016
Earlier work this paper cites.
Examples are not enough, learn to criticize! criticism for interpretability
Kim, Been, Oluwasanmi Koyejo, and Rajiv Khanna. 2016b · 2016
Earlier work this paper cites.
"why should I trust you?": Explaining the predictions of any classifier
Ribeiro, Marco Túlio, Sameer Singh, and Carlos Guestrin. 2016 · 2016
Earlier work this paper cites.
Rethinking the inception architecture for computer vision
Szegedy, Christian, Vincent Vanhoucke, Sergey Ioffe, Jonathon Shlens, and Zbigniew Wojna. 2016 · 2016
Earlier work this paper cites.
Google’s neural machine translation system: Bridging the gap between human and machine translation
Wu, Yonghui, Mike Schuster, Zhifeng Chen, Quoc V. Le, Mohammad Norouzi, Wolfgang Macherey, Maxim Krikun, Yuan Cao, Qin Gao, Klaus Macherey, Jeff Klingner, Apurva Shah, Melvin Johnson, Xiaobing Liu, Lukasz Kaiser, Stephan Gouws, Yoshikiyo Kato, Taku Kudo, Hideto Kazawa, Keith Stevens, George Kurian, Nishant Patil, Wei Wang, Cliff Young, Jason Smith, Jason Riesa, Alex Rudnick, Oriol Vinyals, Greg Corrado, Macduff Hughes, and Jeffrey Dean. 2016 · 2016
Earlier work this paper cites.
Fine-grained analysis of sentence embeddings using auxiliary prediction tasks
Adi, Yossi, Einat Kermany, Yonatan Belinkov, Ofer Lavi, and Yoav Goldberg. 2017 · 2017
Earlier work this paper cites.
Estimating average treatment effects: Supplementary analyses and remaining challenges
Athey, Susan, Guido Imbens, Thai Pham, and Stefan Wager. 2017 · 2017
Earlier work this paper cites.
Applications of topic models
Boyd-Graber, Jordan L., Yuening Hu, and David M. Mimno. 2017 · 2017
Earlier work this paper cites.
Semantics derived automatically from language corpora contain human-like biases
Caliskan, Aylin, Joanna J Bryson, and Arvind Narayanan. 2017 · 2017
Earlier work this paper cites.
Maximum-likelihood augmented discrete generative adversarial networks
Che, Tong, Yanran Li, Ruixiang Zhang, R. Devon Hjelm, Wenjie Li, Yangqiu Song, and Yoshua Bengio. 2017 · 2017
Earlier work this paper cites.
Overlap in observational studies with high-dimensional covariates
D’Amour, Alexander, Peng Ding, Avi Feller, Lihua Lei, and Jasjeet Sekhon. 2017 · 2017
Earlier work this paper cites.
Towards a rigorous science of interpretable machine learning
Doshi-Velez, Finale and Been Kim. 2017 · 2017
Cited alongside, same era.
A challenge set approach to evaluating machine translation
Isabelle, Pierre, Colin Cherry, and George F. Foster. 2017 · 2017
Cited alongside, same era.
Discourse-based objectives for fast unsupervised sentence representation learning
Jernite, Yacine, Samuel R. Bowman, and David A. Sontag. 2017 · 2017
Cited alongside, same era.
Adversarial ranking for language generation
Lin, Kevin, Dianqi Li, Xiaodong He, Ming-Ting Sun, and Zhengyou Zhang. 2017 · 2017
Cited alongside, same era.
A unified approach to interpreting model predictions
Lundberg, Scott M. and Su-In Lee. 2017 · 2017
Cited alongside, same era.
Automatic differentiation in pytorch
Paszke, Adam, Sam Gross, Soumith Chintala, Gregory Chanan, Edward Yang, Zachary DeVito, Zeming Lin, Alban Desmaison, Luca Antiga, and Adam Lerer. 2017 · 2017
Unified language model pre-training for natural language understanding and generation
Dong, Li, Nan Yang, Wenhui Wang, Furu Wei, Xiaodong Liu, Yu Wang, Jianfeng Gao, Ming Zhou, and Hsiao-Wuen Hon. 2019 · 2019
Later among the works it cites.
Automated versus do-it-yourself methods for causal inference: Lessons learned from a data analysis competition
Dorie, Vincent, Jennifer Hill, Uri Shalit, Marc Scott, Dan Cervone, et al. 2019 · 2019
Later among the works it cites.
Pytorch lightning
Falcon, WA and .al. 2019 · 2019
Later among the works it cites.
The case for evaluating causal models using interventional measures and empirical data
Gentzel, Amanda, Dan Garant, and David Jensen. 2019 · 2019
Later among the works it cites.
Lipstick on a pig: Debiasing methods cover up systematic gender biases in word embeddings but do not remove them
Gonen, Hila and Yoav Goldberg. 2019 · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Social bias in elicited natural language inferences
Rudinger, Rachel, Chandler May, and Benjamin Van Durme. 2017 · 2017
Cited alongside, same era.
A hybrid convolutional variational autoencoder for text generation
Semeniuta, Stanislau, Aliaksei Severyn, and Erhardt Barth. 2017 · 2017
Cited alongside, same era.
How grammatical is character-level neural machine translation? assessing MT quality with contrastive translation pairs
Sennrich, Rico. 2017 · 2017
Cited alongside, same era.
Adversarial generation of natural language
Subramanian, Sandeep, Sai Rajeswar, Francis Dutil, Chris Pal, and Aaron C. Courville. 2017 · 2017
Cited alongside, same era.
Attention is all you need
Vaswani, Ashish, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N. Gomez, Lukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Cited alongside, same era.
Controllable invariance through adversarial feature learning
Xie, Qizhe, Zihang Dai, Yulun Du, Eduard H. Hovy, and Graham Neubig. 2017 · 2017
Cited alongside, same era.
Goyal, Yash, Ziyan Wu, Jan Ernst, Dhruv Batra, Devi Parikh, and Stefan Lee. 2019 · 2019
Later among the works it cites.
Attention is not explanation
Jain, Sarthak and Byron C. Wallace. 2019 · 2019
Later among the works it cites.
Proceedings of the 2019 ACL Workshop BlackboxNLP: Analyzing and Interpreting Neural Networks for NLP . Association for Computational Linguistics, Florence, Italy
Linzen, Tal, Grzegorz Chrupała, Yonatan Belinkov, and Dieuwke Hupkes, editors. 2019 · 2019
Later among the works it cites.
Direct optimization through arg max for discrete variational auto-encoder
Lorberbom, Guy, Tommi S. Jaakkola, Andreea Gane, and Tamir Hazan. 2019 · 2019
Later among the works it cites.
Exploring numeracy in word embeddings
Naik, Aakanksha, Abhilasha Ravichander, Carolyn Penstein Rosé, and Eduard H. Hovy. 2019 · 2019
Later among the works it cites.
Language models are unsupervised multitask learners
Radford, Alec, Jeffrey Wu, Rewon Child, David Luan, Dario Amodei, and Ilya Sutskever. 2019 · 2019
Later among the works it cites.
Deep contextualized self-training for low resource dependency parsing
Rotman, Guy and Roi Reichart. 2019 · 2019
Later among the works it cites.
A social media study on the effects of psychiatric medication use
Saha, Koustuv, Benjamin Sugar, John Torous, Bruno D. Abrahao, Emre Kiciman, and Munmun De Choudhury. 2019 · 2019
Later among the works it cites.
How to fine-tune BERT for text classification?
Sun, Chi, Xipeng Qiu, Yige Xu, and Xuanjing Huang. 2019 · 2019
Later among the works it cites.
What are the biases in my word embedding?
Swinger, Nathaniel, Maria De-Arteaga, Neil Thomas Heffernan IV, Mark D. M. Leiserson, and Adam Tauman Kalai. 2019 · 2019
Later among the works it cites.
Unsupervised word embeddings capture latent knowledge from materials science literature
Tshitoyan, Vahe, John Dagdelen, Leigh Weston, Alexander Dunn, Ziqin Rong, Olga Kononova, Kristin A. Persson, Gerbrand Ceder, and Anubhav Jain. 2019 · 2019
Later among the works it cites.
Attention is not not explanation
Wiegreffe, Sarah and Yuval Pinter. 2019 · 2019
Later among the works it cites.
Xlnet: Generalized autoregressive pretraining for language understanding
Yang, Zhilin, Zihang Dai, Yiming Yang, Jaime G. Carbonell, Ruslan Salakhutdinov, and Quoc V. Le. 2019 · 2019
Later among the works it cites.
Active deep learning to detect demographic traits in free-form clinical notes
Feder, Amir, Danny Vainstein, Roni Rosenfeld, Tzvika Hartman, Avinatan Hassidim, and Yossi Matias. 2020 · 2020
Closest in time.
Evaluating models’ local decision boundaries via contrast sets
Gardner, Matt, Yoav Artzi, Victoria Basmova, Jonathan Berant, Ben Bogin, Sihao Chen, Pradeep Dasigi, Dheeru Dua, Yanai Elazar, Ananth Gottumukkala, Nitish Gupta, Hannaneh Hajishirzi, Gabriel Ilharco, Daniel Khashabi, Kevin Lin, Jiangming Liu, Nelson F. Liu, Phoebe Mulcaire, Qiang Ning, Sameer Singh, Noah A. Smith, Sanjay Subramanian, Reut Tsarfaty, Eric Wallace, Ally Zhang, and Ben Zhou. 2020 · 2020
Closest in time.
Visual interaction with deep learning models through collaborative semantic inference
Gehrmann, Sebastian, Hendrik Strobelt, Robert Krüger, Hanspeter Pfister, and Alexander M. Rush. 2020 · 2020
Closest in time.
Kingdom: Knowledge-guided domain adaptation for sentiment analysis
Ghosal, Deepanway, Devamanyu Hazarika, Abhinaba Roy, Navonil Majumder, Rada Mihalcea, and Soujanya Poria. 2020 · 2020
Closest in time.
Don’t stop pretraining: Adapt language models to domains and tasks
Gururangan, Suchin, Ana Marasovic, Swabha Swayamdipta, Kyle Lo, Iz Beltagy, Doug Downey, and Noah A. Smith. 2020 · 2020
Closest in time.
Reducing sentiment bias in language models via counterfactual evaluation
Huang, Po-Sen, Huan Zhang, Ray Jiang, Robert Stanforth, Johannes Welbl, Jack Rae, Vishal Maini, Dani Yogatama, and Pushmeet Kohli. 2020 · 2020
Closest in time.
Learning the difference that makes A difference with counterfactually-augmented data
Kaushik, Divyansh, Eduard H. Hovy, and Zachary Chase Lipton. 2020 · 2020
Closest in time.
Text and causal inference: A review of using text to remove confounding from causal estimates
Keith, Katherine A., David Jensen, and Brendan O’Connor. 2020 · 2020
Closest in time.
Biobert: a pre-trained biomedical language representation model for biomedical text mining
Lee, Jinhyuk, Wonjin Yoon, Sungdong Kim, Donghyeon Kim, Sunkyu Kim, Chan Ho So, and Jaewoo Kang. 2020 · 2020
Closest in time.
Predicting in-game actions from interviews of NBA players
Oved, Nadav, Amir Feder, and Roi Reichart. 2020 · 2020
Closest in time.
Neural unsupervised domain adaptation in NLP - A survey
Ramponi, Alan and Barbara Plank. 2020 · 2020
Closest in time.
Null it out: Guarding protected attributes by iterative nullspace projection
Ravfogel, Shauli, Yanai Elazar, Hila Gonen, Michael Twiton, and Yoav Goldberg. 2020 · 2020
Closest in time.
Adjusting for confounding with text matching
Roberts, M. E., Brandon M Stewart, and Richard A. Nielsen. 2020 · 2020
Closest in time.
Investigating gender bias in language models using causal mediation analysis
Vig, Jesse, Sebastian Gehrmann, Yonatan Belinkov, Sharon Qian, Daniel Nevo, Yaron Singer, and Stuart M. Shieber. 2020 · 2020
Closest in time.
Discovering shifts to suicidal ideation from mental health content in social media
Choudhury, Munmun De, Emre Kiciman, Mark Dredze, Glen Coppersmith, and Mrinal Kumar. 2016 · 2098
Closest in time.