Fetching the paper…
Reading the bibliography…
Selecting input features of top relevance has become a popular method for building self-explaining models.
Sarthak Jain and Byron C. Wallace. 2019 · 1902
Earlier work this paper cites.
Roberta: A robustly optimized bert pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019 · 1907
Earlier work this paper cites.
Eraser: A benchmark to evaluate rationalized nlp models
Jay DeYoung, Sarthak Jain, Nazneen Fatema Rajani, Eric Lehman, Caiming Xiong, Richard Socher, and Byron C. Wallace. 2019 · 1911
Earlier work this paper cites.
On the transfer of masses (in russian)
Leonid Kantorovich. 1942 · 1942
Earlier work this paper cites.
Tres observaciones sobre el algebra lineal
Garrett Birkhoff. 1946 · 1946
Earlier work this paper cites.
Concerning nonnegative matrices and doubly stochastic matrices
Richard Sinkhorn and Paul Knopp. 1967 · 1967
Earlier work this paper cites.
Notes of the birkhoff algorithm for doubly stochastic matrices
Richard A. Brualdi. 1982 · 1982
Earlier work this paper cites.
Décomposition polaire et réarrangement monotone des champs de vecteurs
Yann Brenier. 1987 · 1987
Earlier work this paper cites.
Combinatorial Matrix Classes , volume 108
Richard A Brualdi. 2006 · 2006
Earlier work this paper cites.
Natural Language Processing with Python
Steven Bird, Edward Loper, and Ewan Klein. 2009 · 2009
Earlier work this paper cites.
Free boundaries in optimal transport and monge-ampère obstacle problems
Luis A. Caffarelli and Robert J. McCann. 2010 · 2010
Earlier work this paper cites.
The optimal partial transport problem
Alessio Figalli. 2010 · 2010
Earlier work this paper cites.
Sinkhorn distances: Lightspeed computation of optimal transport
Marco Cuturi. 2013 · 2013
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P. Kingma and Jimmy Ba. 2014 · 2014
Earlier work this paper cites.
Neural machine translation by jointly learning to align and translate
Dzmitry Bahdanau, Kyunghyun Cho, and Yoshua Bengio. 2015 · 2015
Earlier work this paper cites.
A large annotated corpus for learning natural language inference
Samuel R. Bowman, Gabor Angeli, Christopher Potts, and Christopher D. Manning. 2015 · 2015
Earlier work this paper cites.
Abc-cnn: An attention based convolutional neural network for visual question answering
Kan Chen, Jiang Wang, Liang-Chieh Chen, Haoyuan Gao, Wei Xu, and Ram Nevatia. 2015 · 2015
Earlier work this paper cites.
From word embeddings to document distances
Matt Kusner, Yu Sun, Nicholas Kolkin, and Kilian Weinberger. 2015 · 2015
Earlier work this paper cites.
Reasoning about entailment with neural attention
Tim Rocktäschel, Edward Grefenstette, Karl Moritz Hermann, Tomáš Kočiskỳ, and Phil Blunsom. 2015 · 2015
Earlier work this paper cites.
A neural attention model for abstractive sentence summarization
Alexander M. Rush, Sumit Chopra, and Jason Weston. 2015 · 2015
Earlier work this paper cites.
Learning hybrid representations to retrieve semantically equivalent questions
Cícero dos Santos, Luciano Barbosa, Dasha Bogdanova, and Bianca Zadrozny. 2015 · 2015
Cited alongside, same era.
Long short-term memory-networks for machine reading
Jianpeng Cheng, Li Dong, and Mirella Lapata. 2016 · 2016
Cited alongside, same era.
Rationalizing neural predictions
Tao Lei, Regina Barzilay, and Tommi Jaakkola. 2016 · 2016
Cited alongside, same era.
Understanding neural networks through representation erasure
Jiwei Li, Will Monroe, and Dan Jurafsky. 2016 · 2016
Cited alongside, same era.
From softmax to sparsemax: A sparse model of attention and multi-label classification
Andre Martins and Ramon Astudillo. 2016 · 2016
Cited alongside, same era.
A decomposable attention model for natural language inference
Simple recurrent units for highly parallelizable recurrence
Tao Lei, Yu Zhang, Sida I. Wang, Hui Dai, and Yoav Artzi. 2018 · 2018
Later among the works it cites.
Learning when to concentrate or divert attention: Self-adaptive attention temperature for neural machine translation
Junyang Lin, Xu Sun, Xuancheng Ren, Muyu Li, and Qi Su. 2018 · 2018
Later among the works it cites.
Sparse and constrained attention for neural machine translation
Chaitanya Malaviya, Pedro Ferreira, and André F. T. Martins. 2018 · 2018
Later among the works it cites.
Learning latent permutations with gumbel-sinkhorn networks
Gonzalo Mena, David Belanger, Scott Linderman, and Jasper Snoek. 2018 · 2018
Later among the works it cites.
Sparsemap: Differentiable sparse structured inference
Vlad Niculae, André F. T. Martins, Mathieu Blondel, and Claire Cardie. 2018 · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Ankur Parikh, Oscar Täckström, Dipanjan Das, and Jakob Uszkoreit. 2016 · 2016
Cited alongside, same era.
“why should i trust you?”: Explaining the predictions of any classifier
Marco Tulio Ribeiro, Sameer Singh, and Carlos Guestrin. 2016 · 2016
Cited alongside, same era.
Stabilized sparse scaling algorithms for entropy regularized transport problems
Bernhard Schmitzer. 2016 · 2016
Cited alongside, same era.
Rationale-augmented convolutional neural networks for text classification
Ye Zhang, Iain Marshall, and Byron C. Wallace. 2016 · 2016
Cited alongside, same era.
Enriching word vectors with subword information
Piotr Bojanowski, Edouard Grave, Armand Joulin, and Tomas Mikolov. 2017 · 2017
Cited alongside, same era.
A regularized framework for sparse and structured neural attention
Vlad Niculae and Mathieu Blondel. 2017 · 2017
Cited alongside, same era.
Automatic differentiation in pytorch
Adam Paszke, Sam Gross, Soumith Chintala, Gregory Chanan, Edward Yang, Zachary DeVito, Zeming Lin, Alban Desmaison, Luca Antiga, and Adam Lerer. 2017 · 2017
Cited alongside, same era.
Adversarial domain adaptation for duplicate question detection
Darsh Shah, Tao Lei, Alessandro Moschitti, Salvatore Romeo, and Preslav Nakov. 2018 · 2018
Later among the works it cites.
FEVER: a large-scale dataset for fact extraction and VERification
James Thorne, Andreas Vlachos, Christos Christodoulopoulos, and Arpit Mittal. 2018 · 2018
Later among the works it cites.
Interpretable neural predictions with differentiable binary variables
Joost Bastings, Wilker Aziz, and Ivan Titov. 2019 · 2019
Later among the works it cites.
A game theoretic approach to class-wise selective rationalization
Shiyu Chang, Yang Zhang, Mo Yu, and Tommi Jaakkola. 2019 · 2019
Later among the works it cites.
Multi-news: A large-scale multi-document summarization dataset and abstractive hierarchical model
Alexander Fabbri, Irene Li, Tianwei She, Suyi Li, and Dragomir Radev. 2019 · 2019
Later among the works it cites.
Latent retrieval for weakly supervised open domain question answering
Kenton Lee, Ming-Wei Chang, and Kristina Toutanova. 2019 · 2019
Later among the works it cites.
Inferring which medical treatments work from reports of clinical trials
Eric Lehman, Jay DeYoung, Regina Barzilay, and Byron C. Wallace. 2019 · 2019
Later among the works it cites.
CNM: An interpretable complex-valued network for matching
Qiuchi Li, Benyou Wang, and Massimo Melucci. 2019 · 2019
Later among the works it cites.
Dialog intent induction with deep multi-view clustering
Hugh Perkins and Yi Yang. 2019 · 2019
Later among the works it cites.
Computational optimal transport
Gabriel Peyré and Marco Cuturi. 2019 · 2019
Later among the works it cites.
Generating token-level explanations for natural language inference
James Thorne, Andreas Vlachos, Christos Christodoulopoulos, and Arpit Mittal. 2019 · 2019
Later among the works it cites.
Attention is not not explanation
Sarah Wiegreffe and Yuval Pinter. 2019 · 2019
Later among the works it cites.
Gromov-Wasserstein learning for graph matching and node embedding
Hongteng Xu, Dixin Luo, Hongyuan Zha, and Lawrence Carin Duke. 2019 · 2019
Later among the works it cites.
Rethinking cooperative rationalization: Introspective extraction and complement control
Mo Yu, Shiyu Chang, Yang Zhang, and Tommi Jaakkola. 2019 · 2019
Later among the works it cites.
Show, attend and tell: Neural image caption generation with visual attention
Kelvin Xu, Jimmy Lei Ba, Ryan Kiros, Kyunghyun Cho, Aaron Courville, Ruslan Salakhutdinov, Richard S. Zemel, and Yoshua Bengio. 2015 · 2057
Closest in time.