Fetching the paper…
Reading the bibliography…
Machine Learning (ML) models are increasingly used to make critical decisions in real-world applications, yet they have become more complex, making them harder to understand.
Sentence-bert: Sentence embeddings using siamese bert-networks
N. Reimers and I. Gurevych · 1908
Earlier work this paper cites.
A survey on dialogue systems: Recent advances and new frontiers
H. Chen, X. Liu, D. Yin, and J. Tang · 1931
Earlier work this paper cites.
Providing a unified account of definite noun phrases in discourse
B. J. Grosz, A. K. Joshi, and S. Weinstein · 1983
Earlier work this paper cites.
Generating patient-specific interactive natural language explanations
G. Carenini, V. O. Mittal, and J. D. Moore · 1994
Earlier work this paper cites.
Psychological aspects of natural language. use: our words, our selves
J. W. Pennebaker, M. R. Mehl, and K. G. Niederhoffer · 2002
Earlier work this paper cites.
Radar: A personal assistant that learns to reduce email overload
M. Freed, J. Carbonell, G. Gordon, J. Hayes, B. A. Myers, D. Siewiorek, S. Smith, A. Steinfeld, and A. Tomasic · 2008
Earlier work this paper cites.
Toward establishing trust in adaptive agents
A. Glass, D. L. McGuinness, and M. Wolverton · 2008
Earlier work this paper cites.
Scikit-learn: Machine learning in Python
F. Pedregosa, G. Varoquaux, A. Gramfort, V. Michel, B. Thirion, O. Grisel, M. Blondel, P. Prettenhofer, R. Weiss, V. Dubourg, J. Vanderplas, A. Passos, D. Cournapeau, M. Brucher, M. Perrot, and E. Duchesnay · 2011
Earlier work this paper cites.
Developing Dialogue Managers from Limited Amounts of Data , pages 5–17
V. Rieser and O. Lemon · 2012
Earlier work this paper cites.
Accurate intelligible models with pairwise interactions
Y. Lou, R. Caruana, J. Gehrke, and G. Hooker · 2013
Earlier work this paper cites.
Deep inside convolutional networks: Visualising image classification models and saliency maps
K. Simonyan, A. Vedaldi, and A. Zisserman · 2014
Earlier work this paper cites.
Sequence to sequence learning with neural networks
I. Sutskever, O. Vinyals, and Q. V. Le · 2014
Earlier work this paper cites.
Ten challenges in highly-interactive dialog systems
N. G. Ward and D. DeVault · 2015
Earlier work this paper cites.
Interpretable classification models for recidivism prediction
J. Zeng, B. Ustun, and C. Rudin · 2015
Earlier work this paper cites.
Machine bias
J. Angwin, J. Larson, S. Mattu, and L. Kirchner · 2016
Earlier work this paper cites.
Interpretable decision sets: A joint framework for description and prediction
H. Lakkaraju, S. H. Bach, and J. Leskovec · 2016
Earlier work this paper cites.
"why should I trust you?": Explaining the predictions of any classifier
M. T. Ribeiro, S. Singh, and C. Guestrin · 2016
Earlier work this paper cites.
Supersparse linear integer models for optimized medical scoring systems
B. Ustun and C. Rudin · 2016
Earlier work this paper cites.
Learning certifiably optimal rule lists for categorical data
E. Angelino, N. Larus-Stone, D. Alabi, M. Seltzer, and C. Rudin · 2017
Earlier work this paper cites.
UCI machine learning repository, 2017
D. Dua and C. Graff · 2017
Earlier work this paper cites.
End-to-end task-completion neural dialogue systems
X. Li, Y.-N. Chen, L. Li, J. Gao, and A. Celikyilmaz · 2017
Earlier work this paper cites.
Using context information for dialog act classification in DNN framework
Y. Liu, K. Han, Z. Tan, and Y. Lei · 2017
Earlier work this paper cites.
A unified approach to interpreting model predictions
S. M. Lundberg and S.-I. Lee · 2017
Earlier work this paper cites.
Prolific.ac—a subject pool for online experiments
S. Palan and C. Schitter · 2017
Earlier work this paper cites.
Grad-cam: Visual explanations from deep networks via gradient-based localization
R. R. Selvaraju, M. Cogswell, A. Das, R. Vedantam, D. Parikh, and D. Batra · 2017
Earlier work this paper cites.
Generating high-quality and informative conversation responses with sequence-to-sequence models
Y. Shao, S. Gouws, D. Britz, A. Goldie, B. Strope, and R. Kurzweil · 2017
Earlier work this paper cites.
Smoothgrad: removing noise by adding noise
D. Smilkov, N. Thorat, B. Kim, F. Viégas, and M. Wattenberg · 2017
Cited alongside, same era.
Axiomatic attribution for deep networks
M. Sundararajan, A. Taly, and Q. Yan · 2017
Cited alongside, same era.
Evaluating semantic parsing against a simple web-based question answering model
A. Talmor, M. Geva, and J. Berant · 2017
Cited alongside, same era.
On the robustness of interpretability methods
D. Alvarez-Melis and T. S. Jaakkola · 2018
Cited alongside, same era.
Neural approaches to conversational AI
J. Gao, M. Galley, and L. Li · 2018
Cited alongside, same era.
A simple and effective model-based variable importance measure
B. M. Greenwell, B. C. Boehmke, and A. J. McCarthy · 2018
Cited alongside, same era.
From local explanations to global understanding with explainable ai for trees
S. M. Lundberg, G. Erion, H. Chen, A. DeGrave, J. M. Prutkin, B. Nair, R. Katz, J. Himmelfarb, N. Bansal, and S.-I. Lee · 2020
Later among the works it cites.
Explaining machine learning classifiers through diverse counterfactual explanations
R. K. Mothilal, A. Sharma, and C. Tan · 2020
Later among the works it cites.
Improving compositional generalization in semantic parsing
I. Oren, J. Herzig, N. Gupta, M. Gardner, and J. Berant · 2020
Later among the works it cites.
Exploring the limits of transfer learning with a unified text-to-text transformer
C. Raffel, N. Shazeer, A. Roberts, K. Lee, S. Narang, M. Matena, Y. Zhou, W. Li, and P. J. Liu · 2020
Later among the works it cites.
The language interpretability tool: Extensible, interactive visualizations and analysis for NLP models, 2020
I. Tenney, J. Wexler, J. Bastings, T. Bolukbasi, A. Coenen, S. Gehrmann, E. Jiang, M. Pushkarna, C. Radebaugh, E. Reif, and A. Yuan · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Dialog-to-action: Conversational question answering over a large-scale knowledge base
D. Guo, D. Tang, N. Duan, M. Zhou, and J. Yin · 2018
Cited alongside, same era.
Model agnostic supervised local explanations
G. Plumb, D. Molitor, and A. Talwalkar · 2018
Cited alongside, same era.
Anchors: High-precision model-agnostic explanations
M. T. Ribeiro, S. Singh, and C. Guestrin · 2018
Cited alongside, same era.
Glass-box: Explaining ai decisions with counterfactual statements through conversation with a voice-enabled virtual assistant
K. Sokol and P. A. Flach · 2018
Cited alongside, same era.
Spider: A large-scale human-labeled dataset for complex and cross-domain semantic parsing and text-to-sql task
T. Yu, R. Zhang, K. Yang, M. Yasunaga, D. Wang, Z. Li, J. Ma, I. Li, Q. Yao, S. Roman, Z. Zhang, and D. Radev · 2018
Cited alongside, same era.
L-shapley and c-shapley: Efficient model interpretation for structured data
J. Chen, L. Song, M. J. Wainwright, and M. I. Jordan · 2019
Cited alongside, same era.
Recent advances and challenges in task-oriented dialog systems
Z. Zhang, R. Takanobu, Q. Zhu, M. Huang, and X. Zhu · 2020
Later among the works it cites.
Neural additive models: Interpretable machine learning with neural nets
R. Agarwal, L. Melnick, N. Frosst, X. Zhang, B. Lengerich, R. Caruana, and G. E. Hinton · 2021
Later among the works it cites.
How interpretable and trustworthy are gams?
C.-H. Chang, S. Tan, B. Lengerich, A. Goldenberg, and R. Caruana · 2021
Later among the works it cites.
The out-of-distribution problem in explainability and search methods for feature importance explanations
P. Hase, H. Xie, and M. Bansal · 2021
Later among the works it cites.
Constrained language models yield few-shot semantic parsers
R. Shin, C. H. Lin, S. Thomson, C. Chen, S. Roy, E. A. Platanios, A. Pauls, D. Klein, J. Eisner, and B. V. Durme · 2021
Later among the works it cites.
Reliable post hoc explanations: Modeling uncertainty in explainability
D. Slack, A. Hilgard, S. Singh, and H. Lakkaraju · 2021
Later among the works it cites.
CREAD: Combined resolution of ellipses and anaphora in dialogues
B.-H. Tseng, S. Bhargava, J. Lu, J. R. A. Moniz, D. Piraviperumal, L. Li, and H. Yu · 2021
Later among the works it cites.
GPT-J-6B: A 6 Billion Parameter Autoregressive Language Model
B. Wang and A. Komatsuzaki · 2021
Later among the works it cites.
Compositional generalization for neural semantic parsing via span-level supervised attention
P. Yin, H. Fang, G. Neubig, A. Pauls, E. A. Platanios, Y. Su, S. Thomson, and J. Andreas · 2021
Later among the works it cites.
Calibrate before use: Improving few-shot performance of language models
T. Z. Zhao, E. Wallace, S. Feng, D. Klein, and S. Singh · 2021
Later among the works it cites.
Node-gam: Neural generalized additive model for interpretable deep learning
C.-H. Chang, R. Caruana, and A. Goldenberg · 2022
Closest in time.
Hint: Integration testing for ai-based features with humans in the loop
Q. Chen, T. Schnabel, B. Nushi, and S. Amershi · 2022
Closest in time.
oegedijk/explainerdashboard: v0.3.8.2: reverses set_shap_values bug introduced in 0.3.8.1
O. Dijk, oegesam, R. Bell, Lily, Simon-Free, B. Serna, rajgupt, yanhong-zhao ef, A. Gädke, Hugo, and T. Okumus · 2022
Closest in time.
Mediators: Conversational agents explaining nlp model behavior
N. Feldhus, A. M. Ravichandran, and S. Möller · 2022
Closest in time.
Structurally diverse sampling reduces spurious correlations in semantic parsing datasets
S. Gupta, S. Singh, and M. Gardner · 2022
Closest in time.
The disagreement problem in explainable machine learning: A practitioner’s perspective
S. Krishna, T. Han, A. Gu, J. Pombra, S. Jabbari, S. Wu, and H. Lakkaraju · 2022
Closest in time.
Rethinking explainability as a dialogue: A practitioner’s perspective
H. Lakkaraju, D. Slack, Y. Chen, C. Tan, and S. Sing · 2022
Closest in time.
Rethinking the role of demonstrations: What makes in-context learning work?
S. Min, X. Lyu, A. Holtzman, M. Artetxe, M. Lewis, H. Hajishirzi, and L. Zettlemoyer · 2022
Closest in time.
Emergent abilities of large language models
J. Wei, Y. Tay, R. Bommasani, C. Raffel, B. Zoph, S. Borgeaud, D. Yogatama, M. Bosma, D. Zhou, D. Metzler, E. H. Chi, T. Hashimoto, O. Vinyals, P. Liang, J. Dean, and W. Fedus · 2022
Closest in time.
An explanation of in-context learning as implicit bayesian inference
S. M. Xie, A. Raghunathan, P. Liang, and T. Ma · 2022
Closest in time.
Interpretability and fairness evaluation of deep learning models on mimic-iv dataset
C. Meng, L. Trinh, N. Xu, J. Enouen, and Y. Liu · 2045
Closest in time.