Fetching the paper…
Reading the bibliography…
While large language models (LLMs) have taken great strides towards helping humans with a plethora of tasks, hallucinations remain a major impediment towards gaining user trust.
“Captum: A unified and generic model interpretability library for PyTorch”, 2020
Narine Kokhlikyan et al · 2009
Earlier work this paper cites.
“Towards a rigorous science of interpretable machine learning”
Finale Doshi-Velez and Been Kim · 2017
Earlier work this paper cites.
“TriviaQA: A Large Scale Distantly Supervised Challenge Dataset for Reading Comprehension”
Mandar Joshi, Eunsol Choi, Daniel Weld and Luke Zettlemoyer · 2017
Earlier work this paper cites.
“A unified approach to interpreting model predictions”
Scott Lundberg and Su-In Lee · 2017
Earlier work this paper cites.
“Axiomatic attribution for deep networks”
Mukund Sundararajan, Ankur Taly and Qiqi Yan · 2017
Earlier work this paper cites.
“Why Is Google Translate Spitting Out Sinister Religious Prophecies?” Accessed: 2023-08-16, https://www.vice.com/en/article/j5npeg/why-is-google-translate-spitting-out-sinister-religious-prophecies , 2018
Jon Christian · 2018
Earlier work this paper cites.
“T-REx: A Large Scale Alignment of Natural Language with Knowledge Base Triples”
Hady Elsahar et al · 2018
Earlier work this paper cites.
“Explaining explanations: An overview of interpretability of machine learning”
Leilani Gilpin et al · 2018
Earlier work this paper cites.
“A survey of methods for explaining black box models”
Riccardo Guidotti et al · 2018
Earlier work this paper cites.
“The mythos of model interpretability: In machine learning, the concept of interpretability is both important and slippery.”
Zachary Lipton · 2018
Earlier work this paper cites.
“Layer-wise relevance propagation: an overview”
Grégoire Montavon et al · 2019
Earlier work this paper cites.
“Language Models as Knowledge Bases?”
Fabio Petroni et al · 2019
Earlier work this paper cites.
“Unsupervised quality estimation for neural machine translation”
Marina Fomicheva et al · 2020
Earlier work this paper cites.
“Retrieval-augmented generation for knowledge-intensive nlp tasks”
Patrick Lewis et al · 2020
Earlier work this paper cites.
“Open-domain conversational agents: Current progress, open problems, and future directions”
Stephen Roller et al · 2020
Cited alongside, same era.
Artidoro Pagnoni, Vidhisha Balachandran and Yulia Tsvetkov · 2021
Cited alongside, same era.
“Amazon SageMaker Debugger: A system for real-time insights into machine learning model training”
Nathalie Rauschmayr et al · 2021
Cited alongside, same era.
“Analyzing the Source and Target Contributions to Predictions in Neural Machine Translation”
Elena Voita, Rico Sennrich and Ivan Titov · 2021
Cited alongside, same era.
“On the lack of robust interpretability of neural text classifiers”
Muhammad Zafar et al · 2021
“Detecting and Mitigating Hallucinations in Machine Translation: Model Internal Workings Alone Do Well, Sentence Similarity Even Better”
David Dale, Elena Voita, Loic Barrault and Marta. Costa-jussà · 2023
Closest in time.
“Looking for a Needle in a Haystack: A Comprehensive Study of Hallucinations in Neural Machine Translation”
Nuno. Guerreiro, Elena Voita and André Martins · 2023
Closest in time.
Lei Huang et al · 2023
Closest in time.
“Survey of hallucination in natural language generation”
Ziwei Ji et al · 2023
Closest in time.
“Evaluating Open-Domain Question Answering in the Era of Large Language Models”
Ehsan Kamalloo, Nouha Dziri, Charles Clarke and Davood Rafiei · 2023
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
“Ist-unbabel 2021 submission for the quality estimation shared task”
Chrysoula Zerva et al · 2021
Cited alongside, same era.
“Towards Opening the Black Box of Neural Machine Translation: Source and Target Interpretations of the Transformer”
Javier Ferrando et al · 2022
Cited alongside, same era.
“Language models (mostly) know what they know”
Saurav Kadavath et al · 2022
Cited alongside, same era.
“Holistic evaluation of language models”
Percy Liang et al · 2022
Cited alongside, same era.
“TruthfulQA: Measuring How Models Mimic Human Falsehoods”
Stephanie Lin, Jacob Hilton and Owain Evans · 2022
Cited alongside, same era.
“Locating and editing factual associations in GPT”
Kevin Meng, David Bau, Alex Andonian and Yonatan Belinkov · 2022
Cited alongside, same era.
“Training language models to follow instructions with human feedback”
Long Ouyang et al · 2022
Cited alongside, same era.
“SelfCheckGPT: Zero-Resource Black-Box Hallucination Detection for Generative Large Language Models”
Potsawee Manakul, Adian Liusie and Mark Gales · 2023
Closest in time.
“Need a Last Minute Mother’s Day Gift? AI Is Here to Help.” Accessed: 2023-10-11, https://about.you.com/need-a-last-minute-mothers-day-gift-ai-is-here-to-help-d363b17e76b4/ , 2023
2023
Closest in time.
OpenAI · 2023
Closest in time.
“An important next step on our AI journey” Accessed: 2023-10-11, https://blog.google/technology/ai/bard-google-ai-search-updates/ , 2023
Sundar Pichai · 2023
Closest in time.
“Alpaca: A strong, replicable instruction-following model”
Rohan Taori et al · 2023
Closest in time.
“Welcome to the era of chatgpt et al. the prospects of large language models”
Timm Teubner et al · 2023
Closest in time.
“How language model hallucinations can snowball”
Muru Zhang et al · 2023
Closest in time.
“Evaluating Correctness and Faithfulness of Instruction-Following Models for Question Answering”
Vaibhav Adlakha et al · 2024
Closest in time.
Qinyuan Wu et al · 2024
Closest in time.