Fetching the paper…
Reading the bibliography…
Retrieval augmented generation (RAG) is frequently used to mitigate hallucinations and provide up-to-date knowledge for large language models (LLMs).
Retrieval-augmented generation for knowledge-intensive nlp tasks
Patrick Lewis, Ethan Perez, Aleksandra Piktus, Fabio Petroni, Vladimir Karpukhin, Naman Goyal, Heinrich Küttler, Mike Lewis, Wen-Tau Yih, Tim Rocktäschel, and Others · 2020
Earlier work this paper cites.
Generation-Augmented retrieval for open-domain question answering
Yuning Mao, Pengcheng He, Xiaodong Liu, Yelong Shen, Jianfeng Gao, Jiawei Han, and Weizhu Chen · 2020
Earlier work this paper cites.
Entity-based knowledge conflicts in question answering
Shayne Longpre, Kartik Perisetla, Anthony Chen, Nikhil Ramesh, Chris DuBois, and Sameer Singh · 2021
Earlier work this paper cites.
Retrieval augmentation reduces hallucination in conversation
Kurt Shuster, Spencer Poff, Moya Chen, Douwe Kiela, and Jason Weston · 2021
Earlier work this paper cites.
Creating trustworthy LLMs: Dealing with hallucinations in healthcare AI
Muhammad Aurangzeb Ahmad, Ilker Yaramis, and Taposh Dutta Roy · 2023
Earlier work this paper cites.
Evaluation of GPT-3.5 and GPT-4 for supporting real-world information needs in healthcare delivery
Debadutta Dash, Rahul Thapa, Juan M Banda, Akshay Swaminathan, Morgan Cheatham, Mehr Kashyap, Nikesh Kotecha, Jonathan H Chen, Saurabh Gombar, Lance Downing, Rachel Pedreira, Ethan Goh, Angel Arnaout, Garret Kenn Morris, Honor Magon, Matthew P Lungren, Eric Horvitz, and Nigam H Shah · 2023
Earlier work this paper cites.
Gemini: A family of highly capable multimodal models
Gemini Team · 2023
Earlier work this paper cites.
RaLLe: A framework for developing and evaluating Retrieval-Augmented large language models
Yasuto Hoshi, Daisuke Miyashita, Youyang Ng, Kento Tatsuno, Yasuhiro Morioka, Osamu Torii, and Jun Deguchi · 2023
Earlier work this paper cites.
Survey of hallucination in natural language generation
Ziwei Ji, Nayeon Lee, Rita Frieske, Tiezheng Yu, Dan Su, Yan Xu, Etsuko Ishii, Ye Jin Bang, Andrea Madotto, and Pascale Fung · 2023
Earlier work this paper cites.
Challenges and applications of large language models
Jean Kaddour, Joshua Harris, Maximilian Mozes, Herbie Bradley, Roberta Raileanu, and Robert McHardy · 2023
Cited alongside, same era.
Haoqiang Kang, Juntong Ni, and Huaxiu Yao · 2023
Cited alongside, same era.
DetectGPT: Zero-shot machine-generated text detection using probability curvature
E Mitchell, Yoonho Lee, Alexander Khazatsky, Christopher D Manning, and Chelsea Finn · 2023
Cited alongside, same era.
Does ChatGPT provide appropriate and equitable medical advice?: A vignette-based, clinical evaluation across care contexts
Anthony J Nastasi, Katherine R Courtright, Scott D Halpern, and Gary E Weissman · 2023
Cited alongside, same era.
GPT-4 technical report
OpenAI · 2023
Cited alongside, same era.
Ragged edges: The double-edged sword of retrieval-augmented chatbots
Philip Feldman Foulds, R James, and Shimei Pan · 2024
Closest in time.
Google restricts ai search tool after “nonsensical” answers told people to eat rocks and put glue on pizza, May 2024
Robert Hart · 2024
Closest in time.
Mosh Levy, Alon Jacoby, and Yoav Goldberg · 2024
Closest in time.
TrustLLM: Trustworthiness in large language models
Lichao Sun, Yue Huang, Haoran Wang, Siyuan Wu, Qihui Zhang, Chujie Gao, Yixin Huang, Wenhan Lyu, Yixuan Zhang, Xiner Li, Zhengliang Liu, Yixin Liu, Yijue Wang, Zhikun Zhang, Bhavya Kailkhura, Caiming Xiong, Chaowei Xiao, Chunyuan Li, Eric Xing, Furong Huang, Hao Liu, Heng Ji, Hongyi Wang, Huan Zhang, Huaxiu Yao, Manolis Kellis, Marinka Zitnik, Meng Jiang, Mohit Bansal, James Zou, Jian Pei, Jian Liu, Jianfeng Gao, Jiawei Han, Jieyu Zhao, Jiliang Tang, Jindong Wang, John Mitchell, Kai Shu, Kaidi Xu, Kai-Wei Chang, Lifang He, Lifu Huang, Michael Backes, Neil Zhenqiang Gong, Philip S Yu, Pin-Yu Chen, Quanquan Gu, Ran Xu, Rex Ying, Shuiwang Ji, Suman Jana, Tianlong Chen, Tianming Liu, Tianyi Zhou, Willian Wang, Xiang Li, Xiangliang Zhang, Xiao Wang, Xing Xie, Xun Chen, Xuyu Wang, Yan Liu, Yanfang Ye, Yinzhi Cao, Yong Chen, and Yue Zhao · 2024
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Med-HALT: Medical domain hallucination test for large language models
Ankit Pal, Logesh Kumar Umapathi, and Malaikannan Sankarasubbu · 2023
Cited alongside, same era.
Jian Xie, Kai Zhang, Jiangjie Chen, Renze Lou, and Yu Su · 2023
Cited alongside, same era.
Medical chatbot using OpenAI’s GPT-3 told a fake patient to kill themselves
Ryan Daws · 2024
Cited alongside, same era.
Benchmarking large language models in Retrieval-Augmented generation
Jiawei Chen, Hongyu Lin, Xianpei Han, and Le Sun
Cited in the paper.
Benchmarking large language models in retrieval-augmented generation
Jiawei Chen, Hongyu Lin, Xianpei Han, and Le Sun
Cited in the paper.
RAGAS: Automated evaluation of retrieval augmented generation
Shahul Es, Jithin James, Luis Espinosa-Anke, and Steven Schockaert
Cited in the paper.
Ragas: Automated evaluation of retrieval augmented generation
Shahul Es, Jithin James, Luis Espinosa-Anke, and Steven Schockaert
Cited in the paper.
Closest in time.
Why google’s AI overviews gets things wrong
Rhiannon Williams · 2024
Closest in time.
Certifiably robust rag against retrieval corruption
Chong Xiang, Tong Wu, Zexuan Zhong, David Wagner, Danqi Chen, and Prateek Mittal · 2024
Closest in time.
RetrievalQA: Assessing adaptive Retrieval-Augmented generation for short-form Open-Domain question answering
Zihan Zhang, Meng Fang, and Ling Chen · 2024
Closest in time.
The first to know: How token distributions reveal hidden knowledge in large Vision-Language models?
Qinyu Zhao, Ming Xu, Kartik Gupta, Akshay Asthana, Liang Zheng, and Stephen Gould · 2024
Closest in time.