Fetching the paper…
Reading the bibliography…
Trustworthiness of generative language models (GLMs) is crucial in their deployment to critical decision making systems.
Determination of sample sizes for setting tolerance limits
Samuel S Wilks · 1941
Earlier work this paper cites.
The comparison and evaluation of forecasters
Morris H DeGroot and Stephen E Fienberg · 1983
Earlier work this paper cites.
A theory of the learnable
Leslie G Valiant · 1984
Earlier work this paper cites.
A new algorithm for data compression
Philip Gage · 1994
Earlier work this paper cites.
Inductive confidence machines for regression
Harris Papadopoulos, Kostas Proedrou, Volodya Vovk, and Alex Gammerman · 2002
Earlier work this paper cites.
Transforming classifier scores into accurate multiclass probability estimates
Bianca Zadrozny and Charles Elkan · 2002
Earlier work this paper cites.
Algorithmic learning in a random world
Vladimir Vovk, Alex Gammerman, and Glenn Shafer · 2005
Earlier work this paper cites.
Conditional validity of inductive conformal predictors
Vladimir Vovk · 2013
Earlier work this paper cites.
A large annotated corpus for learning natural language inference
Samuel Bowman, Gabor Angeli, Christopher Potts, and Christopher D Manning · 2015
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Lukasz Kaiser, and Illia Polosukhin · 2017
Earlier work this paper cites.
Selective classification for deep neural networks
Yonatan Geifman and Ran El-Yaniv · 2017
Earlier work this paper cites.
A broad-coverage challenge corpus for sentence understanding through inference
Adina Williams, Nikita Nangia, and Samuel R Bowman · 2018
Earlier work this paper cites.
Transforming question answering datasets into natural language inference datasets
Dorottya Demszky, Kelvin Guu, and Percy Liang · 2018
Earlier work this paper cites.
Language models are unsupervised multitask learners
Alec Radford, Jeffrey Wu, Rewon Child, David Luan, Dario Amodei, Ilya Sutskever, et al · 2019
Cited alongside, same era.
Natural questions: a benchmark for question answering research
Tom Kwiatkowski, Jennimaria Palomaki, Olivia Redfield, Michael Collins, Ankur Parikh, Chris Alberti, Danielle Epstein, Illia Polosukhin, Jacob Devlin, Kenton Lee, et al · 2019
Cited alongside, same era.
Language models are few-shot learners, 2020
Tom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, Sandhini Agarwal, Ariel Herbert-Voss, Gretchen Krueger, Tom Henighan, Rewon Child, Aditya Ramesh, Daniel M. Ziegler, Jeffrey Wu, Clemens Winter, Christopher Hesse, Mark Chen, Eric Sigler, Mateusz Litwin, Scott Gray, Benjamin Chess, Jack Clark, Christopher Berner, Sam McCandlish, Alec Radford, Ilya Sutskever, and Dario Amodei · 2020
Cited alongside, same era.
Don’t say that! making inconsistent dialogue unlikely with unlikelihood training
Margaret Li, Stephen Roller, Ilia Kulikov, Sean Welleck, Y-Lan Boureau, Kyunghyun Cho, and Jason Weston · 2020
Cited alongside, same era.
Pac confidence sets for deep neural networks via calibrated prediction
Calibrating sequence likelihood improves conditional language generation
Yao Zhao, Mikhail Khalman, Rishabh Joshi, Shashi Narayan, Mohammad Saleh, and Peter J Liu · 2022
Later among the works it cites.
Llama: Open and efficient foundation language models, 2023
Hugo Touvron, Thibaut Lavril, Gautier Izacard, Xavier Martinet, Marie-Anne Lachaux, Timothée Lacroix, Baptiste Rozière, Naman Goyal, Eric Hambro, Faisal Azhar, Aurelien Rodriguez, Armand Joulin, Edouard Grave, and Guillaume Lample · 2023
Closest in time.
Stanford alpaca: An instruction-following llama model
Rohan Taori, Ishaan Gulrajani, Tianyi Zhang, Yann Dubois, Xuechen Li, Carlos Guestrin, Percy Liang, and Tatsunori B. Hashimoto · 2023
Closest in time.
Acon 2 : Adaptive conformal consensus for provable blockchain oracles, 2023
Sangdon Park, Osbert Bastani, and Taesoo Kim · 2023
Closest in time.
Selfcheckgpt: Zero-resource black-box hallucination detection for generative large language models
Potsawee Manakul, Adian Liusie, and Mark Gales · 2023
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Sangdon Park, Osbert Bastani, Nikolai Matni, and Insup Lee · 2020
Cited alongside, same era.
Distribution-free, risk-controlling prediction sets
Stephen Bates, Anastasios Angelopoulos, Lihua Lei, Jitendra Malik, and Michael I Jordan · 2021
Cited alongside, same era.
Adaptive conformal inference under distribution shift, 2021
Isaac Gibbs and Emmanuel Candès · 2021
Cited alongside, same era.
How Can We Know When Language Models Know? On the Calibration of Language Models for Question Answering
Zhengbao Jiang, Jun Araki, Haibo Ding, and Graham Neubig · 2021
Cited alongside, same era.
Few-shot conformal prediction with auxiliary tasks, 2021
Adam Fisch, Tal Schuster, Tommi Jaakkola, and Regina Barzilay · 2021
Cited alongside, same era.
Can NLI Models Verify QA Systems’ Predictions?, September 2021
Jifan Chen, Eunsol Choi, and Greg Durrett · 2021
Cited alongside, same era.
Training language models to follow instructions with human feedback, 2022
Long Ouyang, Jeff Wu, Xu Jiang, Diogo Almeida, Carroll L. Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, John Schulman, Jacob Hilton, Fraser Kelton, Luke Miller, Maddie Simens, Amanda Askell, Peter Welinder, Paul Christiano, Jan Leike, and Ryan Lowe · 2022
Cited alongside, same era.
PAC prediction sets for meta-learning
Sangdon Park, Edgar Dobriban, Insup Lee, and Osbert Bastani · 2022
Cited alongside, same era.
Christopher Mohri, Daniel Andor, Eunsol Choi, and Michael Collins · 2023
Closest in time.
Selection by Prediction with Conformal p-values, May 2023
Ying Jin and Emmanuel J. Candès · 2023
Closest in time.
Conformal Language Modeling, June 2024
Victor Quach, Adam Fisch, Tal Schuster, Adam Yala, Jae Ho Sohn, Tommi S. Jaakkola, and Regina Barzilay · 2024
Closest in time.
Confidence on the Focal: Conformal Prediction with Selection-Conditional Coverage, March 2024
Ying Jin and Zhimei Ren · 2024
Closest in time.
Language models with conformal factuality guarantees
Christopher Mohri and Tatsunori Hashimoto · 2024
Closest in time.
Large language model validity via enhanced conformal prediction methods, June 2024
John J. Cherian, Isaac Gibbs, and Emmanuel J. Candès · 2024
Closest in time.
Conformal Alignment: Knowing When to Trust Foundation Models with Guarantees, May 2024
Yu Gui, Ying Jin, and Zhimei Ren · 2024
Closest in time.