Fetching the paper…
Reading the bibliography…
We fine-tune GPT-3 to answer long-form questions using a text-based web-browsing environment, which allows the model to search and navigate the web.
Acceleration of stochastic approximation by averaging
B. T. Polyak and A. B. Juditsky · 1992
Earlier work this paper cites.
Retrieval-augmented generation for knowledge-intensive NLP tasks
P. Lewis, E. Perez, A. Piktus, F. Petroni, V. Karpukhin, N. Goyal, H. Küttler, M. Lewis, W.-t. Yih, T. Rocktäschel, et al · 2005
Earlier work this paper cites.
Framing theory
D. Chong and J. N. Druckman · 2007
Earlier work this paper cites.
Question and answer test-train overlap in open-domain question answering datasets
P. Lewis, P. Stenetorp, and S. Riedel · 2008
Earlier work this paper cites.
Building watson: An overview of the deepqa project
D. Ferrucci, E. Brown, J. Chu-Carroll, J. Fan, D. Gondek, A. A. Kalyanpur, A. Lally, J. W. Murdock, E. Nyberg, J. Prager, et al · 2010
Earlier work this paper cites.
Automation bias: a systematic review of frequency, effect mediators, and mitigators
K. Goddard, A. Roudsari, and J. C. Wyatt · 2012
Earlier work this paper cites.
Superintelligence: Paths, Dangers, Strategies
N. Bostrom · 2014
Earlier work this paper cites.
Crystal Society
M. Harms · 2016
Earlier work this paper cites.
TriviaQA: A large scale distantly supervised challenge dataset for reading comprehension
M. Joshi, E. Choi, D. S. Weld, and L. Zettlemoyer · 2017
Earlier work this paper cites.
Proximal policy optimization algorithms
J. Schulman, F. Wolski, P. Dhariwal, A. Radford, and O. Klimov · 2017
Earlier work this paper cites.
World of bits: An open-domain platform for web-based agents
T. Shi, A. Karpathy, L. Fan, J. Hernandez, and P. Liang · 2017
Cited alongside, same era.
Supervising strong learners by amplifying weak experts
P. Christiano, B. Shlegeris, and D. Amodei · 2018
Cited alongside, same era.
I. Gur, U. Rueckert, A. Faust, and D. Hakkani-Tur · 2018
Cited alongside, same era.
G. Irving, P. Christiano, and D. Amodei · 2018
Cited alongside, same era.
Scalable agent alignment via reward modeling: a research direction
J. Leike, D. Krueger, T. Everitt, M. Martic, V. Maini, and S. Legg · 2018
On faithfulness and factuality in abstractive summarization
J. Maynez, S. Narayan, B. Bohnet, and R. McDonald · 2020
Later among the works it cites.
Learning to summarize from human feedback
N. Stiennon, L. Ouyang, J. Wu, D. M. Ziegler, R. Lowe, C. Voss, A. Radford, D. Amodei, and P. Christiano · 2020
Later among the works it cites.
Boosting search engines with interactive agents
L. Adolphs, B. Boerschinger, C. Buck, M. C. Huebscher, M. Ciaramita, L. Espeholt, T. Hofmann, and Y. Kilcher · 2021
Closest in time.
S. Bhakthavatsalam, D. Khashabi, T. Khot, B. D. Mishra, K. Richardson, A. Sabharwal, C. Schoenick, O. Tafjord, and P. Clark · 2021
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
ELI5: Long form question answering
A. Fan, Y. Jernite, E. Perez, D. Grangier, J. Weston, and M. Auli · 2019
Cited alongside, same era.
Interactive machine comprehension with information seeking agents
X. Yuan, J. Fu, M.-A. Cote, Y. Tay, C. Pal, and A. Trischler · 2019
Cited alongside, same era.
Language models are few-shot learners
T. B. Brown, B. Mann, N. Ryder, M. Subbiah, J. Kaplan, P. Dhariwal, A. Neelakantan, P. Shyam, G. Sastry, A. Askell, et al · 2020
Cited alongside, same era.
REALM: Retrieval-augmented language model pre-training
K. Guu, K. Lee, Z. Tung, P. Pasupat, and M.-W. Chang · 2020
Cited alongside, same era.
Dense passage retrieval for open-domain question answering
V. Karpukhin, B. Oğuz, S. Min, P. Lewis, L. Wu, S. Edunov, D. Chen, and W.-t. Yih · 2020
Cited alongside, same era.
H. Cheng, Y. Shen, X. Liu, P. He, W. Chen, and J. Gao · 2021
Closest in time.
Truthful AI: Developing and governing AI that does not lie
O. Evans, O. Cotton-Barratt, L. Finnveden, A. Bales, A. Balwit, P. Wills, L. Righetti, and W. Saunders · 2021
Closest in time.
Hurdles to progress in long-form question answering
K. Krishna, A. Roy, and M. Iyyer · 2021
Closest in time.
TruthfulQA: Measuring how models mimic human falsehoods
S. Lin, J. Hilton, and O. Evans · 2021
Closest in time.
Rethinking search: Making experts out of dilettantes
D. Metzler, Y. Tay, D. Bahri, and M. Najork · 2021
Closest in time.
Retrieval augmentation reduces hallucination in conversation
K. Shuster, S. Poff, M. Chen, D. Kiela, and J. Weston · 2021
Closest in time.