Fetching the paper…
Reading the bibliography…
This paper presents first successful steps in designing search agents that learn meta-strategies for iterative query refinement in information-seeking tasks.
Three models for the description of language
N. Chomsky · 1956
Earlier work this paper cites.
Review of B. F. Skinner, Verbal Behavior
N. Chomsky · 1959
Earlier work this paper cites.
Probabilistic and weighted grammars
Arto Salomaa · 1969
Earlier work this paper cites.
Relevance feedback in information retrieval
J. J. Rocchio · 1971
Earlier work this paper cites.
Ask for Information Retrieval: Part I. Background and Theory
N.J. Belkin, R.N. Oddy, and H.M. Brooks · 1982
Earlier work this paper cites.
Cognitive models from subcognitive skills
D Michie, M Bain, and J Hayes-Miches · 1990
Earlier work this paper cites.
Orienteering in an information landscape: How information seekers get from here to there
Vicki L. O’Day and Robin Jeffries · 1993
Earlier work this paper cites.
Context and page analysis for improved web search
Steve Lawrence and C. Lee. Giles · 1998
Earlier work this paper cites.
Patterns of Search: Analyzing and Modeling Web Query Refinement
T. Lau and E. Horvitz · 1999
Earlier work this paper cites.
The trec-8 question answering track report
Ellen Voorhees · 2000
Earlier work this paper cites.
Learning search engine specific query transformations for question answering
Eugene Agichtein, Steve Lawrence, and Luis Gravano · 2001
Earlier work this paper cites.
Eye-tracking analysis of user behavior in www search
Laura A. Granka, Thorsten Joachims, and Geri Gay · 2004
Earlier work this paper cites.
Umass at trec 2004: Novelty and hard
Nasreen Jaleel, James Allan, W. Croft, Fernando Diaz, Leah Larkey, Xiaoyan Li, Mark Smucker, and Courtney Wade · 2004
Earlier work this paper cites.
The Perfect Search Engine is Not Enough: A Study of Orienteering Behavior in Directed Search
Jaime Teevan, Christine Alvarado, Mark S. Ackerman, and David R. Karger · 2004
Earlier work this paper cites.
Accurately interpreting clickthrough data as implicit feedback
Thorsten Joachims, Laura Granka, Bing Pan, Helene Hembrooke, and Geri Gay · 2005
Earlier work this paper cites.
Understanding the Relationship between Searchers’ Queries and Information Goals
Doug Downey, Susan Dumais, Dan Liebling, and Eric Horvitz · 2008
Earlier work this paper cites.
Search Engines Information Retrieval in Practice
W.B. Croft, Donald Metzler, and Trevor Strohman · 2009
Earlier work this paper cites.
Search user interfaces
Marti Hearst · 2009
Earlier work this paper cites.
Analyzing and evaluating query reformulation strategies in web search logs
Jeff Huang and Efthimis Efthimiadis · 2009
Earlier work this paper cites.
Patterns of query reformulation during web searching
B. J. Jansen, D. L. Booth, and A. Spink · 2009
Earlier work this paper cites.
Training parsers by inverse reinforcement learning
Gergely Neu and Csaba Szepesvári · 2009
Earlier work this paper cites.
Eyetracking Web Usability
Jakob Nielsen and Kara Pernice · 2009
Earlier work this paper cites.
The Probabilistic Relevance Framework: BM25 and Beyond
Stephen Robertson and Hugo Zaragoza · 2009
Earlier work this paper cites.
Interactively optimizing information retrieval systems as a dueling bandits problem
Y. Yue and T. Joachims · 2009
Earlier work this paper cites.
A contextual-bandit approach to personalized news article
L. Li, W. Chu, J. Langford, and R.E. Schapire · 2010
Earlier work this paper cites.
A reduction of imitation learning and structured prediction to no-regret online learning
Stephane Ross, Geoffrey Gordon, and Drew Bagnell · 2011
Earlier work this paper cites.
Playing atari with deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Alex Graves, Ioannis Antonoglou, Daan Wierstra, and Martin Riedmiller · 2013
Cited alongside, same era.
Query expansion using term distribution and term association
Dipasree Pal, Mandar Mitra, and Kalyankumar Datta · 2013
Cited alongside, same era.
A theoretical analysis of ndcg type ranking measures
Yining Wang, Liwei Wang, Yuanzhi Li, Di He, and Tie-Yan Liu · 2013
Cited alongside, same era.
How do children reformulate their search queries?
Sophie Rutter, Nigel Ford, and Paul Clough · 2015
Cited alongside, same era.
Generative adversarial imitation learning
Jonathan Ho and Stefano Ermon · 2016
Cited alongside, same era.
Improving information extraction by acquiring external evidence with reinforcement learning
Karthik Narasimhan, Adam Yala, and Regina Barzilay · 2016
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, Sandhini Agarwal, Ariel Herbert-Voss, Gretchen Krueger, Tom Henighan, Rewon Child, Aditya Ramesh, Daniel Ziegler, Jeffrey Wu, Clemens Winter, Chris Hesse, Mark Chen, Eric Sigler, Mateusz Litwin, Scott Gray, Benjamin Chess, Jack Clark, Christopher Berner, Sam McCandlish, Alec Radford, Ilya Sutskever, and Dario Amodei · 2020
Later among the works it cites.
On the weaknesses of reinforcement learning for neural machine translation
Leshem Choshen, Lior Fox, Zohar Aizenbud, and Omri Abend · 2020
Later among the works it cites.
Overview of the trec 2020 deep learning track
Nick Craswell, Bhaskar Mitra, Emine Yilmaz, Daniel Fernando Campos, and Ellen M. Voorhees · 2020
Later among the works it cites.
Imitation Learning , pp. 273–306
Hao Dong, Zihan Ding, and Shanghang Zhang (eds.) · 2020
Later among the works it cites.
Seed rl: Scalable and efficient deep-rl with accelerated central inference
Lasse Espeholt, Raphaël Marinier, Piotr Stanczyk, Ke Wang, and Marcin Michalski · 2020
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
SQuAD: 100,000+ questions for machine comprehension of text
Pranav Rajpurkar, Jian Zhang, Konstantin Lopyrev, and Percy Liang · 2016
Cited alongside, same era.
Mastering the game of Go with deep neural networks and tree search
David Silver, Aja Huang, Chris J. Maddison, Arthur Guez, Laurent Sifre, George van den Driessche, Julian Schrittwieser, Ioannis Antonoglou, Veda Panneershelvam, Marc Lanctot, Sander Dieleman, Dominik Grewe, John Nham, Nal Kalchbrenner, Ilya Sutskever, Timothy Lillicrap, Madeleine Leach, Koray Kavukcuoglu, Thore Graepel, and Demis Hassabis · 2016
Cited alongside, same era.
Reading Wikipedia to answer open-domain questions
Danqi Chen, Adam Fisch, Jason Weston, and Antoine Bordes · 2017
Cited alongside, same era.
Task-oriented query reformulation with reinforcement learning
Rodrigo Nogueira and Kyunghyun Cho · 2017
Cited alongside, same era.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Lukasz Kaiser, and Illia Polosukhin · 2017
Cited alongside, same era.
Ask the right questions: Active question reformulation with reinforcement learning
Christian Buck, Jannis Bulian, Massimiliano Ciaramita, Andrea Gesmundo, Neil Houlsby, Wojciech Gajewski, and Wei Wang · 2018
Cited alongside, same era.
Later among the works it cites.
Dense passage retrieval for open-domain question answering
Vladimir Karpukhin, Barlas Oguz, Sewon Min, Patrick Lewis, Ledell Wu, Sergey Edunov, Danqi Chen, and Wen-tau Yih · 2020
Later among the works it cites.
BART: Denoising sequence-to-sequence pre-training for natural language generation, translation, and comprehension
Mike Lewis, Yinhan Liu, Naman Goyal, Marjan Ghazvininejad, Abdelrahman Mohamed, Omer Levy, Veselin Stoyanov, and Luke Zettlemoyer · 2020
Later among the works it cites.
A reinforcement learning framework for relevance feedback
Ali Montazeralghaem, Hamed Zamani, and James Allan · 2020
Later among the works it cites.
Exploring the limits of transfer learning with a unified text-to-text transformer
Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, and Peter J. Liu · 2020
Later among the works it cites.
How much knowledge can you pack into the parameters of a language model?
Adam Roberts, Colin Raffel, and Noam Shazeer · 2020
Later among the works it cites.
Mastering atari, go, chess and shogi by planning with a learned model
Julian Schrittwieser, Ioannis Antonoglou, Thomas Hubert, Karen Simonyan, Laurent Sifre, Simon Schmitt, Arthur Guez, Edward Lockhart, Demis Hassabis, Thore Graepel, Timothy Lillicrap, and David Silver · 2020
Later among the works it cites.
Eye-tracking studies of web search engines: A systematic literature review
Artur Strzelecki · 2020
Later among the works it cites.
Measuring and reducing gendered correlations in pre-trained models
Kellie Webster, Xuezhi Wang, Ian Tenney, Alex Beutel, Emily Pitler, Ellie Pavlick, Jilin Chen, and Slav Petrov · 2020
Later among the works it cites.
Interactive machine comprehension with information seeking agents
Xingdi Yuan, Jie Fu, Marc-Alexandre Côté, Yi Tay, Chris Pal, and Adam Trischler · 2020
Later among the works it cites.
Rtfm: Generalising to new environment dynamics via reading
Victor Zhong, Tim Rocktäschel, and Edward Grefenstette · 2020
Later among the works it cites.
On the dangers of stochastic parrots: Can language models be too big?
Emily M. Bender, Timnit Gebru, Angelina McMillan-Major, and Shmargaret Shmitchell · 2021
Closest in time.
Decision transformer: Reinforcement learning via sequence modeling
Lili Chen, Kevin Lu, Aravind Rajeswaran, Kimin Lee, Aditya Grover, Michael Laskin, Pieter Abbeel, Aravind Srinivas, and Igor Mordatch · 2021
Closest in time.
Leveraging passage retrieval with generative models for open domain question answering
Gautier Izacard and Edouard Grave · 2021
Closest in time.
Offline reinforcement learning as one big sequence modeling problem
Michael Janner, Qiyang Li, and Sergey Levine · 2021
Closest in time.
Revisiting the weaknesses of reinforcement learning for neural machine translation
Samuel Kiegeland and Julia Kreutzer · 2021
Closest in time.
NeurIPS 2020 EfficientQA Competition: Systems, Analyses and Lessons Learned, 2021
Sewon Min, Jordan Boyd-Graber, Chris Alberti, Danqi Chen, Eunsol Choi, Michael Collins, Kelvin Guu, Hannaneh Hajishirzi, Kenton Lee, Jennimaria Palomaki, Colin Raffel, Adam Roberts, Tom Kwiatkowski, Patrick Lewis, Yuxiang Wu, Heinrich Küttler, Linqing Liu, Pasquale Minervini, Pontus Stenetorp, Sebastian Riedel, Sohee Yang, Minjoon Seo, Gautier Izacard, Fabio Petroni, Lucas Hosseini, Nicola De Cao, Edouard Grave, Ikuya Yamada, Sonse Shimaoka, Masatoshi Suzuki, Shumpei Miyawaki, Shun Sato, Ryo Takahashi, Jun Suzuki, Martin Fajcik, Martin Docekal, Karel Ondrej, Pavel Smrz, Hao Cheng, Yelong Shen, Xiaodong Liu, Pengcheng He, Weizhu Chen, Jianfeng Gao, Barlas Oguz, Xilun Chen, Vladimir Karpukhin, Stan Peshterliev, Dmytro Okhonko, Michael Schlichtkrull, Sonal Gupta, Yashar Mehdad, and Wen tau Yih · 2021
Closest in time.
Webgpt: Browser-assisted question-answering with human feedback
Reiichiro Nakano, Jacob Hilton, S. Arun Balaji, Jeff Wu, Long Ouyang, Christina Kim, Christopher Hesse, Shantanu Jain, Vineet Kosaraju, William Saunders, Xu Jiang, Karl Cobbe, Tyna Eloundou, Gretchen Krueger, Kevin Button, Matthew Knight, Benjamin Chess, and John Schulman · 2021
Closest in time.
RocketQA: An optimized training approach to dense passage retrieval for open-domain question answering
Yingqi Qu, Yuchen Ding, Jing Liu, Kai Liu, Ruiyang Ren, Wayne Xin Zhao, Daxiang Dong, Hua Wu, and Haifeng Wang · 2021
Closest in time.
Answering complex open-domain questions with multi-hop dense retrieval
Wenhan Xiong, Xiang Li, Srini Iyer, Jingfei Du, Patrick Lewis, William Yang Wang, Yashar Mehdad, Scott Yih, Sebastian Riedel, Douwe Kiela, and Barlas Oguz · 2021
Closest in time.
Drl4ir: 2nd workshop on deep reinforcement learning for information retrieval
Weinan Zhang, Xiangyu Zhao, Li Zhao, Dawei Yin, and Grace Hui Yang · 2021
Closest in time.