Understand
We propose a goal-driven web navigation as a benchmark task for evaluating an agent with abilities to understand natural language and plan on partially observed environments.
- In this challenging task, an agent navigates through a website, which is represented as a graph consisting of web pages as nodes and hyperlinks as directed edges, to find a web page in which a query appears.
- The agent is required to have sophisticated high-level reasoning based on natural languages and efficient sequential decision-making capability to succeed.
- We release a software tool, called WebNav, that automatically transforms a website into this goal-driven web navigation task, and as an example, we make WikiNav, a dataset constructed from the English Wikipedia.
Built on
Learning representations by back-propagating errors
David Rumelhart, Geoffrey Hinton, and Ronald Williams · 1986
Earlier work this paper cites.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber · 1997
Earlier work this paper cites.
Focused crawling: a new approach to topic-specific web resource discovery
Soumen Chakrabarti, Martin Van den Berg, and Byron Dom · 1999
Earlier work this paper cites.
Deepbot: a focused crawler for accessing hidden web content
Manuel Álvarez, Juan Raposo, Alberto Pan, Fidel Cacheda, Fernando Bellas, and Víctor Carneiro · 2007
Earlier work this paper cites.
Similar
Wikispeedia: An online game for inferring semantic distances between concepts
Robert West, Joelle Pineau, and Doina Precup · 2009
Cited alongside, same era.
Efficient estimation of word representations in vector space
Tomas Mikolov, Kai Chen, Greg Corrado, and Jeffrey Dean · 2013
Cited alongside, same era.
Neural machine translation by jointly learning to align and translate
Dzmitry Bahdanau, Kyunghyun Cho, and Yoshua Bengio · 2014
Cited alongside, same era.
Evolving deep unsupervised convolutional networks for vision-based reinforcement learning
Jan Koutník, Jürgen Schmidhuber, and Faustino Gomez · 2014
Cited alongside, same era.
Automatic versus human navigation in information networks
Robert West and Jure Leskovec
Cited in the paper.
Human wayfinding in information networks
Robert West and Jure Leskovec
Cited in the paper.
Then
Neuroevolution in games: State of the art and open challenges
Sebastian Risi and Julian Togelius · 2014
Later among the works it cites.
Deep reinforcement learning with an unbounded action space
Ji He, Jianshu Chen, Xiaodong He, Jianfeng Gao, Lihong Li, Li Deng, and Mari Ostendorf · 2015
Later among the works it cites.
Human-level control through deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A Rusu, Joel Veness, Marc G Bellemare, Alex Graves, Martin Riedmiller, Andreas K Fidjeland, Georg Ostrovski, et al · 2015
Later among the works it cites.
Language understanding for text-based games using deep reinforcement learning
Karthik Narasimhan, Tejas Kulkarni, and Regina Barzilay · 2015
Later among the works it cites.
Beyond the bibliography
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…