Fetching the paper…
Reading the bibliography…
In reinforcement learning, the state of the real world is often represented by feature vectors.
Reinforcement learning with selective perception and hidden state
McCallum, Andrew Kachites · 1996
Earlier work this paper cites.
Efficient reinforcement learning in factored mdps
Kearns, Michael and Koller, Daphne · 1999
Earlier work this paper cites.
Near-optimal reinforcement learning in polynomial time
Kearns, Michael and Singh, Satinder · 2002
Earlier work this paper cites.
Logarithmic online regret bounds for undiscounted reinforcement learning
Ortner, P and Auer, R · 2007
Earlier work this paper cites.
The adaptive k-meteorologists problem and its application to structure learning and feature selection in reinforcement learning
Diuk, Carlos, Li, Lihong, and Leffler, Bethany R · 2009
Earlier work this paper cites.
Automatic feature selection for model-based reinforcement learning in factored mdps
Kroon, Mark and Whiteson, Shimon · 2009
Earlier work this paper cites.
Reinforcement learning in finite mdps: Pac analysis
Strehl, Alexander L, Li, Lihong, and Littman, Michael L · 2009
Cited alongside, same era.
Structure learning in ergodic factored mdps without knowledge of the transition function’s in-degree
Chakraborty, Doran and Stone, Peter · 2011
Cited alongside, same era.
Online discovery of feature dependencies
Geramifard, Alborz, Doshi, Finale, Redding, Joshua, Roy, Nicholas, and How, Jonathan · 2011
Cited alongside, same era.
Online-to-confidence-set conversions and application to sparse stochastic bandits
Abbasi-Yadkori, Yasin, Pal, David, and Szepesvari, Csaba · 2012
Cited alongside, same era.
Greedy algorithms for sparse reinforcement learning
Painter-Wakefield, Christopher and Parr, Ronald · 2012
Cited alongside, same era.
Playing atari with deep reinforcement learning
Mnih, Volodymyr, Kavukcuoglu, Koray, Silver, David, Graves, Alex, Antonoglou, Ioannis, Wierstra, Daan, and Riedmiller, Martin · 2013
Later among the works it cites.
Online feature selection for model-based reinforcement learning
Nguyen, Trung Thanh, Li, Zhuoru, Silander, Tomi, and Leong, Tze-Yun · 2013
Later among the works it cites.
Selecting near-optimal approximate state representations in reinforcement learning
Ortner, Ronald, Maillard, Odalric-Ambrym, and Ryabko, Daniil · 2014
Later among the works it cites.
Concurrent pac rl
Guo, Zhaohan and Brunskill, Emma · 2015
Later among the works it cites.
Off-policy model-based learning under unknown factored dynamics
Hallak, Assaf, Schnitzler, François, Mann, Timothy Arthur, and Mannor, Shie · 2015
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…