Fetching the paper…
Reading the bibliography…
Reinforcement learning algorithms usually assume that all actions are always available to an agent.
Dynamic programming princeton university press
Bellmann, R · 1957
Earlier work this paper cites.
STRIPS: A new approach to the application of theorem proving to problem solving
Fikes, R. E. and Nilsson, N. J · 1971
Earlier work this paper cites.
The theory of affordances
Gibson, J. J · 1977
Earlier work this paper cites.
Affordances and the body: An intentional analysis of gibson’s ecological approach to visual perception
Heft, H · 1989
Earlier work this paper cites.
Between mdps and semi-mdps: A framework for temporal abstraction in reinforcement learning
Sutton, R. S., Precup, D., and Singh, S · 1999
Earlier work this paper cites.
Further experiments in the evolution of minimally cognitive behavior: From perceiving affordances to selective attention
Slocum, A. C., Downey, D. C., and Beer, R. D · 2000
Earlier work this paper cites.
Automatic discovery of subgoals in reinforcement learning using diverse density
McGovern, A. and Barto, A. G · 2001
Earlier work this paper cites.
An outline of a theory of affordances
Chemero, A · 2003
Earlier work this paper cites.
Action elimination and stopping conditions for reinforcement learning
Even-Dar, E., Mannor, S., and Mansour, Y · 2003
Earlier work this paper cites.
Learning about objects through action-initial steps towards artificial cognition
Fitzpatrick, P., Metta, G., Natale, L., Rao, S., and Sandini, G · 2003
Earlier work this paper cites.
Improving action selection in mdp’s via knowledge transfer
Sherstov, A. A. and Stone, P · 2005
Earlier work this paper cites.
Identifying useful subgoals in reinforcement learning by local graph partitioning
Şimşek, Ö., Wolfe, A. P., and Barto, A. G · 2005
Earlier work this paper cites.
Effective control knowledge transfer through learning skill and representation hierarchies
Asadi, M. and Huber, M · 2007
Earlier work this paper cites.
Affordance-based imitation learning in robots
Lopes, M., Melo, F. S., and Montesano, L · 2007
Earlier work this paper cites.
An object-oriented representation for efficient reinforcement learning
Diuk, C., Cohen, A., and Littman, M. L · 2008
Earlier work this paper cites.
Learning object affordances: from sensory–motor coordination to imitation
Montesano, L., Lopes, M., Bernardino, A., and Santos-Victor, J · 2008
Earlier work this paper cites.
Simple local models for complex dynamical systems
Talvitie, E. and Singh, S. P · 2009
Cited alongside, same era.
Rectified linear units improve restricted boltzmann machines
Nair, V. and Hinton, G. E · 2010
Cited alongside, same era.
What good are actions? accelerating learning using learned action priors
Rosman, B. and Ramamoorthy, S · 2012
Cited alongside, same era.
Detecting affordances by mental imagery
Schenck, W., Hasenbein, H., and Möller, R · 2012
Cited alongside, same era.
Toward affordance-aware planning
Abel, D., Barth-Maron, G., MacGlashan, J., and Tellex, S · 2014
Cited alongside, same era.
Improving reinforcement learning with interactive feedback and affordances
Cruz, F., Magg, S., Weber, C., and Wermter, S · 2014
Cited alongside, same era.
Multi-modal feedback for affordance-driven interactive reinforcement learning
Cruz, F., Parisi, G. I., and Wermter, S · 2018
Later among the works it cites.
Neural predictive belief representations
Guo, Z. D., Azar, M. G., Piot, B., Pires, B. A., and Munos, R · 2018
Later among the works it cites.
PAC reinforcement learning with an imperfect model
Jiang, N · 2018
Later among the works it cites.
Data-efficient hierarchical reinforcement learning
Nachum, O., Gu, S. S., Lee, H., and Levine, S · 2018
Later among the works it cites.
Reinforcement learning: An introduction
Sutton, R. S. and Barto, A. G · 2018
Later among the works it cites.
Infobot: Transfer and exploration via the information bottleneck
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Kingma, D. P. and Ba, J · 2014
Cited alongside, same era.
Empowerment–an introduction
Salge, C., Glackin, C., and Polani, D · 2014
Cited alongside, same era.
Goal-based action priors
Abel, D., Hershkowitz, D. E., Barth-Maron, G., Brawner, S., O’Farrell, K., MacGlashan, J., and Tellex, S · 2015
Cited alongside, same era.
The dependence of effective planning horizon on model accuracy
Jiang, N., Kulesza, A., Singh, S., and Lewis, R · 2015
Cited alongside, same era.
Universal value function approximators
Schaul, T., Horgan, D., Gregor, K., and Silver, D · 2015
Cited alongside, same era.
Learning to detect visual grasp affordance
Song, H. O., Fritz, M., Goehring, D., and Darrell, T · 2015
Cited alongside, same era.
Goyal, A., Islam, R., Strouse, D., Ahmed, Z., Botvinick, M., Larochelle, H., Bengio, Y., and Levine, S · 2019
Later among the works it cites.
Shaping belief states with generative environment models for rl
Gregor, K., Rezende, D. J., Besse, F., Wu, Y., Merzic, H., and van den Oord, A · 2019
Later among the works it cites.
Harutyunyan, A., Dabney, W., Borsa, D., Heess, N., Munos, R., and Precup, D · 2019
Later among the works it cites.
Hierarchical affordance discovery using intrinsic motivation
Manoury, A., Nguyen, S. M., and Buche, C · 2019
Later among the works it cites.
Hierarchical foresight: Self-supervised learning of long-horizon tasks via visual subgoal generation
Nair, S. and Finn, C · 2019
Later among the works it cites.
The machine learning reproducibility checklist
Pineau, J · 2019
Later among the works it cites.
Mastering atari, go, chess and shogi by planning with a learned model
Schrittwieser, J., Antonoglou, I., Hubert, T., Simonyan, K., Sifre, L., Schmitt, S., Guez, A., Lockhart, E., Hassabis, D., Graepel, T., et al · 2019
Later among the works it cites.
Grandmaster level in starcraft ii using multi-agent reinforcement learning
Vinyals, O., Babuschkin, I., Czarnecki, W. M., Mathieu, M., Dudzik, A., Chung, J., Choi, D. H., Powell, R., Ewalds, T., Georgiev, P., et al · 2019
Later among the works it cites.
Learn what not to learn: Action elimination with deep reinforcement learning
Zahavy, T., Haroush, M., Merlis, N., Mankowitz, D. J., and Mannor, S · 2019
Later among the works it cites.
Model-based reinforcement learning for atari
Kaiser, L., Babaeizadeh, M., Milos, P., Osinski, B., Campbell, R. H., Czechowski, K., Erhan, D., Finn, C., Kozakowski, P., Levine, S., et al · 2020
Closest in time.
Options of interest: Temporal abstraction with interest functions
Khetarpal, K., Klissarov, M., Chevalier-Boisvert, M., Bacon, P.-L., and Precup, D · 2020
Closest in time.