Fetching the paper…
Reading the bibliography…
Inverse reinforcement learning (IRL) enables an agent to learn complex behavior by observing demonstrations from a (near-)optimal policy.
Machine teaching for bayesian learners in the exponential family
Zhu, X. (2013) · 1913
Earlier work this paper cites.
Les problemes de decisions sequentielles. cahiers du centre d’etudes de recherche operationnelle vol. 2, pp. 161-179
De, G. G. (1960) · 1960
Earlier work this paper cites.
The principle of maximum causal entropy for estimating interacting processes
Ziebart, B. D., Bagnell, J. A., and Dey, A. K. (2013) · 1980
Earlier work this paper cites.
On the complexity of teaching
Goldman, S. A. and Kearns, M. J. (1995) · 1995
Earlier work this paper cites.
Constrained Markov decision processes
Altman, E. (1999) · 1999
Earlier work this paper cites.
Apprenticeship learning via inverse reinforcement learning
Abbeel, P. and Ng, A. Y. (2004) · 2004
Earlier work this paper cites.
Convex optimization
Boyd, S. and Vandenberghe, L. (2004) · 2004
Earlier work this paper cites.
Maximum entropy models with inequality constraints: A case study on text categorization
Kazama, J. and Tsujii, J. (2005) · 2005
Earlier work this paper cites.
Maximum margin planning
Ratliff, N. D., Bagnell, J. A., and Zinkevich, M. A. (2006) · 2006
Earlier work this paper cites.
Maximum entropy density estimation with generalized regularization and an application to species distribution modeling
Dudík, M., Phillips, S. J., and Schapire, R. E. (2007) · 2007
Earlier work this paper cites.
Maximum entropy inverse reinforcement learning
Ziebart, B. D., Maas, A. L., Bagnell, J. A., and Dey, A. K. (2008) · 2008
Earlier work this paper cites.
Feature construction for inverse reinforcement learning
Levine, S., Popovic, Z., and Koltun, V. (2010) · 2010
Earlier work this paper cites.
Modeling purposeful adaptive behavior with the principle of maximum causal entropy
Ziebart, B. D. (2010) · 2010
Earlier work this paper cites.
Relative entropy inverse reinforcement learning
Boularias, A., Kober, J., and Peters, J. (2011) · 2011
Earlier work this paper cites.
Comparing Action-query Strategies in Semi-autonomous Agents
Cohn, R., Durfee, E., and Singh, S. (2011) · 2011
Earlier work this paper cites.
Algorithmic and human teaching of sequential decision tasks
Cakmak, M. and Lopes, M. (2012) · 2012
Cited alongside, same era.
Revisiting Frank-Wolfe: Projection-free sparse convex optimization
Jaggi, M. (2013) · 2013
Cited alongside, same era.
On actively teaching the crowd to classify
Singla, A., Bogunovic, I., Bartók, G., Karbasi, A., and Krause, A. (2013) · 2013
Cited alongside, same era.
Near-optimally teaching the crowd to classify
Singla, A., Bogunovic, I., Bartók, G., Karbasi, A., and Krause, A. (2014) · 2014
Cited alongside, same era.
Machine teaching: An inverse problem to machine learning and an approach toward optimal education
Zhu, X. (2015) · 2015
Cited alongside, same era.
Cooperative inverse reinforcement learning
Hadfield-Menell, D., Russell, S. J., Abbeel, P., and Dragan, A. (2016) · 2016
Cited alongside, same era.
Teaching Inverse Reinforcement Learners via Features and Demonstrations
Haug, L., Tschiatschek, S., and Singla, A. (2018) · 2018
Later among the works it cites.
Towards black-box iterative machine teaching
Liu, W., Dai, B., li, X., Rehg, J. M., and Song, L. (2018) · 2018
Later among the works it cites.
Teaching categories to human learners with visual explanations
Mac Aodha, O., Su, S., Chen, Y., Perona, P., and Yue, Y. (2018) · 2018
Later among the works it cites.
Interactive optimal teaching with unknown learners
Melo, F. S., Guerra, C., and Lopes, M. (2018) · 2018
Later among the works it cites.
Lifelong inverse reinforcement learning
Mendez, J. A. M., Shivkumar, S., and Eaton, E. (2018) · 2018
Later among the works it cites.
An algorithmic perspective on imitation learning
Osa, T., Pajarinen, J., Neumann, G., Bagnell, J. A., Abbeel, P., Peters, J., et al. (2018) · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
The teaching dimension of linear learners
Liu, J. and Zhu, X. (2016) · 2016
Cited alongside, same era.
Repeated inverse reinforcement learning
Amin, K., Jiang, N., and Singh, S. P. (2017) · 2017
Cited alongside, same era.
Multi-view decision processes: the helper-ai problem
Dimitrakakis, C., Parkes, D. C., Radanovic, G., and Tylkin, P. (2017) · 2017
Cited alongside, same era.
Multi-agent reinforcement learning in sequential social dilemmas
Leibo, J. Z., Zambaldi, V., Lanctot, M., Marecki, J., and Graepel, T. (2017) · 2017
Cited alongside, same era.
Risk-Aware Active Inverse Reinforcement Learning
Brown, D. S., Cui, Y., and Niekum, S. (2018) · 2018
Cited alongside, same era.
Efficient probabilistic performance bounds for inverse reinforcement learning
Brown, D. S. and Niekum, S. (2018) · 2018
Cited alongside, same era.
A review on bilevel optimization: from classical to evolutionary approaches and applications
Sinha, A., Malo, P., and Deb, K. (2018) · 2018
Later among the works it cites.
Infinite time horizon maximum causal entropy inverse reinforcement learning
Zhou, Z., Bloem, M., and Bambos, N. (2018) · 2018
Later among the works it cites.
An Overview of Machine Teaching
Zhu, X., Singla, A., Zilles, S., and Rafferty, A. N. (2018) · 2018
Later among the works it cites.
Machine teaching for inverse reinforcement learning: Algorithms and applications
Brown, D. S. and Niekum, S. (2019) · 2019
Closest in time.
Towards deployment of robust AI agents for human-machine partnerships
Ghosh, A., Tschiatschek, S., Mahdavi, H., and Singla, A. (2019) · 2019
Closest in time.
Teaching multiple concepts to a forgetful learner
Hunziker, A., Chen, Y., Mac Aodha, O., Rodriguez, M. G., Krause, A., Perona, P., Yue, Y., and Singla, A. (2019) · 2019
Closest in time.
Interactive teaching algorithms for inverse reinforcement learning
Kamalaruban, P., Devidze, R., Cevher, V., and Singla, A. (2019) · 2019
Closest in time.
Iterative classroom teaching
Yeo, T., Kamalaruban, P., Singla, A., Merchant, A., Asselborn, T., Faucon, L., Dillenbourg, P., and Cevher, V. (2019) · 2019
Closest in time.