Fetching the paper…
Reading the bibliography…
We propose Convex Constraint Learning for Reinforcement Learning (CoCoRL), a novel approach for inferring shared constraints in a Constrained Markov Decision Process (CMDP) from a set of safe demonstrations with possibly different reward functions.
The maximum numbers of faces of a convex polytope
P. McMullen · 1970
Earlier work this paper cites.
Polyhedral combinatorics
W. R. Pulleyblank · 1983
Earlier work this paper cites.
Markov decision processes
M. L. Puterman · 1990
Earlier work this paper cites.
The quickhull algorithm for convex hulls
C. B. Barber, D. P. Dobkin, and H. Huhdanpaa · 1996
Earlier work this paper cites.
Constrained Markov decision processes , volume 7
E. Altman · 1999
Earlier work this paper cites.
Algorithms for inverse reinforcement learning
A. Y. Ng, S. Russell, et al · 2000
Earlier work this paper cites.
The convex hull in a new model of computation
A. Edalat, A. Lieutier, and E. Kashefi · 2001
Earlier work this paper cites.
Apprenticeship learning via inverse reinforcement learning
P. Abbeel and A. Y. Ng · 2004
Earlier work this paper cites.
The cross-entropy method: a unified approach to combinatorial optimization, Monte-Carlo simulation, and machine learning , volume 133
R. Y. Rubinstein and D. P. Kroese · 2004
Earlier work this paper cites.
Computability in computational geometry
A. Edalat, A. A. Khanban, and A. Lieutier · 2005
Earlier work this paper cites.
Random polytopes, convex bodies, and approximation
A. Baddeley, I. Bárány, and R. Schneider · 2007
Earlier work this paper cites.
Maximum entropy inverse reinforcement learning
B. D. Ziebart, A. L. Maas, J. A. Bagnell, A. K. Dey, et al · 2008
Earlier work this paper cites.
Apprenticeship learning about multiple intentions
M. Babes, V. Marivate, K. Subramanian, and M. L. Littman · 2011
Earlier work this paper cites.
Guarantees concerning geometric objects with imprecise points
J. Sember · 2011
Earlier work this paper cites.
Nonparametric Bayesian inverse reinforcement learning for multiple reward functions
J. Choi and K.-E. Kim · 2012
Earlier work this paper cites.
Bayesian multitask inverse reinforcement learning
C. Dimitrakakis and C. A. Rothkopf · 2012
Cited alongside, same era.
One-class classification: taxonomy of study and review of techniques
S. S. Khan and M. G. Madden · 2014
Cited alongside, same era.
A comprehensive survey on safe reinforcement learning
J. Garcıa and F. Fernández · 2015
Cited alongside, same era.
Repeated inverse reinforcement learning
K. Amin, N. Jiang, and S. Singh · 2017
Cited alongside, same era.
Imitation learning: A survey of learning methods
A. Hussein, M. M. Gaber, E. Elyan, and C. Jayne · 2017
Cited alongside, same era.
Learning convex polytopes with margin
L.-A. Gottlieb, E. Kaufman, A. Kontorovich, and G. Nivasch · 2018
Cited alongside, same era.
Maximum likelihood constraint inference for inverse reinforcement learning
D. R. Scobee and S. S. Sastry · 2020
Later among the works it cites.
Making human-like trade-offs in constrained environments by learning from demonstrations
A. Glazier, A. Loreggia, N. Mattei, T. Rahgooy, F. Rossi, and K. B. Venable · 2021
Later among the works it cites.
Inverse constrained reinforcement learning
S. Malik, U. Anwar, A. Aghasi, and A. Ahmed · 2021
Later among the works it cites.
Maximum likelihood constraint inference from stochastic demonstrations
D. L. McPherson, K. C. Stocking, and S. S. Sastry · 2021
Later among the works it cites.
Discretizing dynamics for maximum likelihood constraint inference
K. C. Stocking, D. L. McPherson, R. P. Matthew, and C. J. Tomlin · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
J. Leike, D. Krueger, T. Everitt, M. Martic, V. Maini, and S. Legg · 2018
Cited alongside, same era.
An environment for autonomous driving decision-making
E. Leurent · 2018
Cited alongside, same era.
Approximate robust control of uncertain dynamical systems
E. Leurent, Y. Blanco, D. Efimov, and O.-A. Maillard · 2018
Cited alongside, same era.
Constrained cross-entropy method for safe reinforcement learning
M. Wen and U. Topcu · 2018
Cited alongside, same era.
Learning a prior over intent via meta-inverse reinforcement learning
K. Xu, E. Ratner, A. Dragan, S. Levine, and C. Finn · 2019
Cited alongside, same era.
Meta-inverse reinforcement learning with probabilistic context variables
L. Yu, T. Yu, C. Finn, and S. Ermon · 2019
Cited alongside, same era.
Meta-adversarial inverse reinforcement learning for decision-making tasks
P. Wang, H. Li, and C.-Y. Chan · 2021
Later among the works it cites.
Active learning with safety constraints
R. Camilleri, A. Wagenmaker, J. Morgenstern, L. Jain, and K. Jamieson · 2022
Later among the works it cites.
Interactively learning preference constraints in linear bandits
D. Lindner, S. Tschiatschek, K. Hofmann, and A. Krause · 2022
Later among the works it cites.
Bayesian inverse constrained reinforcement learning
D. Papadimitriou, U. Anwar, and D. S. Brown · 2022
Later among the works it cites.
How to not drive: Learning driving constraints from demonstration
K. Rezaee and P. Yadmellat · 2022
Later among the works it cites.
Maximum causal entropy inverse constrained reinforcement learning
M. Baert, P. Mazzaglia, S. Leroux, and P. Simoens · 2023
Closest in time.
Learning soft constraints from constrained expert demonstrations
A. Gaurav, K. Rezaee, G. Liu, and P. Poupart · 2023
Closest in time.
Reward (mis) design for autonomous driving
W. B. Knox, A. Allievi, H. Banzhaf, F. Schmitt, and P. Stone · 2023
Closest in time.
Benchmarking constraint inference in inverse reinforcement learning
G. Liu, Y. Luo, A. Gaurav, K. Rezaee, and P. Poupart · 2023
Closest in time.