Fetching the paper…

Constrained Variational Policy Optimization for Safe Reinforcement Learning · Around