Fetching the paper…

Policy Mirror Descent for Regularized Reinforcement Learning: A Generalized Framework with Linear Convergence · Around