Fetching the paper…

Offline RL with No OOD Actions: In-Sample Learning via Implicit Value Regularization · Around