Fetching the paper…

Policy Regularization via Noisy Advantage Values for Cooperative Multi-agent Actor-Critic methods · Around