Fetching the paper…

Policy Gradient in Partially Observable Environments: Approximation and Convergence · Around