Fetching the paper…

Maximizing Information Gain in Partially Observable Environments via Prediction Reward · Around