2019

Self-Supervised Exploration via Disagreement

Pathak, Deepak, Gandhi, Dhiraj, Gupta, Abhinav

Understand

Efficient exploration is a long-standing problem in sensorimotor learning.

  • Major advances have been demonstrated in noise-free, non-stochastic domains such as video games and simulation.
  • However, most of these formulations either get stuck in environments with stochastic dynamics or are too inefficient to be scalable to real robotics setups.
  • In this paper, we propose a formulation for exploration inspired by the work in active learning literature.

Reading the bibliography…