2019

HIGhER : Improving instruction following with Hindsight Generation for Experience Replay

Cideron, Geoffrey, Seurin, Mathieu, Strub, Florian et al.

Understand

Language creates a compact representation of the world and allows the description of unlimited situations and objectives through compositionality.

  • While these characterizations may foster instructing, conditioning or structuring interactive agent behavior, it remains an open-problem to correctly relate language understanding and reinforcement learning in even simple instruction following scenarios.
  • This joint learning problem is alleviated through expert demonstrations, auxiliary losses, or neural inductive biases.
  • In this paper, we propose an orthogonal approach called Hindsight Generation for Experience Replay (HIGhER) that extends the Hindsight Experience Replay (HER) approach to the language-conditioned policy setting.

Reading the bibliography…