2023

Vision-Language Models as Success Detectors

Du, Yuqing, Konyushkova, Ksenia, Denil, Misha et al.

Understand

Detecting successful behaviour is crucial for training intelligent agents.

  • As such, generalisable reward models are a prerequisite for agents that can learn to generalise their behaviour.
  • In this work we focus on developing robust success detectors that leverage large, pretrained vision-language models (Flamingo, Alayrac et al.
  • (2022)) and human reward annotations.

Reading the bibliography…