2018

Interactive Visual Grounding of Referring Expressions for Human-Robot Interaction

Shridhar, Mohit, Hsu, David

Understand

This paper presents INGRESS, a robot system that follows human natural language instructions to pick and place everyday objects.

  • The core issue here is the grounding of referring expressions: infer objects and their relationships from input images and language expressions.
  • INGRESS allows for unconstrained object categories and unconstrained language expressions.
  • Further, it asks questions to disambiguate referring expressions interactively.

Reading the bibliography…