Fetching the paper…

Multimodal Attention Networks for Low-Level Vision-and-Language Navigation · Around