Resources for building applications with Dependency Minimal Recursion Semantics
Ann Copestake, Guy Emerson, Michael W. Goodman, Matic Horvat, Alexander Kuhnle, and Ewa Muszyńska. 2016 · 2016
Later among the works it cites.
Multimodal compact bilinear pooling for visual question answering and visual grounding
Akira Fukui, Dong Huk Park, Daylen Yang, Anna Rohrbach, Trevor Darrell, and Marcus Rohrbach. 2016 · 2016
Later among the works it cites.
Making the V in VQA matter: Elevating the role of image understanding in Visual Question Answering
Original
Yash Goyal, Tejas Khot, Douglas Summers-Stay, Dhruv Batra, and Devi Parikh. 2016 · 2016
Later among the works it cites.
Focused evaluation for image description with binary forced-choice tasks
Micah Hodosh and Julia Hockenmaier. 2016 · 2016
Later among the works it cites.
Revisiting Visual Question Answering Baselines
Allan Jabri, Armand Joulin, and Laurens van der Maaten. 2016 · 2016
Later among the works it cites.
The Malmo platform for artificial intelligence experimentation
Matthew Johnson, Katja Hofmann, Tim Hutton, and David Bignell. 2016 · 2016
Later among the works it cites.
Virtual embodiment: A scalable long-term strategy for artificial intelligence research
Original
Douwe Kiela, Luana Bulat, Anita L. Vero, and Stephen Clark. 2016 · 2016
Later among the works it cites.
Hierarchical question-image co-attention for visual question answering
Jiasen Lu, Jianwei Yang, Dhruv Batra, and Devi Parikh. 2016 · 2016
Later among the works it cites.
“Look, some green circles!”: Learning to quantify from images
Ionut Sorodoc, Angeliki Lazaridou, Gemma Boleda, Aurélie Herbelot, Sandro Pezzelle, and Raffaella Bernardi. 2016 · 2016
Later among the works it cites.
RNN Approaches to Text Normalization: A Challenge
Original
Richard Sproat and Navdeep Jaitly. 2016 · 2016
Later among the works it cites.
Yin and Yang: Balancing and answering binary visual questions
Peng Zhang, Yash Goyal, Douglas Summers-Stay, Dhruv Batra, and Devi Parikh. 2016 · 2016
Later among the works it cites.
Adopting abstract images for semantic scene understanding
C. Lawrence Zitnick, Ramakrishna Vedantam, and Devi Parikh. 2016 · 2016
Later among the works it cites.
CLEVR: A diagnostic dataset for compositional language and elementary visual reasoning
Justin Johnson, Bharath Hariharan, Laurens van der Maaten, Li Fei-Fei, C. Lawrence Zitnick, and Ross Girshick. 2017 · 2017
Closest in time.
Discovering objects and their relations from entangled scene representations
David Raposo, Adam Santoro, David Barrett, Razvan Pascanu, Timothy Lillicrap, and Peter W. Battaglia. 2017 · 2017
Closest in time.
Understanding deep learning requires rethinking generalization
Chiyuan Zhang, Samy Bengio, Moritz Hardt, Benjamin Recht, and Oriol Vinyals. 2017 · 2017
Closest in time.