Fetching the paper…
Reading the bibliography…
We introduce "Talk The Walk", the first large-scale dialogue dataset grounded in action and perception.
Saying what you mean in dialogue: A study in conceptual and semantic co-ordination
Simon Garrod and Anthony Anderson · 1987
Earlier work this paper cites.
The hcrc map task corpus
Anne H. Anderson, Miles Bader, Ellen Gurman Bard, Elizabeth Boyle, Gwyneth Doherty, Simon Garrod, Stephen Isard, Jacqueline Kowtko, Jan McAllister, Jim Miller, Catherine Sotillo, Henry S. Thompson, and Regina Weinert · 1991
Earlier work this paper cites.
Simple statistical gradient-following algorithms for connectionist reinforcement learning
Ronald J Williams · 1992
Earlier work this paper cites.
Introduction to Reinforcement Learning
Richard S. Sutton and Andrew G. Barto · 1998
Earlier work this paper cites.
Grounding words in perception and action: computational insights
Deb Roy · 2005
Earlier work this paper cites.
The development of embodied cognition: Six lessons from babies
Linda Smith and Michael Gasser · 2005
Earlier work this paper cites.
Walk the talk: Connecting language, knowledge, and action in route instructions
Matt MacMahon, Brian Stankiewicz, and Benjamin Kuipers · 2006
Earlier work this paper cites.
Online learning for offroad robots: Using spatial label propagation to learn long-range traversability
Raia Hadsell, Pierre Sermanet, Jeff Han, Beat Flepp, Urs Muller, and Yann LeCun · 2007
Earlier work this paper cites.
Grounded cognition
Lawrence W. Barsalou · 2008
Earlier work this paper cites.
Learning to interpret natural language navigation instructions fro mobservations
David L. Chen and Raymond J. Mooney · 2011
Earlier work this paper cites.
Language grounding in robots
Luc Steels and Manfred Hild · 2012
Earlier work this paper cites.
Weakly supervised learning of semantic parsers for mapping instructions to actions
Yoav Artzi and Luke Zettlemoyer · 2013
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P. Kingma and Jimmy Ba · 2014
Earlier work this paper cites.
Microsoft coco: Common objects in context
Tsung-Yi Lin, Michael Maire, Serge Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Dollár, and C Lawrence Zitnick · 2014
Earlier work this paper cites.
Response-based learning for grounded machine translation
Stefan Riezler, Patrick Simianer, and Carolin Haas · 2014
Earlier work this paper cites.
Vqa: Visual question answering
Stanislaw Antol, Aishwarya Agrawal, Jiasen Lu, Margaret Mitchell, Dhruv Batra, C Lawrence Zitnick, and Devi Parikh · 2015
Cited alongside, same era.
Show, attend and tell: Neural image caption generation with visual attention
Kelvin Xu, Jimmy Ba, Ryan Kiros, Kyunghyun Cho, Aaron Courville, Ruslan Salakhudinov, Rich Zemel, and Yoshua Bengio · 2015
Cited alongside, same era.
Grounding distributional semantics in the visual world
Marco Baroni · 2016
Cited alongside, same era.
Enriching word vectors with subword information
Piotr Bojanowski, Edouard Grave, Armand Joulin, and Tomas Mikolov · 2016
Cited alongside, same era.
Abhishek Das, Satwik Kottur, Khushi Gupta, Avi Singh, Deshraj Yadav, José M. F. Moura, Devi Parikh, and Dhruv Batra · 2016
Cited alongside, same era.
Emergent language in a multi-modal, multi-step referential game
Katrina Evtimova, Andrew Drozdov, Douwe Kiela, and Kyunghyun Cho · 2017
Later among the works it cites.
Learning symmetric collaborative dialogue agents with dynamic knowledge graph embeddings
He He, Anusha Balakrishnan, Mihail Eric, and Percy Liang · 2017
Later among the works it cites.
Grounded language learning in a simulated 3d world
Karl Moritz Hermann, Felix Hill, Simon Green, Fumin Wang, Ryan Faulkner, Hubert Soyer, David Szepesvari, Wojtek Czarnecki, Max Jaderberg, Denis Teplyashin, et al · 2017
Later among the works it cites.
Clevr: A diagnostic dataset for compositional language and elementary visual reasoning
Justin Johnson, Bharath Hariharan, Laurens van der Maaten, Li Fei-Fei, C Lawrence Zitnick, and Ross Girshick · 2017
Later among the works it cites.
Deep embodiment: grounding semantics in perceptual modalities (PhD thesis)
Douwe Kiela · 2017
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Harm de Vries, Florian Strub, Sarath Chandar, Olivier Pietquin, Hugo Larochelle, and Aaron C. Courville · 2016
Cited alongside, same era.
Multi30k: Multilingual english-german image descriptions
Desmond Elliott, Stella Frank, Khalil Sima’an, and Lucia Specia · 2016
Cited alongside, same era.
Synthetic data for text localisation in natural images
A. Gupta, A. Vedaldi, and A. Zisserman · 2016
Cited alongside, same era.
Deep residual learning for image recognition
K. He, X. Zhang, S. Ren, and J. Sun · 2016
Cited alongside, same era.
Virtual embodiment: A scalable long-term strategy for artificial intelligence research
Douwe Kiela, Luana Bulat, Anita L. Vero, and Stephen Clark · 2016
Cited alongside, same era.
Multi-agent cooperation and the emergence of (natural) language
Angeliki Lazaridou, Alexander Peysakhovich, and Marco Baroni · 2016
Cited alongside, same era.
Listen, attend, and walk: Neural mapping of navigational instructions to action sequences
Hongyuan Mei, Mohit Bansal, and Matthew R Walter · 2016
Cited alongside, same era.
Later among the works it cites.
Learning visually grounded sentence representations
Douwe Kiela, Alexis Conneau, Allan Jabri, and Maximilian Nickel · 2017
Later among the works it cites.
Natural Language Does Not Emerge ’Naturally’ in Multi-Agent Dialog
Satwik Kottur, José M.F. Moura, Stefan Lee, and Dhruv Batra · 2017
Later among the works it cites.
Deal or no deal? end-to-end learning for negotiation dialogues
Mike Lewis, Denis Yarats, Yann N Dauphin, Devi Parikh, and Dhruv Batra · 2017
Later among the works it cites.
Parlai: A dialog research software platform
Alexander H Miller, Will Feng, Adam Fisch, Jiasen Lu, Dhruv Batra, Antoine Bordes, Devi Parikh, and Jason Weston · 2017
Later among the works it cites.
Emergence of grounded compositional language in multi-agent populations
Igor Mordatch and Pieter Abbeel · 2017
Later among the works it cites.
End-to-end optimization of goal-driven and visually grounded dialogue systems
Florian Strub, Harm De Vries, Jeremie Mary, Bilal Piot, Aaron Courville, and Olivier Pietquin · 2017
Later among the works it cites.
Revisiting im2gps in the deep learning era
Nam Vo, Nathan Jacobs, and James Hays · 2017
Later among the works it cites.
A deep compositional framework for human-like language acquisition in virtual environment
Haonan Yu, Haichao Zhang, and Wei Xu · 2017
Later among the works it cites.
Learning to navigate in cities without a map
Piotr Mirowski, Matthew Koichi Grimes, Mateusz Malinowski, Karl Moritz Hermann, Keith Anderson, Denis Teplyashin, Karen Simonyan, Koray Kavukcuoglu, Andrew Zisserman, and Raia Hadsell · 2018
Closest in time.
Film: Visual reasoning with a general conditioning layer
Ethan Perez, Florian Strub, Harm De Vries, Vincent Dumoulin, and Aaron Courville · 2018
Closest in time.