Fetching the paper…
Reading the bibliography…
Understanding multimodal perception for embodied AI is an open question because such inputs may contain highly complementary as well as redundant information for the task.
Playing atari with deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Alex Graves, Ioannis Antonoglou, Daan Wierstra, and Martin A. Riedmiller · 2013
Earlier work this paper cites.
Deep inside convolutional networks: Visualising image classification models and saliency maps
K. Simonyan, A. Vedaldi, and Andrew Zisserman · 2014
Earlier work this paper cites.
Hierarchical latent semantic mapping for automated topic generation
Guorui Zhou and Guang Chen · 2015
Earlier work this paper cites.
Charlie Beattie, Joel Z. Leibo, Denis Teplyashin, Tom Ward, Marcus Wainwright, Heinrich Küttler, Andrew Lefrancq, Simon Green, Víctor Valdés, Amir Sadik, Julian Schrittwieser, Keith Anderson, Sarah York, Max Cant, Adam Cain, Adrian Bolton, Stephen Gaffney, Helen King, Demis Hassabis, Shane Legg, and Stig Petersen · 2016
Earlier work this paper cites.
Openai gym, 2016
Greg Brockman, Vicki Cheung, Ludwig Pettersson, Jonas Schneider, John Schulman, Jie Tang, and Wojciech Zaremba · 2016
Earlier work this paper cites.
Not just a black box: Learning important features through propagating activation differences
Avanti Shrikumar, Peyton Greenside, Anna Shcherbina, and Anshul Kundaje · 2016
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks
Alex Krizhevsky, Ilya Sutskever, and Geoffrey E Hinton · 2017
Earlier work this paper cites.
A unified view of gradient-based attribution methods for deep neural networks
Marco Ancona, Enea Ceolini, A. Cengiz Öztireli, and Markus H. Gross · 2017
Earlier work this paper cites.
Smoothgrad: removing noise by adding noise
Daniel Smilkov, Nikhil Thorat, Been Kim, Fernanda B. Viégas, and Martin Wattenberg · 2017
Earlier work this paper cites.
Axiomatic attribution for deep networks
Mukund Sundararajan, Ankur Taly, and Qiqi Yan · 2017
Earlier work this paper cites.
Neural Modular Control for Embodied Question Answering
Abhishek Das, Georgia Gkioxari, Stefan Lee, Devi Parikh, and Dhruv Batra · 2018
Earlier work this paper cites.
Minimalistic gridworld environment for openai gym
Maxime Chevalier-Boisvert, Lucas Willems, and Suman Pal · 2018
Cited alongside, same era.
Visual semantic navigation using scene priors
Wei Yang, X. Wang, Ali Farhadi, Abhinav Kumar Gupta, and Roozbeh Mottaghi · 2019
Cited alongside, same era.
Xrai: Better attributions through regions
Andrei Kapishnikov, Tolga Bolukbasi, Fernanda Vi’egas, and Michael Terry · 2019
Cited alongside, same era.
Learning to explore using active neural slam
Devendra Singh Chaplot, Dhiraj Gandhi, Saurabh Gupta, Abhinav Gupta, and Ruslan Salakhutdinov · 2020
Cited alongside, same era.
Object goal navigation using goal-oriented semantic exploration
Devendra Singh Chaplot, Dhiraj Gandhi, Abhinav Gupta, and Ruslan Salakhutdinov · 2020
Cited alongside, same era.
ALFRED: A Benchmark for Interpreting Grounded Instructions for Everyday Tasks
Mohit Shridhar, Jesse Thomason, Daniel Gordon, Yonatan Bisk, Winson Han, Roozbeh Mottaghi, Luke Zettlemoyer, and Dieter Fox · 2020
Factorizing perception and policy for interactive instruction following
Kunal Pratap Singh, Suvaansh Bhambri, Byeonghwi Kim, Roozbeh Mottaghi, and Jonghyun Choi · 2021
Later among the works it cites.
Episodic Transformer for Vision-and-Language Navigation
Alexander Pashevich, Cordelia Schmid, and Chen Sun · 2021
Later among the works it cites.
Explaining multimodal errors in autonomous vehicles
Leilani H. Gilpin, Vishnu Penubarthi, and Lalana Kagal · 2021
Later among the works it cites.
Guided integrated gradients: an adaptive path method for removing noise
Andrei Kapishnikov, Subhashini Venugopalan, Besim Avci, Benjamin D. Wedin, Michael Terry, and Tolga Bolukbasi · 2021
Later among the works it cites.
Hierarchical task learning from language instructions with unified transformers and self-monitoring
Yichi Zhang and Joyce Chai · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Visualizing the impact of feature attribution baselines
Pascal Sturmfels, Scott Lundberg, and Su-In Lee · 2020
Cited alongside, same era.
Attribution in scale and space
Shawn Xu, Subhashini Venugopalan, and Mukund Sundararajan · 2020
Cited alongside, same era.
End-to-end grasping policies for human-in-the-loop robots via deep reinforcement learning*
Mohammadreza Sharif, Deniz Erdogmus, Chris Amato, and Taskin Padir · 2021
Cited alongside, same era.
Gpt-3: What’s it good for?
Robert Dale · 2021
Cited alongside, same era.
Coupling vision and proprioception for navigation of legged robots
Zipeng Fu, Ashish Kumar, Ananye Agarwal, Haozhi Qi, Jitendra Malik, and Deepak Pathak · 2022
Later among the works it cites.
FILM: Following instructions in language with modular methods
So Yeon Min, Devendra Singh Chaplot, Pradeep Kumar Ravikumar, Yonatan Bisk, and Ruslan Salakhutdinov · 2022
Later among the works it cites.
Teach: Task-driven embodied agents that chat
Aishwarya Padmakumar, Jesse Thomason, Ayush Shrivastava, Patrick Lange, Anjali Narayan-Chen, Spandana Gella, Robinson Piramuthu, Gokhan Tur, and Dilek Hakkani-Tur · 2022
Later among the works it cites.
Vision-and-language navigation: A survey of tasks, methods, and future directions
Jing Gu, Eliana Stefani, Qi Wu, Jesse Thomason, and Xin Eric Wang · 2022
Later among the works it cites.
Zero-shot reward specification via grounded natural language
Parsa Mahmoudieh, Deepak Pathak, and Trevor Darrell · 2022
Later among the works it cites.