Receptive fields and functional architecture in two nonstriate visual areas (18 and 19) of the cat
D. H. Hubel and T. N. Wiesel · 1965
Earlier work this paper cites.
Spatial and temporal contrast sensitivities of neurones in lateral geniculate nucleus of macaque
A. Derrington and P. Lennie · 1984
Earlier work this paper cites.
Segregation of form, color, movement, and depth: anatomy, physiology, and perception
M. Livingstone and D. Hubel · 1988
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database
J. Deng, W. Dong, R. Socher, L.-J. Li, K. Li, and L. Fei-Fei · 2009
Earlier work this paper cites.
Understanding natural language commands for robotic navigation and mobile manipulation
S. Tellex, T. Kollar, S. Dickerson, M. Walter, A. Banerjee, S. Teller, and N. Roy · 2011
Earlier work this paper cites.
Interpreting and executing recipes with a cooking robot
M. Bollini, S. Tellex, T. Thompson, N. Roy, and D. Rus · 2013
Earlier work this paper cites.
Learning from unscripted deictic gesture and language for human-robot interactions
C. Matuszek, L. Bo, L. Zettlemoyer, and D. Fox · 2014
Earlier work this paper cites.
The ecological approach to visual perception: classic edition
J. J. Gibson · 2014
Earlier work this paper cites.
Single image 3d object detection and pose estimation for grasping
M. Zhu, K. G. Derpanis, Y. Yang, S. Brahmbhatt, M. Zhang, C. Phillips, M. Lecce, and K. Daniilidis · 2014
Earlier work this paper cites.
Two-stream convolutional networks for action recognition in videos
Original
K. Simonyan and A. Zisserman · 2014
Earlier work this paper cites.
Learning to interpret natural language commands through human-robot dialog
J. Thomason, S. Zhang, R. J. Mooney, and P. Stone · 2015
Earlier work this paper cites.
Tell me dave: Context-sensitive grounding of natural language to manipulation instructions
D. K. Misra, J. Sung, K. Lee, and A. Saxena · 2016
Earlier work this paper cites.
Natural language communication with robots
Y. Bisk, D. Yuret, and D. Marcu · 2016
Earlier work this paper cites.
Convolutional two-stream network fusion for video action recognition
C. Feichtenhofer, A. Pinz, and A. Zisserman · 2016
Earlier work this paper cites.
Group equivariant convolutional networks
T. Cohen and M. Welling · 2016
Earlier work this paper cites.
Pybullet, a python module for physics simulation for games, robotics and machine learning
E. Coumans and Y. Bai · 2016
Earlier work this paper cites.
Deep residual learning for image recognition
K. He, X. Zhang, S. Ren, and J. Sun · 2016
Earlier work this paper cites.
Mask r-cnn
K. He, G. Gkioxari, P. Dollár, and R. Girshick · 2017
Earlier work this paper cites.
Multi-view self-supervised deep learning for 6d pose estimation in the amazon picking challenge
A. Zeng, K.-T. Yu, S. Song, D. Suo, E. Walker, A. Rodriguez, and J. Xiao · 2017
Earlier work this paper cites.
Attention is all you need
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, L. Kaiser, and I. Polosukhin · 2017
Earlier work this paper cites.
End-to-end learning of semantic grasping
E. Jang, S. Vijayanarasimhan, P. Pastor, J. Ibarz, and S. Levine · 2017
Earlier work this paper cites.
Qt-opt: Scalable deep reinforcement learning for vision-based robotic manipulation
D. Kalashnikov, A. Irpan, P. Pastor, J. Ibarz, A. Herzog, E. Jang, D. Quillen, E. Holly, M. Kalakrishnan, V. Vanhoucke, et al · 2018
Earlier work this paper cites.
Interactive visual grounding of referring expressions for human-robot interaction
M. Shridhar and D. Hsu · 2018
Earlier work this paper cites.