Fetching the paper…
Reading the bibliography…
Learning to solve precision-based manipulation tasks from visual feedback using Reinforcement Learning (RL) could drastically reduce the engineering efforts required by traditional robot systems.
R. Bajcsy, “Active perception,” Proceedings of the IEEE , 1988
1988
Earlier work this paper cites.
D. Wilkes and J. K. Tsotsos, Active object recognition . University of Toronto, 1994
1994
Earlier work this paper cites.
L. P. Kaelbling, M. L. Littman, and A. R. Cassandra, “Planning and acting in partially observable stochastic domains,” Artificial Intelligence , 1998
1998
Earlier work this paper cites.
R. Sutton, D. A. McAllester, S. Singh, and Y. Mansour, “Policy gradient methods for reinforcement learning with function approximation,” in NIPS , 1999
1999
Earlier work this paper cites.
M. Land, N. Mennie, and J. Rusted, “The roles of vision and eye movements in the control of activities of daily living,” Perception , 1999
1999
Earlier work this paper cites.
R. J. Williams, “Simple statistical gradient-following algorithms for connectionist reinforcement learning,” Machine Learning , 2004
2004
Earlier work this paper cites.
R. Sutton, “Learning to predict by the methods of temporal differences,” Machine Learning , 2005
2005
Earlier work this paper cites.
B. D. Ziebart, A. L. Maas, J. Bagnell, and A. Dey, “Maximum entropy inverse reinforcement learning,” in AAAI , 2008
2008
Earlier work this paper cites.
J. Bagnell and B. D. Ziebart, “Modeling purposeful adaptive behavior with the principle of maximum causal entropy,” 2010
2010
Earlier work this paper cites.
S. Chen, Y. Li, and N. M. Kwok, “Active vision in robotic systems: A survey of recent developments,” The International Journal of Robotics Research , 2011
2011
Earlier work this paper cites.
F. Fraundorfer, L. Heng, D. Honegger, G. H. Lee, L. Meier, P. Tanskanen, and M. Pollefeys, “Vision-based autonomous mapping and exploration using a quadrotor mav,” in 2012 IEEE/RSJ International Conference on Intelligent Robots and Systems , 2012
2012
Earlier work this paper cites.
E. Todorov, T. Erez, and Y. Tassa, “Mujoco: A physics engine for model-based control,” in 2012 IEEE/RSJ International Conference on Intelligent Robots and Systems , 2012
2012
Earlier work this paper cites.
I. Lenz, H. Lee, and A. Saxena, “Deep learning for detecting robotic grasps,” The International Journal of Robotics Research , 2015
2015
Earlier work this paper cites.
S. Kriegel, C. Rink, T. Bodenmüller, and M. Suppa, “Efficient next-best-scan planning for autonomous 3d surface reconstruction of unknown objects,” Journal of Real-Time Image Processing , 2015
2015
Earlier work this paper cites.
T. Lillicrap, J. Hunt, A. Pritzel, N. Heess, T. Erez, Y. Tassa, D. Silver, and D. Wierstra, “Continuous control with deep reinforcement learning,” CoRR , 2016
2016
Earlier work this paper cites.
S. Levine, C. Finn, T. Darrell, and P. Abbeel, “End-to-end training of deep visuomotor policies,” The Journal of Machine Learning Research , 2016
2016
Earlier work this paper cites.
L. Pinto and A. Gupta, “Supersizing self-supervision: Learning to grasp from 50k tries and 700 robot hours,” 2016 IEEE International Conference on Robotics and Automation (ICRA) , 2016
2016
Earlier work this paper cites.
J. Ba, J. Kiros, and G. E. Hinton, “Layer normalization,” ArXiv , vol. abs/1607.06450, 2016
2016
Earlier work this paper cites.
S. Gu, E. Holly, T. Lillicrap, and S. Levine, “Deep reinforcement learning for robotic manipulation with asynchronous off-policy updates,” 2017 IEEE International Conference on Robotics and Automation (ICRA) , 2017
2017
Earlier work this paper cites.
A. Vaswani, N. M. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, L. Kaiser, and I. Polosukhin, “Attention is all you need,” 2017
2017
Cited alongside, same era.
S. Levine, P. Pastor, A. Krizhevsky, and D. Quillen, “Learning hand-eye coordination for robotic grasping with deep learning and large-scale data collection,” The International Journal of Robotics Research , 2018
2018
Cited alongside, same era.
Y. Xiang, T. Schmidt, V. Narayanan, and D. Fox, “Posecnn: A convolutional neural network for 6d object pose estimation in cluttered scenes,” 2018
2018
Cited alongside, same era.
T. Haarnoja, A. Zhou, P. Abbeel, and S. Levine, “Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor,” 2018
2018
Cited alongside, same era.
A. Nair, V. H. Pong, M. Dalal, S. Bahl, S. Lin, and S. Levine, “Visual reinforcement learning with imagined goals,” in NeurIPS , 2018
A. Zhan, P. Zhao, L. Pinto, P. Abbeel, and M. Laskin, “A framework for efficient robotic manipulation,” 2020
2020
Later among the works it cites.
O. M. Andrychowicz, B. Baker, M. Chociej, R. Józefowicz, B. McGrew, J. W. Pachocki, A. Petron, M. Plappert, G. Powell, A. Ray, J. Schneider, S. Sidor, J. Tobin, P. Welinder, L. Weng, and W. Zaremba, “Learning dexterous in-hand manipulation,” The International Journal of Robotics Research , vol. 39, pp. 20 – 3, 2020
2020
Later among the works it cites.
G. Shi, Y. Zhu, J. Tremblay, S. Birchfield, F. Ramos, A. Anandkumar, and Y. Zhu, “Fast uncertainty quantification for deep object pose estimation,” 2020
2020
Later among the works it cites.
W. Yan, A. Vangipuram, P. Abbeel, and L. Pinto, “Learning predictive representations for deformable objects using contrastive estimation,” 2020
2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2018
Cited alongside, same era.
L. Pinto, M. Andrychowicz, P. Welinder, W. Zaremba, and P. Abbeel, “Asymmetric actor critic for image-based robot learning,” 2018
2018
Cited alongside, same era.
C. Li, J. Bai, and G. Hager, “A unified framework for multi-view multi-class object pose estimation,” 2018
2018
Cited alongside, same era.
B. Hepp, D. Dey, S. N. Sinha, A. Kapoor, N. Joshi, and O. Hilliges, “Learn-to-score: Efficient 3d scene exploration by predicting view utility,” in Proceedings of the European conference on computer vision (ECCV) , 2018
2018
Cited alongside, same era.
R. Cheng, A. Agarwal, and K. Fragkiadaki, “Reinforcement learning of active vision for manipulating objects under occlusions,” in Conference on Robot Learning . PMLR, 2018
2018
Cited alongside, same era.
X. Wang, R. B. Girshick, A. Gupta, and K. He, “Non-local neural networks,” 2018 IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2018
2018
Cited alongside, same era.
J. Jiang and Z. Lu, “Learning attentional communication for multi-agent cooperation,” in NeurIPS , 2018
2018
Cited alongside, same era.
A. Zeng, S. Song, J. Lee, A. Rodriguez, and T. Funkhouser, “Tossingbot: Learning to throw arbitrary objects with residual physics,” 2019
2019
Cited alongside, same era.
I. Akinola, J. Varley, and D. Kalashnikov, “Learning precise 3d manipulation from multiple uncalibrated cameras,” in 2020 IEEE International Conference on Robotics and Automation (ICRA) , 2020
2020
Later among the works it cites.
Y. Zaky, G. Paruthi, B. Tripp, and J. Bergstra, “Active perception and representation for robotic manipulation,” 2020
2020
Later among the works it cites.
2020
Later among the works it cites.
2021
Later among the works it cites.
I. Kostrikov, D. Yarats, and R. Fergus, “Image augmentation is all you need: Regularizing deep reinforcement learning from pixels,” 2021
2021
Later among the works it cites.
2021
Later among the works it cites.
2021
Later among the works it cites.
B. Chen, P. Abbeel, and D. Pathak, “Unsupervised learning of visual 3d keypoints for control,” 2021
2021
Later among the works it cites.
2021
Later among the works it cites.
L. Chen, K. Lu, A. Rajeswaran, K. Lee, A. Grover, M. Laskin, P. Abbeel, A. Srinivas, and I. Mordatch, “Decision transformer: Reinforcement learning via sequence modeling,” 2021
2021
Later among the works it cites.
2021
Later among the works it cites.
2021
Later among the works it cites.
S. James, K. Wada, T. Laidlow, and A. Davison, “Coarse-to-fine q-attention: Efficient learning for visual robotic manipulation via discretisation,” 2021
2021
Later among the works it cites.
N. Hansen and X. Wang, “Generalization in reinforcement learning by soft data augmentation,” in International Conference on Robotics and Automation , 2021
2021
Later among the works it cites.