Fetching the paper…
Reading the bibliography…
Visual control policies can encounter significant performance degradation when visual conditions like lighting or camera position differ from those seen during training -- often exhibiting sharp declines in capability even for minor differences.
R. S. Sutton, “Dyna, an integrated architecture for learning, planning, and reacting,” ACM Sigart Bulletin
1991
Earlier work this paper cites.
A. R. Cassandra, L. P. Kaelbling, and M. L. Littman, “Acting optimally in partially observable stochastic domains,” in AAAI Conference on Artificial Intelligence
1994
Earlier work this paper cites.
J. Pineau, G. Gordon, S. Thrun, et al
2003
Earlier work this paper cites.
T. Walsh, S. Goschin, and M. Littman, “Integrating sample-based planning and model-based reinforcement learning,” in Proceedings of the AAAI Conference on Artificial Intelligence
2010
Earlier work this paper cites.
D. P. Kingma and J. Ba, “Adam: A method for stochastic optimization,” arXiv preprint arXiv:1412.6980
2014
Earlier work this paper cites.
V. Mnih, K. Kavukcuoglu, D. Silver, A. A. Rusu, J. Veness, M. G. Bellemare, A. Graves, M. A. Riedmiller, A. K. Fidjeland, G. Ostrovski, S. Petersen, C. Beattie, A. Sadik, I. Antonoglou, H. King, D. Kumaran, D. Wierstra, S. Legg, and D. Hassabis, “Human-level control through deep reinforcement learning,” Nature
2015
Earlier work this paper cites.
H. V. Hasselt, A. Guez, and D. Silver, “Deep reinforcement learning with double q-learning,” in AAAI Conference on Artificial Intelligence
2015
Earlier work this paper cites.
J. Schulman, P. Moritz, S. Levine, M. I. Jordan, and P. Abbeel, “High-dimensional continuous control using generalized advantage estimation,” CoRR
2015
Earlier work this paper cites.
S. Levine, C. Finn, T. Darrell, and P. Abbeel, “End-to-end training of deep visuomotor policies,” JMLR
2016
Earlier work this paper cites.
V. Mnih, A. P. Badia, M. Mirza, A. Graves, T. P. Lillicrap, T. Harley, D. Silver, and K. Kavukcuoglu, “Asynchronous methods for deep reinforcement learning,” in International Conference on Machine Learning
2016
Earlier work this paper cites.
J. Ba, J. R. Kiros, and G. E. Hinton, “Layer normalization,” ArXiv
2016
Earlier work this paper cites.
C. Qi, H. Su, K. Mo, and L. J. Guibas, “Pointnet: Deep learning on point sets for 3d classification and segmentation,” 2017 IEEE Conference on Computer Vision and Pattern Recognition (CVPR)
2016
Earlier work this paper cites.
J. Oh, S. Singh, and H. Lee, “Value prediction network,” in NeurIPS
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
S. Elfwing, E. Uchibe, and K. Doya, “Sigmoid-weighted linear units for neural network function approximation in reinforcement learning,” Neural networks : the official journal of the International Neural Network Society
2017
Earlier work this paper cites.
T. Bhattacharjee, G. Lee, H. Song, and S. S. Srinivasa, “Towards robotic feeding: Role of haptics in fork-based food manipulation,” IEEE Robotics and Automation Letters
2018
Earlier work this paper cites.
Y. Zhu, Z. Wang, J. Merel, A. Rusu, T. Erez, S. Cabi, S. Tunyasuvunakool, J. Kramár, R. Hadsell, N. de Freitas, and N. Heess, “Reinforcement and imitation learning for diverse visuomotor skills,” in Proceedings of Robotics: Science and Systems
2018
Earlier work this paper cites.
K. Cobbe, O. Klimov, C. Hesse, T. Kim, and J. Schulman, “Quantifying generalization in reinforcement learning,” in International Conference on Machine Learning
2018
Earlier work this paper cites.
K. Chua, R. Calandra, R. McAllister, and S. Levine, “Deep reinforcement learning in a handful of trials using probabilistic dynamics models,” in Proceedings of the 32nd International Conference on Neural Information Processing Systems
2018
Earlier work this paper cites.
D. Ha and J. Schmidhuber, “World models,” arXiv preprint arXiv:1803.10122
2018
Earlier work this paper cites.
P. Battaglia, J. B. C. Hamrick, V. Bapst, A. Sanchez, V. Zambaldi, M. Malinowski, A. Tacchetti, D. Raposo, A. Santoro, R. Faulkner, C. Gulcehre, F. Song, A. Ballard, J. Gilmer, G. E. Dahl, A. Vaswani, K. Allen, C. Nash, V. J. Langston, C. Dyer, N. Heess, D. Wierstra, P. Kohli, M. Botvinick, O. Vinyals, Y. Li, and R. Pascanu, “Relational inductive biases, deep learning, and graph networks,” arXiv
2018
Earlier work this paper cites.
E. Imani and M. White, “Improving regression performance with distributional losses,” in International Conference on Machine Learning
2018
Earlier work this paper cites.
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
D. Yarats, A. Zhang, I. Kostrikov, B. Amos, J. Pineau, and R. Fergus, “Improving sample efficiency in model-free reinforcement learning from images,” in AAAI Conference on Artificial Intelligence
2019
Earlier work this paper cites.
A. Mandlekar, J. Booher, M. Spero, A. Tung, A. Gupta, Y. Zhu, A. Garg, S. Savarese, and L. Fei-Fei, “Scaling robot supervision to hundreds of hours with roboturk: Robotic manipulation dataset through human reasoning and dexterity,” in 2019 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS)
2019
Cited alongside, same era.
D. Hafner, T. Lillicrap, I. Fischer, R. Villegas, D. Ha, H. Lee, and J. Davidson, “Learning latent dynamics for planning from pixels,” in International conference on machine learning
2019
Cited alongside, same era.
Y. Li, J. Wu, R. Tedrake, J. B. Tenenbaum, and A. Torralba, “Learning particle dynamics for manipulating rigid bodies, deformable objects, and fluids,” in ICLR
2019
Cited alongside, same era.
W. Wu, Z. Qi, and L. Fuxin, “Pointconv: Deep convolutional networks on 3d point clouds,” in Proceedings of the IEEE/CVF Conference on computer vision and pattern recognition
2019
Cited alongside, same era.
J. Krantz and S. Lee, “Sim-2-sim transfer for vision-and-language navigation in continuous environments,” in European Conference on Computer Vision (ECCV)
2022
Later among the works it cites.
2022
Later among the works it cites.
F. Ebert, Y. Yang, K. Schmeckpeper, B. Bucher, G. Georgakis, K. Daniilidis, C. Finn, and S. Levine, “Bridge data: Boosting generalization of robotic skills with cross-domain datasets,” RSS
2022
Later among the works it cites.
N. Hansen, X. Wang, and H. Su, “Temporal difference learning for model predictive control,” in International Conference on Machine Learning
2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
L. Kaiser, M. Babaeizadeh, P. Milos, B. Osinski, R. H. Campbell, K. Czechowski, D. Erhan, C. Finn, P. Kozakowski, S. Levine, A. Mohiuddin, R. Sepassi, G. Tucker, and H. Michalewski, “Model-based reinforcement learning for atari,” International Conference on Learning Representations
2020
Cited alongside, same era.
2020
Cited alongside, same era.
S. Dasari, F. Ebert, S. Tian, S. Nair, B. Bucher, K. Schmeckpeper, S. Singh, S. Levine, and C. Finn, “Robonet: Large-scale multi-robot learning,” in Proceedings of the Conference on Robot Learning
2020
Cited alongside, same era.
R. Veerapaneni, J. D. Co-Reyes, M. Chang, M. Janner, C. Finn, J. Wu, J. Tenenbaum, and S. Levine, “Entity abstraction in visual model-based reinforcement learning,” in Conference on Robot Learning
2020
Cited alongside, same era.
A. Sanchez-Gonzalez, J. Godwin, T. Pfaff, R. Ying, J. Leskovec, and P. W. Battaglia, “Learning to simulate complex physics with graph networks,” in ICML
2020
Cited alongside, same era.
F. Xiang, Y. Qin, K. Mo, Y. Xia, H. Zhu, F. Liu, M. Liu, H. Jiang, Y. Yuan, H. Wang, L. Yi, A. X. Chang, L. J. Guibas, and H. Su, “SAPIEN: A simulated part-based interactive environment,” in The IEEE Conference on Computer Vision and Pattern Recognition (CVPR)
2020
Cited alongside, same era.
N. Hansen and X. Wang, “Generalization in reinforcement learning by soft data augmentation,” 2021 IEEE International Conference on Robotics and Automation (ICRA)
2020
Cited alongside, same era.
D. Yarats, I. Kostrikov, and R. Fergus, “Image augmentation is all you need: Regularizing deep reinforcement learning from pixels,” in International Conference on Learning Representations
2021
Cited alongside, same era.
2022
Later among the works it cites.
N. Hansen, Z. Yuan, Y. Ze, T. Mu, A. Rajeswaran, H. Su, H. Xu, and X. Wang, “On pre-training for visuo-motor control: Revisiting a learning-from-scratch baseline,” in International Conference on Machine Learning (ICML)
2023
Later among the works it cites.
2023
Later among the works it cites.
I. Radosavovic, T. Xiao, S. James, P. Abbeel, J. Malik, and T. Darrell, “Real-world robot learning with masked visual pre-training,” in Conference on Robot Learning
2023
Later among the works it cites.
2023
Later among the works it cites.
X. Li, W. Wu, X. Z. Fern, and L. Fuxin, “Improving the robustness of point convolution on k-nearest neighbor neighborhoods with a viewpoint-invariant coordinate transform,” in Proceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision
2023
Later among the works it cites.
A. Handa, A. Allshire, V. Makoviychuk, A. Petrenko, R. Singh, J. Liu, D. Makoviichuk, K. V. Wyk, A. Zhurkevich, B. Sundaralingam, Y. S. Narang, J.-F. Lafleche, D. Fox, and G. State, “Dextreme: Transfer of agile in-hand manipulation from simulation to reality,” 2023 IEEE International Conference on Robotics and Automation (ICRA)
2023
Later among the works it cites.
I. Singh, A. Liang, M. Shridhar, and J. Thomason, “Self-supervised 3d representation learning for robotics,” in ICRA2023 Workshop on Pretraining for Robotics (PT4R)
2023
Later among the works it cites.
2023
Later among the works it cites.
Y. Zhu, Z. Jiang, P. Stone, and Y. Zhu, “Learning generalizable manipulation policies with object-centric 3d representations,” in Conference on Robot Learning
2023
Later among the works it cites.
R. Yang, Y. Lin, X. Ma, H. Hu, C. Zhang, and T. Zhang, “What is essential for unseen goal generalization of offline goal-conditioned rl?,” ICML
2023
Later among the works it cites.
2023
Later among the works it cites.
S. Rajeswar, P. Mazzaglia, T. Verbelen, A. Piché, B. Dhoedt, A. Courville, and A. Lacoste, “Mastering the unsupervised reinforcement learning benchmark from pixels,” in 40th International Conference on Machine Learning
2023
Later among the works it cites.
N. Hansen, Y. Lin, H. Su, X. Wang, V. Kumar, and A. Rajeswaran, “Modem: Accelerating visual model-based reinforcement learning with demonstrations,” in International Conference on Learning Representations (ICLR)
2023
Later among the works it cites.
2023
Later among the works it cites.
Y. Huang, A. Conkey, and T. Hermans, “Planning for Multi-Object Manipulation with Graph Neural Network Relational Classifiers,” in IEEE International Conference on Robotics and Automation (ICRA)
2023
Later among the works it cites.
J. Gu, F. Xiang, X. Li, Z. Ling, X. Liu, T. Mu, Y. Tang, S. Tao, X. Wei, Y. Yao, X. Yuan, P. Xie, Z. Huang, R. Chen, and H. Su, “Maniskill2: A unified benchmark for generalizable manipulation skills,” in International Conference on Learning Representations
2023
Later among the works it cites.
C. Bao, H. Xu, Y. Qin, and X. Wang, “Dexart: Benchmarking generalizable dexterous manipulation with articulated objects,” in Conference on Computer Vision and Pattern Recognition 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
https://github.com/NM512/dreamerv3 torch, “Dreamerv3-torch,” Github
2023
Later among the works it cites.