Fetching the paper…
Reading the bibliography…
A prominent approach to visual Reinforcement Learning (RL) is to learn an internal state representation using self-supervised methods, which has the potential benefit of improved sample-efficiency and generalization through additional learning signal and inductive biases.
S. Umeyama, “Least-squares estimation of transformation parameters between two point patterns,” IEEE Transactions on Pattern Analysis & Machine Intelligence , vol. 13, no. 04, pp. 376–380, 1991
1991
Earlier work this paper cites.
A. Dobbins, R. Jeo, J. Fiser, and J. Allman, “Distance modulation of neural activity in the visual cortex,” Science , 1998
1998
Earlier work this paper cites.
V. Mnih, K. Kavukcuoglu, D. Silver, A. Graves, I. Antonoglou, D. Wierstra, and M. Riedmiller, “Playing atari with deep reinforcement learning,” arXiv , 2013
2013
Earlier work this paper cites.
T. Lillicrap, J. Hunt, A. Pritzel, N. Heess, T. Erez, Y. Tassa, D. Silver, and D. Wierstra, “Continuous control with deep reinforcement learning,” arXiv , 2016
2016
Earlier work this paper cites.
S. Levine, C. Finn, T. Darrell, and P. Abbeel, “End-to-end training of deep visuomotor policies,” JMLR , 2016
2016
Earlier work this paper cites.
D. Silver, T. Hubert, J. Schrittwieser, I. Antonoglou, M. Lai, A. Guez, M. Lanctot, L. Sifre, D. Kumaran, T. Graepel et al. , “Mastering chess and shogi by self-play with a general reinforcement learning algorithm,” arXiv , 2017
2017
Earlier work this paper cites.
I. Higgins, A. Pal, A. A. Rusu, L. Matthey, C. P. Burgess, A. Pritzel, M. M. Botvinick, C. Blundell, and A. Lerchner, “Darla: Improving zero-shot transfer in reinforcement learning,” arXiv , 2017
2017
Earlier work this paper cites.
A. Nair, V. H. Pong, M. Dalal, S. Bahl, S. Lin, and S. Levine, “Visual reinforcement learning with imagined goals,” in NeurIPS , 2018
2018
Earlier work this paper cites.
R. Cheng, A. Agarwal, and K. Fragkiadaki, “Reinforcement learning of active vision for manipulating objects under occlusions,” in CoRL , 2018
2018
Earlier work this paper cites.
T. Haarnoja, A. Zhou, P. Abbeel, and S. Levine, “Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor,” arXiv , 2018
2018
Earlier work this paper cites.
D. Yarats, A. Zhang, I. Kostrikov, B. Amos, J. Pineau, and R. Fergus, “Improving sample efficiency in model-free reinforcement learning from images,” 2019
2019
Earlier work this paper cites.
H.-Y. F. Tung, R. Cheng, and K. Fragkiadaki, “Learning spatial common sense with geometry-aware recurrent networks,” in CVPR , 2019
2019
Earlier work this paper cites.
T. Yu, D. Quillen, Z. He, R. Julian, K. Hausman, C. Finn, and S. Levine, “Meta-world: A benchmark and evaluation for multi-task and meta reinforcement learning,” in CoRL , 2019
2019
Earlier work this paper cites.
M. Jaritz, J. Gu, and H. Su, “Multi-view PointNet for 3D Scene Understanding,” arXiv , 2019
2019
Earlier work this paper cites.
C. P. Burgess, L. Matthey, N. Watters, R. Kabra, I. Higgins, M. M. Botvinick, and A. Lerchner, “Monet: Unsupervised scene decomposition and representation,” arXiv , 2019
2019
Earlier work this paper cites.
O. M. Andrychowicz, B. Baker, M. Chociej, R. Józefowicz, B. McGrew, J. W. Pachocki, A. Petron, M. Plappert, G. Powell, A. Ray, J. Schneider, S. Sidor, J. Tobin, P. Welinder, L. Weng, and W. Zaremba, “Learning dexterous in-hand manipulation,” IJRR , 2020
2020
Cited alongside, same era.
M. Laskin, A. Srinivas, and P. Abbeel, “Curl: Contrastive unsupervised representations for reinforcement learning,” in ICML , 2020
2020
Cited alongside, same era.
I. Akinola, J. Varley, and D. Kalashnikov, “Learning precise 3d manipulation from multiple uncalibrated cameras,” in ICRA . IEEE, 2020, pp. 4616–4622
2020
Cited alongside, same era.
C. Wang, R. Martín-Martín, D. Xu, J. Lv, C. Lu, L. Fei-Fei, S. Savarese, and Y. Zhu, “6-pack: Category-level 6d pose tracker with anchor-based keypoints,” in ICRA . IEEE, 2020, pp. 10 059–10 066
2020
Cited alongside, same era.
W. Ye, S. Liu, T. Kurutach, P. Abbeel, and Y. Gao, “Mastering atari games with limited data,” arXiv , 2021
2021
Later among the works it cites.
N. Hansen, H. Su, and X. Wang, “Stabilizing deep q-learning with convnets and vision transformers under data augmentation,” arXiv , 2021
2021
Later among the works it cites.
2021
Later among the works it cites.
H. Qi, X. Wang, D. Pathak, Y. Ma, and J. Malik, “Learning long-term visual dynamics with region proposal interaction networks,” in ICLR , 2021
2021
Later among the works it cites.
Y. Li, S. Li, V. Sitzmann, P. Agrawal, and A. Torralba, “3d neural scene representations for visuomotor control,” arXiv , 2021
2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
D. Hafner, T. P. Lillicrap, J. Ba, and M. Norouzi, “Dream to control: Learning behaviors by latent imagination,” arXiv , 2020
2020
Cited alongside, same era.
M. Laskin, K. Lee, A. Stooke, L. Pinto, P. Abbeel, and A. Srinivas, “Reinforcement learning with augmented data,” arXiv , 2020
2020
Cited alongside, same era.
I. Kostrikov, D. Yarats, and R. Fergus, “Image augmentation is all you need: Regularizing deep reinforcement learning from pixels,” ICLR , 2020
2020
Cited alongside, same era.
H.-Y. F. Tung, X. Zhou, M. Prabhudesai, S. Lal, and K. Fragkiadaki, “3d-oes: Viewpoint-invariant object-factorized environment simulators,” in CoRL , 2020
2020
Cited alongside, same era.
K. Wang, B. Kang, J. Shao, and J. Feng, “Improving generalization in reinforcement learning with mixture regularization,” arXiv , 2020
2020
Cited alongside, same era.
R. Julian, B. Swanson, G. Sukhatme, S. Levine, C. Finn, and K. Hausman, “Never stop learning: The effectiveness of fine-tuning in robotic reinforcement learning,” arXiv , 2020
2020
Cited alongside, same era.
X. Chen, H. Fan, R. Girshick, and K. He, “Improved baselines with momentum contrastive learning,” arXiv , 2020
2020
Cited alongside, same era.
N. Hansen and X. Wang, “Generalization in reinforcement learning by soft data augmentation,” in ICRA , 2021
2021
Cited alongside, same era.
J. Ichnowski, Y. Avigal, J. Kerr, and K. Goldberg, “Dex-neRF: Using a neural radiance field to grasp transparent objects,” in CoRL , 2021
2021
Later among the works it cites.
Y.-Y. Tsai, H. Xu, Z. Ding, C. Zhang, E. Johns, and B. Huang, “Droid: Minimizing the reality gap using single-shot human demonstration,” RA-L , 2021
2021
Later among the works it cites.
Y. Du, O. Watkins, T. Darrell, P. Abbeel, and D. Pathak, “Auto-tuned sim-to-real transfer,” ICRA , 2021
2021
Later among the works it cites.
A. Kumar, Z. Fu, D. Pathak, and J. Malik, “Rma: Rapid motor adaptation for legged robots,” arXiv , 2021
2021
Later among the works it cites.
T. Xiao, I. Radosavovic, T. Darrell, and J. Malik, “Masked visual pre-training for motor control,” arXiv , 2022
2022
Closest in time.
S. Parisi, A. Rajeswaran, S. Purushwalkam, and A. K. Gupta, “The unsurprising effectiveness of pre-trained vision models for control,” arXiv , 2022
2022
Closest in time.
S. Nair, A. Rajeswaran, V. Kumar, C. Finn, and A. Gupta, “R3m: A universal visual representation for robot manipulation,” arXiv , 2022
2022
Closest in time.
D. Driess, I. Schubert, P. Florence, Y. Li, and M. Toussaint, “Reinforcement learning with neural radiance fields,” arXiv , 2022
2022
Closest in time.
R. Jangir, N. Hansen, S. Ghosal, M. Jain, and X. Wang, “Look closer: Bridging egocentric and third-person views with transformers for robotic manipulation,” RA-L , 2022
2022
Closest in time.