Fetching the paper…
Reading the bibliography…
Visual model-based reinforcement learning (RL) has the potential to enable sample-efficient robot learning from visual observations.
Imagenet: A large-scale hierarchical image database
J. Deng, W. Dong, R. Socher, L.-J. Li, K. Li, and L. Fei-Fei · 2009
Earlier work this paper cites.
Pilco: A model-based and data-efficient approach to policy search
M. Deisenroth and C. E. Rasmussen · 2011
Earlier work this paper cites.
Synthesis and stabilization of complex behaviors through online trajectory optimization
Y. Tassa, T. Erez, and E. Todorov · 2012
Earlier work this paper cites.
Auto-encoding variational bayes
D. P. Kingma and M. Welling · 2014
Earlier work this paper cites.
Dropout: a simple way to prevent neural networks from overfitting
N. Srivastava, G. Hinton, A. Krizhevsky, I. Sutskever, and R. Salakhutdinov · 2014
Earlier work this paper cites.
Deepmpc: Learning deep latent features for model predictive control
I. Lenz, R. A. Knepper, and A. Saxena · 2015
Earlier work this paper cites.
Embed to control: A locally linear latent dynamics model for control from raw images
M. Watter, J. Springenberg, J. Boedecker, and M. Riedmiller · 2015
Earlier work this paper cites.
Distilling the knowledge in a neural network
G. Hinton, O. Vinyals, J. Dean, et al · 2015
Earlier work this paper cites.
High-dimensional continuous control using generalized advantage estimation
J. Schulman, P. Moritz, S. Levine, M. Jordan, and P. Abbeel · 2015
Earlier work this paper cites.
Convolutional feature masking for joint object and stuff segmentation
J. Dai, K. He, and J. Sun · 2015
Earlier work this paper cites.
Efficient object localization using convolutional networks
J. Tompson, R. Goroshin, A. Jain, Y. LeCun, and C. Bregler · 2015
Earlier work this paper cites.
Deep spatial autoencoders for visuomotor learning
C. Finn, X. Y. Tan, Y. Duan, T. Darrell, S. Levine, and P. Abbeel · 2016
Earlier work this paper cites.
Deep residual learning for image recognition
K. He, X. Zhang, S. Ren, and J. Sun · 2016
Earlier work this paper cites.
Loss is its own reward: Self-supervision for reinforcement learning
E. Shelhamer, P. Mahmoudieh, M. Argus, and T. Darrell · 2016
Earlier work this paper cites.
Deep visual foresight for planning robot motion
C. Finn and S. Levine · 2017
Earlier work this paper cites.
Attention is all you need
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin · 2017
Earlier work this paper cites.
Neural discrete representation learning
A. Van Den Oord, O. Vinyals, et al · 2017
Earlier work this paper cites.
Reinforcement learning with unsupervised auxiliary tasks
M. Jaderberg, V. Mnih, W. M. Czarnecki, T. Schaul, J. Z. Leibo, D. Silver, and K. Kavukcuoglu · 2017
Earlier work this paper cites.
Deep reinforcement learning in a handful of trials using probabilistic dynamics models
K. Chua, R. Calandra, R. McAllister, and S. Levine · 2018
Earlier work this paper cites.
Model-ensemble trust-region policy optimization
T. Kurutach, I. Clavera, Y. Duan, A. Tamar, and P. Abbeel · 2018
Earlier work this paper cites.
Visual foresight: Model-based deep reinforcement learning for vision-based robotic control
F. Ebert, C. Finn, S. Dasari, A. Xie, A. Lee, and S. Levine · 2018
Earlier work this paper cites.
World models
D. Ha and J. Schmidhuber · 2018
Earlier work this paper cites.
Reinforcement learning: An introduction
R. S. Sutton and A. G. Barto · 2018
Earlier work this paper cites.
Fixing a broken elbo
A. Alemi, B. Poole, I. Fischer, J. Dillon, R. A. Saurous, and K. Murphy · 2018
Earlier work this paper cites.
Dropblock: A regularization method for convolutional networks
G. Ghiasi, T.-Y. Lin, and Q. V. Le · 2018
Earlier work this paper cites.
Representation learning with contrastive predictive coding
A. v. d. Oord, Y. Li, and O. Vinyals · 2018
Earlier work this paper cites.
When to trust your model: Model-based policy optimization
M. Janner, J. Fu, M. Zhang, and S. Levine · 2019
Earlier work this paper cites.
Learning latent dynamics for planning from pixels
D. Hafner, T. Lillicrap, I. Fischer, R. Villegas, D. Ha, H. Lee, and J. Davidson · 2019
Earlier work this paper cites.
Solar: Deep structured representations for model-based reinforcement learning
M. Zhang, S. Vikram, L. Smith, P. Abbeel, M. Johnson, and S. Levine · 2019
Earlier work this paper cites.
Model-based reinforcement learning for atari
L. Kaiser, M. Babaeizadeh, P. Milos, B. Osinski, R. H. Campbell, K. Czechowski, D. Erhan, C. Finn, P. Kozakowski, S. Levine, et al · 2019
Cited alongside, same era.
Deepmdp: Learning continuous latent space models for representation learning
C. Gelada, S. Kumar, J. Buckman, O. Nachum, and M. G. Bellemare · 2019
Cited alongside, same era.
Unsupervised state representation learning in atari
A. Anand, E. Racah, S. Ozair, Y. Bengio, M.-A. Côté, and R. D. Hjelm · 2019
Cited alongside, same era.
Unsupervised learning of object structure and dynamics from videos
M. Minderer, C. Sun, R. Villegas, F. Cole, K. P. Murphy, and H. Lee · 2019
Cited alongside, same era.
Meta-world: A benchmark and evaluation for multi-task and meta reinforcement learning
T. Yu, D. Quillen, Z. He, R. Julian, K. Hausman, C. Finn, and S. Levine · 2020
Cited alongside, same era.
RLBench: The robot learning benchmark & learning environment
Levit: a vision transformer in convnet’s clothing for faster inference
B. Graham, A. El-Nouby, H. Touvron, P. Stock, A. Joulin, H. Jégou, and M. Douze · 2021
Later among the works it cites.
Vivit: A video vision transformer
A. Arnab, M. Dehghani, G. Heigold, C. Sun, M. Lučić, and C. Schmid · 2021
Later among the works it cites.
Cvt: Introducing convolutions to vision transformers
H. Wu, B. Xiao, N. Codella, M. Liu, X. Dai, L. Yuan, and L. Zhang · 2021
Later among the works it cites.
Object-aware regularization for addressing causal confusion in imitation learning
J. Park, Y. Seo, C. Liu, L. Zhao, T. Qin, J. Shin, and T.-Y. Liu · 2021
Later among the works it cites.
Data-efficient reinforcement learning with self-predictive representations
M. Schwarzer, A. Anand, R. Goel, R. D. Hjelm, A. Courville, and P. Bachman · 2021
Later among the works it cites.
Playvirtual: Augmenting cycle-consistent virtual trajectories for reinforcement learning
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
S. James, Z. Ma, D. R. Arrojo, and A. J. Davison · 2020
Cited alongside, same era.
Dream to control: Learning behaviors by latent imagination
D. Hafner, T. Lillicrap, J. Ba, and M. Norouzi · 2020
Cited alongside, same era.
dm_control: Software and tasks for continuous control
Y. Tassa, S. Tunyasuvunakool, A. Muldal, Y. Doron, S. Liu, S. Bohez, J. Merel, T. Erez, T. Lillicrap, and N. Heess · 2020
Cited alongside, same era.
Scalable methods for computing state similarity in deterministic markov decision processes
P. S. Castro · 2020
Cited alongside, same era.
Deep reinforcement and infomax learning
B. Mazoure, R. T. d. Combes, T. Doan, P. Bachman, and R. D. Hjelm · 2020
Cited alongside, same era.
Curl: Contrastive unsupervised representations for reinforcement learning
A. Srinivas, M. Laskin, and P. Abbeel · 2020
Cited alongside, same era.
Keypoints into the future: Self-supervised correspondence in model-based reinforcement learning
L. Manuelli, Y. Li, P. Florence, and R. Tedrake · 2020
Cited alongside, same era.
T. Yu, C. Lan, W. Zeng, M. Feng, Z. Zhang, and Z. Chen · 2021
Later among the works it cites.
Learning invariant representations for reinforcement learning without reconstruction
A. Zhang, R. McAllister, R. Calandra, Y. Gal, and S. Levine · 2021
Later among the works it cites.
Behavior from the void: Unsupervised active pre-training
H. Liu and P. Abbeel · 2021
Later among the works it cites.
Reinforcement learning with prototypical representations
D. Yarats, R. Fergus, A. Lazaric, and L. Pinto · 2021
Later among the works it cites.
Model-based inverse reinforcement learning from visual demonstrations
N. Das, S. Bechtle, T. Davchev, D. Jayaraman, A. Rai, and F. Meier · 2021
Later among the works it cites.
Improving sample efficiency in model-free reinforcement learning from images
D. Yarats, A. Zhang, I. Kostrikov, B. Amos, J. Pineau, and R. Fergus · 2021
Later among the works it cites.
Mastering visual continuous control: Improved data-augmented reinforcement learning
D. Yarats, R. Fergus, A. Lazaric, and L. Pinto · 2021
Later among the works it cites.
Image augmentation is all you need: Regularizing deep reinforcement learning from pixels
I. Kostrikov, D. Yarats, and R. Fergus · 2021
Later among the works it cites.
Three things everyone should know about vision transformers
H. Touvron, M. Cord, A. El-Nouby, J. Verbeek, and H. Jégou · 2022
Closest in time.
Maskvit: Masked visual pre-training for video prediction
A. Gupta, S. Tian, Y. Zhang, J. Wu, R. Martín-Martín, and L. Fei-Fei · 2022
Closest in time.
Masked autoencoders as spatiotemporal learners
C. Feichtenhofer, H. Fan, Y. Li, and K. He · 2022
Closest in time.
Simmim: A simple framework for masked image modeling
Z. Xie, Z. Zhang, Y. Cao, Y. Lin, J. Bao, Z. Yao, Q. Dai, and H. Hu · 2022
Closest in time.
Masked feature prediction for self-supervised visual pre-training
C. Wei, H. Fan, S. Xie, C.-Y. Wu, A. Yuille, and C. Feichtenhofer · 2022
Closest in time.
Q-attention: Enabling efficient learning for vision-based robotic manipulation
S. James and A. J. Davison · 2022
Closest in time.
Coarse-to-Fine Q-attention: Efficient learning for visual robotic manipulation via discretisation
S. James, K. Wada, T. Laidlow, and A. J. Davison · 2022
Closest in time.
Videomae: Masked autoencoders are data-efficient learners for self-supervised video pre-training
Z. Tong, Y. Song, J. Wang, and L. Wang · 2022
Closest in time.
Multimodal masked autoencoders learn transferable representations
X. Geng, H. Liu, L. Lee, D. Schuurams, S. Levine, and P. Abbeel · 2022
Closest in time.
Multimae: Multi-modal multi-task masked autoencoders
R. Bachmann, D. Mizrahi, A. Atanov, and A. Zamir · 2022
Closest in time.
Reinforcement learning with action-free pre-training from videos
Y. Seo, K. Lee, S. James, and P. Abbeel · 2022
Closest in time.
Mask-based latent reconstruction for reinforcement learning
T. Yu, Z. Zhang, C. Lan, Z. Chen, and Y. Lu · 2022
Closest in time.
R3m: A universal visual representation for robot manipulation
S. Nair, A. Rajeswaran, V. Kumar, C. Finn, and A. Gupta · 2022
Closest in time.
The unsurprising effectiveness of pre-trained vision models for control
S. Parisi, A. Rajeswaran, S. Purushwalkam, and A. Gupta · 2022
Closest in time.
Masked visual pre-training for motor control
T. Xiao, I. Radosavovic, T. Darrell, and J. Malik · 2022
Closest in time.