Fetching the paper…
Reading the bibliography…
Visual model-based RL methods typically encode image observations into low-dimensional representations in a manner that does not eliminate redundant information.
Convex Optimization
S. Boyd and L. Vandenberghe · 2004
Earlier work this paper cites.
Adapting component analysis
F. Dorri and A. Ghodsi · 2012
Earlier work this paper cites.
Playing atari with deep reinforcement learning
V. Mnih, K. Kavukcuoglu, D. Silver, A. Graves, I. Antonoglou, D. Wierstra, and M. A. Riedmiller · 2013
Earlier work this paper cites.
Bisimulation metrics are optimal value functions
N. Ferns and D. Precup · 2014
Earlier work this paper cites.
Generative adversarial nets
I. Goodfellow, J. Pouget-Abadie, M. Mirza, B. Xu, D. Warde-Farley, S. Ozair, A. Courville, and Y. Bengio · 2014
Earlier work this paper cites.
Transfer joint matching for unsupervised domain adaptation
M. Long, J. Wang, G. Ding, J. Sun, and P. S. Yu · 2014
Earlier work this paper cites.
Deep domain confusion: Maximizing for domain invariance
E. Tzeng, J. Hoffman, N. Zhang, K. Saenko, and T. Darrell · 2014
Earlier work this paper cites.
Domain-adversarial training of neural networks
Y. Ganin, E. Ustinova, H. Ajakan, P. Germain, H. Larochelle, F. Laviolette, M. Marchand, and V. S. Lempitsky · 2015
Earlier work this paper cites.
Adam: A method for stochastic optimization
D. P. Kingma and J. Ba · 2015
Earlier work this paper cites.
Deep variational information bottleneck
A. A. Alemi, I. Fischer, J. V. Dillon, and K. Murphy · 2017
Earlier work this paper cites.
Matterport3D: Learning from RGB-D data in indoor environments
A. Chang, A. Dai, T. Funkhouser, M. Halber, M. Niessner, M. Savva, S. Song, A. Zeng, and Y. Zhang · 2017
Earlier work this paper cites.
Learning from Conditional Distributions via Dual Embeddings
B. Dai, N. He, Y. Pan, B. Boots, and L. Song · 2017
Earlier work this paper cites.
Simultaneous deep transfer across domains and tasks
J. Hoffman, E. Tzeng, T. Darrell, and K. Saenko · 2017
Earlier work this paper cites.
Cycada: Cycle-consistent adversarial domain adaptation
J. Hoffman, E. Tzeng, T. Park, J. Zhu, P. Isola, K. Saenko, A. A. Efros, and T. Darrell · 2017
Earlier work this paper cites.
The kinetics human action video dataset
W. Kay, J. Carreira, K. Simonyan, B. Zhang, C. Hillier, S. Vijayanarasimhan, F. Viola, T. Green, T. Back, A. Natsev, M. Suleyman, and A. Zisserman · 2017
Earlier work this paper cites.
Adversarial discriminative domain adaptation (workshop extended abstract)
E. Tzeng, J. Hoffman, K. Saenko, and T. Darrell · 2017
Earlier work this paper cites.
Neural discrete representation learning
A. van den Oord, O. Vinyals, and K. Kavukcuoglu · 2017
Earlier work this paper cites.
Model predictive path integral control: From theory to parallel computation
G. Williams, A. Aldrich, and E. A. Theodorou · 2017
Earlier work this paper cites.
Addressing function approximation error in actor-critic methods
S. Fujimoto, H. van Hoof, and D. Meger · 2018
Earlier work this paper cites.
Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor
T. Haarnoja, A. Zhou, P. Abbeel, and S. Levine · 2018
Earlier work this paper cites.
State representation learning for control: An overview
T. Lesort, N. D. Rodríguez, J. Goudou, and D. Filliat · 2018
Earlier work this paper cites.
Adversarial teacher-student learning for unsupervised domain adaptation
Z. Meng, J. Li, Y. Gong, and B. Juang · 2018
Earlier work this paper cites.
Neural network dynamics for model-based deep reinforcement learning with model-free fine-tuning
A. Nagabandi, G. Kahn, R. S. Fearing, and S. Levine · 2018
Earlier work this paper cites.
Learning from synthetic data: Addressing domain shift for semantic segmentation
S. Sankaranarayanan, Y. Balaji, A. Jain, S. N. Lim, and R. Chellappa · 2018
Cited alongside, same era.
Time-contrastive networks: Self-supervised learning from video
P. Sermanet, C. Lynch, Y. Chebotar, J. Hsu, E. Jang, S. Schaal, and S. Levine · 2018
Cited alongside, same era.
Deepmind control suite, 2018
Y. Tassa, Y. Doron, A. Muldal, T. Erez, Y. Li, D. de Las Casas, D. Budden, A. Abdolmaleki, J. Merel, A. Lefrancq, T. Lillicrap, and M. Riedmiller · 2018
Cited alongside, same era.
Representation learning with contrastive predictive coding
A. van den Oord, Y. Li, and O. Vinyals · 2018
Cited alongside, same era.
Natural environment benchmarks for reinforcement learning
A. Zhang, Y. Wu, and J. Pineau · 2018
Cited alongside, same era.
Offline reinforcement learning from images with latent space models
R. Rafailov, T. Yu, A. Rajeswaran, and C. Finn · 2021
Later among the works it cites.
Habitat-matterport 3d dataset (HM3d): 1000 large-scale 3d environments for embodied AI
S. K. Ramakrishnan, A. Gokaslan, E. Wijmans, O. Maksymets, A. Clegg, J. M. Turner, E. Undersander, W. Galuba, A. Westbury, A. X. Chang, M. Savva, Y. Zhao, and D. Batra · 2021
Later among the works it cites.
Model-based reinforcement learning via latent-space collocation
O. Rybkin, C. Zhu, A. Nagabandi, K. Daniilidis, I. Mordatch, and S. Levine · 2021
Later among the works it cites.
The distracting control suite - A challenging benchmark for reinforcement learning from pixels
A. Stone, O. Ramirez, K. Konolige, and R. Jonschkowski · 2021
Later among the works it cites.
Decoupling representation learning from reinforcement learning
A. Stooke, K. Lee, P. Abbeel, and M. Laskin · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Supervised representation learning with double encoding-layer autoencoder for transfer learning
F. Zhuang, X. Cheng, P. Luo, S. J. Pan, and Q. He · 2018
Cited alongside, same era.
DeepMDP: Learning continuous latent space models for representation learning
C. Gelada, S. Kumar, J. Buckman, O. Nachum, and M. G. Bellemare · 2019
Cited alongside, same era.
Learning actionable representations with goal conditioned policies
D. Ghosh, A. Gupta, and S. Levine · 2019
Cited alongside, same era.
Learning latent dynamics for planning from pixels
D. Hafner, T. P. Lillicrap, I. Fischer, R. Villegas, D. Ha, H. Lee, and J. Davidson · 2019
Cited alongside, same era.
When to trust your model: Model-based policy optimization
M. Janner, J. Fu, M. Zhang, and S. Levine · 2019
Cited alongside, same era.
Variational discriminator bottleneck: Improving imitation learning, inverse RL, and GANs by constraining information flow
X. B. Peng, A. Kanazawa, S. Toyer, P. Abbeel, and S. Levine · 2019
Cited alongside, same era.
Dream to control: Learning behaviors by latent imagination
D. Hafner, T. P. Lillicrap, J. Ba, and M. Norouzi · 2020
Cited alongside, same era.
Invariance through latent alignment, 2021
T. Yoneda, G. Yang, M. R. Walter, and B. Stadie · 2021
Later among the works it cites.
Learning invariant representations for reinforcement learning without reconstruction
A. Zhang, R. T. McAllister, R. Calandra, Y. Gal, and S. Levine · 2021
Later among the works it cites.
A survey of unsupervised domain adaptation for visual recognition
Y. Zhang · 2021
Later among the works it cites.
Sample-efficient reinforcement learning in the presence of exogenous information
Y. Efroni, D. J. Foster, D. Misra, A. Krishnamurthy, and J. Langford · 2022
Later among the works it cites.
Provably filtering exogenous distractors using multistep inverse dynamics
Y. Efroni, D. Misra, A. Krishnamurthy, A. Agarwal, and J. Langford · 2022
Later among the works it cites.
Modem: Accelerating visual model-based reinforcement learning with demonstrations
N. Hansen, Y. Lin, H. Su, X. Wang, V. Kumar, and A. Rajeswaran · 2022
Later among the works it cites.
Temporal difference learning for model predictive control
N. Hansen, X. Wang, and H. Su · 2022
Later among the works it cites.
Masked autoencoders are scalable vision learners
K. He, X. Chen, S. Xie, Y. Li, P. Dollár, and R. B. Girshick · 2022
Later among the works it cites.
R3M: A universal visual representation for robot manipulation
S. Nair, A. Rajeswaran, V. Kumar, C. Finn, and A. Gupta · 2022
Later among the works it cites.
Real-world robot learning with masked visual pre-training
I. Radosavovic, T. Xiao, S. James, P. Abbeel, J. Malik, and T. Darrell · 2022
Later among the works it cites.
Approximate information state for approximate planning and reinforcement learning in partially observed systems
J. Subramanian, A. Sinha, R. Seraj, and A. Mahajan · 2022
Later among the works it cites.
Denoised mdps: Learning world models better than the world itself
T. Wang, S. S. Du, A. Torralba, P. Isola, A. Zhang, and Y. Tian · 2022
Later among the works it cites.
Daydreamer: World models for physical robot learning
P. Wu, A. Escontrela, D. Hafner, P. Abbeel, and K. Goldberg · 2022
Later among the works it cites.
Dichotomy of control: Separating what you can control from what you cannot
M. Yang, D. Schuurmans, P. Abbeel, and O. Nachum · 2022
Later among the works it cites.
Maniskill2: A unified benchmark for generalizable manipulation skills
J. Gu, F. Xiang, X. Li, Z. Ling, X. Liu, T. Mu, Y. Tang, S. Tao, X. Wei, Y. Yao, X. Yuan, P. Xie, Z. Huang, R. Chen, and H. Su · 2023
Closest in time.
Where are we in the search for an artificial visual cortex for embodied intelligence?
A. Majumdar, K. Yadav, S. Arnaud, Y. J. Ma, C. Chen, S. Silwal, A. Jain, V. Berges, P. Abbeel, J. Malik, D. Batra, Y. Lin, O. Maksymets, A. Rajeswaran, and F. Meier · 2023
Closest in time.
Dinov2: Learning robust visual features without supervision
M. Oquab, T. Darcet, T. Moutakanni, H. Vo, M. Szafraniec, V. Khalidov, P. Fernandez, D. Haziza, F. Massa, A. El-Nouby, M. Assran, N. Ballas, W. Galuba, R. Howes, P. Huang, S. Li, I. Misra, M. G. Rabbat, V. Sharma, G. Synnaeve, H. Xu, H. Jégou, J. Mairal, P. Labatut, A. Joulin, and P. Bojanowski · 2023
Closest in time.