Fetching the paper…
Reading the bibliography…
The visual world provides an abundance of information, but many input pixels received by agents often contain distracting stimuli.
Minimalistic Attacks: How Little it Takes to Fool a Deep Reinforcement Learning Policy
Xinghua Qu, Zhu Sun, Yew-Soon Ong, Abhishek Gupta, and Pengfei Wei. 2020 · 1911
Earlier work this paper cites.
Dream to Control: Learning Behaviors by Latent Imagination
Danijar Hafner, Timothy Lillicrap, Jimmy Ba, and Mohammad Norouzi. 2020 · 1912
Earlier work this paper cites.
Planning and acting in partially observable stochastic domains
Leslie Pack Kaelbling, Michael L Littman, and Anthony R Cassandra. 1998 · 1998
Earlier work this paper cites.
Robust Deep Reinforcement Learning against Adversarial Perturbations on State Observations
Huan Zhang, Hongge Chen, Chaowei Xiao, Bo Li, Mingyan Liu, Duane Boning, and Cho-Jui Hsieh. 2020 · 2003
Earlier work this paper cites.
Reinforcement Learning with Augmented Data
Misha Laskin, Kimin Lee, Adam Stooke, Lerrel Pinto, Pieter Abbeel, and Aravind Srinivas. 2020a · 2004
Earlier work this paper cites.
Michael Laskin, Aravind Srinivas, and Pieter Abbeel. 2020b · 2004
Earlier work this paper cites.
Learning Invariant Representations for Reinforcement Learning without Reconstruction
Amy Zhang, Rowan McAllister, Roberto Calandra, Yarin Gal, and Sergey Levine. 2021 · 2006
Earlier work this paper cites.
Self-Supervised Policy Adaptation during Deployment
Nicklas Hansen, Rishabh Jangir, Yu Sun, Guillem Alenyà, Pieter Abbeel, Alexei A Efros, Lerrel Pinto, and Xiaolong Wang. 2020 · 2007
Earlier work this paper cites.
An image is worth 16x16 words: Transformers for image recognition at scale
Alexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn, Xiaohua Zhai, Thomas Unterthiner, Mostafa Dehghani, Matthias Minderer, Georg Heigold, Sylvain Gelly, et al · 2010
Earlier work this paper cites.
Improving Generalization in Reinforcement Learning with Mixture Regularization
Kaixin Wang, Bingyi Kang, Jie Shao, and Jiashi Feng. 2020 · 2010
Earlier work this paper cites.
Bisimulation Metrics for Continuous Markov Decision Processes
Norm Ferns, Prakash Panangaden, and Doina Precup. 2011 · 2011
Earlier work this paper cites.
Nicklas Hansen and Xiaolong Wang. 2021 · 2011
Earlier work this paper cites.
MuJoCo: A physics engine for model-based control. In 2012 IEEE/RSJ International Conference on Intelligent Robots and Systems . IEEE, 5026–5033
Emanuel Todorov, Tom Erez, and Yuval Tassa. 2012 · 2012
Earlier work this paper cites.
Adam: A Method for Stochastic Optimization
Diederik Kingma and Jimmy Ba. 2015 · 2015
Earlier work this paper cites.
Human-level control through deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A Rusu, Joel Veness, Marc G Bellemare, Alex Graves, Martin Riedmiller, Andreas K Fidjeland, Georg Ostrovski, et al · 2015
Earlier work this paper cites.
Olaf Ronneberger, Philipp Fischer, and Thomas Brox. 2015 · 2015
Earlier work this paper cites.
ImageNet Large Scale Visual Recognition Challenge
Olga Russakovsky, Jia Deng, Hao Su, Jonathan Krause, Sanjeev Satheesh, Sean Ma, Zhiheng Huang, Andrej Karpathy, Aditya Khosla, Michael Bernstein, Alexander C. Berg, and Li Fei-Fei. 2015 · 2015
Earlier work this paper cites.
Jimmy Lei Ba, Jamie Ryan Kiros, and Geoffrey E Hinton. 2016 · 2016
Earlier work this paper cites.
Greg Brockman, Vicki Cheung, Ludwig Pettersson, Jonas Schneider, John Schulman, Jie Tang, and Wojciech Zaremba. 2016 · 2016
Earlier work this paper cites.
Chelsea Finn, Xin Yu Tan, Yan Duan, Trevor Darrell, Sergey Levine, and Pieter Abbeel. 2016 · 2016
Earlier work this paper cites.
Deep Residual Learning for Image Recognition. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition . 770–778
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun. 2016 · 2016
Cited alongside, same era.
Gaussian Error Linear Units (GELUs)
Dan Hendrycks and Kevin Gimpel. 2016 · 2016
Cited alongside, same era.
Learning to Navigate in Complex Environments
Piotr Mirowski, Razvan Pascanu, Fabio Viola, Hubert Soyer, Andrew J Ballard, Andrea Banino, Misha Denil, Ross Goroshin, Laurent Sifre, Koray Kavukcuoglu, et al · 2016
Cited alongside, same era.
Learning to reinforcement learn
Jane X Wang, Zeb Kurth-Nelson, Dhruva Tirumala, Hubert Soyer, Joel Z Leibo, Remi Munos, Charles Blundell, Dharshan Kumaran, and Matt Botvinick. 2016 · 2016
Cited alongside, same era.
Automatic Data Augmentation for Generalization in Reinforcement Learning. In Advances in Neural Information Processing Systems , M. Ranzato, A. Beygelzimer, Y. Dauphin, P.S. Liang, and J. Wortman Vaughan (Eds.), Vol. 34. Curran Associates, Inc., 5402–5415
Roberta Raileanu, Maxwell Goldstein, Denis Yarats, Ilya Kostrikov, and Rob Fergus. 2021 · 2021
Later among the works it cites.
The Distracting Control Suite – A Challenging Benchmark for Reinforcement Learning from Pixels
Austin Stone, Oscar Ramirez, Kurt Konolige, and Rico Jonschkowski. 2021 · 2021
Later among the works it cites.
Jinwei Xing, Takashi Nagata, Kexin Chen, Xinyun Zou, Emre Neftci, and Jeffrey L Krichmar. 2021 · 2021
Later among the works it cites.
Image Augmentation Is All You Need: Regularizing Deep Reinforcement Learning from Pixels
Denis Yarats, Ilya Kostrikov, and Rob Fergus. 2021 · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
John Schulman, Filip Wolski, Prafulla Dhariwal, Alec Radford, and Oleg Klimov. 2017 · 2017
Cited alongside, same era.
Josh Tobin, Rachel Fong, Alex Ray, Jonas Schneider, Wojciech Zaremba, and Pieter Abbeel. 2017 · 2017
Cited alongside, same era.
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Cited alongside, same era.
Places: A 10 million Image Database for Scene Recognition
Bolei Zhou, Agata Lapedriza, Aditya Khosla, Aude Oliva, and Antonio Torralba. 2017 · 2017
Cited alongside, same era.
Target-driven Visual Navigation in Indoor Scenes using Deep Reinforcement Learning
Y Zhu, R Mottaghi, E Kolve, JJ Lim, A Gupta, L Fei-Fei, and A Farhadi. 2017 · 2017
Cited alongside, same era.
Generalization and Regularization in DQN
Jesse Farebrother, Marlos C Machado, and Michael Bowling. 2018 · 2018
Cited alongside, same era.
Tuomas Haarnoja, Aurick Zhou, Pieter Abbeel, and Sergey Levine. 2018 · 2018
Cited alongside, same era.
A Rupam Mahmood, Dmytro Korenkevych, Gautham Vasan, William Ma, and James Bergstra. 2018 · 2018
Cited alongside, same era.
Look where you look! Saliency-guided Q-networks for generalization in visual Reinforcement Learning
David Bertoin, Adil Zouitine, Mehdi Zouitine, and Emmanuel Rachelson. 2022 · 2022
Later among the works it cites.
Magnetic control of tokamak plasmas through deep reinforcement learning
Jonas Degrave, Federico Felici, Jonas Buchli, Michael Neunert, Brendan Tracey, Francesco Carpanese, Timo Ewalds, Roland Hafner, Abbas Abdolmaleki, Diego de Las Casas, et al · 2022
Later among the works it cites.
Laura Graesser, Utku Evci, Erich Elsen, and Pablo Samuel Castro. 2022 · 2022
Later among the works it cites.
Evaluating Vision Transformer Methods for Deep Reinforcement Learning from Pixels
Tianxin Tao, Daniele Reda, and Michiel van de Panne. 2022 · 2022
Later among the works it cites.
Mask-based Latent Reconstruction for Reinforcement Learning
Tao Yu, Zhizheng Zhang, Cuiling Lan, Yan Lu, and Zhibo Chen. 2022 · 2022
Later among the works it cites.
Yufeng Yuan and A Rupam Mahmood. 2022 · 2022
Later among the works it cites.
Don’t Touch What Matters: Task-Aware Lipschitz Data Augmentation for Visual Reinforcement Learning
Zhecheng Yuan, Guozheng Ma, Yao Mu, Bo Xia, Bo Yuan, Xueqian Wang, Ping Luo, and Huazhe Xu. 2022a · 2022
Later among the works it cites.
Pre-Trained Image Encoder for Generalizable Visual Reinforcement Learning
Zhecheng Yuan, Zhengrong Xue, Bo Yuan, Xueqian Wang, Yi Wu, Yang Gao, and Huazhe Xu. 2022b · 2022
Later among the works it cites.
Open X-Embodiment: Robotic Learning Datasets and RT-X Models
Open X-Embodiment Collaboration et al. 2023 · 2023
Closest in time.
Automatic Noise Filtering with Dynamic Sparse Training in Deep Reinforcement Learning
Bram Grooten, Ghada Sokar, Shibhansh Dohare, Elena Mocanu, Matthew E. Taylor, Mykola Pechenizkiy, and Decebal Constantin Mocanu. 2023 · 2023
Closest in time.
Learning Agile Soccer Skills for a Bipedal Robot with Deep Reinforcement Learning
Tuomas Haarnoja, Ben Moran, Guy Lever, Sandy H Huang, Dhruva Tirumala, Markus Wulfmeier, Jan Humplik, Saran Tunyasuvunakool, Noah Y Siegel, Roland Hafner, et al · 2023
Closest in time.
A Survey of Zero-shot Generalisation in Deep Reinforcement Learning
Robert Kirk, Amy Zhang, Edward Grefenstette, and Tim Rocktäschel. 2023 · 2023
Closest in time.
Ignorance is Bliss: Robust Control via Information Gating
Manan Tomar, Riashat Islam, Sergey Levine, and Philip Bachman. 2023 · 2023
Closest in time.
Yan Wang, Gautham Vasan, and A Rupam Mahmood. 2023 · 2023
Closest in time.
Augmentation Curriculum Learning For Generalization in RL
Dylan Yung, Andrew Szot, Prithvijit Chattopadhyay, Judy Hoffman, and Zsolt Kira. 2023 · 2023
Closest in time.