Fetching the paper…
Reading the bibliography…
Training robots to perform complex control tasks from high-dimensional pixel input using reinforcement learning (RL) is sample-inefficient, because image observations are comprised primarily of task-irrelevant information.
L. Itti, C. Koch, and E. Niebur, “A model of saliency-based visual attention for rapid scene analysis,” IEEE Transactions on Pattern Analysis and Machine Intelligence , 1998
1998
Earlier work this paper cites.
Y. Tong, H. Konik, F. Cheikh, and A. Tremeau, “Full Reference Image Quality Assessment Based on Saliency Map Analysis,” Journal of Imaging Science and Technology , 2010
2010
Earlier work this paper cites.
Q. Li, Y. Zhou, and J. Yang, “Saliency Based Image Segmentation,” in International Conference on Information and Multimedia Technology (ICIMT) , 2011
2011
Earlier work this paper cites.
R. Achanta, A. Shaji, K. Smith, A. Lucchi, P. Fua, and S. Süsstrunk, “SLIC Superpixels Compared to State-of-the-Art Superpixel Methods,” IEEE Transactions on Pattern Analysis and Machine Intelligence , 2012
2012
Earlier work this paper cites.
2013
Earlier work this paper cites.
D. Rudoy, D. B. Goldman, E. Shechtman, and L. Zelnik-Manor, “Learning video saliency from human gaze using candidate selection,” in Computer Vision and Pattern Recognition (CVPR) , 2013
2013
Earlier work this paper cites.
D. P. Papadopoulos, A. D. Clarke, F. Keller, and V. Ferrari, “Training Object Class Detectors from Eye Tracking Data,” in European Conference on Computer Vision (ECCV) , 2014
2014
Earlier work this paper cites.
X. Wang, L. Gao, J. Song, and H. Shen, “Beyond Frame-level CNN: Saliency-Aware 3-D CNN With LSTM for Video Action Recognition,” IEEE Signal Processing Letters , 2016
2016
Earlier work this paper cites.
R. Zhao, W. Oyang, and X. Wang, “Person Re-Identification by Saliency Learning,” IEEE Transactions on Pattern Analysis and Machine Intelligence , 2016
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
A. Das, H. Agrawal, L. Zitnick, D. Parikh, and D. Batra, “Human Attention in Visual Question Answering: Do Humans and Deep Networks Look at the Same Regions?” Computer Vision and Image Understanding , 2017
2017
Earlier work this paper cites.
L. Wang, H. Lu, Y. Wang, M. Feng, D. Wang, B. Yin, and X. Ruan, “Learning to Detect Salient Objects with Image-level Supervision,” in Computer Vision and Pattern Recognition (CVPR) , 2017
2017
Earlier work this paper cites.
N. Liu, J. Han, and M.-H. Yang, “PiCANet: Learning Pixel-wise Contextual Attention for Saliency Detection,” in Computer Vision and Pattern Recognition (CVPR) , 2018
2018
Earlier work this paper cites.
T. Haarnoja, A. Zhou, K. Hartikainen, G. Tucker, S. Ha, J. Tan, V. Kumar, H. Zhu, A. Gupta, P. Abbeel, et al. , “Soft Actor-Critic Algorithms and Applications,” International Conference on Machine Learning (ICML) , 2018
2018
Earlier work this paper cites.
W. Wang, J. Shen, F. Guo, M.-M. Cheng, and A. Borji, “Revisiting Video Saliency: A Large-scale Benchmark and a New Model,” in Computer Vision and Pattern Recognition (CVPR) , 2018
2018
Cited alongside, same era.
2019
Cited alongside, same era.
A. Sax, B. Emi, A. R. Zamir, L. Guibas, S. Savarese, and J. Malik, “Mid-Level Visual Representations Improve Generalization and Sample Efficiency for Learning Visuomotor Policies,” Conference on Robot Learning (CoRL) , 2019
2019
Cited alongside, same era.
M. Laskin, A. Srinivas, and P. Abbeel, “CURL: Contrastive Unsupervised Representations for Reinforcement Learning,” in International Conference on Machine Learning (ICML) , 2020
2020
Cited alongside, same era.
N. Hansen and X. Wang, “Generalization in reinforcement learning by soft data augmentation,” in 2021 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2021, pp. 13 611–13 617
2021
Later among the works it cites.
A. Linardos, M. Kümmerer, O. Press, and M. Bethge, “DeepGaze IIE: Calibrated prediction in and out-of-domain for state-of-the-art saliency modeling,” in International Conference on Computer Vision (ICCV) , 2021
2021
Later among the works it cites.
N. Wilde, E. Biyik, D. Sadigh, and S. L. Smith, “Learning Reward Functions from Scale Feedback,” in Conference on Robot Learning (CoRL) , 2022
2022
Later among the works it cites.
S. Tao, X. Li, T. Mu, Z. Huang, Y. Qin, and H. Su, “Abstract-to-Executable Trajectory Translation for One-Shot Task Generalization,” in Neural Information Processing Systems (NeurIPS) Deep Reinforcement Learning Workshop , 2022
2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
M. Laskin, K. Lee, A. Stooke, L. Pinto, P. Abbeel, and A. Srinivas, “Reinforcement Learning with Augmented Data,” Neural Information Processing Systems (NeurIPS) , 2020
2020
Cited alongside, same era.
S. Tunyasuvunakool, A. Muldal, Y. Doron, S. Liu, S. Bohez, J. Merel, T. Erez, T. Lillicrap, N. Heess, and Y. Tassa, “dm_control: Software and Tasks for Continuous Control,” Software Impacts , 2020
2020
Cited alongside, same era.
T. Yu, D. Quillen, Z. He, R. Julian, K. Hausman, C. Finn, and S. Levine, “Meta-World: A Benchmark and Evaluation for Multi-Task and Meta Reinforcement Learning,” in Conference on Robot Learning (CoRL) , 2020
2020
Cited alongside, same era.
S. Cabi, S. G. Colmenarejo, A. Novikov, K. Konyushkova, S. Reed, R. Jeong, K. Zolna, Y. Aytar, D. Budden, M. Vecerik, et al. , “Scaling data-driven robotics with reward sketching and batch reinforcement learning,” Robotics: Science and Systems (RSS) , 2020
2020
Cited alongside, same era.
A. Atrey, K. Clary, and D. Jensen, “Exploratory Not Explanatory: Counterfactual Analysis of Saliency Maps for Deep Reinforcement Learning,” International Conference on Learning Representations (ICLR) , 2020
2020
Cited alongside, same era.
2020
Cited alongside, same era.
K. P. Darby, S. W. Deng, D. B. Walther, and V. M. Sloutsky, “The development of attention to objects and scenes: From object-biased to unbiased,” Child development , 2021
2021
Cited alongside, same era.
A. Bobu, M. Wiggert, C. Tomlin, and A. D. Dragan, “Feature Expansive Reward Learning: Rethinking Human Input,” in Human-Robot Interaction (HRI) , 2021
2021
Cited alongside, same era.
D. Bertoin, A. Zouitine, M. Zouitine, and E. Rachelson, “Look where you look! Saliency-guided Q-networks for generalization in visual Reinforcement Learning,” Neural Information Processing Systems (NeurIPS) , 2022
2022
Later among the works it cites.
A. Boyd, K. W. Bowyer, and A. Czajka, “Human-Aided Saliency Maps Improve Generalization of Deep Learning,” in Winter Conference on Applications of Computer Vision (WACV) , 2022
2022
Later among the works it cites.
S. Nair, A. Rajeswaran, V. Kumar, C. Finn, and A. Gupta, “R3M: A Universal Visual Representation for Robot Manipulation,” Conference on Robot Learning (CoRL) , 2022
2022
Later among the works it cites.
R. Bachmann, D. Mizrahi, A. Atanov, and A. Zamir, “MultiMAE: Multi-modal Multi-task Masked Autoencoders,” European Conference on Computer Vision (ECCV) , 2022
2022
Later among the works it cites.
K. He, X. Chen, S. Xie, Y. Li, P. Dollár, and R. Girshick, “Masked Autoencoders Are Scalable Vision Learners,” in Computer Vision and Pattern Recognition (CVPR) , 2022
2022
Later among the works it cites.
M. Kümmerer, M. Bethge, and T. S. Wallis, “Deepgaze iii: Modeling free-viewing human scanpaths with deep learning,” Journal of Vision , 2022
2022
Later among the works it cites.
A. Boyd, P. Tinsley, K. W. Bowyer, and A. Czajka, “CYBORG: Blending Human Saliency Into the Loss Improves Deep Learning,” in Winter Conference on Applications of Computer Vision (WACV) , 2023
2023
Later among the works it cites.
2023
Later among the works it cites.