Fetching the paper…
Reading the bibliography…
Visual augmentation has become a crucial technique for enhancing the visual robustness of imitation learning.
M. A. Goodale, “Visuomotor control: Where does vision end and action begin?” Current Biology , vol. 8, no. 14, pp. R489–R491, 1998
1998
Earlier work this paper cites.
J. Deng, W. Dong, R. Socher, L.-J. Li, K. Li, and L. Fei-Fei, “Imagenet: A large-scale hierarchical image database,” in 2009 IEEE conference on computer vision and pattern recognition . Ieee, 2009, pp. 248–255
2009
Earlier work this paper cites.
S. Levine, C. Finn, T. Darrell, and P. Abbeel, “End-to-end training of deep visuomotor policies,” Journal of Machine Learning Research , vol. 17, no. 39, pp. 1–40, 2016
2016
Earlier work this paper cites.
C. Finn, T. Yu, T. Zhang, P. Abbeel, and S. Levine, “One-shot visual imitation learning via meta-learning,” in Conference on robot learning . PMLR, 2017, pp. 357–368
2017
Earlier work this paper cites.
P. Abolghasemi, A. Mazaheri, M. Shah, and L. Boloni, “Pay attention!-robustifying a deep visuomotor policy through task-focused visual attention,” in Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2019, pp. 4254–4262
2019
Earlier work this paper cites.
H. Rezatofighi, N. Tsoi, J. Gwak, A. Sadeghian, I. Reid, and S. Savarese, “Generalized intersection over union: A metric and a loss for bounding box regression,” in Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2019, pp. 658–666
2019
Earlier work this paper cites.
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
T. Lüddecke and A. Ecker, “Image segmentation using text and image prompts,” in Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2022, pp. 7086–7096
2022
Earlier work this paper cites.
C. Chi, Z. Xu, S. Feng, E. Cousineau, Y. Du, B. Burchfiel, R. Tedrake, and S. Song, “Diffusion policy: Visuomotor policy learning via action diffusion,” The International Journal of Robotics Research , p. 02783649241273668, 2023
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2024
Cited alongside, same era.
A. Xie, L. Lee, T. Xiao, and C. Finn, “Decomposing the generalization gap in imitation learning for visual robotic manipulation,” in 2024 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2024, pp. 3153–3160
2024
Cited alongside, same era.
2024
Later among the works it cites.
2024
Later among the works it cites.
C. Wang, H. Fang, H.-S. Fang, and C. Lu, “Rise: 3d perception makes real-world robot imitation simple and effective,” in 2024 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) . IEEE, 2024, pp. 2870–2877
2024
Later among the works it cites.
2024
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2024
Cited alongside, same era.
2024
Cited alongside, same era.
2024
Cited alongside, same era.
Z. Chen, Z. Mandi, H. Bharadhwaj, M. Sharma, S. Song, A. Gupta, and V. Kumar, “Semantically controllable augmentations for generalizable robot learning,” The International Journal of Robotics Research , p. 02783649241273686, 2024
2024
Cited alongside, same era.
L. Fan, K. Chen, D. Krishnan, D. Katabi, P. Isola, and Y. Tian, “Scaling laws of synthetic images for model training… for now,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2024, pp. 7382–7392
2024
Cited alongside, same era.
2024
Cited alongside, same era.
2024
Cited alongside, same era.
A. E. Eshratifar, J. V. Soares, K. Thadani, S. Mishra, M. Kuznetsov, Y.-N. Ku, and P. De Juan, “Salient object-aware background generation using text-guided diffusion models,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2024, pp. 7489–7499
2024
Cited alongside, same era.
J. Wang, Y. Qin, K. Kuang, Y. Korkmaz, A. Gurumoorthy, H. Su, and X. Wang, “Cyberdemo: Augmenting simulated human demonstration for real-world dexterous manipulation,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2024, pp. 17 952–17 963
2024
Later among the works it cites.
Z. Zhuang, R. Wang, N. Ingelhag, V. Kyrki, and D. Kragic, “Enhancing visual domain robustness in behaviour cloning via saliency-guided augmentation,” in CoRL Workshop on Safe and Robust Robot Learning for Operation in the Real World , 2024
2024
Later among the works it cites.
2024
Later among the works it cites.
X. Lai, Z. Tian, Y. Chen, Y. Li, Y. Yuan, S. Liu, and J. Jia, “Lisa: Reasoning segmentation via large language model,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2024, pp. 9579–9589
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
Z. Zhang, B. Wu, X. Wang, Y. Luo, L. Zhang, Y. Zhao, P. Vajda, D. Metaxas, and L. Yu, “Avid: Any-length video inpainting with diffusion model,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2024, pp. 7162–7172
2024
Later among the works it cites.