Fetching the paper…
Reading the bibliography…
Generalization to unseen real-world scenarios for robot manipulation requires exposure to diverse datasets during training.
arXiv preprint arXiv:1903.00374
Kaiser L, Babaeizadeh M, Milos P, Osinski B, Campbell RH, Czechowski K, Erhan D, Finn C, Kozakowski P, Levine S et al. (2019) Model-based reinforcement learning for atari · 1903
Earlier work this paper cites.
arXiv preprint arXiv:1909.12238
Song HF, Abdolmaleki A, Springenberg JT, Clark A, Soyer H, Rae JW, Noury S, Ahuja A, Liu S, Tirumala D et al. (2019) V-mpo: On-policy maximum a posteriori policy optimization for discrete and continuous control · 1909
Earlier work this paper cites.
arXiv preprint arXiv:1912.01603
Hafner D, Lillicrap T, Ba J and Norouzi M (2019) Dream to control: Learning behaviors by latent imagination · 1912
Earlier work this paper cites.
Advances in neural information processing systems 14
Kakade SM (2001) A natural policy gradient · 2001
Earlier work this paper cites.
arXiv preprint arXiv:2004.13649
Kostrikov I, Yarats D and Fergus R (2020) Image augmentation is all you need: Regularizing deep reinforcement learning from pixels · 2004
Earlier work this paper cites.
arXiv preprint arXiv:2005.07648
Lynch C and Sermanet P (2020) Language conditioned imitation learning over unstructured data · 2005
Earlier work this paper cites.
In: 2009 IEEE Computer Society Conference on Computer Vision and Pattern Recognition (CVPR 2009), 20-25 June 2009, Miami, Florida, USA . IEEE Computer Society, pp. 248–255
Deng J, Dong W, Socher R, Li L, Li K and Fei-Fei L (2009a) Imagenet: A large-scale hierarchical image database · 2009
Earlier work this paper cites.
In: 2009 IEEE conference on computer vision and pattern recognition . Ieee, pp. 248–255
Deng J, Dong W, Socher R, Li LJ, Li K and Fei-Fei L (2009b) Imagenet: A large-scale hierarchical image database · 2009
Earlier work this paper cites.
arXiv preprint arXiv:2009.12293
Zhu Y, Wong J, Mandlekar A, Martín-Martín R, Joshi A, Nasiriany S and Zhu Y (2020) robosuite: A modular simulation framework and benchmark for robot learning · 2009
Earlier work this paper cites.
arXiv preprint arXiv:2010.14406
Zeng A, Florence P, Tompson J, Welker S, Chien J, Attarian M, Armstrong T, Krasin I, Duong D, Sindhwani V et al. (2020) Transporter networks: Rearranging the visual world for robotic manipulation · 2010
Earlier work this paper cites.
In: Proceedings of the AAAI Conference on Artificial Intelligence , volume 25. pp. 1507–1514
Tellex S, Kollar T, Dickerson S, Walter M, Banerjee A, Teller S and Roy N (2011) Understanding natural language commands for robotic navigation and mobile manipulation · 2011
Earlier work this paper cites.
arXiv preprint arXiv:1312.6114
Kingma DP and Welling M (2013) Auto-encoding variational bayes · 2013
Earlier work this paper cites.
In: 2015 IEEE-RAS 15th International Conference on Humanoid Robots (Humanoids) . IEEE, pp. 657–663
Kumar V and Todorov E (2015) Mujoco haptix: A virtual reality system for hand manipulation · 2015
Earlier work this paper cites.
Levine S, Finn C, Darrell T and Abbeel P (2015) End-to-end training of deep visuomotor policies · 2015
Earlier work this paper cites.
In: Conference on robot learning . PMLR, pp. 357–368
Finn C, Yu T, Zhang T, Abbeel P and Levine S (2017) One-shot visual imitation learning via meta-learning · 2017
Earlier work this paper cites.
arXiv preprint arXiv:1712.04621
Perez L and Wang J (2017) The effectiveness of data augmentation in image classification using deep learning · 2017
Earlier work this paper cites.
arXiv preprint arXiv:1710.06542
Pinto L, Andrychowicz M, Welinder P, Zaremba W and Abbeel P (2017) Asymmetric actor critic for image-based robot learning · 2017
Earlier work this paper cites.
In: 2017 IEEE international conference on robotics and automation (ICRA) . IEEE, pp. 2161–2168
Pinto L and Gupta A (2017) Learning to push by grasping: Using multiple tasks for effective learning · 2017
Earlier work this paper cites.
In: 2017 IEEE/RSJ international conference on intelligent robots and systems (IROS) . IEEE, pp. 23–30
Tobin J, Fong R, Ray A, Schneider J, Zaremba W and Abbeel P (2017) Domain randomization for transferring deep neural networks from simulation to the real world · 2017
Earlier work this paper cites.
Advances in neural information processing systems 30
Vaswani A, Shazeer N, Parmar N, Uszkoreit J, Jones L, Gomez AN, Kaiser Ł and Polosukhin I (2017) Attention is all you need · 2017
Earlier work this paper cites.
arXiv preprint arXiv:1805.09501
Cubuk ED, Zoph B, Mane D, Vasudevan V and Le QV (2018) Autoaugment: Learning augmentation policies from data · 2018
Earlier work this paper cites.
arXiv preprint arXiv:1802.01561
Espeholt L, Soyer H, Munos R, Simonyan K, Mnih V, Ward T, Doron Y, Firoiu V, Harley T, Dunning I et al. (2018) Impala: Scalable distributed deep-rl with importance weighted actor-learner architectures · 2018
Earlier work this paper cites.
arXiv preprint arXiv:1803.10122
Ha D and Schmidhuber J (2018) World models · 2018
Earlier work this paper cites.
In: Conference on Robot Learning . PMLR, pp. 879–893
Mandlekar A, Zhu Y, Garg A, Booher J, Spero M, Tung A, Gao J, Emmons J, Gupta A, Orbay E et al. (2018) Roboturk: A crowdsourcing platform for robotic skill learning through imitation · 2018
Earlier work this paper cites.
Advances in neural information processing systems 31
Nair AV, Pong V, Dalal M, Bahl S, Lin S and Levine S (2018) Visual reinforcement learning with imagined goals · 2018
Earlier work this paper cites.
In: 2018 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, pp. 3782–3788
Nguyen A, Kanoulas D, Muratore L, Caldwell DG and Tsagarakis NG (2018) Translating videos to commands for robotic manipulation with deep recurrent neural networks · 2018
Earlier work this paper cites.
In: Proceedings of the AAAI Conference on Artificial Intelligence , volume 32
Perez E, Strub F, De Vries H, Dumoulin V and Courville A (2018) Film: Visual reasoning with a general conditioning layer · 2018
Earlier work this paper cites.
Qureshi AH, Bency MJ and Yip MC (2018) Motion planning networks · 2018
Earlier work this paper cites.
In: 2019 International Conference on Robotics and Automation (ICRA) . IEEE, pp. 2125–2131
Berscheid L, Rühr T and Kröger T (2019) Improving data efficiency of self-supervised learning for robotic grasping · 2019
Earlier work this paper cites.
In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition
Gupta A, Dollar P and Girshick R (2019) LVIS: A dataset for large vocabulary instance segmentation · 2019
Cited alongside, same era.
In: CoRL
Nagabandi A, Konoglie K, Levine S and Kumar V (2019) Deep dynamics models for learning dexterous manipulation · 2019
Cited alongside, same era.
Advances in Neural Information Processing Systems 33: 17605–17616
Benton G, Finzi M, Izmailov P and Wilson AG (2020) Learning invariances in neural networks from training data · 2020
Cited alongside, same era.
IEEE Robotics and Automation Letters 5(2): 3019–3026
James S, Ma Z, Arrojo DR and Davison AJ (2020) Rlbench: The robot learning benchmark & learning environment · 2020
Cited alongside, same era.
In: Conference on robot learning . PMLR, pp. 1113–1132
Lynch C, Khansari M, Xiao T, Kumar V, Tompson J, Levine S and Sermanet P (2020) Learning latent plans from play · 2020
Cited alongside, same era.
arXiv preprint arXiv:2204.06125
Ramesh A, Dhariwal P, Nichol A, Chu C and Chen M (2022) Hierarchical text-conditional image generation with clip latents · 2022
Later among the works it cites.
arXiv preprint arXiv:2205.06175
Reed S, Zolna K, Parisotto E, Colmenarejo SG, Novikov A, Barth-Maron G, Gimenez M, Sulsky Y, Kay J, Springenberg JT et al. (2022) A generalist agent · 2022
Later among the works it cites.
arXiv preprint arXiv:2205.11487
Saharia C, Chan W, Saxena S, Li L, Whang J, Denton E, Ghasemipour SKS, Ayan BK, Mahdavi SS, Lopes RG et al. (2022) Photorealistic text-to-image diffusion models with deep language understanding · 2022
Later among the works it cites.
arXiv preprint arXiv:2206.11251
Shafiullah NMM, Cui ZJ, Altanzaya A and Pinto L (2022) Behavior transformers: Cloning k k modes with one stone · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Rao K, Harris C, Irpan A, Levine S, Ibarz J and Khansari M (2020) Rl-cyclegan: Reinforcement learning aware simulation-to-real · 2020
Cited alongside, same era.
Nature 588(7839): 604–609
Schrittwieser J, Antonoglou I, Hubert T, Simonyan K, Sifre L, Schmitt S, Guez A, Lockhart E, Hassabis D, Graepel T et al. (2020) Mastering atari, go, chess and shogi by planning with a learned model · 2020
Cited alongside, same era.
Advances in Neural Information Processing Systems 33: 13139–13150
Stepputtis S, Campbell J, Phielipp M, Lee S, Baral C and Ben Amor H (2020) Language-conditioned imitation learning for robot manipulation tasks · 2020
Cited alongside, same era.
In: Conference on Robot Learning (CoRL)
Young S, Gandhi D, Tulsiani S, Gupta A, Abbeel P and Pinto L (2020) Visual imitation made easy · 2020
Cited alongside, same era.
In: Proceedings of the IEEE/CVF International Conference on Computer Vision . pp. 12200–12209
Deng C, Litany O, Duan Y, Poulenard A, Tagliasacchi A and Guibas LJ (2021) Vector neurons: A general framework for so (3)-equivariant networks · 2021
Cited alongside, same era.
In: 2021 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, pp. 6664–6671
Gupta A, Yu J, Zhao TZ, Kumar V, Rovinsky A, Xu K, Devlin T and Levine S (2021) Reset-free reinforcement learning via multi-task learning: Learning dexterous manipulation behaviors without human intervention · 2021
Cited alongside, same era.
Rombach R, Blattmann A, Lorenz D, Esser P and Ommer B (2021) High-resolution image synthesis with latent diffusion models
2021
Cited alongside, same era.
In: Conference on Robot Learning . PMLR, pp. 894–906
Shridhar M, Manuelli L and Fox D (2022) Cliport: What and where pathways for robotic manipulation · 2022
Later among the works it cites.
arXiv preprint arXiv:2209.14792
Singer U, Polyak A, Hayes T, Yin X, An J, Zhang S, Hu Q, Yang H, Ashual O, Gafni O et al. (2022) Make-a-video: Text-to-video generation without text-video data · 2022
Later among the works it cites.
Wang D, Jia M, Zhu X, Walters R and Platt R (2022) On-robot policy learning with o(2)-equivariant SAC · 2022
Later among the works it cites.
Applied Sciences 12(17): 8643
Zafar A, Aamir M, Mohd Nawi N, Arshad A, Riaz S, Alruban A, Dutta AK and Almotairi S (2022) A comparison of pooling methods for convolutional neural networks · 2022
Later among the works it cites.
In: arXiv preprint arXiv:2212.04501
Zhao Y, Misra I, Krähenbühl P and Girdhar R (2022) Learning video representations from large language models · 2022
Later among the works it cites.
arXiv preprint arXiv:2304.08488
Bahl S, Mendonca R, Chen L, Jain U and Pathak D (2023) Affordances from human videos as a versatile representation for robotics · 2023
Later among the works it cites.
In: 2023 IEEE International Conference on Robotics and Automation (ICRA)
Bharadhwaj H, Gupta A and Tulsiani S (2023b) Visual affordance prediction for guiding robot exploration · 2023
Later among the works it cites.
arXiv preprint arXiv:2310.10639
Black K, Nakamoto M, Atreya P, Walke H, Finn C, Kumar A and Levine S (2023) Zero-shot robotic manipulation with pretrained image-editing diffusion models · 2023
Later among the works it cites.
arXiv preprint arXiv:2306.11706
Bousmalis K, Vezzani G, Rao D, Devin C, Lee AX, Bauza M, Davchev T, Zhou Y, Gupta A, Raju A et al. (2023) Robocat: A self-improving foundation agent for robotic manipulation · 2023
Later among the works it cites.
In: Conference on Robot Learning . PMLR, pp. 287–318
Brohan A, Chebotar Y, Finn C, Hausman K, Herzog A, Ho D, Ibarz J, Irpan A, Jang E, Julian R et al. (2023) Do as i can, not as i say: Grounding language in robotic affordances · 2023
Later among the works it cites.
arXiv preprint arXiv:2302.06671
Chen Z, Kiami S, Gupta A and Kumar V (2023) Genaug: Retargeting behaviors to unseen situations via generative augmentation · 2023
Later among the works it cites.
In: 2023 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, pp. 5977–5984
Handa A, Allshire A, Makoviychuk V, Petrenko A, Singh R, Liu J, Makoviichuk D, Van Wyk K, Zhurkevich A, Sundaralingam B et al. (2023) Dextreme: Transfer of agile in-hand manipulation from simulation to reality · 2023
Later among the works it cites.
IEEE Robotics and Automation Letters 8(7): 3956–3963
Kapelyukh I, Vosylius V and Johns E (2023) Dall-e-bot: Introducing web-scale diffusion models to robotics · 2023
Later among the works it cites.
arXiv preprint arXiv:2303.18240
Majumdar A, Yadav K, Arnaud S, Ma YJ, Chen C, Silwal S, Jain A, Berges VP, Abbeel P, Malik J et al. (2023) Where are we in the search for an artificial visual cortex for embodied intelligence? · 2023
Later among the works it cites.
IEEE Robotics and Automation Letters
Mittal M, Yu C, Yu Q, Liu J, Rudin N, Hoeller D, Yuan JL, Singh R, Guo Y, Mazhar H et al. (2023) Orbit: A unified simulation framework for interactive robot learning environments · 2023
Later among the works it cites.
In: Proceedings of the IEEE/CVF International Conference on Computer Vision . pp. 15579–15591
Momeni L, Caron M, Nagrani A, Zisserman A and Schmid C (2023) Verbs in action: Improving verb understanding in video-language models · 2023
Later among the works it cites.
In: Conference on Robot Learning . PMLR, pp. 416–426
Radosavovic I, Xiao T, James S, Abbeel P, Malik J and Darrell T (2023) Real-world robot learning with masked visual pre-training · 2023
Later among the works it cites.
arXiv preprint arXiv:2304.06600
Sharma M, Fantacci C, Zhou Y, Koppula S, Heess N, Scholz J and Aytar Y (2023) Lossless adaptation of pretrained vision models for robotic manipulation · 2023
Later among the works it cites.
In: Conference on Robot Learning . PMLR, pp. 654–665
Shaw K, Bahl S and Pathak D (2023) Videodex: Learning dexterity from internet videos · 2023
Later among the works it cites.
arXiv preprint arXiv:2302.12422
Wang C, Fan L, Sun J, Zhang R, Fei-Fei L, Xu D, Zhu Y and Anandkumar A (2023) Mimicplay: Long-horizon imitation learning by watching human play · 2023
Later among the works it cites.
arXiv preprint arXiv:2302.11550
Yu T, Xiao T, Stone A, Tompson J, Brohan A, Wang S, Singh J, Tan C, Peralta J, Ichter B et al. (2023) Scaling robot learning with semantically imagined experience · 2023
Later among the works it cites.
arXiv preprint arXiv:2304.13705
Zhao TZ, Kumar V, Levine S and Finn C (2023) Learning fine-grained bimanual manipulation with low-cost hardware · 2023
Later among the works it cites.
In: Conference on Robot Learning . PMLR, pp. 2165–2183
Zitkovich B, Yu T, Xu S, Xu P, Xiao T, Xia F, Wu J, Wohlhart P, Welker S, Wahid A et al. (2023) Rt-2: Vision-language-action models transfer web knowledge to robotic control · 2023
Later among the works it cites.
arXiv preprint arXiv:2405.01527
Bharadhwaj H, Mottaghi R, Gupta A and Tulsiani S (2024) Track2act: Predicting point tracks from internet videos enables diverse zero-shot robot manipulation · 2024
Closest in time.