Fetching the paper…
Reading the bibliography…
We introduce the first work to explore web-scale diffusion models for robotics.
H. W. Kuhn and B. Yaw, “The Hungarian method for the assignment problem,” Naval Res. Logist. Quart , 1955
1955
Earlier work this paper cites.
P. Besl and N. D. McKay, “A method for registration of 3-d shapes,” Transactions on Pattern Analysis and Machine Intelligence , 1992
1992
Earlier work this paper cites.
Y. Jiang, M. Lim, and A. Saxena, “Learning object arrangements in 3D scenes using human context,” International Conference on Machine Learning, ICML , 2012
2012
Earlier work this paper cites.
M. J. Schuster, D. Jain, M. Tenorth, and M. Beetz, “Learning organizational principles in human environments,” in International Conference on Robotics and Automation , 2012, pp. 3867–3874
2012
Earlier work this paper cites.
N. Abdo, C. Stachniss, L. Spinello, and W. Burgard, “Robot, organize my shelves! Tidying up objects by predicting user preferences,” in International Conference on Robotics and Automation , 2015
2015
Earlier work this paper cites.
J. Sohl-Dickstein, E. Weiss, N. Maheswaranathan, and S. Ganguli, “Deep unsupervised learning using nonequilibrium thermodynamics,” in International Conference on Machine Learning , 2015
2015
Earlier work this paper cites.
E. Johns, S. Leutenegger, and A. J. Davison, “Deep learning a grasp function for grasping under gripper pose uncertainty,” in International Conference on Intelligent Robots and Systems (IROS) , 2016
2016
Earlier work this paper cites.
C. Finn and S. Levine, “Deep visual foresight for planning robot motion,” in International Conference on Robotics and Automation , 2017
2017
Earlier work this paper cites.
K. He, G. Gkioxari, P. Dollár, and R. Girshick, “Mask R-CNN,” in International Conference on Computer Vision (ICCV) , 2017
2017
Earlier work this paper cites.
M. Kang, Y. Kwon, and S.-E. Yoon, “Automated task planning using object arrangement optimization,” in International Conference on Ubiquitous Robots (UR) , 2018
2018
Earlier work this paper cites.
A. Nair, V. Pong, M. Dalal, S. Bahl, S. Lin, and S. Levine, “Visual reinforcement learning with imagined goals,” in nternational Conference on Neural Information Processing Systems , 2018
2018
Earlier work this paper cites.
A. Akbik, D. Blythe, and R. Vollgraf, “Contextual string embeddings for sequence labeling,” in COLING , 2018
2018
Earlier work this paper cites.
C. Lynch, M. Khansari, T. Xiao, V. Kumar, J. Tompson, S. Levine, and P. Sermanet, “Learning latent plans from play,” Conference on Robot Learning (CoRL) , 2019
2019
Earlier work this paper cites.
Y. Wu, A. Kirillov, F. Massa, W.-Y. Lo, and R. Girshick, “Detectron2,” https://github.com/facebookresearch/detectron2 , 2019
2019
Earlier work this paper cites.
A. Gupta, P. Dollar, and R. Girshick, “LVIS: A dataset for large vocabulary instance segmentation,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2019
2019
Earlier work this paper cites.
A. Akbik, T. Bergmann, D. Blythe, K. Rasul, S. Schweter, and R. Vollgraf, “FLAIR: An easy-to-use framework for state-of-the-art NLP,” in NAACL , 2019
2019
Earlier work this paper cites.
A. Mousavian, C. Eppner, and D. Fox, “6-DOF GraspNet: Variational grasp generation for object manipulation,” in International Conference on Computer Vision (ICCV) , 2019
2019
Earlier work this paper cites.
D. Batra, A. X. Chang, S. Chernova, A. J. Davison, J. Deng, V. Koltun, S. Levine, J. Malik, I. Mordatch, R. Mottaghi, M. Savva, and H. Su, “Rearrangement: A challenge for embodied AI,” arXiv , 2020
2020
Cited alongside, same era.
A. Zeng, P. Florence, J. Tompson, S. Welker, J. Chien, M. Attarian, T. Armstrong, I. Krasin, D. Duong, V. Sindhwani, and J. Lee, “Transporter networks: Rearranging the visual world for robotic manipulation,” Conference on Robot Learning (CoRL) , 2020
2020
Cited alongside, same era.
I. Kapelyukh and E. Johns, “My house, my rules: Learning tidying preferences with graph neural networks,” in Conference on Robot Learning (CoRL) , 2021
2021
Cited alongside, same era.
M. Shridhar, L. Manuelli, and D. Fox, “CLIPort: What and where pathways for robotic manipulation,” in Conference on Robot Learning (CoRL) , 2021
2021
Cited alongside, same era.
M. Wu, fangwei zhong, Y. Xia, and H. Dong, “TarGF: Learning target gradient field for object rearrangement,” in Conference on Neural Information Processing Systems , 2022
2022
Closest in time.
A. Goyal, A. Mousavian, C. Paxton, Y.-W. Chao, B. Okorn, J. Deng, and D. Fox, “IFOR: Iterative flow minimization for robotic object rearrangement,” in Conference on Computer Vision and Pattern Recognition (CVPR) , 2022
2022
Closest in time.
M. Janner, Y. Du, J. Tenenbaum, and S. Levine, “Planning with diffusion for flexible behavior synthesis,” in International Conference on Machine Learning , 2022
2022
Closest in time.
Z. Mandi, H. Bharadhwaj, V. Moens, S. Song, A. Rajeswaran, and V. Kumar, “CACTI: A framework for scalable multi-task multi-scene visual imitation learning,” arXiv , 2022
2022
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
R. Rombach, A. Blattmann, D. Lorenz, P. Esser, and B. Ommer, “High-resolution image synthesis with latent diffusion models,” arXiv , 2021
2021
Cited alongside, same era.
A. S. Chen, S. Nair, and C. Finn, “Learning generalizable robotic reward functions from “in-the-wild” human videos,” Robotics: Science and Systems , 2021
2021
Cited alongside, same era.
A. Radford, J. W. Kim, C. Hallacy, A. Ramesh, G. Goh, S. Agarwal, G. Sastry, A. Askell, P. Mishkin, J. Clark, G. Krueger, and I. Sutskever, “Learning transferable visual models from natural language supervision,” in International Conference on Machine Learning, ICML , 2021
2021
Cited alongside, same era.
T. Ridnik, E. Ben-Baruch, A. Noy, and L. Zelnik-Manor, “ImageNet-21K pretraining for the masses,” arXiv , 2021
2021
Cited alongside, same era.
T. Ridnik, H. Lawen, A. Noy, E. Ben, B. G. Sharir, and I. Friedman, “TResNet: High performance gpu-dedicated architecture,” in Winter Conference on Applications of Computer Vision (WACV) , 2021
2021
Cited alongside, same era.
E. Johns, “Coarse-to-fine imitation learning: Robot manipulation from a single demonstration,” in International Conference on Robotics and Automation , 2021
2021
Cited alongside, same era.
Y. Kant, A. Ramachandran, S. Yenamandra, I. Gilitschenski, D. Batra, A. Szot, and H. Agrawal, “Housekeep: Tidying virtual households using commonsense reasoning,” arXiv , 2022
2022
Cited alongside, same era.
W. Liu, C. Paxton, T. Hermans, and D. Fox, “StructFormer: Learning spatial structure for language-guided semantic rearrangement of novel objects,” International Conference on Robotics and Automation , 2022
2022
Cited alongside, same era.
P. Wang, A. Yang, R. Men, J. Lin, S. Bai, Z. Li, J. Ma, C. Zhou, J. Zhou, and H. Yang, “OFA: Unifying architectures, tasks, and modalities through a simple sequence-to-sequence learning framework,” in International Conference on Machine Learning , 2022
2022
Closest in time.
W. Goodwin, S. Vaze, I. Havoutis, and I. Posner, “Semantically grounded object matching for robust robotic scene rearrangement,” in International Conference on Robotics and Automation , 2022
2022
Closest in time.
V. Vosylius and E. Johns, “Where to start? Transferring simple skills to complex environments,” in Conference on Robot Learning , 2022
2022
Closest in time.
K. Mazur, E. Sucar, and A. J. Davison, “Feature-realistic neural fusion for real-time, open set scene understanding,” arXiv , 2022
2022
Closest in time.
R. Gal, Y. Alaluf, Y. Atzmon, O. Patashnik, A. H. Bermano, G. Chechik, and D. Cohen-Or, “An image is worth one word: Personalizing text-to-image generation using textual inversion,” arXiv , 2022
2022
Closest in time.
T. Kojima, S. S. Gu, M. Reid, Y. Matsuo, and Y. Iwasawa, “Large language models are zero-shot reasoners,” ArXiv , 2022
2022
Closest in time.
C. Conwell and T. Ullman, “Testing relational understanding in text-guided image generation,” arXiv , 2022
2022
Closest in time.
Q. A. Wei, S. Ding, J. J. Park, R. Sajnani, A. Poulenard, S. Sridhar, and L. Guibas, “Lego-net: Learning regular rearrangements of objects in rooms,” arXiv , 2023
2023
Closest in time.
C. Chi, S. Feng, Y. Du, Z. Xu, E. Cousineau, B. Burchfiel, and S. Song, “Diffusion policy: Visuomotor policy learning via action diffusion,” arXiv , 2023
2023
Closest in time.
Z. Chen, S. Kiami, A. Gupta, and V. Kumar, “GenAug: Retargeting behaviors to unseen situations via generative augmentation,” arXiv , 2023
2023
Closest in time.
T. Yu, T. Xiao, A. Stone, J. Tompson, A. Brohan, S. Wang, J. Singh, C. Tan, D. M, J. Peralta, B. Ichter, K. Hausman, and F. Xia, “Scaling robot learning with semantically imagined experience,” arXiv , 2023
2023
Closest in time.
Y. Du, M. Yang, B. Dai, H. Dai, O. Nachum, J. B. Tenenbaum, D. Schuurmans, and P. Abbeel, “Learning universal policies via text-guided video generation,” arXiv , 2023
2023
Closest in time.