Fetching the paper…
Reading the bibliography…
We present Generative Predictive Control (GPC), an inference-time method for improving pretrained behavior-cloning policies without retraining.
J. Carvalho, A. T. Le, M. Baierl, D. Koert, and J. Peters, “Motion planning diffusion: Learning and planning of robot motions with diffusion models.” IEEE, 2023, pp. 1916–1923
1923
Earlier work this paper cites.
L. Lennart, “System identification: theory for the user,” PTR Prentice Hall, Upper Saddle River, NJ , vol. 28, p. 540, 1999
1999
Earlier work this paper cites.
M. Diehl and S. Gros, “Numerical optimal control,” Optimization in Engineering Center (OPTEC) , 2011
2011
Earlier work this paper cites.
E. Olson, “AprilTag: A robust and flexible visual fiducial system,” in Proc. IEEE Int. Conf. Robot. Autom. IEEE, 2011, pp. 3400–3407
2011
Earlier work this paper cites.
J. Oh, X. Guo, H. Lee, R. L. Lewis, and S. Singh, “Action-conditional video prediction using deep networks in atari games,” in Adv. Neural Inf. Process. Syst. , vol. 28, 2015
2015
Earlier work this paper cites.
2017
Earlier work this paper cites.
C. Finn and S. Levine, “Deep visual foresight for planning robot motion,” in Proc. IEEE Int. Conf. Robot. Autom. IEEE, 2017, pp. 2786–2793
2017
Earlier work this paper cites.
D. Ha and J. Schmidhuber, “World models,” arXiv preprint arXiv:1803.10122 , 2018
2018
Earlier work this paper cites.
S. Oprea, P. Martinez-Gonzalez, A. Garcia-Garcia, J. A. Castro-Vargas, S. Orts-Escolano, J. Garcia-Rodriguez, and A. Argyros, “A review on deep learning techniques for video prediction,” in IEEE Trans. Pattern Anal. Mach. Intell. , vol. 44, no. 6. IEEE, 2020, pp. 2806–2826
2020
Earlier work this paper cites.
J. Ho, A. Jain, and P. Abbeel, “Denoising diffusion probabilistic models,” in Adv. Neural Inf. Process. Syst. , vol. 33, 2020, pp. 6840–6851
2020
Earlier work this paper cites.
J. Ho, T. Salimans, A. Gritsenko, W. Chan, M. Norouzi, and D. J. Fleet, “Video diffusion models,” in Adv. Neural Inf. Process. Syst. , vol. 35, 2022, pp. 8633–8646
2022
Earlier work this paper cites.
W. Liu, T. Hermans, S. Chernova, and C. Paxton, “Structdiffusion: Object-centric diffusion for semantic rearrangement of novel objects,” in Workshop on Language and Robotics at CoRL 2022 , 2022
2022
Earlier work this paper cites.
C. Chi, Z. Xu, S. Feng, E. Cousineau, Y. Du, B. Burchfiel, R. Tedrake, and S. Song, “Diffusion policy: Visuomotor policy learning via action diffusion,” in Proc. Robot.: Sci. Syst. , 2023
2023
Earlier work this paper cites.
R. Firoozi, J. Tucker, S. Tian, A. Majumdar, J. Sun, W. Liu, Y. Zhu, S. Song, A. Kapoor, K. Hausman, et al. , “Foundation models in robotics: Applications, challenges, and the future,” The International Journal of Robotics Research , p. 02783649241281508, 2023
2023
Earlier work this paper cites.
L. Wang, Y. Ling, Z. Yuan, M. Shridhar, C. Bao, Y. Qin, B. Wang, H. Xu, and X. Wang, “GenSim: Generating robotic simulation tasks via large language models,” in Proc. Int. Conf. Learn. Represent. , 2023
2023
Earlier work this paper cites.
Z. Chen, S. Kiami, A. Gupta, and V. Kumar, “GenAug: Retargeting behaviors to unseen situations via generative augmentation,” in Proc. Robot.: Sci. Syst. , 2023
2023
Cited alongside, same era.
V. Micheli, E. Alonso, and F. Fleuret, “Transformers are sample-efficient world models,” in Proc. Int. Conf. Learn. Represent. , 2023
2023
Cited alongside, same era.
J. Urain, A. Mandlekar, Y. Du, M. Shafiullah, D. Xu, K. Fragkiadaki, G. Chalvatzaki, and J. Peters, “Deep generative models in robotics: A survey on learning from multimodal demonstrations,” IEEE Trans. Robot. , 2024
2024
Cited alongside, same era.
M. J. Kim, K. Pertsch, S. Karamcheti, T. Xiao, A. Balakrishna, S. Nair, R. Rafailov, E. Foster, G. Lam, P. Sanketi, et al. , “OpenVLA: An open-source vision-language-action model,” in Proc. Conf. Robot. Learn. , 2024
2024
Cited alongside, same era.
P.-C. Ko, J. Mao, Y. Du, S.-H. Sun, and J. B. Tenenbaum, “Learning to act from actionless videos through dense correspondences,” in Proc. Int. Conf. Learn. Represent. , 2024
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
Y. Du, M. Yang, P. Florence, F. Xia, A. Wahid, B. Ichter, P. Sermanet, T. Yu, P. Abbeel, J. B. Tenenbaum, et al. , “Video language planning,” in Proc. Int. Conf. Learn. Represent. , 2024
2024
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
X. Li, K. Hsu, J. Gu, K. Pertsch, O. Mees, H. R. Walke, C. Fu, I. Lunawat, I. Sieh, S. Kirmani, et al. , “Evaluating real-world robot manipulation policies in simulation,” in Proc. Conf. Robot. Learn. , 2024
2024
Cited alongside, same era.
E. Alonso, A. Jelley, V. Micheli, A. Kanervisto, A. J. Storkey, T. Pearce, and F. Fleuret, “Diffusion for world modeling: Visual details matter in atari,” in Adv. Neural Inf. Process. Syst. , vol. 37, 2024, pp. 58 757–58 791
2024
Cited alongside, same era.
D. Hafner, J. Pasukonis, J. Ba, and T. Lillicrap, “Mastering diverse domains through world models,” in Proc. Int. Conf. Learn. Represent. , 2024
2024
Cited alongside, same era.
2024
Cited alongside, same era.
S. Lee, Y. Wang, H. Etukuru, H. J. Kim, N. M. M. Shafiullah, and L. Pinto, “Behavior generation with latent actions,” in Proc. Int. Conf. Mach. Learn. , 2024
2024
Cited alongside, same era.
Y. Wang, Z. Xian, F. Chen, T.-H. Wang, Y. Wang, K. Fragkiadaki, Z. Erickson, D. Held, and C. Gan, “Robogen: Towards unleashing infinite data for automated robot learning via generative simulation,” in Proc. Int. Conf. Mach. Learn. , 2024
2024
Cited alongside, same era.
M. Yang, Y. Du, K. Ghasemipour, J. Tompson, D. Schuurmans, and P. Abbeel, “Learning interactive real-world simulators,” in Proc. Int. Conf. Learn. Represent. , 2024
2024
Cited alongside, same era.
B. Yang, H. Su, N. Gkanatsios, T.-W. Ke, A. Jain, J. Schneider, and K. Fragkiadaki, “Diffusion-es: Gradient-free planning with diffusion for autonomous driving and zero-shot instruction following,” in Proc. IEEE Conf. Comput. Vis. Pattern Recognit. , 2024
2024
Cited alongside, same era.
N. Hansen, H. Su, and X. Wang, “Td-mpc2: Scalable, robust world models for continuous control,” in Proc. Int. Conf. Learn. Represent. , 2024
2024
Later among the works it cites.
G. Zhou, H. Pan, Y. LeCun, and L. Pinto, “Dino-wm: World models on pre-trained visual features enable zero-shot planning,” in Proc. Int. Conf. Mach. Learn. , 2024
2024
Later among the works it cites.
2024
Later among the works it cites.
E. Alonso, A. Jelley, V. Micheli, A. Kanervisto, A. Storkey, T. Pearce, and F. Fleuret, “Diffusion for world modeling: Visual details matter in atari,” in Adv. Neural Inf. Process. Syst. , 2024
2024
Later among the works it cites.
OpenAI, “Chatgpt-4o,” 2024. [Online]. Available: https://openai.com/chatgpt
2024
Later among the works it cites.
M. Nakamoto, O. Mees, A. Kumar, and S. Levine, “Steering your generalists: Improving robotic foundation models via value guidance,” in Proc. Conf. Robot. Learn. , 2024
2024
Later among the works it cites.
Y. Huang, J. Zhang, S. Zou, X. Liu, R. Hu, and K. Xu, “Ladi-wm: A latent diffusion-based world model for predictive manipulation,” in Proc. Conf. Robot. Learn. , 2025
2025
Closest in time.
H. Qi, H. Yin, and H. Yang, “Control-oriented clustering of visual latent representation,” in Proc. Int. Conf. Learn. Represent. , 2025
2025
Closest in time.
2025
Closest in time.
B. Wang, N. Sridhar, C. Feng, M. Van der Merwe, A. Fishman, N. Fazeli, and J. J. Park, “This&that: Language-gesture controlled video generation for robot planning,” in Proc. IEEE Int. Conf. Robot. Autom. , 2025
2025
Closest in time.