Fetching the paper…
Reading the bibliography…
Learning rewards from expert videos offers an affordable and effective solution to specify the intended behaviors for reinforcement learning (RL) tasks.
Ziebart, B.D., Maas, A.L., Bagnell, J.A., Dey, A.K., et al.: Maximum entropy inverse reinforcement learning. In: Association for the Advancement of Artificial Intelligence (AAAI) (2008)
2008
Earlier work this paper cites.
Kingma, D.P., Welling, M.: Auto-encoding variational bayes. arXiv preprint arXiv:1312.6114 (2013)
2013
Earlier work this paper cites.
Sohl-Dickstein, J., Weiss, E., Maheswaranathan, N., Ganguli, S.: Deep unsupervised learning using nonequilibrium thermodynamics. In: International Conference on Machine Learning (ICML) (2015)
2015
Earlier work this paper cites.
2016
Earlier work this paper cites.
Sermanet, P., Lynch, C., Chebotar, Y., Hsu, J., Jang, E., Schaal, S., Levine, S.: Time-contrastive networks: Self-supervised learning from video. In: IEEE International Conference on Robotics and Automation (ICRA) (2017)
2017
Earlier work this paper cites.
Zhu, J.Y., Zhang, R., Pathak, D., Darrell, T., Efros, A.A., Wang, O., Shechtman, E.: Toward multimodal image-to-image translation. Advances in Neural Information Processing Systems (NeurIPS) (2017)
2017
Earlier work this paper cites.
Rajeswaran, A., Kumar, V., Gupta, A., Schulman, J., Todorov, E., Levine, S.: Learning complex dexterous manipulation with deep reinforcement learning and demonstrations. In: Robotics: Science and Systems (RSS) (2018)
2018
Earlier work this paper cites.
Torabi, F., Warnell, G., Stone, P.: Behavioral cloning from observation. In: International Joint Conference on Artificial Intelligence (IJCAI) (2018)
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
Zhang, R., Isola, P., Efros, A.A., Shechtman, E., Wang, O.: The unreasonable effectiveness of deep features as a perceptual metric. In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) (2018)
2018
Earlier work this paper cites.
Burda, Y., Edwards, H., Storkey, A.J., Klimov, O.: Exploration by random network distillation. In: International Conference on Learning Representations (ICLR) (2019)
2019
Earlier work this paper cites.
Nasiriany, S., Pong, V., Lin, S., Levine, S.: Planning with goal-conditioned policies. In: Advances in Neural Information Processing Systems (NeurIPS) (2019)
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
Ho, J., Jain, A., Abbeel, P.: Denoising diffusion probabilistic models. In: Advances in Neural Information Processing Systems (NeurIPS) (2020)
2020
Earlier work this paper cites.
Lynch, C., Sermanet, P.: Language conditioned imitation learning over unstructured data. In: Robotics: Science and Systems (RSS) (2020)
2020
Earlier work this paper cites.
2020
Earlier work this paper cites.
Singh, A., Liu, H., Zhou, G., Yu, A., Rhinehart, N., Levine, S.: Parrot: Data-driven behavioral priors for reinforcement learning. In: International Conference on Learning Representations (ICLR) (2020)
2020
Earlier work this paper cites.
2020
Earlier work this paper cites.
Yu, T., Quillen, D., He, Z., Julian, R.C., Hausman, K., Finn, C., Levine, S.: Meta-world: A benchmark and evaluation for multi-task and meta reinforcement learning. In: Conference on Robot Learning (CoRL) (2020)
2020
Earlier work this paper cites.
Zhu, Z., Lin, K., Dai, B., Zhou, J.: Off-policy imitation learning from observations. Advances in Neural Information Processing Systems (NeurIPS) (2020)
2020
Earlier work this paper cites.
Chen, A.S., Nair, S., Finn, C.: Learning generalizable robotic reward functions from "in-the-wild" human videos. In: Robotics: Science and Systems (RSS) (2021)
2021
Cited alongside, same era.
Esser, P., Rombach, R., Ommer, B.: Taming transformers for high-resolution image synthesis. In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) (2021)
2021
Cited alongside, same era.
Peng, X.B., Ma, Z., Abbeel, P., Levine, S., Kanazawa, A.: Amp: Adversarial motion priors for stylized physics-based character control. ACM Trans. Graph (ToG). (2021)
2021
Cited alongside, same era.
Pertsch, K., Lee, Y., Lim, J.: Accelerating reinforcement learning with learned skill priors. In: Conference on Robot Learning (CoRL) (2021)
2021
Cited alongside, same era.
Radosavovic, I., Wang, X., Pinto, L., Malik, J.: State-only imitation learning for dexterous manipulation. In: IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) (2021)
2023
Closest in time.
Ajay, A., Du, Y., Gupta, A., Tenenbaum, J., Jaakkola, T., Agrawal, P.: Is conditional generative modeling all you need for decision-making? In: International Conference on Learning Representations (ICLR) (2023)
2023
Closest in time.
Ceylan, D., Huang, C.H.P., Mitra, N.J.: Pix2video: Video editing using image diffusion. In: Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV) (2023)
2023
Closest in time.
2023
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2021
Cited alongside, same era.
Rao, D., Sadeghi, F., Hasenclever, L., Wulfmeier, M., Zambelli, M., Vezzani, G., Tirumala, D., Aytar, Y., Merel, J., Heess, N., et al.: Learning transferable motor skills with hierarchical latent mixture policies. In: International Conference on Learning Representations (ICLR) (2021)
2021
Cited alongside, same era.
2021
Cited alongside, same era.
Alakuijala, M., Dulac-Arnold, G., Mairal, J., Ponce, J., Schmid, C.: Learning reward functions for robotic manipulation by observing humans. In: IEEE International Conference on Robotics and Automation (ICRA) (2022)
2022
Cited alongside, same era.
Gu, S., Chen, D., Bao, J., Wen, F., Zhang, B., Chen, D., Yuan, L., Guo, B.: Vector quantized diffusion model for text-to-image synthesis. In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) (2022)
2022
Cited alongside, same era.
2022
Cited alongside, same era.
2022
Cited alongside, same era.
Janner, M., Du, Y., Tenenbaum, J.B., Levine, S.: Planning with diffusion for flexible behavior synthesis. In: International Conference on Machine Learning (ICML) (2022)
2022
Cited alongside, same era.
Chi, C., Feng, S., Du, Y., Xu, Z., Cousineau, E., Burchfiel, B., Song, S.: Diffusion policy: Visuomotor policy learning via action diffusion. In: Robotics: Science and Systems (RSS) (2023)
2023
Closest in time.
Du, Y., Yang, M., Dai, B., Dai, H., Nachum, O., Tenenbaum, J.B., Schuurmans, D., Abbeel, P.: Learning universal policies via text-guided video generation. In: Advances in Neural Information Processing Systems (NeurIPS) (2023)
2023
Closest in time.
Escontrela, A., Adeniji, A., Yan, W., Jain, A., Peng, X.B., Goldberg, K., Lee, Y., Hafner, D., Abbeel, P.: Video prediction models as rewards for reinforcement learning. In: Advances in Neural Information Processing Systems (NeurIPS) (2023)
2023
Closest in time.
Esser, P., Chiu, J., Atighehchian, P., Granskog, J., Germanidis, A.: Structure and content-guided video synthesis with diffusion models. In: Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV) (2023)
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
Kun, L., He, Z., Lu, C., Hu, K., Gao, Y., Xu, H.: Uni-o4: Unifying online and offline deep reinforcement learning with multi-step on-policy optimization. In: International Conference on Learning Representations (ICLR) (2023)
2023
Closest in time.
Ma, Y.J., Sodhani, S., Jayaraman, D., Bastani, O., Kumar, V., Zhang, A.: Vip: Towards universal visual reward and representation via value-implicit pre-training. In: International Conference on Learning Representations (ICLR) (2023)
2023
Closest in time.
2023
Closest in time.
Nair, S., Rajeswaran, A., Kumar, V., Finn, C., Gupta, A.: R3m: A universal visual representation for robot manipulation. In: Conference on Robot Learning (CoRL) (2023)
2023
Closest in time.
Nuti, F., Franzmeyer, T., Henriques, J.F.: Extracting reward functions from diffusion models. In: Advances in Neural Information Processing Systems (NeurIPS) (2023)
2023
Closest in time.
Wang, Z., Hunt, J.J., Zhou, M.: Diffusion policies as an expressive policy class for offline reinforcement learning. In: International Conference on Learning Representations (ICLR) (2023)
2023
Closest in time.
2023
Closest in time.