Fetching the paper…
Reading the bibliography…
Generative modeling of complex behaviors from labeled datasets has been a longstanding problem in decision making.
Differential equations with a discontinuous forcing term
Bushaw, D. W · 1952
Earlier work this paper cites.
On the “bang-bang” control problem
Bellman, R., Glicksberg, I., and Gross, O · 1956
Earlier work this paper cites.
Between mdps and semi-mdps: A framework for temporal abstraction in reinforcement learning
Sutton, R. S., Precup, D., and Singh, S · 1999
Earlier work this paper cites.
Learning options in reinforcement learning
Stolle, M. and Precup, D · 2002
Earlier work this paper cites.
A review of vector quantization techniques
Vasuki, A. and Vanathi, P · 2006
Earlier work this paper cites.
A reduction of imitation learning and structured prediction to no-regret online learning
Ross, S., Gordon, G., and Bagnell, D · 2011
Earlier work this paper cites.
Learning objective functions for manipulation
Kalakrishnan, M., Pastor, P., Righetti, L., and Schaal, S · 2013
Earlier work this paper cites.
Deep reinforcement learning in parameterized action space
Hausknecht, M. and Stone, P · 2015
Earlier work this paper cites.
Maximum entropy deep inverse reinforcement learning
Wulfmeier, M., Ondruska, P., and Posner, I · 2015
Earlier work this paper cites.
Brockman, G., Cheung, V., Pettersson, L., Schneider, J., Schulman, J., Tang, J., and Zaremba, W · 2016
Earlier work this paper cites.
Guided cost learning: Deep inverse optimal control via policy optimization
Finn, C., Levine, S., and Abbeel, P · 2016
Earlier work this paper cites.
Generative adversarial imitation learning
Ho, J. and Ermon, S · 2016
Earlier work this paper cites.
Focal loss for dense object detection
Lin, T.-Y., Goyal, P., Girshick, R., He, K., and Dollár, P · 2017
Earlier work this paper cites.
Discrete sequential prediction of continuous actions for deep rl
Metz, L., Ibarz, J., Jaitly, N., and Davidson, J · 2017
Earlier work this paper cites.
Neural discrete representation learning
Van Den Oord, A., Vinyals, O., et al · 2017
Earlier work this paper cites.
Roboturk: A crowdsourcing platform for robotic skill learning through imitation
Mandlekar, A., Zhu, Y., Garg, A., Booher, J., Spero, M., Tung, A., Gao, J., Emmons, J., Gupta, A., Orbay, E., et al · 2018
Earlier work this paper cites.
Action branching architectures for deep reinforcement learning
Tavakoli, A., Pardo, F., and Kormushev, P · 2018
Earlier work this paper cites.
Relay policy learning: Solving long-horizon tasks via imitation and reinforcement learning
Gupta, A., Kumar, V., Lynch, C., Levine, S., and Hausman, K · 2019
Earlier work this paper cites.
End-to-end interpretable neural motion planner
Zeng, W., Luo, W., Suo, S., Sadat, A., Yang, B., Casas, S., and Urtasun, R · 2019
Earlier work this paper cites.
Language models are few-shot learners
Brown, T., Mann, B., Ryder, N., Subbiah, M., Kaplan, J. D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al · 2020
Cited alongside, same era.
nuscenes: A multimodal dataset for autonomous driving
Caesar, H., Bankiti, V., Lang, A. H., Vora, S., Liong, V. E., Xu, Q., Krishnan, A., Pan, Y., Baldan, G., and Beijbom, O · 2020
Cited alongside, same era.
Jukebox: A generative model for music
Dhariwal, P., Jun, H., Payne, C., Kim, J. W., Radford, A., and Sutskever, I · 2020
Cited alongside, same era.
Denoising diffusion probabilistic models
Ho, J., Jain, A., and Abbeel, P · 2020
Cited alongside, same era.
Learning latent plans from play
Lynch, C., Khansari, M., Xiao, T., Kumar, V., Tompson, J., Levine, S., and Sermanet, P · 2020
Cited alongside, same era.
Toward the fundamental limits of imitation learning
Differentiable raycasting for self-supervised occupancy forecasting
Khurana, T., Hu, P., Dave, A., Ziglar, J., Held, D., and Ramanan, D · 2022
Later among the works it cites.
Automating reinforcement learning with example-based resets
Kim, J., hyeon Park, J., Cho, D., and Kim, H. J · 2022
Later among the works it cites.
Choreographer: Learning and adapting skills in imagination
Mazzaglia, P., Verbelen, T., Dhoedt, B., Lacoste, A., and Rajeswar, S · 2022
Later among the works it cites.
High-resolution image synthesis with latent diffusion models
Rombach, R., Blattmann, A., Lorenz, D., Esser, P., and Ommer, B · 2022
Later among the works it cites.
Behavior transformers: Cloning k k modes with one stone
Shafiullah, N. M., Cui, Z., Altanzaya, A. A., and Pinto, L · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Rajaraman, N., Yang, L., Jiao, J., and Ramchandran, K · 2020
Cited alongside, same era.
Parrot: Data-driven behavioral priors for reinforcement learning
Singh, A., Liu, H., Zhou, G., Yu, A., Rhinehart, N., and Levine, S · 2020
Cited alongside, same era.
Continuous control with action quantization from demonstrations
Dadashi, R., Hussenot, L., Vincent, D., Girgin, S., Raichuk, A., Geist, M., and Pietquin, O · 2021
Cited alongside, same era.
Safe local motion planning with self-supervised freespace forecasting
Hu, P., Huang, A., Dolan, J., Held, D., and Ramanan, D · 2021
Cited alongside, same era.
Accelerating reinforcement learning with learned skill priors
Pertsch, K., Lee, Y., and Lim, J · 2021
Cited alongside, same era.
Perceive, attend, and drive: Learning spatial attention for safe self-driving
Wei, B., Ren, M., Zeng, W., Liang, M., Yang, B., and Urtasun, R · 2021
Cited alongside, same era.
Godiva: Generating open-domain videos from natural descriptions
Wu, C., Huang, L., Zhang, Q., Li, B., Ji, L., Yang, F., Sapiro, G., and Duan, N · 2021
Cited alongside, same era.
Achiam, J., Adler, S., Agarwal, S., Ahmad, L., Akkaya, I., Aleman, F. L., Almeida, D., Altenschmidt, J., Altman, S., Anadkat, S., et al · 2023
Later among the works it cites.
Playfusion: Skill acquisition via diffusion from language-annotated play
Chen, L., Bahl, S., and Pathak, D · 2023
Later among the works it cites.
Diffusion policy: Visuomotor policy learning via action diffusion
Chi, C., Feng, S., Du, Y., Xu, Z., Cousineau, E., Burchfiel, B., and Song, S · 2023
Later among the works it cites.
Planning-oriented autonomous driving
Hu, Y., Yang, J., Chen, L., Li, K., Sima, C., Zhu, X., Chai, S., Du, S., Lin, T., Wang, W., et al · 2023
Later among the works it cites.
Vad: Vectorized scene representation for efficient autonomous driving
Jiang, B., Chen, S., Xu, Q., Liao, B., Chen, J., Zhou, H., Zhang, Q., Liu, W., Huang, C., and Wang, X · 2023
Later among the works it cites.
Action-quantized offline reinforcement learning for robotic skill learning
Luo, J., Dong, P., Wu, J., Kumar, A., Geng, X., and Levine, S · 2023
Later among the works it cites.
Imitating human behaviour with diffusion models
Pearce, T., Rashid, T., Kanervisto, A., Bignell, D., Sun, M., Georgescu, R., Macua, S. V., Tan, S. Z., Momennejad, I., Hofmann, K., et al · 2023
Later among the works it cites.
Sdxl: Improving latent diffusion models for high-resolution image synthesis
Podell, D., English, Z., Lacey, K., Blattmann, A., Dockhorn, T., Müller, J., Penna, J., and Rombach, R · 2023
Later among the works it cites.
Goal-conditioned imitation learning using score-based diffusion policies
Reuss, M., Li, M., Jia, X., and Lioutikov, R · 2023
Later among the works it cites.
Shafiullah, N. M. M., Rai, A., Etukuru, H., Liu, Y., Misra, I., Chintala, S., and Pinto, L · 2023
Later among the works it cites.
Learning fine-grained bimanual manipulation with low-cost hardware
Zhao, T. Z., Kumar, V., Levine, S., and Finn, C · 2023
Later among the works it cites.
Lumiere: A space-time diffusion model for video generation
Bar-Tal, O., Chefer, H., Tov, O., Herrmann, C., Paiss, R., Zada, S., Ephrat, A., Hur, J., Li, Y., Michaeli, T., et al · 2024
Closest in time.
Masked audio generation using a single non-autoregressive transformer
Ziv, A., Gat, I., Lan, G. L., Remez, T., Kreuk, F., Défossez, A., Copet, J., Synnaeve, G., and Adi, Y · 2024
Closest in time.