Fetching the paper…
Reading the bibliography…
Recent advancements in LLMs have revolutionized motion generation models in embodied applications.
Problems of monetary management: the UK experience
Charles AE Goodhart and CAE Goodhart · 1984
Earlier work this paper cites.
Apprenticeship learning via inverse reinforcement learning
Pieter Abbeel and Andrew Y Ng · 2004
Earlier work this paper cites.
Optimal transport: old and new , volume 338
Cédric Villani et al · 2009
Earlier work this paper cites.
Generative adversarial imitation learning
Jonathan Ho and Stefano Ermon · 2016
Earlier work this paper cites.
On convergence and stability of gans
Naveen Kodali, Jacob Abernethy, James Hays, and Zsolt Kira · 2017
Earlier work this paper cites.
Multi-agent generative adversarial imitation learning
Jiaming Song, Hongyu Ren, Dorsa Sadigh, and Stefano Ermon · 2018
Earlier work this paper cites.
Probabilistic prediction of interactive driving behavior via hierarchical inverse reinforcement learning
Liting Sun, Wei Zhan, and Masayoshi Tomizuka · 2018
Earlier work this paper cites.
Wasserstein adversarial imitation learning
Huang Xiao, Michael Herman, Joerg Wagner, Sebastian Ziesche, Jalal Etesami, and Thai Hong Linh · 2019
Earlier work this paper cites.
Learning invariant representations for reinforcement learning without reconstruction
Amy Zhang, Rowan McAllister, Roberto Calandra, Yarin Gal, and Sergey Levine · 2020
Earlier work this paper cites.
Constitutional ai: Harmlessness from ai feedback
Yuntao Bai, Saurav Kadavath, Sandipan Kundu, Amanda Askell, Jackson Kernion, Andy Jones, Anna Chen, Anna Goldie, Azalia Mirhoseini, Cameron McKinnon, et al · 2022
Earlier work this paper cites.
Training language models to follow instructions with human feedback
Long Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida, Carroll Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, et al · 2022
Earlier work this paper cites.
Improving gans with a dynamic discriminator
Ceyuan Yang, Yujun Shen, Yinghao Xu, Deli Zhao, Bo Dai, and Bolei Zhou · 2022
Cited alongside, same era.
Josh Achiam, Steven Adler, Sandhini Agarwal, Lama Ahmad, Ilge Akkaya, Florencia Leoni Aleman, Diogo Almeida, Janko Altenschmidt, Sam Altman, Shyamal Anadkat, et al · 2023
Cited alongside, same era.
Rt-2: Vision-language-action models transfer web knowledge to robotic control
Anthony Brohan, Noah Brown, Justice Carbajal, Yevgen Chebotar, Xi Chen, Krzysztof Choromanski, Tianli Ding, Danny Driess, Avinava Dubey, Chelsea Finn, et al · 2023
Cited alongside, same era.
Contrastive prefence learning: Learning from human feedback without rl
Joey Hejna, Rafael Rafailov, Harshit Sikchi, Chelsea Finn, Scott Niekum, W Bradley Knox, and Dorsa Sadigh · 2023
Cited alongside, same era.
Behaviorgpt: Smart agent simulation for autonomous driving with next-patch prediction
Zikang Zhou, Haibo Hu, Xinhong Chen, Jianping Wang, Nan Guan, Kui Wu, Yung-Hui Li, Yu-Kai Huang, and Chun Jason Xue · 2023
Later among the works it cites.
Reinforcement learning with human feedback for realistic traffic simulation
Yulong Cao, Boris Ivanovic, Chaowei Xiao, and Marco Pavone · 2024
Later among the works it cites.
Jiaxiang Li, Siliang Zeng, Hoi-To Wai, Chenliang Li, Alfredo Garcia, and Mingyi Hong · 2024
Later among the works it cites.
Direct large language model alignment through self-rewarding contrastive prompt distillation
Aiwei Liu, Haoping Bai, Zhiyun Lu, Xiang Kong, Simon Wang, Jiulong Shan, Meng Cao, and Lijie Wen · 2024
Later among the works it cites.
The waymo open sim agents challenge
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Siddharth Karamcheti, Suraj Nair, Annie S Chen, Thomas Kollar, Chelsea Finn, Dorsa Sadigh, and Percy Liang · 2023
Cited alongside, same era.
Optimal transport for offline imitation learning
Yicheng Luo, zhengyao jiang, Samuel Cohen, Edward Grefenstette, and Marc Peter Deisenroth · 2023
Cited alongside, same era.
Wayformer: Motion forecasting via simple & efficient attention networks
Nigamaa Nayakanti, Rami Al-Rfou, Aurick Zhou, Kratarth Goel, Khaled S Refaat, and Benjamin Sapp · 2023
Cited alongside, same era.
Motionlm: Multi-agent motion forecasting as language modeling
Ari Seff, Brian Cera, Dian Chen, Mason Ng, Aurick Zhou, Nigamaa Nayakanti, Khaled S Refaat, Rami Al-Rfou, and Benjamin Sapp · 2023
Cited alongside, same era.
What matters to you? towards visual representation alignment for robot learning
Ran Tian, Chenfeng Xu, Masayoshi Tomizuka, Jitendra Malik, and Andrea Bajcsy · 2023
Cited alongside, same era.
Zephyr: Direct distillation of lm alignment
Lewis Tunstall, Edward Beeching, Nathan Lambert, Nazneen Rajani, Kashif Rasul, Younes Belkada, Shengyi Huang, Leandro von Werra, Clémentine Fourrier, Nathan Habib, et al · 2023
Cited alongside, same era.
Rlcd: Reinforcement learning from contrast distillation for language model alignment
Kevin Yang, Dan Klein, Asli Celikyilmaz, Nanyun Peng, and Yuandong Tian · 2023
Cited alongside, same era.
Escirl: Evolving self-contrastive irl for trajectory prediction in autonomous driving
Zhaorun Chen, Siyue Wang, Zhuokai Zhao, Chaoli Mao, Yiyang Zhou, Jiayu He, and Albert Sibo Hu
Cited in the paper.
Nico Montali, John Lambert, Paul Mougin, Alex Kuefler, Nicholas Rhinehart, Michelle Li, Cole Gulino, Tristan Emrich, Zoey Yang, Shimon Whiteson, et al · 2024
Later among the works it cites.
Trajeglish: Traffic modeling as next-token prediction
Jonah Philion, Xue Bin Peng, and Sanja Fidler · 2024
Later among the works it cites.
Humanoid locomotion as next token prediction
Ilija Radosavovic, Bike Zhang, Baifeng Shi, Jathushan Rajasegaran, Sarthak Kamat, Trevor Darrell, Koushil Sreenath, and Jitendra Malik · 2024
Later among the works it cites.
Inverse-rlignment: Inverse reinforcement learning from demonstrations for llm alignment
Hao Sun and Mihaela van der Schaar · 2024
Later among the works it cites.
Understanding the performance gap between online and offline alignment algorithms
Yunhao Tang, Daniel Zhaohan Guo, Zeyu Zheng, Daniele Calandriello, Yuan Cao, Eugene Tarassov, Rémi Munos, Bernardo Ávila Pires, Michal Valko, Yong Cheng, et al · 2024
Later among the works it cites.
Tokenize the world into object-level knowledge to address long-tail events in autonomous driving
Thomas Tian, Boyi Li, Xinshuo Weng, Yuxiao Chen, Edward Schmerling, Yue Wang, Boris Ivanovic, and Marco Pavone · 2024
Later among the works it cites.
Smart: Scalable multi-agent real-time simulation via next-token prediction
Wei Wu, Xiaoxin Feng, Ziyan Gao, and Yuheng Kan · 2024
Later among the works it cites.