The nature of explanation
Kenneth James Williams Craik · 1943
Earlier work this paper cites.
Cognitive maps in rats and men
Edward C Tolman · 1948
Earlier work this paper cites.
Dynamic models of segregation
Thomas C Schelling · 1971
Earlier work this paper cites.
A framework for representing knowledge, 1974
Marvin Minsky · 1974
Earlier work this paper cites.
Does the chimpanzee have a theory of mind?
David Premack and Guy Woodruff · 1978
Earlier work this paper cites.
Mental models: Towards a cognitive science of language, inference, and consciousness
Philip Nicholas Johnson-Laird · 1983
Earlier work this paper cites.
Integrated architectures for learning, planning, and reacting based on approximating dynamic programming
Richard S Sutton · 1990
Earlier work this paper cites.
Collective intelligence: Mankind’s emerging world in cyberspace
Pierre Lévy · 1997
Earlier work this paper cites.
Heterogeneous beliefs and routes to chaos in a simple asset pricing model
William A Brock and Cars H Hommes · 1998
Earlier work this paper cites.
Gradient-based learning applied to document recognition
Yann LeCun, Léon Bottou, Yoshua Bengio, and Patrick Haffner · 1998
Earlier work this paper cites.
Simultaneous localization and mapping: part i
Hugh Durrant-Whyte and Tim Bailey · 2006
Earlier work this paper cites.
Causal inference in statistics: An overview
Judea Pearl · 2009
Earlier work this paper cites.
Generative social science: Studies in agent-based computational modeling
Joshua M Epstein · 2012
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks
Alex Krizhevsky, Ilya Sutskever, and Geoffrey E Hinton · 2012
Earlier work this paper cites.
Rich feature hierarchies for accurate object detection and semantic segmentation
Ross Girshick, Jeff Donahue, Trevor Darrell, and Jitendra Malik · 2014
Earlier work this paper cites.
Handbook of collective intelligence
Thomas W Malone and Michael Bernstein · 2015
Earlier work this paper cites.
Human-level control through deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A Rusu, Joel Veness, Marc G Bellemare, Alex Graves, Martin Riedmiller, Andreas K Fidjeland, Georg Ostrovski, et al · 2015
Earlier work this paper cites.
Faster r-cnn: Towards real-time object detection with region proposal networks
Shaoqing Ren, Kaiming He, Ross Girshick, and Jian Sun · 2015
Earlier work this paper cites.
Unsupervised learning for physical interaction through video prediction
Chelsea Finn, Ian Goodfellow, and Sergey Levine · 2016
Earlier work this paper cites.
Model predictive control
Basil Kouvaritakis and Mark Cannon · 2016
Earlier work this paper cites.
Mastering the game of go with deep neural networks and tree search
David Silver, Aja Huang, Chris J Maddison, Arthur Guez, Laurent Sifre, George Van Den Driessche, Julian Schrittwieser, Ioannis Antonoglou, Veda Panneershelvam, Marc Lanctot, et al · 2016
Earlier work this paper cites.
Derivative-free optimization via classification
Yang Yu, Hong Qian, and Yi-Qi Hu · 2016
Earlier work this paper cites.
An lstm network for highway trajectory prediction
Florent Altché and Arnaud de La Fortelle · 2017
Earlier work this paper cites.
Matterport3d: Learning from rgb-d data in indoor environments
Angel Chang, Angela Dai, Thomas Funkhouser, Maciej Halber, Matthias Niessner, Manolis Savva, Shuran Song, Andy Zeng, and Yinda Zhang · 2017
Earlier work this paper cites.
Carla: An open urban driving simulator
Alexey Dosovitskiy, German Ros, Felipe Codevilla, Antonio Lopez, and Vladlen Koltun · 2017
Earlier work this paper cites.
Deep visual foresight for planning robot motion
Chelsea Finn and Sergey Levine · 2017
Earlier work this paper cites.
Sequential classification-based optimization for direct policy search
Yi-Qi Hu, Hong Qian, and Yang Yu · 2017
Earlier work this paper cites.
Uncertainty-aware reinforcement learning for collision avoidance
Original
Gregory Kahn, Adam Villaflor, Vitchyr Pong, Pieter Abbeel, and Sergey Levine · 2017
Earlier work this paper cites.
Ai2-thor: An interactive 3d environment for visual ai
Original
Eric Kolve, Roozbeh Mottaghi, Winson Han, Eli VanderBilt, Luca Weihs, Alvaro Herrasti, Matt Deitke, Kiana Ehsani, Daniel Gordon, Yuke Zhu, et al · 2017
Earlier work this paper cites.
Value prediction network
Junhyuk Oh, Satinder Singh, and Honglak Lee · 2017
Earlier work this paper cites.
Pointnet: Deep learning on point sets for 3d classification and segmentation
Charles R Qi, Hao Su, Kaichun Mo, and Leonidas J Guibas · 2017
Earlier work this paper cites.
Pointnet++: Deep hierarchical feature learning on point sets in a metric space, 2017
Charles R. Qi, Li Yi, Hao Su, and Leonidas J. Guibas · 2017
Earlier work this paper cites.
Proximal policy optimization algorithms
Original
John Schulman, Filip Wolski, Prafulla Dhariwal, Alec Radford, and Oleg Klimov · 2017
Earlier work this paper cites.
Mastering the game of go without human knowledge
David Silver, Julian Schrittwieser, Karen Simonyan, Ioannis Antonoglou, Aja Huang, Arthur Guez, Thomas Hubert, Lucas Baker, Matthew Lai, Adrian Bolton, et al · 2017
Earlier work this paper cites.
Attention is all you need
A Vaswani · 2017
Earlier work this paper cites.
Deep reinforcement learning in a handful of trials using probabilistic dynamics models
Kurtland Chua, Roberto Calandra, Rowan McAllister, and Sergey Levine · 2018
Earlier work this paper cites.
Recurrent world models facilitate policy evolution
David Ha and Jürgen Schmidhuber · 2018
Earlier work this paper cites.
World models
Original
David Ha and Jürgen Schmidhuber · 2018
Earlier work this paper cites.
Model-ensemble trust-region policy optimization
Original
Thanard Kurutach, Ignasi Clavera, Yan Duan, Aviv Tamar, and Pieter Abbeel · 2018
Earlier work this paper cites.
Microscopic traffic simulation using sumo
Pablo Alvarez Lopez, Michael Behrisch, Laura Bieker-Walz, Jakob Erdmann, Yun-Pang Flötteröd, Robert Hilbrich, Leonhard Lücken, Johannes Rummel, Peter Wagner, and Evamarie Wießner · 2018
Earlier work this paper cites.
Algorithmic framework for model-based deep reinforcement learning with theoretical guarantees
Original
Yuping Luo, Huazhe Xu, Yuanzhi Li, Yuandong Tian, Trevor Darrell, and Tengyu Ma · 2018
Earlier work this paper cites.
Superminds: The surprising power of people and computers thinking together
Thomas W Malone · 2018
Earlier work this paper cites.
A0c: Alpha zero in continuous action space
Original
Thomas M Moerland, Joost Broekens, Aske Plaat, and Catholijn M Jonker · 2018
Earlier work this paper cites.
Neural network dynamics for model-based deep reinforcement learning with model-free fine-tuning
Anusha Nagabandi, Gregory Kahn, Ronald S Fearing, and Sergey Levine · 2018
Earlier work this paper cites.
Virtualhome: Simulating household activities via programs
Xavier Puig, Kevin Ra, Marko Boben, Jiaman Li, Tingwu Wang, Sanja Fidler, and Antonio Torralba · 2018
Earlier work this paper cites.
Multinet: Real-time joint semantic reasoning for autonomous driving, 2018
Marvin Teichmann, Michael Weber, Marius Zoellner, Roberto Cipolla, and Raquel Urtasun · 2018
Earlier work this paper cites.
Solving rubik’s cube with a robot hand
Original
Ilge Akkaya, Marcin Andrychowicz, Maciek Chociej, Mateusz Litwin, Bob McGrew, Arthur Petron, Alex Paino, Matthias Plappert, Glenn Powell, Raphael Ribas, et al · 2019
Earlier work this paper cites.
Model-free deep reinforcement learning for urban autonomous driving, 2019
Jianyu Chen, Bodi Yuan, and Masayoshi Tomizuka · 2019
Earlier work this paper cites.
Multimodal trajectory predictions for autonomous driving using deep convolutional networks
Henggang Cui, Vladan Radosavljevic, Fang-Chieh Chou, Tsung-Han Lin, Thi Nguyen, Tzu-Kuo Huang, Jeff Schneider, and Nemanja Djuric · 2019
Earlier work this paper cites.
Multi-robot grasp planning for sequential assembly operations
Mehmet Dogar, Andrew Spielberg, Stuart Baker, and Daniela Rus · 2019
Earlier work this paper cites.
Dream to control: Learning behaviors by latent imagination
Original
Danijar Hafner, Timothy Lillicrap, Jimmy Ba, and Mohammad Norouzi · 2019
Earlier work this paper cites.
Learning latent dynamics for planning from pixels
Danijar Hafner, Timothy Lillicrap, Ian Fischer, Ruben Villegas, David Ha, Honglak Lee, and James Davidson · 2019
Earlier work this paper cites.
Mgnet: A unified framework of multigrid and convolutional neural network
Juncai He and Jinchao Xu · 2019
Earlier work this paper cites.
When to trust your model: Model-based policy optimization
Michael Janner, Justin Fu, Marvin Zhang, and Sergey Levine · 2019
Earlier work this paper cites.
Habitat: A platform for embodied ai research
Manolis Savva, Abhishek Kadian, Oleksandr Maksymets, Yili Zhao, Erik Wijmans, Bhavana Jain, Julian Straub, Jia Liu, Vladlen Koltun, Jitendra Malik, et al · 2019
Earlier work this paper cites.
Exploring model-based planning with policy networks
Original
Tingwu Wang and Jimmy Ba · 2019
Earlier work this paper cites.
The emergence of deepfake technology: A review
Mika Westerlund · 2019
Earlier work this paper cites.
Naturalistic driver intention and path prediction using recurrent neural networks
Alex Zyner, Stewart Worrall, and Eduardo Nebot · 2019
Earlier work this paper cites.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al · 2020
Earlier work this paper cites.
Predicting motion of vulnerable road users using high-definition maps and efficient convnets
Fang-Chieh Chou, Tsung-Han Lin, Henggang Cui, Vladan Radosavljevic, Thi Nguyen, Tzu-Kuo Huang, Matthew Niedoba, Jeff Schneider, and Nemanja Djuric · 2020
Earlier work this paper cites.
An image is worth 16x16 words: Transformers for image recognition at scale
Original
Alexey Dosovitskiy · 2020
Earlier work this paper cites.
Threedworld: A platform for interactive multi-modal physical simulation
Original
Chuang Gan, Jeremy Schwartz, Seth Alter, Damian Mrowca, Martin Schrimpf, James Traer, Julian De Freitas, Jonas Kubilius, Abhishek Bhandwaldar, Nick Haber, et al · 2020
Earlier work this paper cites.
Learning to walk in the real world with minimal human effort
Original
Sehoon Ha, Peng Xu, Zhenyu Tan, Sergey Levine, and Jie Tan · 2020
Earlier work this paper cites.
Mastering atari with discrete world models
Original
Danijar Hafner, Timothy Lillicrap, Mohammad Norouzi, and Jimmy Ba · 2020
Earlier work this paper cites.
Denoising diffusion probabilistic models
Jonathan Ho, Ajay Jain, and Pieter Abbeel · 2020
Earlier work this paper cites.
Multimodal trajectory predictions for urban environments using geometric relationships between a vehicle and lanes
Atsushi Kawasaki and Akihito Seki · 2020
Earlier work this paper cites.
A survey on learning-based robotic grasping
Kilian Kleeberger, Richard Bormann, Werner Kraus, and Marco F Huber · 2020
Earlier work this paper cites.
Covernet: Multimodal behavior prediction using trajectory sets
Tung Phan-Minh, Elena Corina Grigore, Freddy A Boulton, Oscar Beijbom, and Eric M Wolff · 2020
Earlier work this paper cites.
A game theoretic framework for model based reinforcement learning
Aravind Rajeswaran, Igor Mordatch, and Vikash Kumar · 2020
Earlier work this paper cites.
Learning to simulate complex physics with graph networks
Alvaro Sanchez-Gonzalez, Jonathan Godwin, Tobias Pfaff, Rex Ying, Jure Leskovec, and Peter Battaglia · 2020
Earlier work this paper cites.
Driving in dense traffic with model-free reinforcement learning
Dhruv Mauria Saxena, Sangjae Bae, Alireza Nakhaei, Kikuo Fujimura, and Maxim Likhachev · 2020
Earlier work this paper cites.
Alfworld: Aligning text and embodied environments for interactive learning
Original
Mohit Shridhar, Xingdi Yuan, Marc-Alexandre Côté, Yonatan Bisk, Adam Trischler, and Matthew Hausknecht · 2020
Earlier work this paper cites.
Score-based generative modeling through stochastic differential equations
Original
Yang Song, Jascha Sohl-Dickstein, Diederik P Kingma, Abhishek Kumar, Stefano Ermon, and Ben Poole · 2020
Earlier work this paper cites.
Design principles for the general data protection regulation (gdpr): A formal concept analysis and its evaluation
Damian A Tamburri · 2020
Earlier work this paper cites.
Sapien: A simulated part-based interactive environment
Fanbo Xiang, Yuzhe Qin, Kaichun Mo, Yikuan Xia, Hao Zhu, Fangchen Liu, Minghua Liu, Hanxiao Jiang, Yifu Yuan, He Wang, et al · 2020
Earlier work this paper cites.
Nebula: Quest for robotic autonomy in challenging environments; team costar at the darpa subterranean challenge
Original
Ali Agha, Kyohei Otsu, Benjamin Morrell, David D Fan, Rohan Thakker, Angel Santamaria-Navarro, Sung-Kyun Kim, Amanda Bouman, Xianmei Lei, Jeffrey Edlund, et al · 2021
Earlier work this paper cites.
Sim2real in robotics and automation: Applications and challenges
Sebastian Höfer, Kostas Bekris, Ankur Handa, Juan Camilo Gamboa, Melissa Mozifian, Florian Golemo, Chris Atkeson, Dieter Fox, Ken Goldberg, John Leonard, et al · 2021
Earlier work this paper cites.
Offline reinforcement learning as one big sequence modeling problem
Michael Janner, Qiyang Li, and Sergey Levine · 2021
Earlier work this paper cites.
Rma: Rapid motor adaptation for legged robots
Original
Ashish Kumar, Zipeng Fu, Deepak Pathak, and Jitendra Malik · 2021
Earlier work this paper cites.
Scene transformer: A unified multi-task model for behavior prediction and planning
Original
Jiquan Ngiam, Benjamin Caine, Vijay Vasudevan, Zhengdong Zhang, Hao-Tien Lewis Chiang, Jeffrey Ling, Rebecca Roelofs, Alex Bewley, Chenxi Liu, Ashish Venugopal, et al · 2021
Earlier work this paper cites.
Learning transferable visual models from natural language supervision, 2021
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, Gretchen Krueger, and Ilya Sutskever · 2021
Earlier work this paper cites.
igibson 1.0: A simulation environment for interactive tasks in large realistic scenes
Bokui Shen, Fei Xia, Chengshu Li, Roberto Martín-Martín, Linxi Fan, Guanzhi Wang, Claudia Pérez-D’Arpino, Shyamal Buch, Sanjana Srivastava, Lyne Tchapmi, et al · 2021
Earlier work this paper cites.
Videogpt: Video generation using vq-vae and transformers
Original
Wilson Yan, Yunzhi Zhang, Pieter Abbeel, and Aravind Srinivas · 2021
Earlier work this paper cites.
Transfusion: Robust lidar-camera fusion for 3d object detection with transformers
Xuyang Bai, Zeyu Hu, Xinge Zhu, Qingqiu Huang, Yilun Chen, Hongbo Fu, and Chiew-Lan Tai · 2022
Earlier work this paper cites.
Procthor: Large-scale embodied ai using procedural generation
Matt Deitke, Eli VanderBilt, Alvaro Herrasti, Luca Weihs, Kiana Ehsani, Jordi Salvador, Winson Han, Eric Kolve, Aniruddha Kembhavi, and Roozbeh Mottaghi · 2022
Earlier work this paper cites.
Glam: Efficient scaling of language models with mixture-of-experts
Nan Du, Yanping Huang, Andrew M Dai, Simon Tong, Dmitry Lepikhin, Yuanzhong Xu, Maxim Krikun, Yanqi Zhou, Adams Wei Yu, Orhan Firat, et al · 2022
Earlier work this paper cites.
Minedojo: Building open-ended embodied agents with internet-scale knowledge
Linxi Fan, Guanzhi Wang, Yunfan Jiang, Ajay Mandlekar, Yuncong Yang, Haoyi Zhu, Andrew Tang, De-An Huang, Yuke Zhu, and Anima Anandkumar · 2022
Earlier work this paper cites.
Inner monologue: Embodied reasoning through planning with language models
Original
Wenlong Huang, Fei Xia, Ted Xiao, Harris Chan, Jacky Liang, Pete Florence, Andy Zeng, Jonathan Tompson, Igor Mordatch, Yevgen Chebotar, et al · 2022
Earlier work this paper cites.
A survey on trajectory-prediction methods for autonomous driving
Yanjun Huang, Jiatong Du, Ziru Yang, Zewei Zhou, Lin Zhang, and Hong Chen · 2022
Earlier work this paper cites.
Multi-modal motion prediction with transformer-based neural network for autonomous driving
Zhiyu Huang, Xiaoyu Mo, and Chen Lv · 2022
Earlier work this paper cites.
Bc-z: Zero-shot task generalization with robotic imitation learning
Eric Jang, Alex Irpan, Mohi Khansari, Daniel Kappler, Frederik Ebert, Corey Lynch, Sergey Levine, and Chelsea Finn · 2022
Earlier work this paper cites.
A path towards autonomous machine intelligence version 0.9. 2, 2022-06-27
Yann LeCun · 2022
Earlier work this paper cites.
Metadrive: Composing diverse driving scenarios for generalizable reinforcement learning
Quanyi Li, Zhenghao Peng, Lan Feng, Qihang Zhang, Zhenghai Xue, and Bolei Zhou · 2022
Earlier work this paper cites.
Bevformer: Learning bird’s-eye-view representation from multi-camera images via spatiotemporal transformers
Original
Zhiqi Li, Wenhai Wang, Hongyang Li, Enze Xie, Chonghao Sima, Tong Lu, Yu Qiao, and Jifeng Dai · 2022
Earlier work this paper cites.
Wayformer: Motion forecasting via simple & efficient attention networks, 2022
Nigamaa Nayakanti, Rami Al-Rfou, Aurick Zhou, Kratarth Goel, Khaled S. Refaat, and Benjamin Sapp · 2022
Earlier work this paper cites.
Social simulacra: Creating populated prototypes for social computing systems
Joon Sung Park, Lindsay Popowski, Carrie Cai, Meredith Ringel Morris, Percy Liang, and Michael S Bernstein · 2022
Earlier work this paper cites.
Avlen: Audio-visual-language embodied navigation in 3d environments
Sudipta Paul, Amit Roy-Chowdhury, and Anoop Cherian · 2022
Earlier work this paper cites.
Pointnext: Revisiting pointnet++ with improved training and scaling strategies
Guocheng Qian, Yuchen Li, Houwen Peng, Jinjie Mai, Hasan Hammoud, Mohamed Elhoseiny, and Bernard Ghanem · 2022
Earlier work this paper cites.
Deepfake detection: A systematic literature review
Md Shohel Rana, Mohammad Nur Nobi, Beddhu Murali, and Andrew H Sung · 2022
Earlier work this paper cites.
Learning to walk in minutes using massively parallel deep reinforcement learning
Nikita Rudin, David Hoeller, Philipp Reist, and Marco Hutter · 2022
Earlier work this paper cites.
Neural theory-of-mind? on the limits of social intelligence in large lms
Original
Maarten Sap, Ronan LeBras, Daniel Fried, and Yejin Choi · 2022
Earlier work this paper cites.
Learning symbolic models for graph-structured physical mechanism
Hongzhi Shi, Jingtao Ding, Yufan Cao, Li Liu, Yong Li, et al · 2022
Earlier work this paper cites.
Motion transformer with global intention localization and local movement refinement
Shaoshuai Shi, Li Jiang, Dengxin Dai, and Bernt Schiele · 2022
Earlier work this paper cites.
A walk in the park: Learning to walk in 20 minutes with model-free reinforcement learning
Original
Laura Smith, Ilya Kostrikov, and Sergey Levine · 2022
Earlier work this paper cites.
Hybridnets: End-to-end perception network, 2022
Dat Vu, Bao Ngo, and Hung Phan · 2022
Earlier work this paper cites.
Yolop: You only look once for panoptic driving perception
Dong Wu, Man-Wen Liao, Wei-Tian Zhang, Xing-Gang Wang, Xiang Bai, Wen-Qing Cheng, and Wen-Yu Liu · 2022
Earlier work this paper cites.
Se-resunet: A novel robotic grasp detection method
Sheng Yu, Di-Hua Zhai, Yuanqing Xia, Haoran Wu, and Jun Liao · 2022
Earlier work this paper cites.
The ai economist: Taxation policy design via two-level deep multiagent reinforcement learning
Stephan Zheng, Alexander Trott, Sunil Srinivasa, David C Parkes, and Richard Socher · 2022
Earlier work this paper cites.
Matrixvt: Efficient multi-camera to bev transformation for 3d perception, 2022
Hongyu Zhou, Zheng Ge, Zeming Li, and Xiangyu Zhang · 2022
Earlier work this paper cites.
Gpt-4 technical report
Original
Josh Achiam, Steven Adler, Sandhini Agarwal, Lama Ahmad, Ilge Akkaya, Florencia Leoni Aleman, Diogo Almeida, Janko Altenschmidt, Sam Altman, Shyamal Anadkat, et al · 2023
Earlier work this paper cites.
Zero-shot robotic manipulation with pretrained image-editing diffusion models
Original
Kevin Black, Mitsuhiko Nakamoto, Pranav Atreya, Homer Walke, Chelsea Finn, Aviral Kumar, and Sergey Levine · 2023
Earlier work this paper cites.
Muvo: A multimodal generative world model for autonomous driving with geometric representations
Daniel Bogdoll, Yitian Yang, and J Marius Zöllner · 2023
Earlier work this paper cites.