Fetching the paper…
Reading the bibliography…
To act in the world, robots rely on a representation of salient task aspects: for example, to carry a coffee mug, a robot may consider movement efficiency or mug orientation in its behavior.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al · 1901
Earlier work this paper cites.
What Size Net Gives Valid Generalization?. In Advances in Neural Information Processing Systems , D. Touretzky (Ed.), Vol. 1. Morgan-Kaufmann
Eric Baum and David Haussler. 1988 · 1988
Earlier work this paper cites.
Hierarchical learning of robot skills by reinforcement. In Proceedings of International Conference on Neural Networks (ICNN’88), San Francisco, CA, USA, March 28 - April 1, 1993 . IEEE, 181–186
Long-Ji Lin. 1993 · 1993
Earlier work this paper cites.
Polynomial Bounds for VC Dimension of Sigmoidal and General Pfaffian Neural Networks
Marek Karpinski and Angus Macintyre. 1997 · 1997
Earlier work this paper cites.
Almost Linear VC Dimension Bounds for Piecewise Polynomial Networks. In Advances in Neural Information Processing Systems , M. Kearns, S. Solla, and D. Cohn (Eds.), Vol. 11. MIT Press
Peter Bartlett, Vitaly Maiorov, and Ron Meir. 1998 · 1998
Earlier work this paper cites.
User adaptation of human-robot interaction model based on Bayesian network and introspection of interaction experience. In IEEE/RSJ International Conference on Intelligent Robots and Systems, IROS 2000, October 30 - Novemver 5, 2000, Takamatsu, Japan . IEEE, 2139–2144
Tetsunari Inamura, Masayuki Inaba, and Hirochika Inoue. 2000 · 2000
Earlier work this paper cites.
Algorithms for inverse reinforcement learning
Andrew Ng and Stuart Russell. 2000 · 2000
Earlier work this paper cites.
Few-Shot Learning via Learning the Representation, Provably
Simon S. Du, Wei Hu, Sham M. Kakade, Jason D. Lee, and Qi Lei. 2020 · 2002
Earlier work this paper cites.
Q-Cut - Dynamic Discovery of Sub-goals in Reinforcement Learning. In Machine Learning: ECML 2002, 13th European Conference on Machine Learning, Helsinki, Finland, August 19-23, 2002, Proceedings (Lecture Notes in Computer Science, Vol. 2430) , Tapio Elomaa, Heikki Mannila, and Hannu Toivonen (Eds.). Springer, 295–306
Ishai Menache, Shie Mannor, and Nahum Shimkin. 2002 · 2002
Earlier work this paper cites.
Apprenticeship learning via inverse reinforcement learning. In Machine Learning (ICML), International Conference on . ACM
Pieter Abbeel and Andrew Y Ng. 2004 · 2004
Earlier work this paper cites.
Tractable learning of large Bayes net structures from sparse data. In Machine Learning, Proceedings of the Twenty-first International Conference (ICML 2004), Banff, Alberta, Canada, July 4-8, 2004 (ACM International Conference Proceeding Series, Vol. 69) , Carla E. Brodley (Ed.). ACM
Anna Goldenberg and Andrew W. Moore. 2004 · 2004
Earlier work this paper cites.
Learning Forward Models for Robots. In IJCAI-05, Proceedings of the Nineteenth International Joint Conference on Artificial Intelligence, Edinburgh, Scotland, UK, July 30 - August 5, 2005 , Leslie Pack Kaelbling and Alessandro Saffiotti (Eds.). Professional Book Center, 1440–1445
Anthony M. Dearden and Yiannis Demiris. 2005 · 2005
Earlier work this paper cites.
Using a Planner for Coordination of Multiagent Team Behavior. In Programming Multi-Agent Systems, Third International Workshop, ProMAS 2005, Utrecht, The Netherlands, July 26, 2005, Revised and Invited Papers (Lecture Notes in Computer Science, Vol. 3862) , Rafael H. Bordini, Mehdi Dastani, Jürgen Dix, and Amal El Fallah Seghrouchni (Eds.). Springer, 90–100
Oliver Obst. 2005 · 2005
Earlier work this paper cites.
Simultaneous learning of structure and value in relational reinforcement learning. In Workshop on Rich Representations for Reinforcement Learning . Citeseer, 57
Scott Sanner. 2005 · 2005
Earlier work this paper cites.
On the use of Bayesian Networks to develop behaviours for mobile robots
Elena Lazkano, Basilio Sierra, Aitzol Astigarraga, and José María Martínez-Otzeta. 2007 · 2006
Earlier work this paper cites.
Learning hierarchical task networks by observation. In Machine Learning, Proceedings of the Twenty-Third International Conference (ICML 2006), Pittsburgh, Pennsylvania, USA, June 25-29, 2006 (ACM International Conference Proceeding Series, Vol. 148) , William W. Cohen and Andrew W. Moore (Eds.). ACM, 665–672
Negin Nejati, Pat Langley, and Tolga Könik. 2006 · 2006
Earlier work this paper cites.
Affordance-based imitation learning in robots. In 2007 IEEE/RSJ International Conference on Intelligent Robots and Systems, October 29 - November 2, 2007, Sheraton Hotel and Marina, San Diego, California, USA . IEEE, 1015–1021
Manuel Lopes, Francisco S. Melo, and Luis Montesano. 2007 · 2007
Earlier work this paper cites.
Modeling affordances using Bayesian networks. In 2007 IEEE/RSJ International Conference on Intelligent Robots and Systems, October 29 - November 2, 2007, Sheraton Hotel and Marina, San Diego, California, USA . IEEE, 4102–4107
Luis Montesano, Manuel Lopes, Alexandre Bernardino, and José Santos-Victor. 2007 · 2007
Earlier work this paper cites.
Boosting structured prediction for imitation learning. In Advances in Neural Information Processing Systems . 1153–1160
Nathan Ratliff, David M Bradley, Joel Chestnutt, and J A Bagnell. 2007 · 2007
Earlier work this paper cites.
Feature Dynamic Bayesian Networks
Marcus Hutter. 2008 · 2008
Earlier work this paper cites.
Automatic discovery and transfer of MAXQ hierarchies. In Machine Learning, Proceedings of the Twenty-Fifth International Conference (ICML 2008), Helsinki, Finland, June 5-9, 2008 (ACM International Conference Proceeding Series, Vol. 307) , William W. Cohen, Andrew McCallum, and Sam T. Roweis (Eds.). ACM, 648–655
Neville Mehta, Soumya Ray, Prasad Tadepalli, and Thomas G. Dietterich. 2008 · 2008
Earlier work this paper cites.
Causal Inference. In Causality: Objectives and Assessment (NIPS 2008 Workshop), Whistler, Canada, December 12, 2008 (JMLR Proceedings, Vol. 6) , Isabelle Guyon, Dominik Janzing, and Bernhard Schölkopf (Eds.). JMLR.org, 39–58
Judea Pearl. 2010 · 2008
Earlier work this paper cites.
Maximum Entropy Inverse Reinforcement Learning. In Proceedings of the 23rd National Conference on Artificial Intelligence - Volume 3 (Chicago, Illinois) (AAAI’08) . AAAI Press, 1433–1438
Brian D. Ziebart, Andrew Maas, J. Andrew Bagnell, and Anind K. Dey. 2008 · 2008
Earlier work this paper cites.
Smoothed Sarsa: Reinforcement learning for robot delivery tasks. In 2009 IEEE International Conference on Robotics and Automation, ICRA 2009, Kobe, Japan, May 12-17, 2009 . IEEE, 2125–2132
Deepak Ramachandran and Rakesh Gupta. 2009 · 2009
Earlier work this paper cites.
Feature construction for inverse reinforcement learning. In Advances in Neural Information Processing Systems . 1342–1350
Sergey Levine, Zoran Popovic, and Vladlen Koltun. 2010 · 2010
Earlier work this paper cites.
Artificial intelligence a modern approach
Stuart J Russell. 2010 · 2010
Earlier work this paper cites.
Learning task constraints for robot grasping using graphical models. In 2010 IEEE/RSJ International Conference on Intelligent Robots and Systems, October 18-22, 2010, Taipei, Taiwan . IEEE, 1579–1585
Dan Song, Kai Huebner, Ville Kyrki, and Danica Kragic. 2010 · 2010
Earlier work this paper cites.
Tsz-Chiu Au, Okhtay Ilghami, Ugur Kuter, J. William Murdock, Dana S. Nau, Dan Wu, and Fusun Yaman. 2011 · 2011
Earlier work this paper cites.
Learning the behavior model of a robot
Guillaume Infantes, Malik Ghallab, and Félix Ingrand. 2011 · 2011
Earlier work this paper cites.
Structure Learning of Causal Bayesian Networks: A Survey
Ashique Rupam Mahmood. 2011 · 2011
Earlier work this paper cites.
A reduction of imitation learning and structured prediction to no-regret online learning. In Proceedings of the fourteenth international conference on artificial intelligence and statistics . JMLR Workshop and Conference Proceedings, 627–635
Stéphane Ross, Geoffrey Gordon, and Drew Bagnell. 2011 · 2011
Earlier work this paper cites.
Multivariate discretization for Bayesian Network structure learning in robot grasping. In IEEE International Conference on Robotics and Automation, ICRA 2011, Shanghai, China, 9-13 May 2011 . IEEE, 1944–1950
Dan Song, Carl Henrik Ek, Kai Huebner, and Danica Kragic. 2011 · 2011
Earlier work this paper cites.
Designing robot learners that ask good questions. In International Conference on Human-Robot Interaction, HRI’12, Boston, MA, USA - March 05 - 08, 2012 , Holly A. Yanco, Aaron Steinfeld, Vanessa Evers, and Odest Chadwicke Jenkins (Eds.). ACM, 17–24
Maya Cakmak and Andrea Lockerd Thomaz. 2012 · 2012
Earlier work this paper cites.
Learning Feature Representations with K-Means. In Neural Networks: Tricks of the Trade
Adam Coates and A. Ng. 2012 · 2012
Earlier work this paper cites.
Efficient high dimensional maximum entropy modeling via symmetric partition functions. In Advances in Neural Information Processing Systems . 575–583
Paul Vernaza and Drew Bagnell. 2012 · 2012
Earlier work this paper cites.
Bayesian nonparametric feature construction for inverse reinforcement learning. In Twenty-Third International Joint Conference on Artificial Intelligence
Jaedeug Choi and Kee-Eung Kim. 2013 · 2013
Earlier work this paper cites.
Learning Probabilistic Hierarchical Task Networks as Probabilistic Context-Free Grammars to Capture User Preferences
Nan Li, William Cushing, Subbarao Kambhampati, and Sung Wook Yoon. 2014 · 2014
Earlier work this paper cites.
RoboBrain: Large-Scale Knowledge Engine for Robots
Ashutosh Saxena, Ashesh Jain, Ozan Sener, Aditya Jami, Dipendra Kumar Misra, and Hema Swetha Koppula. 2014 · 2014
Earlier work this paper cites.
A Bayesian Developmental Approach to Robotic Goal-Based Imitation Learning
Michael Jae-Yoon Chung, Abram Friesen, Dieter Fox, Andrew Meltzoff, and Rajesh Rao. 2015 · 2015
Earlier work this paper cites.
Learning Visual Feature Spaces for Robotic Manipulation with Deep Spatial Autoencoders
Chelsea Finn, Xin Yu Tan, Yan Duan, Trevor Darrell, Sergey Levine, and Pieter Abbeel. 2015 · 2015
Earlier work this paper cites.
A comprehensive survey on safe reinforcement learning
Javier Garcıa and Fernando Fernández. 2015 · 2015
Earlier work this paper cites.
Learning preferences for manipulation tasks from online coactive feedback
Ashesh Jain, Shikhar Sharma, Thorsten Joachims, and Ashutosh Saxena. 2015 · 2015
Earlier work this paper cites.
Interactive Hierarchical Task Learning from a Single Demonstration. In Proceedings of the Tenth Annual ACM/IEEE International Conference on Human-Robot Interaction, HRI 2015, Portland, OR, USA, March 2-5, 2015 , Julie A. Adams, William D. Smart, Bilge Mutlu, and Leila Takayama (Eds.). ACM, 205–212
Anahita Mohseni-Kabir, Charles Rich, Sonia Chernova, Candace L. Sidner, and Daniel Miller. 2015 · 2015
Earlier work this paper cites.
A review of relational machine learning for knowledge graphs
Maximilian Nickel, Kevin Murphy, Volker Tresp, and Evgeniy Gabrilovich. 2015 · 2015
Earlier work this paper cites.
Embed to Control: A Locally Linear Latent Dynamics Model for Control from Raw Images. In Advances in Neural Information Processing Systems , C. Cortes, N. Lawrence, D. Lee, M. Sugiyama, and R. Garnett (Eds.), Vol. 28. Curran Associates, Inc
Manuel Watter, Jost Springenberg, Joschka Boedecker, and Martin Riedmiller. 2015 · 2015
Earlier work this paper cites.
Concrete problems in AI safety
Dario Amodei, Chris Olah, Jacob Steinhardt, Paul Christiano, John Schulman, and Dan Mané. 2016 · 2016
Earlier work this paper cites.
Guided Cost Learning: Deep Inverse Optimal Control via Policy Optimization. In Proceedings of the 33rd International Conference on International Conference on Machine Learning - Volume 48 (New York, NY, USA) (ICML’16) . JMLR.org, 49–58
Chelsea Finn, Sergey Levine, and Pieter Abbeel. 2016 · 2016
Earlier work this paper cites.
Autonomously constructing hierarchical task networks for planning and human-robot collaboration. In 2016 IEEE International Conference on Robotics and Automation, ICRA 2016, Stockholm, Sweden, May 16-21, 2016 , Danica Kragic, Antonio Bicchi, and Alessandro De Luca (Eds.). IEEE, 5469–5476
Bradley Hayes and Brian Scassellati. 2016 · 2016
Earlier work this paper cites.
The Variational Fair Autoencoder
Christos Louizos, Kevin Swersky, Yujia Li, Max Welling, and Richard S. Zemel. 2016 · 2016
Earlier work this paper cites.
Understanding intermediate layers using linear classifier probes. In 5th International Conference on Learning Representations, ICLR 2017, Toulon, France, April 24-26, 2017, Workshop Track Proceedings . OpenReview.net
Guillaume Alain and Yoshua Bengio. 2017 · 2017
Earlier work this paper cites.
Learning Robot Objectives from Physical Human Interaction. In Proceedings of the 1st Annual Conference on Robot Learning (Proceedings of Machine Learning Research, Vol. 78) , Sergey Levine, Vincent Vanhoucke, and Ken Goldberg (Eds.). PMLR, 217–226
Andrea Bajcsy, Dylan P. Losey, Marcia K. O’Malley, and Anca D. Dragan. 2017 · 2017
Earlier work this paper cites.
Spectrally-normalized margin bounds for neural networks. In NIPS
Peter L. Bartlett, Dylan J. Foster, and Matus Telgarsky. 2017 · 2017
Earlier work this paper cites.
Deep Reinforcement Learning from Human Preferences. In Advances in Neural Information Processing Systems , I. Guyon, U. V. Luxburg, S. Bengio, H. Wallach, R. Fergus, S. Vishwanathan, and R. Garnett (Eds.), Vol. 30. Curran Associates, Inc
Paul F Christiano, Jan Leike, Tom Brown, Miljan Martic, Shane Legg, and Dario Amodei. 2017 · 2017
Earlier work this paper cites.
Model-Agnostic Meta-Learning for Fast Adaptation of Deep Networks. In Proceedings of the 34th International Conference on Machine Learning - Volume 70 (Sydney, NSW, Australia) (ICML’17) . JMLR.org, 1126–1135
Chelsea Finn, Pieter Abbeel, and Sergey Levine. 2017 · 2017
Earlier work this paper cites.
Inverse reward design
Dylan Hadfield-Menell, Smitha Milli, Pieter Abbeel, Stuart J Russell, and Anca Dragan. 2017a · 2017
Earlier work this paper cites.
Nearly-tight VC-dimension bounds for piecewise linear neural networks. In Proceedings of the 2017 Conference on Learning Theory (Proceedings of Machine Learning Research, Vol. 65) , Satyen Kale and Ohad Shamir (Eds.). PMLR, 1064–1068
Nick Harvey, Christopher Liaw, and Abbas Mehrabian. 2017 · 2017
Cited alongside, same era.
Improving robot controller transparency through autonomous policy explanation. In 2017 12th ACM/IEEE International Conference on Human-Robot Interaction (HRI . IEEE, 303–312
Bradley Hayes and Julie A Shah. 2017 · 2017
Cited alongside, same era.
DARLA: Improving Zero-Shot Transfer in Reinforcement Learning. In ICML
Irina Higgins, Arka Pal, Andrei A. Rusu, Loïc Matthey, Christopher P. Burgess, Alexander Pritzel, Matthew M. Botvinick, Charles Blundell, and Alexander Lerchner. 2017 · 2017
Cited alongside, same era.
Robot life-long task learning from human demonstrations: a Bayesian approach
Nathan P. Koenig and Maja J. Mataric. 2017 · 2017
Cited alongside, same era.
Learning Bayesian Networks with the Saiyan Algorithm
Anthony C. Constantinou. 2020 · 2020
Later among the works it cites.
Dream to Control: Learning Behaviors by Latent Imagination. In 8th International Conference on Learning Representations, ICLR 2020, Addis Ababa, Ethiopia, April 26-30, 2020 . OpenReview.net
Danijar Hafner, Timothy P. Lillicrap, Jimmy Ba, and Mohammad Norouzi. 2020 · 2020
Later among the works it cites.
CURL: Contrastive Unsupervised Representations for Reinforcement Learning. In Proceedings of the 37th International Conference on Machine Learning (Proceedings of Machine Learning Research, Vol. 119) , Hal Daumé III and Aarti Singh (Eds.). PMLR, 5639–5650
Michael Laskin, Aravind Srinivas, and Pieter Abbeel. 2020 · 2020
Later among the works it cites.
Offline reinforcement learning: Tutorial, review, and perspectives on open problems
Sergey Levine, Aviral Kumar, George Tucker, and Justin Fu. 2020 · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Simple and Scalable Predictive Uncertainty Estimation Using Deep Ensembles. In Proceedings of the 31st International Conference on Neural Information Processing Systems (Long Beach, California, USA) (NIPS’17) . Curran Associates Inc., Red Hook, NY, USA, 6405–6416
Balaji Lakshminarayanan, Alexander Pritzel, and Charles Blundell. 2017 · 2017
Cited alongside, same era.
Knowledge graph embedding: A survey of approaches and applications
Quan Wang, Zhendong Mao, Bin Wang, and Li Guo. 2017 · 2017
Cited alongside, same era.
State abstractions for lifelong reinforcement learning. In International Conference on Machine Learning . PMLR, 10–19
David Abel, Dilip Arumugam, Lucas Lehnert, and Michael Littman. 2018 · 2018
Cited alongside, same era.
Playing Hard Exploration Games by Watching YouTube. In Proceedings of the 32nd International Conference on Neural Information Processing Systems (Montréal, Canada) (NIPS’18) . Curran Associates Inc., Red Hook, NY, USA, 2935–2945
Yusuf Aytar, Tobias Pfaff, David Budden, Tom Le Paine, Ziyu Wang, and Nando de Freitas. 2018 · 2018
Cited alongside, same era.
Learning under Misspecified Objective Spaces. In Proceedings of The 2nd Conference on Robot Learning (Proceedings of Machine Learning Research, Vol. 87) , Aude Billard, Anca Dragan, Jan Peters, and Jun Morimoto (Eds.). PMLR, 796–805
Andreea Bobu, Andrea Bajcsy, Jaime F. Fisac, and Anca D. Dragan. 2018 · 2018
Cited alongside, same era.
Human-Driven Feature Selection for a Robotic Agent Learning Classification Tasks from Demonstration. In 2018 IEEE International Conference on Robotics and Automation, ICRA 2018, Brisbane, Australia, May 21-25, 2018 . IEEE, 6923–6930
Kalesha Bullard, Sonia Chernova, and Andrea Lockerd Thomaz. 2018 · 2018
Cited alongside, same era.
Diversity is all you need: Learning skills without a reward function
Benjamin Eysenbach, Abhishek Gupta, Julian Ibarz, and Sergey Levine. 2018 · 2018
Cited alongside, same era.
Multi-task maximum entropy inverse reinforcement learning
Adam Gleave and Oliver Habryka. 2018 · 2018
Cited alongside, same era.
Causal Discovery in Physical Systems from Videos. In Advances in Neural Information Processing Systems 33: Annual Conference on Neural Information Processing Systems 2020, NeurIPS 2020, December 6-12, 2020, virtual , Hugo Larochelle, Marc’Aurelio Ranzato, Raia Hadsell, Maria-Florina Balcan, and Hsuan-Tien Lin (Eds.)
Yunzhu Li, Antonio Torralba, Anima Anandkumar, Dieter Fox, and Animesh Garg. 2020 · 2020
Later among the works it cites.
Fine-grained driving behavior prediction via context-aware multi-task inverse reinforcement learning. In 2020 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2281–2287
Kentaro Nishi and Masamichi Shimosaka. 2020 · 2020
Later among the works it cites.
SQIL: Imitation Learning via Reinforcement Learning with Sparse Rewards. In 8th International Conference on Learning Representations, ICLR 2020, Addis Ababa, Ethiopia, April 26-30, 2020 . OpenReview.net
Siddharth Reddy, Anca D. Dragan, and Sergey Levine. 2020a · 2020
Later among the works it cites.
Learning Human Objectives by Evaluating Hypothetical Behavior. In Proceedings of the 37th International Conference on Machine Learning, ICML 2020, 13-18 July 2020, Virtual Event (Proceedings of Machine Learning Research, Vol. 119) . PMLR, 8020–8029
Siddharth Reddy, Anca D. Dragan, Sergey Levine, Shane Legg, and Jan Leike. 2020b · 2020
Later among the works it cites.
Scalable Multi-Task Imitation Learning with Autonomous Improvement. In 2020 IEEE International Conference on Robotics and Automation, ICRA 2020, Paris, France, May 31 - August 31, 2020 . IEEE, 2167–2173
Avi Singh, Eric Jang, Alexander Irpan, Daniel Kappler, Murtaza Dalal, Sergey Levine, Mohi Khansari, and Chelsea Finn. 2020 · 2020
Later among the works it cites.
Bridging Knowledge Graphs to Generate Scene Graphs. In Computer Vision - ECCV 2020 - 16th European Conference, Glasgow, UK, August 23-28, 2020, Proceedings, Part XXIII (Lecture Notes in Computer Science, Vol. 12368) , Andrea Vedaldi, Horst Bischof, Thomas Brox, and Jan-Michael Frahm (Eds.). Springer, 606–623
Alireza Zareian, Svebor Karaman, and Shih-Fu Chang. 2020 · 2020
Later among the works it cites.
On the expressivity of markov reward
David Abel, Will Dabney, Anna Harutyunyan, Mark K Ho, Michael Littman, Doina Precup, and Satinder Singh. 2021 · 2021
Later among the works it cites.
Tripletree: A versatile interpretable representation of black box agents and their environments. In Proceedings of the AAAI Conference on Artificial Intelligence , Vol. 35. 11415–11422
Tom Bewley and Jonathan Lawry. 2021 · 2021
Later among the works it cites.
Feature Expansive Reward Learning: Rethinking Human Input. In Proceedings of the 2021 ACM/IEEE International Conference on Human-Robot Interaction (Boulder, CO, USA) (HRI ’21) . Association for Computing Machinery, New York, NY, USA, 216–224
Andreea Bobu, Marius Wiggert, Claire Tomlin, and Anca D. Dragan. 2021b · 2021
Later among the works it cites.
When the ventral visual stream is not enough: A deep learning account of medial temporal lobe involvement in perception
Tyler Bonnen, Daniel L.K. Yamins, and Anthony D. Wagner. 2021 · 2021
Later among the works it cites.
Fixation patterns in simple choice reflect optimal information sampling
Frederick Callaway, Antonio Rangel, and Thomas L Griffiths. 2021 · 2021
Later among the works it cites.
Continual learning of knowledge graph embeddings
Angel Daruna, Mehul Gupta, Mohan Sridharan, and Sonia Chernova. 2021a · 2021
Later among the works it cites.
Towards Robust One-shot Task Execution using Knowledge Graph Embeddings. In IEEE International Conference on Robotics and Automation, ICRA 2021, Xi’an, China, May 30 - June 5, 2021 . IEEE, 11118–11124
Angel Andres Daruna, Lakshmi Nair, Weiyu Liu, and Sonia Chernova. 2021b · 2021
Later among the works it cites.
Semantic-Based Explainable AI: Leveraging Semantic Scene Graphs and Pairwise Ranking to Explain Robot Failures. In 2021 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) . IEEE, 3034–3041
Devleena Das and Sonia Chernova. 2021 · 2021
Later among the works it cites.
How Well Do Self-Supervised Models Transfer?. In IEEE Conference on Computer Vision and Pattern Recognition, CVPR 2021, virtual, June 19-25, 2021 . Computer Vision Foundation / IEEE, 5414–5423
Linus Ericsson, Henry Gouk, and Timothy M. Hospedales. 2021 · 2021
Later among the works it cites.
A Survey on Interpretable Reinforcement Learning
Claire Glanois, Paul Weng, Matthieu Zimmer, Dong Li, Tianpei Yang, Jianye Hao, and Wulong Liu. 2021 · 2021
Later among the works it cites.
Learning Representations by Humans, for Humans. In Proceedings of the 38th International Conference on Machine Learning, ICML 2021, 18-24 July 2021, Virtual Event (Proceedings of Machine Learning Research, Vol. 139) , Marina Meila and Tong Zhang (Eds.). PMLR, 4227–4238
Sophie Hilgard, Nir Rosenfeld, Mahzarin R. Banaji, Jack Cao, and David C. Parkes. 2021 · 2021
Later among the works it cites.
Meta Preference Learning for Fast User Adaptation in Human-Supervisory Multi-Robot Deployments. In 2021 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) . IEEE, 5851–5856
Chao Huang, Wenhao Luo, and Rui Liu. 2021 · 2021
Later among the works it cites.
PEBBLE: Feedback-Efficient Interactive Reinforcement Learning via Relabeling Experience and Unsupervised Pre-training. In Proceedings of the 38th International Conference on Machine Learning, ICML 2021, 18-24 July 2021, Virtual Event (Proceedings of Machine Learning Research, Vol. 139) , Marina Meila and Tong Zhang (Eds.). PMLR, 6152–6163
Kimin Lee, Laura M. Smith, and Pieter Abbeel. 2021 · 2021
Later among the works it cites.
EngineKGI: Closed-Loop Knowledge Graph Inference
Guanglin Niu, Bo Li, Yongfei Zhang, and Shiliang Pu. 2021 · 2021
Later among the works it cites.
Predicting Stable Configurations for Semantic Placement of Novel Objects. In Conference on Robot Learning (CoRL)
Chris Paxton, Chris Xie, Tucker Hermans, and Dieter Fox. 2021 · 2021
Later among the works it cites.
Learning transferable visual models from natural language supervision. In International Conference on Machine Learning . PMLR, 8748–8763
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, et al · 2021
Later among the works it cites.
Zero-shot text-to-image generation. In International Conference on Machine Learning . PMLR, 8821–8831
Aditya Ramesh, Mikhail Pavlov, Gabriel Goh, Scott Gray, Chelsea Voss, Alec Radford, Mark Chen, and Ilya Sutskever. 2021 · 2021
Later among the works it cites.
Pragmatic Image Compression for Human-in-the-Loop Decision-Making. In Advances in Neural Information Processing Systems 34: Annual Conference on Neural Information Processing Systems 2021, NeurIPS 2021, December 6-14, 2021, virtual , Marc’Aurelio Ranzato, Alina Beygelzimer, Yann N. Dauphin, Percy Liang, and Jennifer Wortman Vaughan (Eds.). 26499–26510
Sid Reddy, Anca D. Dragan, and Sergey Levine. 2021 · 2021
Later among the works it cites.
Pretraining Representations for Data-Efficient Reinforcement Learning. In Advances in Neural Information Processing Systems 34: Annual Conference on Neural Information Processing Systems 2021, NeurIPS 2021, December 6-14, 2021, virtual , Marc’Aurelio Ranzato, Alina Beygelzimer, Yann N. Dauphin, Percy Liang, and Jennifer Wortman Vaughan (Eds.). 12686–12699
Max Schwarzer, Nitarshan Rajkumar, Michael Noukhovitch, Ankesh Anand, Laurent Charlin, R. Devon Hjelm, Philip Bachman, and Aaron C. Courville. 2021 · 2021
Later among the works it cites.
Decoupling Representation Learning from Reinforcement Learning. In Proceedings of the 38th International Conference on Machine Learning, ICML 2021, 18-24 July 2021, Virtual Event (Proceedings of Machine Learning Research, Vol. 139) , Marina Meila and Tong Zhang (Eds.). PMLR, 9870–9879
Adam Stooke, Kimin Lee, Pieter Abbeel, and Michael Laskin. 2021 · 2021
Later among the works it cites.
On complementing end-to-end human behavior predictors with planning
Liting Sun, Xiaogang Jia, and Anca D. Dragan. 2021 · 2021
Later among the works it cites.
Representation Matters: Offline Pretraining for Sequential Decision Making. In Proceedings of the 38th International Conference on Machine Learning, ICML 2021, 18-24 July 2021, Virtual Event (Proceedings of Machine Learning Research, Vol. 139) , Marina Meila and Tong Zhang (Eds.). PMLR, 11784–11794
Mengjiao Yang and Ofir Nachum. 2021 · 2021
Later among the works it cites.
SORNet: Spatial Object-Centric Representations for Sequential Manipulation. In 5th Annual Conference on Robot Learning . PMLR, 148–157
Wentao Yuan, Chris Paxton, Karthik Desingh, and Dieter Fox. 2021 · 2021
Later among the works it cites.
Learning Invariant Representations for Reinforcement Learning without Reconstruction. In 9th International Conference on Learning Representations, ICLR 2021, Virtual Event, Austria, May 3-7, 2021 . OpenReview.net
Amy Zhang, Rowan Thomas McAllister, Roberto Calandra, Yarin Gal, and Sergey Levine. 2021b · 2021
Later among the works it cites.
Efficient Reinforcement Learning from Demonstration via Bayesian Network-Based Knowledge Extraction
Yichuan Zhang, Yixing Lan, Qiang Fang, Xin Xu, Junxiang Li, and Yujun Zeng. 2021a · 2021
Later among the works it cites.
Situational Confidence Assistance for Lifelong Shared Autonomy. In 2021 IEEE International Conference on Robotics and Automation (ICRA) . 2783–2789
Matthew Zurek, Andreea Bobu, Daniel S. Brown, and Anca D. Dragan. 2021 · 2021
Later among the works it cites.
The Task Specification Problem. In Conference on Robot Learning . PMLR, 1745–1751
Pulkit Agrawal. 2022 · 2022
Later among the works it cites.
Video PreTraining (VPT): Learning to Act by Watching Unlabeled Online Videos
Bowen Baker, Ilge Akkaya, Peter Zhokhov, Joost Huizinga, Jie Tang, Adrien Ecoffet, Brandon Houghton, Raul Sampedro, and Jeff Clune. 2022 · 2022
Later among the works it cites.
Inducing Structure in Reward Learning by Learning Features
Andreea Bobu, Marius Wiggert, Claire Tomlin, and Anca D. Dragan. 2022 · 2022
Later among the works it cites.
An Empirical Investigation of Representation Learning for Imitation
Xin Chen, Sam Toyer, Cody Wild, Scott Emmons, Ian Fischer, Kuang-Huei Lee, Neel Alex, Steven H. Wang, Ping Luo, Stuart Russell, Pieter Abbeel, and Rohin Shah. 2022 · 2022
Later among the works it cites.
Angel Andres Daruna, Devleena Das, and Sonia Chernova. 2022 · 2022
Later among the works it cites.
Bayesian Structure Learning with Generative Flow Networks
Tristan Deleu, António Góis, Chris Emezue, Mansi Rankawat, Simon Lacoste-Julien, Stefan Bauer, and Yoshua Bengio. 2022 · 2022
Later among the works it cites.
A survey of semantic reasoning frameworks for robotic systems
Weiyu Liu. 2022 · 2022
Later among the works it cites.
On the Effectiveness of Fine-tuning Versus Meta-reinforcement Learning
Zhao Mandi, Pieter Abbeel, and Stephen James. 2022 · 2022
Later among the works it cites.
Learning for Robot Decision Making under Distribution Shift: A Survey
Abhishek Paudel. 2022 · 2022
Later among the works it cites.
Self-Supervised Pretraining Improves Self-Supervised Pretraining. In 2022 IEEE/CVF Winter Conference on Applications of Computer Vision (WACV) . IEEE Computer Society, Los Alamitos, CA, USA, 1050–1060
C. J. Reed, X. Yue, A. Nrusimha, S. Ebrahimi, V. Vijaykumar, R. Mao, B. Li, S. Zhang, D. Guillory, S. Metzger, K. Keutzer, and T. Darrell. 2022 · 2022
Later among the works it cites.
Cliport: What and where pathways for robotic manipulation. In Conference on Robot Learning . PMLR, 894–906
Mohit Shridhar, Lucas Manuelli, and Dieter Fox. 2022 · 2022
Later among the works it cites.
Teaching Robots to Span the Space of Functional Expressive Motion
Arjun Sripathy, Andreea Bobu, Zhongyu Li, Koushil Sreenath, Daniel S. Brown, and Anca D. Dragan. 2022 · 2022
Later among the works it cites.
Latent Space Alignment Using Adversarially Guided Self-Play
Mycal Tucker, Yilun Zhou, and Julie Shah. 2022 · 2022
Later among the works it cites.
Task-Induced Representation Learning. In The Tenth International Conference on Learning Representations, ICLR 2022, Virtual Event, April 25-29, 2022 . OpenReview.net
Jun Yamada, Karl Pertsch, Anisha Gunjal, and Joseph J. Lim. 2022 · 2022
Later among the works it cites.
Incremental Object Grounding Using Scene Graphs
John Seon Keun Yi, Yoonwoo Kim, and Sonia Chernova. 2022 · 2022
Later among the works it cites.
SIRL: Similarity-based Implicit Representation Learning
Andreea Bobu, Yi Liu, Rohin Shah, Daniel S. Brown, and Anca D. Dragan. 2023 · 2023
Closest in time.
Diagnosis, Feedback, Adaptation: A Human-in-the-Loop Framework for Test-Time Policy Adaptation
Andi Peng, Aviv Netanyahu, Mark K Ho, Tianmin Shu, Andreea Bobu, Julie Shah, and Pulkit Agrawal. 2023 · 2023
Closest in time.
Watch this: Scalable cost-function learning for path planning in urban environments. In 2016 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) . 2089–2095
M. Wulfmeier, D. Z. Wang, and I. Posner. 2016 · 2095
Closest in time.