Fetching the paper…
Reading the bibliography…
Humans can leverage physical interaction to teach robot arms.
Stimulus and response generalization: A stochastic model relating generalization to distance in psychological space
Roger N Shepard. 1957 · 1957
Earlier work this paper cites.
Impedance control: An approach to manipulation. In American Control Conference . 304–313
Neville Hogan. 1984 · 1984
Earlier work this paper cites.
Elements of Information Theory
Thomas M Cover. 1999 · 1999
Earlier work this paper cites.
Algorithms for inverse reinforcement learning. In International Conference on Machine Learning
Andrew Y Ng and Stuart J Russell. 2000 · 2000
Earlier work this paper cites.
Apprenticeship learning via inverse reinforcement learning. In International Conference on Machine Learning
Pieter Abbeel and Andrew Y Ng. 2004 · 2004
Earlier work this paper cites.
SNOPT: An SQP algorithm for large-scale constrained optimization
Philip E Gill, Walter Murray, and Michael A Saunders. 2005 · 2005
Earlier work this paper cites.
An atlas of physical human–robot interaction
Agostino De Santis, Bruno Siciliano, Alessandro De Luca, and Antonio Bicchi. 2008 · 2008
Earlier work this paper cites.
Collision detection and reaction: A contribution to safe physical human-robot interaction. In IEEE/RSJ International Conference on Intelligent Robots and Systems . 3356–3363
Sami Haddadin, Alin Albu-Schaffer, Alessandro De Luca, and Gerd Hirzinger. 2008 · 2008
Earlier work this paper cites.
A rational model of preference learning and choice prediction by children. In Advances in Neural Information Processing Systems
Christopher Lucas, Thomas Griffiths, Fei Xu, and Christine Fawcett. 2008 · 2008
Earlier work this paper cites.
Maximum entropy inverse reinforcement learning. In AAAI . 1433–1438
Brian D Ziebart, Andrew L Maas, J Andrew Bagnell, and Anind K Dey. 2008 · 2008
Earlier work this paper cites.
A survey of robot learning from demonstration
Brenna D Argall, Sonia Chernova, Manuela Veloso, and Brett Browning. 2009 · 2009
Earlier work this paper cites.
A reduction of imitation learning and structured prediction to no-regret online learning. In International Conference on Artificial Intelligence and Statistics . 627–635
Stéphane Ross, Geoffrey Gordon, and Drew Bagnell. 2011 · 2011
Earlier work this paper cites.
Keyframe-based learning from demonstration
Baris Akgun, Maya Cakmak, Karl Jiang, and Andrea L Thomaz. 2012 · 2012
Earlier work this paper cites.
Individual Choice Behavior: A Theoretical Analysis
R Duncan Luce. 2012 · 2012
Earlier work this paper cites.
The role of roles: Physical cooperation between humans and robots
Alexander Mörtl, Martin Lawitzky, Ayse Kucukyilmaz, Metin Sezgin, Cagatay Basdogan, and Sandra Hirche. 2012 · 2012
Earlier work this paper cites.
Learning objective functions for manipulation. In IEEE International Conference on Robotics and Automation . 1331–1336
Mrinal Kalakrishnan, Peter Pastor, Ludovic Righetti, and Stefan Schaal. 2013 · 2013
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba. 2014 · 2014
Earlier work this paper cites.
Motion planning with sequential convex optimization and convex collision checking
John Schulman, Yan Duan, Jonathan Ho, Alex Lee, Ibrahim Awwal, Henry Bradlow, Jia Pan, Sachin Patil, Ken Goldberg, and Pieter Abbeel. 2014 · 2014
Cited alongside, same era.
Movement primitives via optimization. In IEEE International Conference on Robotics and Automation . 2339–2346
Anca D Dragan, Katharina Muelling, J Andrew Bagnell, and Siddhartha S Srinivasa. 2015 · 2015
Cited alongside, same era.
Learning preferences for manipulation tasks from online coactive feedback
Ashesh Jain, Shikhar Sharma, Thorsten Joachims, and Ashutosh Saxena. 2015 · 2015
Cited alongside, same era.
Physical human–robot interaction
Sami Haddadin and Elizabeth Croft. 2016 · 2016
Cited alongside, same era.
Deep reinforcement learning from human preferences. In Advances in Neural Information Processing Systems
Paul F Christiano, Jan Leike, Tom Brown, Miljan Martic, Shane Legg, and Dario Amodei. 2017 · 2017
Cited alongside, same era.
An ensemble inverse optimal control approach for robotic task learning and adaptation
Hang Yin, Francisco S Melo, Ana Paiva, and Aude Billard. 2019 · 2019
Later among the works it cites.
Quantifying hypothesis space misspecification in learning from human–robot demonstrations and physical corrections
Andreea Bobu, Andrea Bajcsy, Jaime F Fisac, Sampada Deglurkar, and Anca D Dragan. 2020 · 2020
Later among the works it cites.
Better-than-demonstrator imitation learning via automatically-ranked demonstrations. In Conference on robot learning . PMLR, 330–359
Daniel S Brown, Wonjoon Goo, and Scott Niekum. 2020 · 2020
Later among the works it cites.
Reward-rational (implicit) choice: A unifying formalism for reward learning. In Advances in Neural Information Processing Systems . 4415–4426
Hong Jun Jeon, Smitha Milli, and Anca Dragan. 2020 · 2020
Later among the works it cites.
Four years in review: Statistical practices of likert scales in human-robot interaction studies. In Companion of the 2020 ACM/IEEE International Conference on Human-Robot Interaction . 43–52
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Trajectory deformations from physical human–robot interaction
Dylan P Losey and Marcia K O’Malley. 2017 · 2017
Cited alongside, same era.
Control sharing in human-robot team interaction
Selma Musić and Sandra Hirche. 2017 · 2017
Cited alongside, same era.
Learning robust rewards with adversarial inverse reinforcement learning. In International Conference on Learning Representations
Justin Fu, Katie Luo, and Sergey Levine. 2018 · 2018
Cited alongside, same era.
Reward learning from human preferences and demonstrations in Atari. In Advances in Neural Information Processing Systems
Borja Ibarz, Jan Leike, Tobias Pohlen, Geoffrey Irving, Shane Legg, and Dario Amodei. 2018 · 2018
Cited alongside, same era.
A review of intent detection, arbitration, and communication aspects of shared control for physical human–robot interaction
Dylan P Losey, Craig G McDonald, Edoardo Battaglia, and Marcia K O’Malley. 2018 · 2018
Cited alongside, same era.
An algorithmic perspective on imitation learning
Takayuki Osa, Joni Pajarinen, Gerhard Neumann, J Andrew Bagnell, Pieter Abbeel, Jan Peters, et al · 2018
Cited alongside, same era.
Extrapolating beyond suboptimal demonstrations via inverse reinforcement learning from observations. In International Conference on Machine Learning . 783–792
Daniel Brown, Wonjoon Goo, Prabhat Nagarajan, and Scott Niekum. 2019 · 2019
Cited alongside, same era.
Mariah L Schrum, Michael Johnson, Muyleng Ghuy, and Matthew C Gombolay. 2020 · 2020
Later among the works it cites.
Learning reward functions from diverse sources of human feedback: Optimally integrating demonstrations and preferences
Erdem Bıyık, Dylan P Losey, Malayandi Palan, Nicholas C Landolfi, Gleb Shevchuk, and Dorsa Sadigh. 2021 · 2021
Later among the works it cites.
Learning from suboptimal demonstration via self-supervised reward regression. In Conference on Robot Learning . 1262–1277
Letian Chen, Rohan Paleja, and Matthew Gombolay. 2021 · 2021
Later among the works it cites.
Corrective shared autonomy for addressing task variability
Michael Hagenow, Emmanuel Senft, Robert Radwin, Michael Gleicher, Bilge Mutlu, and Michael Zinn. 2021 · 2021
Later among the works it cites.
ThriftyDAgger: Budget-aware novelty and risk gating for interactive imitation learning. In Conference on Robot Learning
Ryan Hoque, Ashwin Balakrishna, Ellen Novoseller, Albert Wilcox, Daniel S Brown, and Ken Goldberg. 2021 · 2021
Later among the works it cites.
PEBBLE: Feedback-efficient interactive reinforcement learning via relabeling experience and unsupervised pre-training. In International Conference on Machine Learning . 6152–6163
Kimin Lee, Laura M Smith, and Pieter Abbeel. 2021 · 2021
Later among the works it cites.
Learning human objectives from sequences of physical corrections. In IEEE International Conference on Robotics and Automation . 2877–2883
Mengxi Li, Alper Canberk, Dylan P Losey, and Dorsa Sadigh. 2021 · 2021
Later among the works it cites.
Physical interaction as communication: Learning robot objectives online from human corrections
Dylan P Losey, Andrea Bajcsy, Marcia K O’Malley, and Anca D Dragan. 2021 · 2021
Later among the works it cites.
Confidence-aware imitation learning from demonstrations with varying optimality. In Advances in Neural Information Processing Systems
Songyuan Zhang, Zhangjie Cao, Dorsa Sadigh, and Yanan Sui. 2021 · 2021
Later among the works it cites.
Inducing Structure in Reward Learning by Learning Features
Andreea Bobu, Marius Wiggert, Claire Tomlin, and Anca D Dragan. 2022 · 2022
Closest in time.
Mind meld: Personalized meta-learning for robot-centric imitation learning. In 2022 17th ACM/IEEE International Conference on Human-Robot Interaction (HRI) . IEEE, 157–165
Mariah L Schrum, Erin Hedlund-Botti, Nina Moorman, and Matthew C Gombolay. 2022 · 2022
Closest in time.
Expert intervention learning
Jonathan Spencer, Sanjiban Choudhury, Matthew Barnes, Matthew Schmittle, Mung Chiang, Peter Ramadge, and Sidd Srinivasa. 2022 · 2022
Closest in time.