Fetching the paper…
Reading the bibliography…
How can we train an assistive human-machine interface (e.g., an electromyography-based limb prosthesis) to translate a user's raw command signals into the actions of a robot or computer when there is no prior mapping, we cannot ask the user for supervision in the form of action labels or reward feedback, and we do not have prior knowledge of the tasks the user is trying to accomplish? The key idea in this paper is that, regardless of the task, when an interface is more intuitive, the user's commands are less noisy.
A mathematical theory of communication
Claude Elwood Shannon · 1948
Earlier work this paper cites.
The decipherment of Linear B
John Chadwick · 1967
Earlier work this paper cites.
Self-organization in a perceptual network
Ralph Linsker · 1988
Earlier work this paper cites.
Glove-Talk: A neural network interface between a data-glove and a speech synthesizer
Sidney S Fels and Geoffrey E Hinton · 1993
Earlier work this paper cites.
Glove-Talk II—a neural-network interface which maps gestures to parallel formant speech synthesizer controls
Sidney S Fels and Geoffrey E Hinton · 1997
Earlier work this paper cites.
Mapping performer parameters to synthesis engines
Andy Hunt and Marcelo M Wanderley · 2002
Earlier work this paper cites.
Snow Crash
Neal Stephenson · 2003
Earlier work this paper cites.
Information theory, inference and learning algorithms
David J C MacKay · 2003
Earlier work this paper cites.
The IM algorithm: a variational approach to information maximization
David Agakov and Felix Barber · 2004
Earlier work this paper cites.
Empowerment: A universal agent-centric measure of control
Alexander S Klyubin, Daniel Polani, and Chrystopher L Nehaniv · 2005
Earlier work this paper cites.
Gaussian Processes for Machine Learning
Carl Edward Rasmussen and Christopher K. I. Williams · 2006
Earlier work this paper cites.
Evaluating the robustness of learning from implicit feedback
Filip Radlinski and Thorsten Joachims · 2006
Earlier work this paper cites.
Elementary cryptanalysis
Abraham Sinkov and Todd Feil · 2009
Earlier work this paper cites.
Exploiting co-adaptation for the design of symbiotic neuroprosthetic assistants
Justin C Sanchez, Babak Mahmoudi, Jack DiGiovanna, and Jose C Principe · 2009
Earlier work this paper cites.
Interactively shaping agents via human reinforcement: The TAMER framework
W Bradley Knox and Peter Stone · 2009
Earlier work this paper cites.
Modeling interaction via the principle of maximum causal entropy
Brian D Ziebart, J Andrew Bagnell, and Anind K Dey · 2010
Earlier work this paper cites.
Scikit-learn: Machine learning in Python
F. Pedregosa, G. Varoquaux, A. Gramfort, V. Michel, B. Thirion, O. Grisel, M. Blondel, P. Prettenhofer, R. Weiss, V. Dubourg, J. Vanderplas, A. Passos, D. Cournapeau, M. Brucher, M. Perrot, and E. Duchesnay · 2011
Earlier work this paper cites.
Online human training of a myoelectric prosthesis controller via actor-critic reinforcement learning
Patrick M Pilarski, Michael R Dawson, Thomas Degris, Farbod Fahimi, Jason P Carey, and Richard S Sutton · 2011
Earlier work this paper cites.
A high-performance neural prosthesis enabled by control algorithm design
Vikash Gilja, Paul Nuyujukian, Cindy A Chestek, John P Cunningham, M Yu Byron, Joline M Fan, Mark M Churchland, Matthew T Kaufman, Jonathan C Kao, Stephen I Ryu, et al · 2012
Earlier work this paper cites.
Hierarchical relative entropy policy search
Christian Daniel, Gerhard Neumann, and Jan Peters · 2012
Earlier work this paper cites.
Design and analysis of closed-loop decoder adaptation algorithms for brain-machine interfaces
Siddharth Dangi, Amy L Orsborn, Helene G Moorman, and Jose M Carmena · 2013
Earlier work this paper cites.
A multi-agent control framework for co-adaptation in brain-computer interfaces
Josh S Merel, Roy Fox, Tony Jebara, and Liam Paninski · 2013
Earlier work this paper cites.
Continuous closed-loop decoder adaptation with a recursive maximum likelihood algorithm allows for rapid performance acquisition in brain-machine interfaces
Siddharth Dangi, Suraj Gowda, Helene G Moorman, Amy L Orsborn, Kelvin So, Maryam Shanechi, and Jose M Carmena · 2014
Earlier work this paper cites.
Autonomy infused teleoperation with application to BCI manipulation
Katharina Muelling, Arun Venkatraman, Jean-Sebastien Valois, John Downey, Jeffrey Weiss, Shervin Javdani, Martial Hebert, Andrew B Schwartz, Jennifer L Collinger, and J Andrew Bagnell · 2015
Earlier work this paper cites.
Contextual Markov decision processes
Assaf Hallak, Dotan Di Castro, and Shie Mannor · 2015
Earlier work this paper cites.
Neuroprosthetic decoder training as imitation learning
Josh Merel, David Carlson, Liam Paninski, and John P Cunningham · 2015
Cited alongside, same era.
Progressive co-adaptation in human-machine interaction
Paolo Gallina, Nicola Bellotto, and Massimiliano Di Luca · 2015
Cited alongside, same era.
Variational information maximisation for intrinsically motivated reinforcement learning
Shakir Mohamed and Danilo Jimenez Rezende · 2015
Cited alongside, same era.
Learning language games through interaction
Sida I Wang, Percy Liang, and Christopher D Manning · 2016
Cited alongside, same era.
Karol Gregor, Danilo Jimenez Rezende, and Daan Wierstra · 2016
Social influence as intrinsic motivation for multi-agent deep reinforcement learning
Natasha Jaques, Angeliki Lazaridou, Edward Hughes, Caglar Gulcehre, Pedro Ortega, DJ Strouse, Joel Z Leibo, and Nando De Freitas · 2019
Later among the works it cites.
Speech synthesis from neural decoding of spoken sentences
Gopala K Anumanchipalli, Josh Chartier, and Edward F Chang · 2019
Later among the works it cites.
A computational model of human decision making and learning for assessment of co-adaptation in neuro-adaptive human-robot interaction
Stefan K Ehrlich and Gordon Cheng · 2019
Later among the works it cites.
Visceral machines: Reinforcement learning with intrinsic physiological rewards
Daniel McDuff and Ashish Kapoor · 2019
Later among the works it cites.
Dynamics-aware unsupervised discovery of skills
Archit Sharma, Shixiang Gu, Sergey Levine, Vikash Kumar, and Karol Hausman · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Greg Brockman, Vicki Cheung, Ludwig Pettersson, Jonas Schneider, John Schulman, Jie Tang, and Wojciech Zaremba · 2016
Cited alongside, same era.
Jacob Andreas, Anca Dragan, and Dan Klein · 2017
Cited alongside, same era.
Learning models for shared control of human-machine systems with unknown dynamics
Alexander Broad, Todd David Murphey, and Brenna Dee Argall · 2017
Cited alongside, same era.
Human-robot mutual adaptation in collaborative tasks: Models and experiments
Stefanos Nikolaidis, David Hsu, and Siddhartha Srinivasa · 2017
Cited alongside, same era.
Interactive learning from policy-dependent human feedback
James MacGlashan, Mark K Ho, Robert Loftin, Bei Peng, David Roberts, Matthew E Taylor, and Michael L Littman · 2017
Cited alongside, same era.
Active preference-based learning of reward functions
Dorsa Sadigh, Anca D Dragan, Shankar Sastry, and Sanjit A Seshia · 2017
Cited alongside, same era.
On variational lower bounds of mutual information
Ben Poole, Sherjil Ozair, Aäron van den Oord, Alexander A Alemi, and George Tucker · 2018
Cited alongside, same era.
Fast task inference with variational intrinsic successor features
Steven Hansen, Will Dabney, Andre Barreto, Tom Van de Wiele, David Warde-Farley, and Volodymyr Mnih · 2019
Later among the works it cites.
Developing embodied familiarity with hyperphysical phenomena
Gray Crawford · 2019
Later among the works it cites.
Putting an end to end-to-end: Gradient-isolated learning of representations
Sindy Löwe, Peter O’Connor, and Bastiaan Veeling · 2019
Later among the works it cites.
AvE: Assistance via empowerment
Yuqing Du, Stas Tiomkin, Emre Kiciman, Daniel Polani, Pieter Abbeel, and Anca Dragan · 2020
Later among the works it cites.
Thread: Circuits
Nick Cammarata, Shan Carter, Gabriel Goh, Chris Olah, Michael Petrov, Ludwig Schubert, Chelsea Voss, Ben Egan, and Swee Kiat Lim · 2020
Later among the works it cites.
High-performance brain-to-text communication via imagined handwriting
Francis R Willett, Donald T Avansino, Leigh R Hochberg, Jaimie M Henderson, and Krishna V Shenoy · 2020
Later among the works it cites.
Learning adaptive language interfaces through decomposition
Siddharth Karamcheti, Dorsa Sadigh, and Percy Liang · 2020
Later among the works it cites.
Residual policy learning for shared autonomy
Charles Schaff and Matthew R Walter · 2020
Later among the works it cites.
Shared autonomy with learned latent actions
Hong Jun Jeon, Dylan P Losey, and Dorsa Sadigh · 2020
Later among the works it cites.
Errors in human-robot interactions and their effects on robot learning
Su Kyoung Kim, Elsa Andrea Kirchner, Lukas Schloßmüller, and Frank Kirchner · 2020
Later among the works it cites.
Accelerating reinforcement learning agent with EEG-based implicit human feedback
Duo Xu, Mohit Agarwal, Faramarz Fekri, and Raghupathy Sivakumar · 2020
Later among the works it cites.
The empathic framework for task learning from implicit human feedback
Yuchen Cui, Qiping Zhang, Alessandro Allievi, Peter Stone, Scott Niekum, and W Bradley Knox · 2020
Later among the works it cites.
MediaPipe hands: On-device real-time hand tracking
Fan Zhang, Valentin Bazarevsky, Andrey Vakunov, Andrei Tkachenka, George Sung, Chuo-Ling Chang, and Matthias Grundmann · 2020
Later among the works it cites.
Jean Harb, Tom Schaul, Doina Precup, and Pierre-Luc Bacon · 2020
Later among the works it cites.
A framework for optimizing co-adaptation in body-machine interfaces
Dalia De Santis · 2021
Later among the works it cites.
X2T: Training an x-to-text typing interface with online learning from user feedback
Jensen Gao, Siddharth Reddy, Glen Berseth, Nicholas Hardy, Nikhilesh Natraj, Karunesh Ganguly, Anca Dragan, and Sergey Levine · 2021
Later among the works it cites.
Randomized ensembled double Q-learning: Learning fast without a model
Xinyue Chen, Che Wang, Zijian Zhou, and Keith Ross · 2021
Later among the works it cites.
CIC: Contrastive intrinsic control for unsupervised skill discovery
Michael Laskin, Hao Liu, Xue Bin Peng, Denis Yarats, Aravind Rajeswaran, and Pieter Abbeel · 2022
Closest in time.
ASHA: Assistive teleoperation via human-in-the-loop reinforcement learning
Sean Chen, Jensen Gao, Siddharth Reddy, Glen Berseth, Anca D. Dragan, and Sergey Levine · 2022
Closest in time.
Low-dimensional embeddings for interaction design
Marius Mihai Rusu, Svenja Yvonne Schött, John H Williamson, Albrecht Schmidt, and Roderick Murray-Smith · 2022
Closest in time.