Fetching the paper…
Reading the bibliography…
Commercial and industrial deployments of robot fleets at Amazon, Nimble, Plus One, Waymo, and Zoox query remote human teleoperators when robots are at risk or unable to make task progress.
A mathematical theory of communication
C. Shannon · 1948
Earlier work this paper cites.
Adaptive allocation of decision making responsibility between supervisor and computer
W. B. Rouse · 1976
Earlier work this paper cites.
Queueing theory
R. B. Cooper · 1981
Earlier work this paper cites.
Efficient training of artificial neural networks for autonomous navigation
D. A. Pomerleau · 1991
Earlier work this paper cites.
Cooperative assistance for remote robot supervision
R. R. Murphy and E. Rogers · 1996
Earlier work this paper cites.
Adjustable control autonomy for manned space flight
D. Kortenkamp, D. Keirn-Schreckenghost, and R. P. Bonasso · 2000
Earlier work this paper cites.
Towards adjustable autonomy for the real world
P. Scerri, D. V. Pynadath, and M. Tambe · 2002
Earlier work this paper cites.
Fan-out: Measuring human control of multiple robots
D. R. Olsen Jr and S. B. Wood · 2004
Earlier work this paper cites.
Real time scheduling theory: A historical perspective
L. Sha, T. Abdelzaher, A. Cervin, T. Baker, A. Burns, G. Buttazzo, M. Caccamo, J. Lehoczky, A. K. Mok, et al · 2004
Earlier work this paper cites.
User modelling for principled sliding autonomy in human-robot teams
B. Sellner, R. Simmons, and S. Singh · 2005
Earlier work this paper cites.
Coordinated multiagent teams and sliding autonomy for large-scale assembly
B. Sellner, F. W. Heger, L. M. Hiatt, R. Simmons, and S. Singh · 2006
Earlier work this paper cites.
Sliding autonomy for peer-to-peer human-robot teams
M. B. Dias, B. Kannan, B. Browning, E. Jones, B. Argall, M. F. Dias, M. Zinck, M. Veloso, and A. Stentz · 2008
Earlier work this paper cites.
A survey of robot learning from demonstration
B. D. Argall, S. Chernova, M. Veloso, and B. Browning · 2009
Earlier work this paper cites.
Interactive policy learning through confidence-based autonomy
S. Chernova and M. Veloso · 2009
Earlier work this paper cites.
Computing the effects of operator attention allocation in human control of multiple robots
J. W. Crandall, M. L. Cummings, M. Della Penna, and P. M. De Jong · 2010
Earlier work this paper cites.
Designing interactions for robot active learners
M. Cakmak, C. Chao, and A. L. Thomaz · 2010
Earlier work this paper cites.
Adaptive automation, level of automation, allocation authority, supervisory control, and adaptive control: Distinctions and modes of adaptation
T. B. Sheridan · 2011
Earlier work this paper cites.
A reduction of imitation learning and structured prediction to no-regret online learning
S. Ross, G. J. Gordon, and J. A. Bagnell · 2011
Earlier work this paper cites.
Supervisory control of multiple robots: Effects of imperfect automation and individual differences
J. Y. Chen and M. J. Barnes · 2012
Earlier work this paper cites.
Maximum mean discrepancy imitation learning
B. Kim and J. Pineau · 2013
Earlier work this paper cites.
Human interaction with multiple remote robots
M. Lewis · 2013
Earlier work this paper cites.
Imperfect automation in scheduling operator attention on control of multi-robots
S.-Y. Chien, M. Lewis, S. Mehrotra, and K. Sycara · 2013
Earlier work this paper cites.
Supervisory control of multiple social robots for navigation
K. Zheng, D. F. Glas, T. Kanda, H. Ishiguro, and N. Hagita · 2013
Earlier work this paper cites.
Power to the people: The role of humans in interactive machine learning
S. Amershi, M. Cakmak, W. B. Knox, and T. Kulesza · 2014
Earlier work this paper cites.
Human–agent teaming for multirobot control: A review of human factors issues
J. Y. Chen and M. J. Barnes · 2014
Cited alongside, same era.
Human supervisory control of robotic teams: Integrating cognitive modeling with engineering design
J. R. Peters, V. Srivastava, G. S. Taylor, A. Surana, M. P. Eckstein, and F. Bullo · 2015
Cited alongside, same era.
An invitation to imitation
J. A. Bagnell · 2015
Cited alongside, same era.
SHIV: Reducing supervisor burden using support vectors for efficient learning from demonstrations in high dimensional state spaces
M. Laskey, S. Staszak, W. Hsieh, J. Mahler, F. Pokorny, A. Dragan, and K. Goldberg · 2016
Cited alongside, same era.
Generative adversarial imitation learning
J. Ho and S. Ermon · 2016
Cited alongside, same era.
Learning from interventions: Human-robot interaction as both explicit and implicit feedback
J. Spencer, S. Choudhury, M. Barnes, M. Schmittle, M. Chiang, P. Ramadge, and S. Srinivasa · 2020
Later among the works it cites.
Human-in-the-loop imitation learning using remote teleoperation
A. Mandlekar, D. Xu, R. Martín-Martín, Y. Zhu, L. Fei-Fei, and S. Savarese · 2020
Later among the works it cites.
Scaled autonomy: Enabling human operators to control robot fleets
G. Swamy, S. Reddy, S. Levine, and A. D. Dragan · 2020
Later among the works it cites.
Thriftydagger: Budget-aware novelty and risk gating for interactive imitation learning
R. Hoque, A. Balakrishna, E. Novoseller, A. Wilcox, D. S. Brown, and K. Goldberg · 2021
Later among the works it cites.
LazyDAgger: Reducing context switching in interactive imitation learning
R. Hoque, A. Balakrishna, C. Putterman, M. Luo, D. S. Brown, D. Seita, B. Thananjeyan, E. Novoseller, and K. Goldberg · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
G. Brockman, V. Cheung, L. Pettersson, J. Schneider, J. Schulman, J. Tang, and W. Zaremba · 2016
Cited alongside, same era.
Federated learning: Strategies for improving communication efficiency
J. Konečný, H. B. McMahan, F. X. Yu, P. Richtarik, A. T. Suresh, and D. Bacon · 2016
Cited alongside, same era.
Query-efficient imitation learning for end-to-end autonomous driving
J. Zhang and K. Cho · 2017
Cited alongside, same era.
Intelligent agent supporting human–multi-robot team collaboration
A. Rosenfeld, N. Agmon, O. Maksimov, and S. Kraus · 2017
Cited alongside, same era.
DART: Noise injection for robust imitation learning
M. Laskey, J. Lee, R. Fox, A. Dragan, and K. Goldberg · 2017
Cited alongside, same era.
Proximal policy optimization algorithms
J. Schulman, F. Wolski, P. Dhariwal, A. Radford, and O. Klimov · 2017
Cited alongside, same era.
Waymo’s Robot Cars, and the Humans Who Tend to Them
A. C. Madrigal · 2018
Cited alongside, same era.
Isaac gym: High performance gpu-based physics simulation for robot learning
V. Makoviychuk, L. Wawrzyniak, Y. Guo, M. Lu, K. Storey, M. Macklin, D. Hoeller, N. Rudin, A. Allshire, A. Handa, and G. State · 2021
Later among the works it cites.
Scheduling and path-planning for operator oversight of multiple robots
S. A. Zanlongo, P. Dirksmeier, P. Long, T. Padir, and L. Bobadilla · 2021
Later among the works it cites.
Land: Learning to navigate from disengagements
G. Kahn, P. Abbeel, and S. Levine · 2021
Later among the works it cites.
Recovery rl: Safe reinforcement learning with learned recovery zones
B. Thananjeyan, A. Balakrishna, S. Nair, M. Luo, K. P. Srinivasan, M. Hwang, J. E. Gonzalez, J. Ibarz, C. Finn, and K. Goldberg · 2021
Later among the works it cites.
Robin deals with a world where things are changing all around it
A. Brown · 2022
Closest in time.
Robotic Arms Are Using Machine Learning to Reach Deeper Into Distribution
J. Smith · 2022
Closest in time.
Parcel Monitor , May 2022
Logistics Automation with Plus One Robotics · 2022
Closest in time.
How Zoox Builds Autonomous Vehicles from the Wheels Up-Blog
E. Chu · 2022
Closest in time.
Traversing supervisor problem: An approximately optimal approach to multi-robot assistance
T. Ji, R. Dong, and K. Driggs-Campbell · 2022
Closest in time.
Scalable operator allocation for multi-robot assistance: A restless bandit approach
A. Dahiya, N. Akbarzadeh, A. Mahajan, and S. L. Smith · 2022
Closest in time.
Interactive reinforcement learning with bayesian fusion of multimodal advice
S. Trick, F. Herbert, C. A. Rothkopf, and D. Koert · 2022
Closest in time.
Synergistic scheduling of learning and allocation of tasks in human-robot teams
S. Vats, O. Kroemer, and M. Likhachev · 2022
Closest in time.
Rapid locomotion via reinforcement learning
G. B. Margolis, G. Yang, K. Paigwar, T. Chen, and P. Agrawal · 2022
Closest in time.
Masked visual pre-training for motor control
T. Xiao, I. Radosavovic, T. Darrell, and J. Malik · 2022
Closest in time.
Adversarial motion priors make good substitutes for complex reward functions
A. Escontrela, X. B. Peng, W. Yu, T. Zhang, A. Iscen, K. Goldberg, and P. Abbeel · 2022
Closest in time.
Ase: Large-scale reusable adversarial skill embeddings for physically simulated characters
X. B. Peng, Y. Guo, L. Halper, S. Levine, and S. Fidler · 2022
Closest in time.
Learning to walk in minutes using massively parallel deep reinforcement learning
N. Rudin, D. Hoeller, P. Reist, and M. Hutter · 2022
Closest in time.
Eliciting compatible demonstrations for multi-human imitation learning
K. Gandhi, S. Karamcheti, M. Liao, and D. Sadigh · 2022
Closest in time.
Data games: A game-theoretic approach to swarm robotic data collection
O. Akcin, P. Li, S. Agarwal, and S. Chinchali · 2022
Closest in time.