Fetching the paper…
Reading the bibliography…
When robots interact with human partners, often these partners change their behavior in response to the robot.
R. Caruana, “Multitask learning,” Machine Learning , 1997
1997
Earlier work this paper cites.
M. Bowling and M. Veloso, “Multiagent learning using a variable learning rate,” Artificial Intelligence , pp. 215–250, 2002
2002
Earlier work this paper cites.
R. Vidal, O. Shakernia, H. J. Kim, D. H. Shim, and S. Sastry, “Probabilistic pursuit-evasion games: Theory, implementation, and experimental evaluation,” IEEE Transactions on Robotics and Automation , vol. 18, no. 5, pp. 662–669, 2002
2002
Earlier work this paper cites.
K. Cho, B. van Merriënboer, D. Bahdanau, and Y. Bengio, “On the properties of neural machine translation: Encoder–decoder approaches,” in Proceedings of Eighth Workshop on Syntax, Semantics and Structure in Statistical Translation , 2014, pp. 103–111
2014
Earlier work this paper cites.
D. Sadigh, S. Sastry, S. A. Seshia, and A. D. Dragan, “Planning for autonomous cars that leverage effects on human actions,” in Robotics: Science and Systems , vol. 2, 2016, pp. 1–9
2016
Earlier work this paper cites.
A. Bestick, R. Bajcsy, and A. D. Dragan, “Implicitly assisting humans to choose good grasps in robot to human handovers,” in International Symposium on Experimental Robotics , 2016, pp. 341–354
2016
Earlier work this paper cites.
F. Doshi-Velez and G. Konidaris, “Hidden parameter markov decision processes: A semiparametric regression approach for discovering latent task parametrizations,” in IJCAI , 2016, p. 1432
2016
Earlier work this paper cites.
2017
Earlier work this paper cites.
J. N. Foerster, R. Y. Chen, M. Al-Shedivat, S. Whiteson, P. Abbeel, and I. Mordatch, “Learning with opponent-learning awareness,” in Int. Conf. Autonomous Agents and MultiAgent Systems , 2017
2017
Cited alongside, same era.
J. Foerster, G. Farquhar, T. Afouras, N. Nardelli, and S. Whiteson, “Counterfactual multi-agent policy gradients,” in AAAI , 2018
2018
Cited alongside, same era.
K. Cao, A. Lazaridou, M. Lanctot, J. Z. Leibo, K. Tuyls, and S. Clark, “Emergent communication through negotiation,” in International Conference on Learning Representations , 2018
2018
Cited alongside, same era.
T. Haarnoja, A. Zhou, P. Abbeel, and S. Levine, “Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor,” in International Conference on Machine Learning , 2018
2018
Cited alongside, same era.
A. Xie, D. P. Losey, R. Tolsma, C. Finn, and D. Sadigh, “Learning latent representations to influence multi-agent interaction,” in Conference on Robot Learning , 2020
2020
Later among the works it cites.
W. Z. Wang, A. Shih, A. Xie, and D. Sadigh, “Influencing towards stable multi-agent interactions,” in Conf. on Robot Learning , 2021
2021
Later among the works it cites.
K. K. Ndousse, D. Eck, S. Levine, and N. Jaques, “Emergent social learning via multi-agent reinforcement learning,” in International Conference on Machine Learning , 2021, pp. 7991–8004
2021
Later among the works it cites.
M. Li, M. Kwon, and D. Sadigh, “Influencing leading and following in human-robot teams,” Autonomous Robots , vol. 45, pp. 959–978, 2021
2021
Later among the works it cites.
A. Shih, A. Sawhney, J. Kondic, S. Ermon, and D. Sadigh, “On the critical role of conventions in adaptive human-AI collaboration,” in International Conference on Learning Representations , 2021
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
D. P. Losey and D. Sadigh, “Robots that take advantage of human trust,” in IEEE/RSJ International Conference on Intelligent Robots and Systems , 2019, pp. 7001–7008
2019
Cited alongside, same era.
S. Saunderson and G. Nejat, “How robots influence humans: A survey of nonverbal communication in social human–robot interaction,” International Journal of Social Robotics , pp. 575–608, 2019
2019
Cited alongside, same era.
M. Carroll, R. Shah, M. K. Ho, T. Griffiths, S. Seshia, P. Abbeel, and A. Dragan, “On the utility of learning about humans for human-AI coordination,” in NeurIPS , 2019
2019
Cited alongside, same era.
2020
Cited alongside, same era.
2021
Later among the works it cites.
A. Lupu, B. Cui, H. Hu, and J. Foerster, “Trajectory diversity for zero-shot coordination,” 2021
2021
Later among the works it cites.
S. Habibian and D. P. Losey, “Encouraging human interaction with robot teams: Legible and fair subtask allocations,” IEEE Robotics and Automation Letters , vol. 7, no. 3, pp. 6685–6692, 2022
2022
Closest in time.
A. Jonnavittula and D. P. Losey, “Communicating robot conventions through shared autonomy,” in IEEE International Conference on Robotics and Automation , 2022
2022
Closest in time.