Fetching the paper…
Reading the bibliography…
Interactive robotic grasping using natural language is one of the most fundamental tasks in human-robot interaction.
L. P. Kaelbling, M. L. Littman, and A. R. Cassandra, “Planning and acting in partially observable stochastic domains,” Artificial Intelligence , vol. 101, no. 1-2, pp. 99–134, 1998
1998
Earlier work this paper cites.
M. A. Goodrich and A. C. Schultz, Human-robot interaction: a survey . Now Publishers Inc, 2008
2008
Earlier work this paper cites.
I. Lutkebohle, J. Peltason, L. Schillingmann, B. Wrede, S. Wachsmuth, C. Elbrechter, and R. Haschke, “The curious robot-structuring interactive robot learning,” in IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2009, pp. 4156–4162
2009
Earlier work this paper cites.
L. Steels and M. Hild, Language grounding in robots . Springer Science & Business Media, 2012
2012
Earlier work this paper cites.
S. Tellex, P. Thakerll, R. Deitsl, D. Simeonovl, T. Kollar, and N. Royl, “Toward information theoretic human-robot dialog,” in Robotics: Science and Systems (RSS) . Robotics: Science and Systems, 2013, p. 409
2013
Earlier work this paper cites.
A. Somani, N. Ye, D. Hsu, and W. S. Lee, “Despot: Online pomdp planning with regularization,” Advances in Neural Information Processing Systems (NIPS) , vol. 26, pp. 1772–1780, 2013
2013
Earlier work this paper cites.
S. Tellex, R. Knepper, A. Li, D. Rus, and N. Roy, “Asking for help using inverse semantics,” in Robotics: Science and Systems (RSS) . Robotics: Science and Systems, 2014
2014
Earlier work this paper cites.
R. Kiros, R. Salakhutdinov, and R. Zemel, “Multimodal neural language models,” in International Conference on Machine Learning (ICML) . PMLR, 2014, pp. 595–603
2014
Earlier work this paper cites.
S. Kazemzadeh, V. Ordonez, M. Matten, and T. Berg, “Referitgame: Referring to objects in photographs of natural scenes,” in Conference on Empirical Methods in Natural Language Processing (EMNLP) , 2014, pp. 787–798
2014
Earlier work this paper cites.
T.-Y. Lin, M. Maire, S. Belongie, J. Hays, P. Perona, D. Ramanan, P. Dollár, and C. L. Zitnick, “Microsoft coco: Common objects in context,” in European Conference on Computer Vision (ECCV) . Springer, 2014, pp. 740–755
2014
Earlier work this paper cites.
S. Li, R. Scalise, H. Admoni, S. Rosenthal, and S. S. Srinivasa, “Spatial references and perspective in natural language instructions for collaborative manipulation,” in IEEE International Symposium on Robot and Human Interactive Communication (RO-MAN) . IEEE, 2016, pp. 44–51
2016
Earlier work this paper cites.
L. Yu, P. Poirson, S. Yang, A. C. Berg, and T. L. Berg, “Modeling context in referring expressions,” in European Conference on Computer Vision (ECCV) . Springer, 2016, pp. 69–85
2016
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2016, pp. 770–778
2016
Cited alongside, same era.
Y. Li, C. Huang, X. Tang, and C. Change Loy, “Learning to disambiguate by asking discriminative questions,” in IEEE International Conference on Computer Vision (ICCV) , 2017, pp. 3419–3428
2017
Cited alongside, same era.
D. Whitney, E. Rosen, J. MacGlashan, L. L. Wong, and S. Tellex, “Reducing errors in object-fetching interactions through social feedback,” in IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2017, pp. 1006–1013
2017
Cited alongside, same era.
J. Yang, J. Lu, S. Lee, D. Batra, and D. Parikh, “Visual curiosity: Learning to ask questions to learn visual recognition,” in Conference on Robot Learning (CoRL) . PMLR, 2018, pp. 63–80
2018
Cited alongside, same era.
S. Tellex, N. Gopalan, H. Kress-Gazit, and C. Matuszek, “Robots that use language,” Annual Review of Control, Robotics, and Autonomous Systems , vol. 3, pp. 25–55, 2020
2020
Later among the works it cites.
M. Forbes, R. P. Rao, L. Zettlemoyer, and M. Cakmak, “Robot programming by demonstration with situated spatial language understanding,” in IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2015, pp. 2014–2020
2020
Later among the works it cites.
Y. Qiao, C. Deng, and Q. Wu, “Referring expression comprehension: A survey of methods and datasets,” IEEE Transactions on Multimedia , 2020
2020
Later among the works it cites.
M. Shridhar, D. Mittal, and D. Hsu, “Ingress: Interactive visual grounding of referring expressions,” The International Journal of Robotics Research , vol. 39, no. 2-3, pp. 217–232, 2020
2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
J. Hatori, Y. Kikuchi, S. Kobayashi, K. Takahashi, Y. Tsuboi, Y. Unno, W. Ko, and J. Tan, “Interactively picking real-world objects with unconstrained spoken language instructions,” in IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2018, pp. 3774–3781
2018
Cited alongside, same era.
H. Ahn, S. Choi, N. Kim, G. Cha, and S. Oh, “Interactive text2pickup networks for natural language-based human–robot collaboration,” IEEE Robotics and Automation Letters , vol. 3, no. 4, pp. 3308–3315, 2018
2018
Cited alongside, same era.
L. Yu, Z. Lin, X. Shen, J. Yang, X. Lu, M. Bansal, and T. L. Berg, “Mattnet: Modular attention network for referring expression comprehension,” in IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2018, pp. 1307–1315
2018
Cited alongside, same era.
M. Z. Hossain, F. Sohel, M. F. Shiratuddin, and H. Laga, “A comprehensive survey of deep learning for image captioning,” ACM Computing Surveys (CsUR) , vol. 51, no. 6, pp. 1–36, 2019
2019
Cited alongside, same era.
J. Thomason, A. Padmakumar, J. Sinapov, N. Walker, Y. Jiang, H. Yedidsion, J. Hart, P. Stone, and R. J. Mooney, “Improving grounded natural language understanding through human-robot dialog,” in International Conference on Robotics and Automation (ICRA) . IEEE, 2019, pp. 6934–6941
2019
Cited alongside, same era.
K. Morohashi and J. Miura, “Query generation for resolving ambiguity in user’s command for a mobile service robot,” in European Conference on Mobile Robots (ECMR) . IEEE, 2019, pp. 1–6
2019
Cited alongside, same era.
Y. Cui, M. Jia, T.-Y. Lin, Y. Song, and S. Belongie, “Class-balanced loss based on effective number of samples,” in IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , 2019, pp. 9268–9277
2019
Cited alongside, same era.
X. Lou, Y. Yang, and C. Choi, “Learning to generate 6-dof grasp poses with reachability awareness,” in IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2020, pp. 1532–1538
2020
Later among the works it cites.
Y. Yang, Y. Liu, H. Liang, X. Lou, and C. Choi, “Attribute-based robotic grasping with one-grasp adaptation,” in IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2021
2021
Later among the works it cites.
2021
Later among the works it cites.
O. Kroemer, S. Niekum, and G. Konidaris, “A review of robot learning for manipulation: Challenges, representations, and algorithms,” Journal of Machine Learning Research , vol. 22, pp. 1–82, 2021
2021
Later among the works it cites.
H. Zhang, Y. Lu, C. Yu, D. Hsu, X. Lan, and N. Zheng, “Invigorate: Interactive visual grounding and grasping in clutter,” in Robotics: Science and Systems (RSS) . Robotics: Science and Systems, 2021
2021
Later among the works it cites.
C. Xie, Y. Xiang, A. Mousavian, and D. Fox, “Unseen object instance segmentation for robotic environments,” IEEE Transactions on Robotics , 2021
2021
Later among the works it cites.
X. Lou, Y. Yang, and C. Choi, “Collision-aware target-driven object grasping in constrained environments,” in IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2021, pp. 6364–6370
2021
Later among the works it cites.