Fetching the paper…
Reading the bibliography…
We introduce a pipeline that enhances a general-purpose Vision Language Model, GPT-4V(ision), to facilitate one-shot visual teaching for robotic manipulation.
H. T. Kuhn and W. L. Inequalities, “Related systems,” Annals of Mathematic Studies, Princeton Univ. Press. EEUU
1956
Earlier work this paper cites.
T. Yoshikawa, “Passive and active closures by constraining mechanisms,” in 1996 ICRA
1996
Earlier work this paper cites.
Psychology press, 2014
J. J. Gibson, The ecological approach to visual perception: classic edition · 2014
Earlier work this paper cites.
R. Goyal, S. Ebrahimi Kahou, V. Michalski, J. Materzynska, S. Westphal, H. Kim, V. Haenel, I. Fruend, P. Yianilos, M. Mueller-Freitag, et al
2017
Earlier work this paper cites.
A. Saudabayev, Z. Rysbek, R. Khassenova, and H. A. Varol, “Human grasping database for activities of daily living with depth, color and kinematic data streams,” Scientific data
2018
Earlier work this paper cites.
P. Pramanick, H. B. Barua, and C. Sarkar, “Decomplex: Task planning from complex natural instructions by a collocating robot,” in 2020 IROS
2020
Earlier work this paper cites.
2020
Earlier work this paper cites.
K. Sasabuchi, N. Wake, and K. Ikeuchi, “Task-oriented motion mapping on robots of various configuration using body role division,” IEEE Robotics and Automation Letters
2020
Earlier work this paper cites.
D. Damen, H. Doughty, G. M. Farinella, S. Fidler, A. Furnari, E. Kazakos, D. Moltisanti, J. Munro, T. Perrett, W. Price, et al
2020
Earlier work this paper cites.
S. G. Venkatesh, R. Upadrashta, and B. Amrutur, “Translating natural language instructions to computer programs for robot manipulation,” in 2021 IROS
2021
Earlier work this paper cites.
C. R. Garrett, R. Chitnis, R. Holladay, B. Kim, T. Silver, L. P. Kaelbling, and T. Lozano-Pérez, “Integrated task and motion planning,” Annual review of control, robotics, and autonomous systems
2021
Earlier work this paper cites.
N. Wake, R. Arakawa, I. Yanokura, T. Kiyokawa, K. Sasabuchi, J. Takamatsu, and K. Ikeuchi, “A learning-from-observation framework: One-shot robot teaching for grasp-manipulation-release household operations,” in 2021 SII
2021
Earlier work this paper cites.
N. Wake, I. Yanokura, K. Sasabuchi, and K. Ikeuchi, “Verbal focus-of-attention system for learning-from-demonstration,” in 2021 ICRA
2021
Earlier work this paper cites.
Y. Jiang, A. Gupta, Z. Zhang, G. Wang, Y. Dou, Y. Chen, L. Fei-Fei, A. Anandkumar, Y. Zhu, and L. Fan, “Vima: General robot manipulation with multimodal prompts,” arXiv
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
W. Huang, P. Abbeel, D. Pathak, and I. Mordatch, “Language models as zero-shot planners: Extracting actionable knowledge for embodied agents,” in 2022 ICML
2022
Earlier work this paper cites.
I. Yanaokura, N. Wake, K. Sasabuchi, R. Arakawa, K. Okada, J. Takamatsu, M. Inaba, and K. Ikeuchi, “A multimodal learning-from-observation towards all-at-once robot teaching using task cohesion,” in 2022 SII
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
X. Zhou, R. Girdhar, A. Joulin, P. Krähenbühl, and I. Misra, “Detecting twenty-thousand classes using image-level supervision,” in 2022 ECCV
2022
Cited alongside, same era.
D. Saito, K. Sasabuchi, N. Wake, J. Takamatsu, H. Koike, and K. Ikeuchi, “Task-grasping from a demonstrated human strategy,” in 2022 Humanoids
2022
Cited alongside, same era.
2022
Cited alongside, same era.
2023
Cited alongside, same era.
X. Li, M. Liu, H. Zhang, C. Yu, J. Xu, H. Wu, C. Cheang, Y. Jing, W. Zhang, H. Liu, et al
2023
Closest in time.
2023
Closest in time.
C. Zhao, S. Yuan, C. Jiang, J. Cai, H. Yu, M. Y. Wang, and Q. Chen, “Erra: An embodied representation and reasoning architecture for long-horizon language-conditioned manipulation tasks,” IEEE Robotics and Automation Letters
2023
Closest in time.
O. Mees, J. Borja-Diaz, and W. Burgard, “Grounding language with visual affordances over unstructured data,” in 2023 ICRA
2023
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Cited alongside, same era.
N. Wake, A. Kanehira, K. Sasabuchi, J. Takamatsu, and K. Ikeuchi, “Chatgpt empowered long-step robot control in various environments: A case application,” IEEE Access
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Closest in time.
2023
Closest in time.
S. S. Raman, V. Cohen, D. Paulius, I. Idrees, E. Rosen, R. Mooney, and S. Tellex, “Cape: Corrective actions from precondition errors using large language models,” in 2nd Workshop on Language and Robot Learning: Language as Grounding
2023
Closest in time.
2023
Closest in time.
N. Wake, A. Kanehira, K. Sasabuchi, J. Takamatsu, and K. Ikeuchi, “Interactive task encoding system for learning-from-observation,” in 2023 AIM
2023
Closest in time.
N. Wake, D. Saito, K. Sasabuchi, H. Koike, and K. Ikeuchi, “Text-driven object affordance for guiding grasp-type recognition in multimodal robot teaching,” Machine Vision and Applications
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
Accessed: 2024-08-17
OpenAI, “Chatgpt.” https://openai.com/blog/chatgpt · 2024
Closest in time.
Accessed: 2024-08-17
OpenAI, “Gpt-4.” https://openai.com/research/gpt-4 · 2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
K. Ikeuchi, J. Takamatsu, K. Sasabuchi, N. Wake, and A. Kanehira, “Applying learning-from-observation to household service robots: three task common-sense formulations,” Frontiers in Computer Science
2024
Closest in time.
K. Ikeuchi, N. Wake, K. Sasabuchi, and J. Takamatsu, “Semantic constraints to represent common sense required in household actions for multimodal learning-from-observation robot,” The International Journal of Robotics Research
2024
Closest in time.
Accessed: 2024-08-17
Ultralytics, “Yolo.” https://www.ultralytics.com/yolo · 2024
Closest in time.
Accessed: 2024-08-17
Video-Annotator, “pythonvideoannotator.” https://github.com/video-annotator/pythonvideoannotator · 2024
Closest in time.