Fetching the paper…
Reading the bibliography…
We present MOSAIC, a modular architecture for coordinating multiple robots to (a) interact with users using natural language and (b) manipulate an open vocabulary of everyday objects.
Answer set programming and plan generation
V. Lifschitz · 2002
Earlier work this paper cites.
Robust real-time face detection
P. Viola and M. J. Jones · 2004
Earlier work this paper cites.
Answer set programming at a glance
G. Brewka, T. Eiter, and M. Truszczyński · 2011
Earlier work this paper cites.
PDDL2.1: an extension to PDDL for expressing temporal planning domains
M. Fox and D. Long · 2011
Earlier work this paper cites.
A human-aware manipulation planner
E. A. Sisbot and R. Alami · 2012
Earlier work this paper cites.
Human3. 6m: Large scale datasets and predictive methods for 3d human sensing in natural environments
C. Ionescu, D. Papava, V. Olaru, and C. Sminchisescu · 2013
Earlier work this paper cites.
URL https://www.prolific.com
Prolific, 2014 · 2014
Earlier work this paper cites.
Predicting human reaching motion in collaborative tasks using inverse optimal control and iterative re-planning
J. Mainprice, R. Hayne, and D. Berenson · 2015
Earlier work this paper cites.
Behavior trees in robotics and AI: an introduction
M. Colledanchise and P. Ögren · 2017
Earlier work this paper cites.
Proximal policy optimization algorithms
J. Schulman, F. Wolski, P. Dhariwal, A. Radford, and O. Klimov · 2017
Earlier work this paper cites.
Human-aware robotic assistant for collaborative assembly: Integrating human motion prediction with planning in time
V. Unhelkar, P. A. Lasota, Q. Tyroller, R.-D. Buhai, L. Marceau, B. Deml, and J. A. Shah · 2018
Earlier work this paper cites.
An empirical comparison of pddl-based and asp-based task planners
Y. Jiang, S. Zhang, P. Khandelwal, and P. Stone · 2018
Earlier work this paper cites.
AMASS: Archive of motion capture as surface shapes
N. Mahmood, N. Ghorbani, N. F. Troje, G. Pons-Moll, and M. J. Black · 2019
Earlier work this paper cites.
Multi-robot grasp planning for sequential assembly operations
M. Dogar, A. Spielberg, S. Baker, and D. Rus · 2019
Earlier work this paper cites.
Robonet: Large-scale multi-robot learning
S. Dasari, F. Ebert, S. Tian, S. Nair, B. Bucher, K. Schmeckpeper, S. Singh, S. Levine, and C. Finn · 2019
Earlier work this paper cites.
A. Raffin, A. Hill, M. Ernestus, A. Gleave, A. Kanervisto, and N. Dormann · 2019
Earlier work this paper cites.
Blazepose: On-device real-time body pose tracking
V. Bazarevsky, I. Grishchenko, K. Raveendran, T. L. Zhu, F. Zhang, and M. Grundmann · 2020
Earlier work this paper cites.
Learning a decentralized multi-arm motion planner
H. Ha, J. Xu, and S. Song · 2020
Earlier work this paper cites.
History repeats itself: Human motion prediction via motion attention
W. Mao, M. Liu, and M. Salzmann · 2020
Earlier work this paper cites.
Learning transferable visual models from natural language supervision
A. Radford, J. W. Kim, C. Hallacy, A. Ramesh, G. Goh, S. Agarwal, G. Sastry, A. Askell, P. Mishkin, J. Clark, et al · 2021
Earlier work this paper cites.
Space-time-separable graph convolutional network for pose forecasting
T. Sofianos, A. Sampieri, L. Franco, and F. Galasso · 2021
Earlier work this paper cites.
Do as i can, not as i say: Grounding language in robotic affordances
M. Ahn, A. Brohan, N. Brown, Y. Chebotar, O. Cortes, B. David, C. Finn, C. Fu, K. Gopalakrishnan, K. Hausman, et al · 2022
Earlier work this paper cites.
Rt-1: Robotics transformer for real-world control at scale
A. Brohan, N. Brown, J. Carbajal, Y. Chebotar, J. Dabis, C. Finn, K. Gopalakrishnan, K. Hausman, A. Herzog, J. Hsu, et al · 2022
Earlier work this paper cites.
Bc-z: Zero-shot task generalization with robotic imitation learning
E. Jang, A. Irpan, M. Khansari, D. Kappler, F. Ebert, C. Lynch, S. Levine, and C. Finn · 2022
Cited alongside, same era.
Cliport: What and where pathways for robotic manipulation
M. Shridhar, L. Manuelli, and D. Fox · 2022
Cited alongside, same era.
R3m: A universal visual representation for robot manipulation
S. Nair, A. Rajeswaran, V. Kumar, C. Finn, and A. Gupta · 2022
Cited alongside, same era.
The design of stretch: A compact, lightweight mobile manipulator for indoor human environments, 2022
C. C. Kemp, A. Edsinger, H. M. Clever, and B. Matulevich · 2022
Cited alongside, same era.
URL https://franka.de/documents
Franka research 3, 2022 · 2022
Cited alongside, same era.
Do as i can, not as i say: Grounding language in robotic affordances, 2022
Diffusion policy: Visuomotor policy learning via action diffusion
C. Chi, S. Feng, Y. Du, Z. Xu, E. Cousineau, B. Burchfiel, and S. Song · 2023
Later among the works it cites.
Open-world object manipulation using pre-trained vision-language models
A. Stone, T. Xiao, Y. Lu, K. Gopalakrishnan, K.-H. Lee, Q. Vuong, P. Wohlhart, B. Zitkovich, F. Xia, C. Finn, et al · 2023
Later among the works it cites.
Rt-2: Vision-language-action models transfer web knowledge to robotic control
A. Brohan, N. Brown, J. Carbajal, Y. Chebotar, X. Chen, K. Choromanski, T. Ding, D. Driess, A. Dubey, C. Finn, et al · 2023
Later among the works it cites.
Video owl-vit: Temporally-consistent open-world localization in video
G. Heigold, M. Minderer, A. Gritsenko, A. Bewley, D. Keysers, M. Lučić, F. Yu, and T. Kipf · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
M. Ahn, A. Brohan, N. Brown, Y. Chebotar, O. Cortes, B. David, C. Finn, C. Fu, K. Gopalakrishnan, K. Hausman, A. Herzog, D. Ho, J. Hsu, J. Ibarz, B. Ichter, A. Irpan, E. Jang, R. J. Ruano, K. Jeffrey, S. Jesmonth, N. J. Joshi, R. Julian, D. Kalashnikov, Y. Kuang, K.-H. Lee, S. Levine, Y. Lu, L. Luu, C. Parada, P. Pastor, J. Quiambao, K. Rao, J. Rettinghouse, D. Reyes, P. Sermanet, N. Sievers, C. Tan, A. Toshev, V. Vanhoucke, F. Xia, T. Xiao, P. Xu, S. Xu, M. Yan, and A. Zeng · 2022
Cited alongside, same era.
Progprompt: Generating situated robot task plans using large language models, 2022
I. Singh, V. Blukis, A. Mousavian, A. Goyal, D. Xu, J. Tremblay, D. Fox, J. Thomason, and A. Garg · 2022
Cited alongside, same era.
Vima: General robot manipulation with multimodal prompts
Y. Jiang, A. Gupta, Z. Zhang, G. Wang, Y. Dou, Y. Chen, L. Fei-Fei, A. Anandkumar, Y. Zhu, and L. Fan · 2022
Cited alongside, same era.
Model predictive control for fluid human-to-robot handovers
W. Yang, B. Sundaralingam, C. Paxton, I. Akinola, Y.-W. Chao, M. Cakmak, and D. Fox · 2022
Cited alongside, same era.
Motion planning combines human motion prediction for human-robot cooperation
H. Ling, G. Liu, L. Zhu, B. Huang, F. Lu, H. Wu, G. Tian, and Z. Ji · 2022
Cited alongside, same era.
Mild: Multimodal interactive latent dynamics for learning human-robot interaction
V. Prasad, D. Koert, R. M. Stock-Homburg, J. Peters, and G. Chalvatzaki · 2022
Cited alongside, same era.
Cape: Corrective actions from precondition errors using large language models
S. S. Raman, V. Cohen, D. Paulius, I. Idrees, E. Rosen, R. Mooney, and S. Tellex · 2022
Cited alongside, same era.
X. Zhao, W. Ding, Y. An, Y. Du, T. Yu, M. Li, M. Tang, and J. Wang · 2023
Later among the works it cites.
Manicast: Collaborative manipulation with cost-aware human forecasting
K. Kedia, P. Dan, A. Bhardwaj, and S. Choudhury · 2023
Later among the works it cites.
Predictive control of cooperative robots sharing common workspace
A. Tika and N. Bajcinca · 2023
Later among the works it cites.
Text2motion: from natural language instructions to feasible plans
K. Lin, C. Agia, T. Migimatsu, M. Pavone, and J. Bohg · 2023
Later among the works it cites.
Demo2code: From summarizing demonstrations to synthesizing code via extended chain-of-thought, 2023
H. Wang, G. Gonzalez-Pumariega, Y. Sharma, and S. Choudhury · 2023
Later among the works it cites.
Tidybot: personalized robot assistance with large language models
J. Wu, R. Antonova, A. Kan, M. Lepert, A. Zeng, S. Song, J. Bohg, S. Rusinkiewicz, and T. Funkhouser · 2023
Later among the works it cites.
Language-driven representation learning for robotics
S. Karamcheti, S. Nair, A. S. Chen, T. Kollar, C. Finn, D. Sadigh, and P. Liang · 2023
Later among the works it cites.
Palm-e: An embodied multimodal language model
D. Driess, F. Xia, M. S. Sajjadi, C. Lynch, A. Chowdhery, B. Ichter, A. Wahid, J. Tompson, Q. Vuong, T. Yu, et al · 2023
Later among the works it cites.
Voxposer: Composable 3d value maps for robotic manipulation with language models
W. Huang, C. Wang, R. Zhang, Y. Li, J. Wu, and L. Fei-Fei · 2023
Later among the works it cites.
Kite: Keypoint-conditioned policies for semantic manipulation
P. Sundaresan, S. Belkhale, D. Sadigh, and J. Bohg · 2023
Later among the works it cites.
Grounding complex natural language commands for temporal tasks in unseen environments, 2023
J. X. Liu, Z. Yang, I. Idrees, S. Liang, B. Schornstein, S. Tellex, and A. Shah · 2023
Later among the works it cites.
Cook2ltl: Translating cooking recipes to ltl formulae using large language models
A. Mavrogiannis, C. Mavrogiannis, and Y. Aloimonos · 2023
Later among the works it cites.
Open x-embodiment: Robotic learning datasets and rt-x models
A. Padalkar, A. Pooley, A. Jain, A. Bewley, A. Herzog, A. Irpan, A. Khazatsky, A. Rai, A. Singh, A. Brohan, et al · 2023
Later among the works it cites.
Scaling open-vocabulary object detection
M. Minderer, A. Gritsenko, and N. Houlsby · 2023
Later among the works it cites.
Ok-robot: What really matters in integrating open-knowledge models for robotics
P. Liu, Y. Orru, C. Paxton, N. M. M. Shafiullah, and L. Pinto · 2024
Closest in time.
Tree of thoughts: Deliberate problem solving with large language models
S. Yao, D. Yu, J. Zhao, I. Shafran, T. Griffiths, Y. Cao, and K. Narasimhan · 2024
Closest in time.
Moka: Open-vocabulary robotic manipulation through mark-based visual prompting
F. Liu, K. Fang, P. Abbeel, and S. Levine · 2024
Closest in time.
Pivot: Iterative visual prompting elicits actionable knowledge for vlms
S. Nasiriany, F. Xia, W. Yu, T. Xiao, J. Liang, I. Dasgupta, A. Xie, D. Driess, A. Wahid, Z. Xu, et al · 2024
Closest in time.