Fetching the paper…
Reading the bibliography…
The proliferation of collaborative robots across diverse tasks and embodiments presents a central challenge: achieving lifelong adaptability, scalable coordination, and robust scheduling in multi-agent systems.
R. Fierro, L. Chaimowicz, and V. Kumar, “Multi-robot cooperation,” in Autonomous Mobile Robots . CRC Press, 2018, pp. 417–460
2018
Earlier work this paper cites.
Y. Rizk, M. Awad, and E. W. Tunstel, “Cooperative heterogeneous multi-robot systems: A survey,” ACM Computing Surveys (CSUR) , vol. 52, no. 2, pp. 1–31, 2019
2019
Earlier work this paper cites.
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
X. An, C. Wu et al. , “Multi-robot systems and cooperative object transport: Communications, platforms, and challenges,” IEEE Open Journal of the Computer Society , vol. 4, pp. 23–36, 2023
2023
Earlier work this paper cites.
W. Huang, C. Wang et al. , “Voxposer: Composable 3d value maps for robotic manipulation with language models,” in Conference on Robot Learning . PMLR, 2023, pp. 540–562
2023
Earlier work this paper cites.
Y. Mu, Q. Zhang et al. , “Embodiedgpt: Vision-language pre-training via embodied chain of thought,” Advances in Neural Information Processing Systems , vol. 36, pp. 25 081–25 094, 2023
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
B. Zitkovich, T. Yu et al. , “Rt-2: Vision-language-action models transfer web knowledge to robotic control,” in Conference on Robot Learning . PMLR, 2023, pp. 2165–2183
2023
Earlier work this paper cites.
Z. Li, L. T. Yang, X. Nie, B. Ren, and X. Deng, “Enhancing sentence representation with visually-supervised multimodal pre-training,” in Proceedings of the 31st ACM International Conference on Multimedia , 2023, pp. 5686–5695
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
A. Agrawal, A. S. Bedi, and D. Manocha, “Rtaw: An attention inspired reinforcement learning method for multi-robot task allocation in warehouse environments,” in 2023 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2023, pp. 1393–1399
2023
Earlier work this paper cites.
H. Guo, Z. Liu et al. , “Cross-entropy regularized policy gradient for multirobot nonadversarial moving target search,” IEEE Transactions on Robotics , vol. 39, no. 4, pp. 2569–2584, 2023
2023
Earlier work this paper cites.
D. Patiño, S. Mayya et al. , “Learning to navigate in turbulent flows with aerial robot swarms: A cooperative deep reinforcement learning approach,” IEEE Robotics and Automation Letters , vol. 8, no. 7, pp. 4219–4226, 2023
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
C. Chi, Z. Xu et al. , “Diffusion policy: Visuomotor policy learning via action diffusion,” The International Journal of Robotics Research , p. 02783649241273668, 2023
2023
Earlier work this paper cites.
Z. Mandi et al. , “Roco: Dialectic multi-robot collaboration with large language models,” in 2024 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2024, pp. 286–299
2024
Earlier work this paper cites.
2024
Earlier work this paper cites.
2024
Earlier work this paper cites.
2024
Earlier work this paper cites.
2024
Earlier work this paper cites.
J. Liu, M. Liu, Z. Wang, L. Lee, K. Zhou, P. An, S. Yang, R. Zhang, Y. Guo, and S. Zhang, “Robomamba: Multimodal state space model for efficient robot reasoning and manipulation,” arXiv e-prints , pp. arXiv–2406, 2024
2024
Earlier work this paper cites.
2024
Earlier work this paper cites.
S. H. Vemprala, R. Bonatti et al. , “Chatgpt for robotics: Design principles and model abilities,” Ieee Access , vol. 12, pp. 55 682–55 696, 2024
2024
Earlier work this paper cites.
2024
Earlier work this paper cites.
Q. Gu et al. , “Conceptgraphs: Open-vocabulary 3d scene graphs for perception and planning,” in 2024 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2024, pp. 5021–5028
2024
Earlier work this paper cites.
2024
Earlier work this paper cites.
A. Hurst, A. Lerer et al. , “Gpt-4o system card,” arXiv preprint arXiv:2410.21276 , 2024
2024
Cited alongside, same era.
Z. Chen, J. Wu et al. , “Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks,” in IEEE/CVF conference on computer vision and pattern recognition , 2024, pp. 24 185–24 198
2024
Cited alongside, same era.
2024
Cited alongside, same era.
Z. Li, L. T. Yang, B. Ren, X. Nie, Z. Gao, C. Tan, and S. Z. Li, “Mlip: Enhancing medical visual representation with divergence encoder and knowledge-guided contrastive learning,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2024, pp. 11 704–11 714
2024
Cited alongside, same era.
OpenAI, “Learning to reason with llms,” https://openai.com/index/learning-to-reason-with-llms/ , 2024, accessed: 2025-03-02
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2024
Cited alongside, same era.
M. Liu, M. Wang, H. Ding, Y. Xu, Y. Zhao, and Y. Wei, “Segment anything with precise interaction,” in Proceedings of the 32nd ACM International Conference on Multimedia , 2024, pp. 3790–3799
2024
Cited alongside, same era.
2024
Cited alongside, same era.
2024
Cited alongside, same era.
Y. Zhu, Z. Ou et al. , “Retrieval-augmented embodied agents,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2024, pp. 17 985–17 995
2024
Cited alongside, same era.
T. Greenawalt, “Amazon has more than 750,000 robots that sort, lift, and carry packages—see them in action,” Amazon News, Mar. 2025, last updated: March 03, 2025. [Online]. Available: https://www.aboutamazon.com/news/operations/amazon-robotics-delivering-the-future
2025
Cited alongside, same era.
2025
Cited alongside, same era.
Figure AI, “Helix: A vision-language-action model for generalist humanoid control,” https://www.figure.ai/news/helix , 2025, accessed: 2025-04-18
2025
Cited alongside, same era.
2025
Closest in time.
Y. Ji, H. Tan, J. Shi, X. Hao, Y. Zhang, H. Zhang, P. Wang, M. Zhao, Y. Mu, P. An et al. , “Robobrain: A unified brain model for robotic manipulation from abstract to concrete,” in Proceedings of the Computer Vision and Pattern Recognition Conference , 2025, pp. 1724–1734
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.
Y. Ji, Y. Liu, Z. Zhang, Z. Zhang, Y. Zhao, X. Hao, G. Zhou, X. Zhang, and X. Zheng, “Enhancing adversarial robustness of vision-language models through low-rank adaptation,” in Proceedings of the 2025 International Conference on Multimedia Retrieval , 2025, pp. 550–559
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.
S. Bai, W. Zhou, P. Ding, W. Zhao, D. Wang, and B. Chen, “Rethinking latent redundancy in behavior cloning: An information bottleneck approach for robot manipulation,” in Forty-second International Conference on Machine Learning , 2025
2025
Closest in time.
2025
Closest in time.
E. Zhou, Q. Su et al. , “Code-as-monitor: Constraint-aware visual programming for reactive and proactive robotic failure detection,” in Proceedings of the Computer Vision and Pattern Recognition Conference , 2025, pp. 6919–6929
2025
Closest in time.
W. Huang, C. Wang et al. , “Rekep: Spatio-temporal reasoning of relational keypoint constraints for robotic manipulation,” in Conference on Robot Learning . PMLR, 2025, pp. 4573–4602
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.
A. A. Adil, S. Sakhrieh et al. , “A multi-robot collaborative manipulation framework for dynamic and obstacle-dense environments: integration of deep learning for real-time task execution,” Frontiers in Robotics and AI , vol. 12, p. 1585544, 2025
2025
Closest in time.
Y. Fan, S. Bai, X. Tong, P. Ding, Y. Zhu, H. Lu, F. Dai, W. Zhao, Y. Liu, S. Huang et al. , “Long-vla: Unleashing long-horizon capability of vision language action model for robot manipulation,” in Conference on Robot Learning . PMLR, 2025, pp. 2018–2037
2037
Closest in time.