Fetching the paper…
Reading the bibliography…
The rapid advancement of generative AI and multi-modal foundation models has shown significant potential in advancing robotic manipulation.
Springer Handbook of Robotics
B Siciliano. 2008 · 2008
Earlier work this paper cites.
The YCB object and Model set: Towards common benchmarks for manipulation research. In 2015 international conference on advanced robotics (ICAR) . IEEE, 510–517
Berk Calli, Arjun Singh, Aaron Walsman, Siddhartha Srinivasa, Pieter Abbeel, and Aaron M Dollar. 2015 · 2015
Earlier work this paper cites.
ImageNet Large Scale Visual Recognition Challenge
Olga Russakovsky, Jia Deng, Hao Su, et al · 2015
Earlier work this paper cites.
A Mathematical Introduction to Robotic Manipulation
Richard M Murray, Zexiang Li, and S Shankar Sastry. 2017 · 2017
Earlier work this paper cites.
FiLM: Visual Reasoning with a General Conditioning Layer. In Proceedings of the AAAI conference on artificial intelligence , Vol. 32
Ethan Perez, Florian Strub, Harm De Vries, Vincent Dumoulin, and Aaron Courville. 2018 · 2018
Earlier work this paper cites.
EfficientNet: Rethinking Model Scaling for Convolutional Neural Networks. In Proceedings of the 36th International Conference on Machine Learning (Proceedings of Machine Learning Research, Vol. 97) , Kamalika Chaudhuri and Ruslan Salakhutdinov (Eds.). PMLR, 6105–6114
Mingxing Tan and Quoc Le. 2019 · 2019
Earlier work this paper cites.
Active Fuzzing for Testing and Securing Cyber-Physical Systems. In Proceedings of the 29th ACM SIGSOFT International Symposium on Software Testing and Analysis . 14–26
Yuqi Chen, Bohan Xuan, Christopher M Poskitt, Jun Sun, and Fan Zhang. 2020 · 2020
Earlier work this paper cites.
The Rise of Automation and Robotics in Warehouse Management
Amandeep Dhaliwal. 2020 · 2020
Earlier work this paper cites.
An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale. In International Conference on Learning Representations
Alexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn, Xiaohua Zhai, Thomas Unterthiner, Mostafa Dehghani, Matthias Minderer, Georg Heigold, Sylvain Gelly, et al · 2020
Earlier work this paper cites.
DeepGini: prioritizing massive tests to enhance the robustness of deep neural networks. In Proceedings of the 29th ACM SIGSOFT International Symposium on Software Testing and Analysis . 177–188
Yang Feng, Qingkai Shi, Xinyu Gao, Jun Wan, Chunrong Fang, and Zhenyu Chen. 2020 · 2020
Earlier work this paper cites.
Robotics and Industry 4.0
Ruchi Goel and Pooja Gupta. 2020 · 2020
Earlier work this paper cites.
CPS-Based Self-Adaptive Collaborative Control for Smart Production-Logistics Systems
Zhengang Guo, Yingfeng Zhang, Xibin Zhao, and Xiaoyu Song. 2020 · 2020
Earlier work this paper cites.
Structure-Invariant Testing for Machine Translation. In Proceedings of the ACM/IEEE 42nd International Conference on Software Engineering . 961–973
Pinjia He, Clara Meister, and Zhendong Su. 2020 · 2020
Earlier work this paper cites.
Approximation-Refinement Testing of Compute-Intensive Cyber-Physical Models: An Approach Based on System Identification. In Proceedings of the ACM/IEEE 42nd International Conference on Software Engineering . 372–384
Claudio Menghi, Shiva Nejati, Lionel Briand, and Yago Isasi Parache. 2020 · 2020
Earlier work this paper cites.
Uncertainty-aware Self-training for Few-shot Text Classification. In Advances in Neural Information Processing Systems , Vol. 33. Curran Associates, Inc., 21199–21212
Subhabrata Mukherjee and Ahmed Awadallah. 2020 · 2020
Earlier work this paper cites.
Automatic Testing and Improvement of Machine Translation. In Proceedings of the ACM/IEEE 42nd International Conference on Software Engineering . 974–985
Zeyu Sun, Jie M Zhang, Mark Harman, Mike Papadakis, and Lu Zhang. 2020 · 2020
Earlier work this paper cites.
Metamorphic Object Insertion for Testing Object Detection Systems. In Proceedings of the 35th IEEE/ACM International Conference on Automated Software Engineering . 1053–1065
Shuai Wang and Zhendong Su. 2020 · 2020
Earlier work this paper cites.
Cyber-Physical Power System (CPPS): A Review on Modeling, Simulation, and Analysis With Cyber Security Applications
Rajaa Vikhram Yohanandhan, Rajvikram Madurai Elavarasan, et al · 2020
Earlier work this paper cites.
Sim-to-Real Transfer in Deep Reinforcement Learning for Robotics: a Survey. In 2020 IEEE Symposium Series on Computational Intelligence (SSCI) . IEEE, 737–744
Wenshuai Zhao, Jorge Peña Queralta, and Tomi Westerlund. 2020 · 2020
Earlier work this paper cites.
Generating metamorphic relations for cyber-physical systems with genetic programming: an industrial case study. In Proceedings of the 29th ACM Joint Meeting on European Software Engineering Conference and Symposium on the Foundations of Software Engineering . 1264–1274
Jon Ayerdi, Valerio Terragni, et al · 2021
Earlier work this paper cites.
Advanced applications of industrial robotics: New trends and possibilities
Andrius Dzedzickis, Jurga Subačiūtė-Žemaitienė, Ernestas Šutinys, Urtė Samukaitė-Bubnienė, and Vytautas Bučinskas. 2021 · 2021
Earlier work this paper cites.
Sim2Real in Robotics and Automation: Applications and Challenges
Sebastian Höfer, Kostas Bekris, Ankur Handa, et al · 2021
Earlier work this paper cites.
Service Robots in the Healthcare Sector
Jane Holland, Liz Kingston, Conor McCarthy, Eddie Armstrong, Peter O’Dwyer, Fionn Merz, and Mark McConnell. 2021 · 2021
Earlier work this paper cites.
Coverage-based Scene Fuzzing for Virtual Autonomous Driving Testing
Zhisheng Hu, Shengjian Guo, Zhenyu Zhong, and Kang Li. 2021 · 2021
Earlier work this paper cites.
A Survey of Robots in Healthcare
Maria Kyrarini, Fotios Lygerakis, Akilesh Rajavenkatanarayanan, Christos Sevastopoulos, Harish Ram Nambiappan, Kodur Krishna Chaitanya, Ashwin Ramesh Babu, Joanne Mathew, and Fillia Makedon. 2021 · 2021
Earlier work this paper cites.
Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback
Yuntao Bai, Andy Jones, Kamal Ndousse, et al · 2022
Earlier work this paper cites.
RT-1: Robotics Transformer for Real-World Control at Scale
Anthony Brohan, Noah Brown, Justice Carbajal, et al · 2022
Earlier work this paper cites.
FaithDial: A Faithful Benchmark for Information-Seeking Dialogue
Nouha Dziri, Ehsan Kamalloo, Sivan Milton, Osmar Zaiane, Mo Yu, Edoardo M Ponti, and Siva Reddy. 2022 · 2022
Earlier work this paper cites.
Adaptive test selection for deep neural networks. In Proceedings of the 44th International Conference on Software Engineering . 73–85
Xinyu Gao, Yang Feng, Yining Yin, Zixi Liu, Zhenyu Chen, and Baowen Xu. 2022 · 2022
Earlier work this paper cites.
TruthfulQA: Measuring How Models Mimic Human Falsehoods. In Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics . 3214–3252
Stephanie Lin, Jacob Hilton, and Owain Evans. 2022 · 2022
Earlier work this paper cites.
A Conceptual Model for the Adoption of Autonomous Robots in the Supply Chain and Logistics Industry
Mohamed Shamout, Rabeb Ben-Abdallah, et al · 2022
Earlier work this paper cites.
Confidence-driven weighted retraining for predicting safety-critical failures in autonomous driving systems
Andrea Stocco and Paolo Tonella. 2022 · 2022
Earlier work this paper cites.
Chain-of-Thought Prompting Elicits Reasoning in Large Language Models. In Advances in Neural Information Processing Systems , Vol. 35. Curran Associates, Inc., 24824–24837
Jason Wei, Xuezhi Wang, Dale Schuurmans, et al · 2022
Earlier work this paper cites.
Automated testing of image captioning systems. In Proceedings of the 31st ACM SIGSOFT International Symposium on Software Testing and Analysis . 467–479
Boxi Yu, Zhiqing Zhong, Xinran Qin, Jiayi Yao, Yuancheng Wang, and Pinjia He. 2022 · 2022
Cited alongside, same era.
FalsifAI: Falsification of AI-Enabled Hybrid Control Systems Guided by Time-Aware Coverage Criteria
Zhenya Zhang, Deyun Lyu, Paolo Arcaini, Lei Ma, Ichiro Hasuo, and Jianjun Zhao. 2022 · 2022
Cited alongside, same era.
Data Augmentation in Classification and Segmentation: A Survey and New Strategies
Khaled Alomar, Halil Ibrahim Aysel, and Xiaohao Cai. 2023 · 2023
Cited alongside, same era.
RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Anthony Brohan, Noah Brown, Justice Carbajal, et al · 2023
Cited alongside, same era.
MultiPL-E: A Scalable and Extensible Approach to Benchmarking Neural Code Generation
Federico Cassano, John Gouwar, Daniel Nguyen, et al · 2023
Tree of Thoughts: Deliberate Problem Solving with Large Language Models. In Advances in Neural Information Processing Systems , Vol. 36. Curran Associates, Inc., 11809–11822
Shunyu Yao, Dian Yu, Jeffrey Zhao, et al · 2023
Later among the works it cites.
Sigmoid Loss for Language Image Pre-Training. In Proceedings of the IEEE/CVF International Conference on Computer Vision . 11975–11986
Xiaohua Zhai, Basil Mustafa, Alexander Kolesnikov, and Lucas Beyer. 2023 · 2023
Later among the works it cites.
Self-Edit: Fault-Aware Code Editor for Code Generation. In Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers . Association for Computational Linguistics, 769–787
Kechi Zhang, Zhuo Li, Jia Li, Ge Li, and Zhi Jin. 2023 · 2023
Later among the works it cites.
Specification-Based Autonomous Driving System Testing
Yuan Zhou, Yang Sun, Yun Tang, et al · 2023
Later among the works it cites.
DeepGD: A Multi-Objective Black-Box Test Selection Approach for Deep Neural Networks
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
BehAVExplor: Behavior Diversity Guided Testing for Autonomous Driving Systems. In Proceedings of the 32nd ACM SIGSOFT International Symposium on Software Testing and Analysis . 488–500
Mingfei Cheng, Yuan Zhou, and Xiaofei Xie. 2023 · 2023
Cited alongside, same era.
PaLM-E: An Embodied Multimodal Language Model. In Proceedings of the 40th International Conference on Machine Learning (Proceedings of Machine Learning Research, Vol. 202) . PMLR, 8469–8488
Danny Driess, Fei Xia, Mehdi S. M. Sajjadi, et al · 2023
Cited alongside, same era.
Vision-Language Models as Success Detectors. In Proceedings of The 2nd Conference on Lifelong Learning Agents (Proceedings of Machine Learning Research, Vol. 232) . PMLR, 120–136
Yuqing Du, Ksenia Konyushkova, Misha Denil, Akhil Raju, Jessica Landon, Felix Hill, Nando de Freitas, and Serkan Cabi. 2023 · 2023
Cited alongside, same era.
ManiSkill2: A Unified Benchmark for Generalizable Manipulation Skills
Jiayuan Gu, Fanbo Xiang, Xuanlin Li, et al · 2023
Cited alongside, same era.
Recent Trends in Task and Motion Planning for Robotics: A Survey
Huihui Guo, Fan Wu, Yunchuan Qin, Ruihui Li, Keqin Li, and Kenli Li. 2023 · 2023
Cited alongside, same era.
GenAIPABench: A Benchmark for Generative AI-based Privacy Assistants
Aamir Hamid, Hemanth Reddy Samidi, Tim Finin, Primal Pappachan, and Roberto Yus. 2023 · 2023
Cited alongside, same era.
A Survey on Deep Reinforcement Learning Algorithms for Robotic Manipulation
Dong Han, Beni Mulyana, Vladimir Stankovic, and Samuel Cheng. 2023 · 2023
Cited alongside, same era.
Zohreh Aghababaeyan, Manel Abdellatif, Mahboubeh Dadkhah, and Lionel Briand. 2024 · 2024
Closest in time.
RoboAgent: Generalization and Efficiency in Robot Manipulation via Semantic Augmentations and Action Chunking. In 2024 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 4788–4795
Homanga Bharadhwaj, Jay Vakil, Mohit Sharma, Abhinav Gupta, Shubham Tulsiani, and Vikash Kumar. 2024 · 2024
Closest in time.
Testing of Deep Reinforcement Learning Agents with Surrogate Models
Matteo Biagiola and Paolo Tonella. 2024 · 2024
Closest in time.
Teaching Large Language Models to Self-Debug. In The Twelfth International Conference on Learning Representations
Xinyun Chen, Maxwell Lin, Nathanael Schärli, and Denny Zhou. 2024 · 2024
Closest in time.
Test Input Prioritization for Machine Learning Classifiers
Xueqi Dang, Yinghua Li, Mike Papadakis, Jacques Klein, Tegawendé F Bissyandé, and Yves Le Traon. 2024 · 2024
Closest in time.
MultiTest: Physical-Aware Object Insertion for Testing Multi-sensor Fusion Perception Systems. In Proceedings of the IEEE/ACM 46th International Conference on Software Engineering . 1–13
Xinyu Gao, Zhijie Wang, Yang Feng, Lei Ma, Zhenyu Chen, and Baowen Xu. 2024 · 2024
Closest in time.
Active Testing of Large Language Model via Multi-Stage Sampling
Yuheng Huang, Jiayang Song, Qiang Hu, Felix Juefei-Xu, and Lei Ma. 2024 · 2024
Closest in time.
SWE-bench: Can Language Models Resolve Real-World GitHub Issues?. In The Twelfth International Conference on Learning Representations
Carlos E Jimenez, John Yang, Alexander Wettig, Shunyu Yao, Kexin Pei, Ofir Press, and Karthik Narasimhan. 2024 · 2024
Closest in time.
OpenVLA: An Open-Source Vision-Language-Action Model
Moo Jin Kim, Karl Pertsch, Siddharth Karamcheti, et al · 2024
Closest in time.
Automated Testing Linguistic Capabilities of NLP Models
Jaeseong Lee, Simin Chen, Austin Mordahl, Cong Liu, Wei Yang, and Shiyi Wei. 2024 · 2024
Closest in time.
Prioritizing test cases for deep learning-based video classifiers
Yinghua Li, Xueqi Dang, Lei Ma, Jacques Klein, and Tegawendé F Bissyandé. 2024a · 2024
Closest in time.
Visual Instruction Tuning
Haotian Liu, Chunyuan Li, Qingyang Wu, and Yong Jae Lee. 2024a · 2024
Closest in time.
Is Your Code Generated by ChatGPT Really Correct? Rigorous Evaluation of Large Language Models for Code Generation
Jiawei Liu, Chunqiu Steven Xia, Yuyao Wang, and Lingming Zhang. 2024b · 2024
Closest in time.
Lost in Translation: A Study of Bugs Introduced by Large Language Models while Translating Code. In Proceedings of the IEEE/ACM 46th International Conference on Software Engineering
Rangeet Pan, Ali Reza Ibrahimzada, et al · 2024
Closest in time.
TrustLLM: Trustworthiness in Large Language Models. In Proceedings of the 41st International Conference on Machine Learning (Proceedings of Machine Learning Research, Vol. 235) . PMLR, 20166–20270
Lichao Sun, Yue Huang, et al · 2024
Closest in time.
Octo: An Open-Source Generalist Robot Policy
Octo Model Team, Dibya Ghosh, Homer Walke, et al · 2024
Closest in time.
MORTAR: A Model-based Runtime Action Repair Framework for AI-enabled Cyber-Physical Systems
Renzhi Wang, Zhehua Zhou, Jiayang Song, Xuan Xie, Xiaofei Xie, and Lei Ma. 2024 · 2024
Closest in time.
Agentless: Demystifying LLM-based Software Engineering Agents
Chunqiu Steven Xia, Yinlin Deng, Soren Dunn, and Lingming Zhang. 2024 · 2024
Closest in time.
Online Safety Analysis for LLMs: a Benchmark, an Assessment, and a Path Forward
Xuan Xie, Jiayang Song, Zhehua Zhou, Yuheng Huang, Da Song, and Lei Ma. 2024 · 2024
Closest in time.
SWE-agent: Agent-Computer Interfaces Enable Automated Software Engineering. In Advances in Neural Information Processing Systems , Vol. 37. Curran Associates, Inc., 50528–50652
John Yang, Carlos E. Jimenez, et al · 2024
Closest in time.
CoderEval: A Benchmark of Pragmatic Code Generation with Generative Pre-trained Models. In Proceedings of the 46th IEEE/ACM International Conference on Software Engineering . 1–12
Hao Yu, Bo Shen, Dezhi Ran, Jiaxin Zhang, Qi Zhang, Yuchi Ma, Guangtai Liang, Ying Li, Qianxiang Wang, and Tao Xie. 2024 · 2024
Closest in time.
AutoCodeRover: Autonomous Program Improvement. In Proceedings of the 33rd ACM SIGSOFT International Symposium on Software Testing and Analysis . 1592–1604
Yuntong Zhang, Haifeng Ruan, Zhiyu Fan, and Abhik Roychoudhury. 2024 · 2024
Closest in time.
PromptRobust: Towards Evaluating the Robustness of Large Language Models on Adversarial Prompts. In Proceedings of the 1st ACM Workshop on Large AI Systems and Models with Privacy and Safety Analysis (Salt Lake City, UT, USA) (LAMPS ’24) . 57–68
Kaijie Zhu, Jindong Wang, Jiaheng Zhou, et al · 2024
Closest in time.
Look Before You Leap: An Exploratory Study of Uncertainty Analysis for Large Language Models
Yuheng Huang, Jiayang Song, Zhijie Wang, et al · 2025
Closest in time.
LLaRA: Supercharging Robot Learning Data for Vision-Language Policy. In The Thirteenth International Conference on Learning Representations
Xiang Li, Cristina Mata, Jongwoo Park, et al · 2025
Closest in time.
Towards Understanding the Characteristics of Code Generation Errors Made by Large Language Models. In Proceedings of the IEEE/ACM 47th International Conference on software Engineering (ICSE ’25)
Zhijie Wang, Zijie Zhou, Da Song, Yuheng Huang, Shengmai Chen, Lei Ma, and Tianyi Zhang. 2025 · 2025
Closest in time.
ISR-LLM: Iterative Self-Refined Large Language Model for Long-Horizon Sequential Task Planning. In 2024 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2081–2088
Zhehua Zhou, Jiayang Song, Kunpeng Yao, Zhan Shu, and Lei Ma. 2024 · 2088
Closest in time.
Task and Motion Planning with Large Language Models for Object Rearrangement. In 2023 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) . IEEE, 2086–2092
Yan Ding, Xiaohan Zhang, Chris Paxton, and Shiqi Zhang. 2023 · 2092
Closest in time.