Fetching the paper…
Reading the bibliography…
Automated planning is traditionally the domain of experts, utilized in fields like manufacturing and healthcare with the aid of expert planning tools.
Thematic analysis
Victoria Clarke and Virginia Braun. 2014 · 1952
Earlier work this paper cites.
Mathematical logic . Vol. 2
Heinz-Dieter Ebbinghaus, Jörg Flum, Wolfgang Thomas, and Ann S Ferebee. 1994 · 1994
Earlier work this paper cites.
Designing web usability: The practice of simplicity
Jakob Nielsen. 1999 · 1999
Earlier work this paper cites.
Measuring Usability with the USE Questionnaire
Arnold Lund. 2001 · 2001
Earlier work this paper cites.
PDDL2. 1: An Extension to PDDL for Expressing Temporal Planning Domains
Maria Fox and Derek Long. 2003 · 2003
Earlier work this paper cites.
A PDDL based tool for automatic web service composition. In International Workshop on Principles and Practice of Semantic Web Reasoning . Springer, 149–163
Joachim Peer. 2004 · 2004
Earlier work this paper cites.
Determinants of customers’ responses to customized offers: Conceptual framework and research propositions
Itamar Simonson. 2005 · 2005
Earlier work this paper cites.
Principles of model checking
Christel Baier and Joost-Pieter Katoen. 2008 · 2008
Earlier work this paper cites.
Concise finite-domain representations for PDDL planning tasks
Malte Helmert. 2009 · 2009
Earlier work this paper cites.
Programming assistance based on contracts and modular verification in the automation domain. In Proceedings of the 2010 ACM Symposium on Applied Computing . 2544–2551
Dominik Hurnaus and Herbert Prähofer. 2010 · 2010
Earlier work this paper cites.
Designing the user interface: strategies for effective human-computer interaction
Ben Shneiderman and Catherine Plaisant. 2010 · 2010
Earlier work this paper cites.
Recent advances and future challenges in automated manufacturing planning
David Bourne, Jonathan Corney, and Satyandra K Gupta. 2011 · 2011
Earlier work this paper cites.
PRISM 4.0: Verification of probabilistic real-time systems. In Computer Aided Verification: 23rd International Conference, CAV 2011, Snowbird, UT, USA, July 14-20, 2011. Proceedings 23 . Springer, 585–591
Marta Kwiatkowska, Gethin Norman, and David Parker. 2011 · 2011
Earlier work this paper cites.
Knowledge engineering tools in planning: State-of-the-art and future challenges
M Shah, Lukás Chrpa, Falilat Jimoh, D Kitchin, T McCluskey, Simon Parkinson, and Mauro Vallati. 2013 · 2013
Earlier work this paper cites.
Statistical power of within and between-subjects designs in economic experiments
Charles Bellemare, Luc Bissonnette, and Sabine Kröger. 2014 · 2014
Earlier work this paper cites.
“The fridge door is open”–Temporal Verification of a Robotic Assistant’s Behaviours. In Advances in Autonomous Robotics Systems: 15th Annual Conference, TAROS 2014, Birmingham, UK, September 1-3, 2014. Proceedings 15 . Springer, 97–108
Clare Dixon, Matt Webster, Joe Saunders, Michael Fisher, and Kerstin Dautenhahn. 2014 · 2014
Earlier work this paper cites.
Putting users in control of their recommendations. In Proceedings of the 9th ACM Conference on Recommender Systems . 3–10
F Maxwell Harper, Funing Xu, Harmanpreet Kaur, Kyle Condiff, Shuo Chang, and Loren Terveen. 2015 · 2015
Earlier work this paper cites.
Moodplay: Interactive mood-based music discovery and recommendation. In Proceedings of the 2016 conference on user modeling adaptation and personalization . 275–279
Ivana Andjelkovic, Denis Parra, and John O’Donovan. 2016 · 2016
Earlier work this paper cites.
Automated Planning and Acting
Malik Ghallab, Dana Nau, and Paolo Traverso. 2016 · 2016
Earlier work this paper cites.
A synthesis of automated planning and reinforcement learning for efficient, robust decision-making
Matteo Leonetti, Luca Iocchi, and Peter Stone. 2016 · 2016
Earlier work this paper cites.
How do different levels of user control affect cognitive load and acceptance of recommendations?. In IntRS@ RecSys . 35–42
Yucheng Jin, Bruno De Lemos Ribeiro Pinto Cardoso, and Katrien Verbert. 2017 · 2017
Earlier work this paper cites.
Artificial intelligence in radiology
Ahmed Hosny, Chintan Parmar, John Quackenbush, Lawrence H Schwartz, and Hugo JWL Aerts. 2018 · 2018
Earlier work this paper cites.
Synthesis for robots: Guarantees and feedback for robot behavior
Hadas Kress-Gazit, Morteza Lahijanian, and Vasumathi Raman. 2018 · 2018
Earlier work this paper cites.
Authoring and verifying human-robot interactions. In Proceedings of the 31st annual acm symposium on user interface software and technology . 75–86
David Porfirio, Allison Sauppé, Aws Albarghouthi, and Bilge Mutlu. 2018 · 2018
Earlier work this paper cites.
Guidelines for Human-AI Interaction. In Proceedings of the 2019 CHI Conference on Human Factors in Computing Systems (Glasgow, Scotland Uk) (CHI ’19) . Association for Computing Machinery, New York, NY, USA, 1–13
Saleema Amershi, Dan Weld, Mihaela Vorvoreanu, Adam Fourney, Besmira Nushi, Penny Collisson, Jina Suh, Shamsi Iqbal, Paul N. Bennett, Kori Inkpen, Jaime Teevan, Ruth Kikin-Gil, and Eric Horvitz. 2019 · 2019
Earlier work this paper cites.
Designing for the better by taking users into account: A qualitative evaluation of user control mechanisms in (news) recommender systems. In Proceedings of the 13th ACM conference on recommender systems . 69–77
Jaron Harambam, Dimitrios Bountouridis, Mykola Makhortykh, and Joris Van Hoboken. 2019 · 2019
Earlier work this paper cites.
Explainable artificial intelligence applications in NLP, biomedical, and malware classification: a literature review. In Intelligent Computing: Proceedings of the 2019 Computing Conference, Volume 2 . Springer, 1269–1292
Sherin Mary Mathews. 2019 · 2019
Earlier work this paper cites.
Reliability and Inter-rater Reliability in Qualitative Research: Norms and Guidelines for CSCW and HCI Practice
Nora McDonald, Sarita Schoenebeck, and Andrea Forte. 2019 · 2019
Earlier work this paper cites.
Retrieval-augmented generation for knowledge-intensive nlp tasks
Patrick Lewis, Ethan Perez, Aleksandra Piktus, Fabio Petroni, Vladimir Karpukhin, Naman Goyal, Heinrich Küttler, Mike Lewis, Wen-tau Yih, Tim Rocktäschel, et al · 2020
Earlier work this paper cites.
Novice-AI music co-creation via AI-steering tools for deep generative models. In Proceedings of the 2020 CHI conference on human factors in computing systems . 1–13
Ryan Louie, Andy Coenen, Cheng Zhi Huang, Michael Terry, and Carrie J Cai. 2020 · 2020
Earlier work this paper cites.
Unintended Consequences of Introducing AI Systems for Decision Making
Anne-Sophie Mayer, Franz Strich, and Marina Fiedler. 2020 · 2020
Earlier work this paper cites.
On faithfulness and factuality in abstractive summarization
Joshua Maynez, Shashi Narayan, Bernd Bohnet, and Ryan McDonald. 2020 · 2020
Earlier work this paper cites.
Transforming robot programs based on social context. In Proceedings of the 2020 CHI conference on human factors in computing systems . 1–12
David Porfirio, Allison Sauppé, Aws Albarghouthi, and Bilge Mutlu. 2020 · 2020
Earlier work this paper cites.
Authr: A task authoring environment for human-robot teams. In Proceedings of the 33rd annual acm symposium on user interface software and technology . 1194–1208
Andrew Schoen, Curt Henrichs, Mathias Strohkirch, and Bilge Mutlu. 2020 · 2020
Earlier work this paper cites.
On the dangers of stochastic parrots: Can language models be too big?. In Proceedings of the 2021 ACM conference on fairness, accountability, and transparency . 610–623
Emily M Bender, Timnit Gebru, Angelina McMillan-Major, and Shmargaret Shmitchell. 2021 · 2021
Earlier work this paper cites.
Uncertainty as a form of transparency: Measuring, communicating, and using uncertainty. In Proceedings of the 2021 AAAI/ACM Conference on AI, Ethics, and Society . 401–413
Umang Bhatt, Javier Antorán, Yunfeng Zhang, Q Vera Liao, Prasanna Sattigeri, Riccardo Fogliato, Gabrielle Melançon, Ranganath Krishnan, Jason Stanley, Omesh Tickoo, et al · 2021
Earlier work this paper cites.
Few-shot self-rationalization with natural language prompts
Ana Marasović, Iz Beltagy, Doug Downey, and Matthew E Peters. 2021 · 2021
Earlier work this paper cites.
GTPyhop: A hierarchical goal+ task planner implemented in Python
Dana Nau, Yash Bansod, Sunandita Patra, Mark Roberts, and Ruoxi Li. [n. d.] · 2021
Earlier work this paper cites.
The effects of explainability and causability on perception, trust, and acceptance: Implications for explainable AI
Donghee Shin. 2021 · 2021
Earlier work this paper cites.
Reframing human-AI collaboration for generating free-text explanations
Sarah Wiegreffe, Jack Hessel, Swabha Swayamdipta, Mark Riedl, and Yejin Choi. 2021 · 2021
Earlier work this paper cites.
A framework for human-computer interactive street network design based on a multi-stage deep learning approach
Zhou Fang, Jiaxin Qi, Lubin Fan, Jianqiang Huang, Ying Jin, and Tianren Yang. 2022 · 2022
Cited alongside, same era.
The probabilistic model checker Storm
Christian Hensel, Sebastian Junges, Joost-Pieter Katoen, Tim Quatmann, and Matthias Volk. 2022 · 2022
Cited alongside, same era.
Explanations from large language models make small reasoners better
Shiyang Li, Jianshu Chen, Yelong Shen, Zhiyu Chen, Xinlu Zhang, Zekun Li, Hong Wang, Jing Qian, Baolin Peng, Yi Mao, et al · 2022
Cited alongside, same era.
Structure synthesis for extended robot state automata. In International Conference on Robotics in Alpe-Adria Danube Region . Springer, 71–79
Lukas Sauer and Dominik Henrich. 2022 · 2022
Cited alongside, same era.
Investigating explainability of generative AI for code through scenario-based design. In Proceedings of the 27th International Conference on Intelligent User Interfaces . 212–228
Planning Domain Simulation: An Interactive System for Plan Visualisation
Emanuele De Pellegrin and Ronald P. A. Petrick. 2024 · 2024
Later among the works it cites.
EvaluLLM: LLM assisted evaluation of generative outputs. In Companion Proceedings of the 29th International Conference on Intelligent User Interfaces . 30–32
Michael Desmond, Zahra Ashktorab, Qian Pan, Casey Dugan, and James M Johnson. 2024 · 2024
Later among the works it cites.
Robust Planning with LLM-Modulo Framework: Case Study in Travel Planning
Atharva Gundawar, Mudit Verma, Lin Guan, Karthik Valmeekam, Siddhant Bhambri, and Subbarao Kambhampati. 2024 · 2024
Later among the works it cites.
Llm-guided formal verification coupled with mutation testing. In 2024 Design, Automation & Test in Europe Conference & Exhibition (DATE) . IEEE, 1–2
Muhammad Hassan, Sallar Ahmadi-Pour, Khushboo Qayyum, Chandan Kumar Jha, and Rolf Drechsler. 2024 · 2024
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Jiao Sun, Q Vera Liao, Michael Muller, Mayank Agarwal, Stephanie Houde, Kartik Talamadupula, and Justin D Weisz. 2022 · 2022
Cited alongside, same era.
Large language models still can’t plan (a benchmark for LLMs on planning and reasoning about change). In NeurIPS 2022 Foundation Models for Decision Making Workshop
Karthik Valmeekam, Alberto Olmo, Sarath Sreedharan, and Subbarao Kambhampati. 2022 · 2022
Cited alongside, same era.
React: Synergizing reasoning and acting in language models
Shunyu Yao, Jeffrey Zhao, Dian Yu, Nan Du, Izhak Shafran, Karthik Narasimhan, and Yuan Cao. 2022 · 2022
Cited alongside, same era.
Josh Achiam, Steven Adler, Sandhini Agarwal, Lama Ahmad, Ilge Akkaya, Florencia Leoni Aleman, Diogo Almeida, Janko Altenschmidt, Sam Altman, Shyamal Anadkat, et al · 2023
Cited alongside, same era.
Chateval: Towards better llm-based evaluators through multi-agent debate
Chi-Min Chan, Weize Chen, Yusheng Su, Jianxuan Yu, Wei Xue, Shanghang Zhang, Jie Fu, and Zhiyuan Liu. 2023 · 2023
Cited alongside, same era.
Teaching large language models to self-debug
Xinyun Chen, Maxwell Lin, Nathanael Schärli, and Denny Zhou. 2023 · 2023
Cited alongside, same era.
Critic: Large language models can self-correct with tool-interactive critiquing
Zhibin Gou, Zhihong Shao, Yeyun Gong, Yelong Shen, Yujiu Yang, Nan Duan, and Weizhu Chen. 2023 · 2023
Cited alongside, same era.
Can large language models explain themselves? a study of llm-generated self-explanations
Shiyuan Huang, Siddarth Mamidanna, Shreedhar Jangam, Yilun Zhou, and Leilani H Gilpin. 2023 · 2023
Cited alongside, same era.
Hui Huang, Yingqi Qu, Jing Liu, Muyun Yang, and Tiejun Zhao. 2024b · 2024
Later among the works it cites.
Understanding the planning of LLM agents: A survey
Xu Huang, Weiwen Liu, Xiaolong Chen, Xingmei Wang, Hao Wang, Defu Lian, Yasheng Wang, Ruiming Tang, and Enhong Chen. 2024a · 2024
Later among the works it cites.
Testing and Understanding Erroneous Planning in LLM Agents through Synthesized User Inputs
Zhenlan Ji, Daoyuan Wu, Pingchuan Ma, Zongjie Li, and Shuai Wang. 2024 · 2024
Later among the works it cites.
LLMs Can’t Plan, But Can Help Planning in LLM-Modulo Frameworks
Subbarao Kambhampati, Karthik Valmeekam, Lin Guan, Kaya Stechly, Mudit Verma, Siddhant Bhambri, Lucas Saldyt, and Anil Murthy. 2024a · 2024
Later among the works it cites.
Why and when llm-based assistants can go wrong: Investigating the effectiveness of prompt-based interactions for software help-seeking. In Proceedings of the 29th International Conference on Intelligent User Interfaces . 288–303
Anjali Khurana, Hariharan Subramonyam, and Parmit K Chilana. 2024 · 2024
Later among the works it cites.
Understanding Large-Language Model (LLM)-powered Human-Robot Interaction
Callie Y Kim, Christine P Lee, and Bilge Mutlu. 2024 · 2024
Later among the works it cites.
Post hoc explanations of language models can improve language models
Satyapriya Krishna, Jiaqi Ma, Dylan Slack, Asma Ghandeharioun, Sameer Singh, and Himabindu Lakkaraju. 2024 · 2024
Later among the works it cites.
Applications, Challenges, and Future Directions of Human-in-the-Loop Learning
Sushant Kumar, Sumit Datta, Vishakha Singh, Deepanwita Datta, Sanjay Kumar Singh, and Ritesh Sharma. 2024 · 2024
Later among the works it cites.
The AI-DEC: A Card-based Design Method for User-centered AI Explanations. In Proceedings of the 2024 ACM Designing Interactive Systems Conference . 1010–1028
Christine P Lee, Min Kyung Lee, and Bilge Mutlu. 2024a · 2024
Later among the works it cites.
Rex: Designing user-centered repair and explanations to address robot failures. In Proceedings of the 2024 ACM designing interactive systems conference . 2911–2925
Christine P Lee, Pragathi Praveena, and Bilge Mutlu. 2024b · 2024
Later among the works it cites.
Improving llm reasoning through scaling inference computation with collaborative verification
Zhenwen Liang, Ye Liu, Tong Niu, Xiangliang Zhang, Yingbo Zhou, and Semih Yavuz. 2024 · 2024
Later among the works it cites.
Speak From Heart: An Emotion-Guided LLM-Based Multimodal Method for Emotional Dialogue Generation. In Proceedings of the 2024 International Conference on Multimedia Retrieval . 533–542
Chenxiao Liu, Zheyong Xie, Sirui Zhao, Jin Zhou, Tong Xu, Minglei Li, and Enhong Chen. 2024c · 2024
Later among the works it cites.
Exploring and evaluating hallucinations in llm-powered code generation
Fang Liu, Yang Liu, Lin Shi, Houkun Huang, Ruifeng Wang, Zhen Yang, Li Zhang, Zhongqi Li, and Yuchi Ma. 2024b · 2024
Later among the works it cites.
Beyond chatbots: Explorellm for structured thoughts and personalized model responses. In Extended Abstracts of the CHI Conference on Human Factors in Computing Systems . 1–12
Xiao Ma, Swaroop Mishra, Ariel Liu, Sophie Ying Su, Jilin Chen, Chinmay Kulkarni, Heng-Tze Cheng, Quoc Le, and Ed Chi. 2024 · 2024
Later among the works it cites.
Large language models: A survey
Shervin Minaee, Tomas Mikolov, Narjes Nikzad, Meysam Chenaghlu, Richard Socher, Xavier Amatriain, and Jianfeng Gao. 2024 · 2024
Later among the works it cites.
Optimization modeling and verification from problem specifications using a multi-agent multi-stage LLM framework
Mahdi Mostajabdaveh, Timothy T Yu, Rindranirina Ramamonjison, Giuseppe Carenini, Zirui Zhou, and Yong Zhang. 2024 · 2024
Later among the works it cites.
User-LLM: Efficient LLM Contextualization with User Embeddings
Lin Ning, Luyang Liu, Jiaxing Wu, Neo Wu, Devora Berlowitz, Sushant Prakash, Bradley Green, Shawn O’Banion, and Jun Xie. 2024 · 2024
Later among the works it cites.
On the prospects of incorporating large language models (llms) in automated planning and scheduling (aps). In Proceedings of the International Conference on Automated Planning and Scheduling , Vol. 34. 432–444
Vishal Pallagani, Bharath Chandra Muppasani, Kaushik Roy, Francesco Fabiano, Andrea Loreggia, Keerthiram Murugesan, Biplav Srivastava, Francesca Rossi, Lior Horesh, and Amit Sheth. 2024 · 2024
Later among the works it cites.
Offsetbias: Leveraging debiased data for tuning evaluators
Junsoo Park, Seungyeon Jwa, Meiying Ren, Daeyoung Kim, and Sanghyuk Choi. 2024 · 2024
Later among the works it cites.
Goal-Oriented End-User Programming of Robots. In Proceedings of the 2024 ACM/IEEE International Conference on Human-Robot Interaction (Boulder, CO, USA) (HRI ’24) . Association for Computing Machinery, New York, NY, USA, 582–591
David Porfirio, Mark Roberts, and Laura M. Hiatt. 2024 · 2024
Later among the works it cites.
Personalized recommendation systems powered by large language models: Integrating semantic understanding and user preferences
Fu Shang, Fanyi Zhao, Mingxuan Zhang, Jun Sun, and Jiatu Shi. 2024 · 2024
Later among the works it cites.
Who validates the validators? aligning llm-assisted evaluation of llm outputs with human preferences. In Proceedings of the 37th Annual ACM Symposium on User Interface Software and Technology . 1–14
Shreya Shankar, JD Zamfirescu-Pereira, Björn Hartmann, Aditya Parameswaran, and Ian Arawjo. 2024 · 2024
Later among the works it cites.
Generalized Planning in PDDL Domains with Pretrained Large Language Models
Tom Silver, Soham Dan, Kavitha Srinivas, Joshua B. Tenenbaum, Leslie Kaelbling, and Michael Katz. 2024 · 2024
Later among the works it cites.
Llm-as-a-judge & reward model: What they can and cannot do
Guijin Son, Hyunwoo Ko, Hoyoung Lee, Yewon Kim, and Seunghyeok Hong. 2024 · 2024
Later among the works it cites.
Bridging the Gulf of Envisioning: Cognitive Challenges in Prompt Based Interactions with LLMs. In Proceedings of the CHI Conference on Human Factors in Computing Systems . 1–19
Hari Subramonyam, Roy Pea, Christopher Pondoc, Maneesh Agrawala, and Colleen Seifert. 2024 · 2024
Later among the works it cites.
The metacognitive demands and opportunities of generative AI. In Proceedings of the CHI Conference on Human Factors in Computing Systems . 1–24
Lev Tankelevitch, Viktor Kewenig, Auste Simkute, Ava Elizabeth Scott, Advait Sarkar, Abigail Sellen, and Sean Rintel. 2024 · 2024
Later among the works it cites.
Lukas Teufelberger, Xintong Liu, Zhipeng Li, Max Moebus, and Christian Holz. 2024 · 2024
Later among the works it cites.
Language models don’t always say what they think: unfaithful explanations in chain-of-thought prompting
Miles Turpin, Julian Michael, Ethan Perez, and Samuel Bowman. 2024 · 2024
Later among the works it cites.
MathChat: Converse to Tackle Challenging Math Problems with LLM Agents. In ICLR 2024 Workshop on Large Language Model (LLM) Agents
Yiran Wu, Feiran Jia, Shaokun Zhang, Hangyu Li, Erkang Zhu, Yue Wang, Yin Tat Lee, Richard Peng, Qingyun Wu, and Chi Wang. 2024 · 2024
Later among the works it cites.
Travelplanner: A benchmark for real-world planning with language agents
Jian Xie, Kai Zhang, Jiangjie Chen, Tinghui Zhu, Renze Lou, Yuandong Tian, Yanghua Xiao, and Yu Su. 2024 · 2024
Later among the works it cites.
Harnessing the power of llms in practice: A survey on chatgpt and beyond
Jingfeng Yang, Hongye Jin, Ruixiang Tang, Xiaotian Han, Qizhang Feng, Haoming Jiang, Shaochen Zhong, Bing Yin, and Xia Hu. 2024a · 2024
Later among the works it cites.
Plug in the safety chip: Enforcing constraints for llm-driven robot agents. In 2024 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 14435–14442
Ziyi Yang, Shreyas S Raman, Ankit Shah, and Stefanie Tellex. 2024b · 2024
Later among the works it cites.
Cfbench: A comprehensive constraints-following benchmark for llms
Tao Zhang, Yanjun Shen, Wenjing Luo, Yan Zhang, Hao Liang, Fan Yang, Mingan Lin, Yujing Qiao, Weipeng Chen, Bin Cui, et al · 2024
Later among the works it cites.
Explainability for large language models: A survey
Haiyan Zhao, Hanjie Chen, Fan Yang, Ninghao Liu, Huiqi Deng, Hengyi Cai, Shuaiqiang Wang, Dawei Yin, and Mengnan Du. 2024 · 2024
Later among the works it cites.
Judging llm-as-a-judge with mt-bench and chatbot arena
Lianmin Zheng, Wei-Lin Chiang, Ying Sheng, Siyuan Zhuang, Zhanghao Wu, Yonghao Zhuang, Zi Lin, Zhuohan Li, Dacheng Li, Eric Xing, et al · 2024
Later among the works it cites.
Mathverse: Does your multi-modal llm truly see the diagrams in visual math problems?. In European Conference on Computer Vision . Springer, 169–186
Renrui Zhang, Dongzhi Jiang, Yichi Zhang, Haokun Lin, Ziyu Guo, Pengshuo Qiu, Aojun Zhou, Pan Lu, Kai-Wei Chang, Yu Qiao, et al · 2025
Closest in time.