Fetching the paper…
Reading the bibliography…
The development of conversational AI assistants is an iterative process with multiple components.
Design and analysis of stepped wedge cluster randomized trials
Michael A Hussey and James P Hughes. 2007 · 2007
Earlier work this paper cites.
A survey of the state of explainable AI for natural language processing
Marina Danilevsky, Kun Qian, Ranit Aharonov, Yannis Katsis, Ban Kawas, and Prithviraj Sen. 2020 · 2010
Earlier work this paper cites.
Modeling user preferences in recommender systems: A classification framework for explicit and implicit user feedback
Gawesh Jawaheer, Peter Weller, and Patty Kostkova. 2014 · 2014
Earlier work this paper cites.
Perspectives for evaluating conversational AI
Mahipal Jadeja and Neelanshi Varia. 2017 · 2017
Earlier work this paper cites.
All that’s ‘human’ is not gold: Evaluating human evaluation of generated text
Elizabeth Clark, Tal August, Sofia Serrano, Nikita Haduong, Suchin Gururangan, and Noah A. Smith. 2021 · 2021
Earlier work this paper cites.
Experts, errors, and context: A large-scale study of human evaluation for machine translation
Markus Freitag, George Foster, David Grangier, Viresh Ratnakar, Qijun Tan, and Wolfgang Macherey. 2021 · 2021
Earlier work this paper cites.
The DevOps handbook: How to create world-class agility, reliability, & security in technology organizations
Gene Kim, Jez Humble, Patrick Debois, John Willis, and Nicole Forsgren. 2021 · 2021
Earlier work this paper cites.
Advances in collaborative filtering
Yehuda Koren, Steffen Rendle, and Robert Bell. 2021 · 2021
Earlier work this paper cites.
Human evaluation of automatically generated text: Current trends and best practice guidelines
Chris van der Lee, Albert Gatt, Emiel van Miltenburg, and Emiel Krahmer. 2021 · 2021
Cited alongside, same era.
Denoising implicit feedback for recommendation
Wenjie Wang, Fuli Feng, Xiangnan He, Liqiang Nie, and Tat-Seng Chua. 2021 · 2021
Cited alongside, same era.
Holistic evaluation of language models
Percy Liang, Rishi Bommasani, Tony Lee, Dimitris Tsipras, Dilara Soylu, Michihiro Yasunaga, Yian Zhang, Deepak Narayanan, Yuhuai Wu, Ananya Kumar, et al. 2022 · 2022
Cited alongside, same era.
Mega: Multilingual evaluation of generative ai
Kabir Ahuja, Harshita Diddee, Rishav Hada, Millicent Ochieng, Krithika Ramesh, Prachi Jain, Akshay Nambi, Tanuja Ganu, Sameer Segal, Mohamed Ahmed, et al. 2023 · 2023
Cited alongside, same era.
Generative ai at work
Erik Brynjolfsson, Danielle Li, and Lindsey R Raymond. 2023 · 2023
Cited alongside, same era.
Ai transparency in the age of llms: A human-centered research roadmap
Q Vera Liao and Jennifer Wortman Vaughan. 2023 · 2023
Later among the works it cites.
First tragedy, then parse: History repeats itself in the new era of large language models
Naomi Saphra, Eve Fleisig, Kyunghyun Cho, and Adam Lopez. 2023 · 2023
Later among the works it cites.
Unveiling security, privacy, and ethical concerns of chatgpt
Xiaodong Wu, Ran Duan, and Jianbing Ni. 2023 · 2023
Later among the works it cites.
Judging llm-as-a-judge with mt-bench and chatbot arena
Lianmin Zheng, Wei-Lin Chiang, Ying Sheng, Siyuan Zhuang, Zhanghao Wu, Yonghao Zhuang, Zi Lin, Zhuohan Li, Dacheng Li, Eric Xing, Hao Zhang, Joseph E Gonzalez, and Ion Stoica. 2023 · 2023
Later among the works it cites.
New ai integrations in adobe experience platform
Anjul Bhambhri. 2024 · 2024
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
A survey on evaluation of large language models
Yupeng Chang, Xu Wang, Jindong Wang, Yuan Wu, Linyi Yang, Kaijie Zhu, Hao Chen, Xiaoyuan Yi, Cunxiang Wang, Yidong Wang, et al. 2023 · 2023
Cited alongside, same era.
Bridging the Gap: A Survey on Integrating (Human) Feedback for Natural Language Generation
Patrick Fernandes, Aman Madaan, Emmy Liu, António Farinhas, Pedro Henrique Martins, Amanda Bertsch, José G. C. de Souza, Shuyan Zhou, Tongshuang Wu, Graham Neubig, and André F. T. Martins. 2023 · 2023
Cited alongside, same era.
Bridging the gap in ai-driven workflows: The case for domain-specific generative bots
Akit Kumar, M.S. Lakshmi Devi, and Jeffrey S. Saltz. 2023 · 2023
Cited alongside, same era.
Natural Language Interfaces to Databases
Yunyao Li, Dragomir R. Radev, and Davood Rafiei. 2024 · 2024
Closest in time.
State of what art? a call for multi-prompt llm evaluation
Moran Mizrahi, Guy Kaplan, Dan Malkin, Rotem Dror, Dafna Shahaf, and Gabriel Stanovsky. 2024 · 2024
Closest in time.
A survey on large language model (llm) security and privacy: The good, the bad, and the ugly
Yifan Yao, Jinhao Duan, Kaidi Xu, Yuanfang Cai, Zhibo Sun, and Yue Zhang. 2024 · 2024
Closest in time.