Fetching the paper…
Reading the bibliography…
Large language models (LLMs) with chat-based capabilities, such as ChatGPT, are widely used in various workflows.
Basics of qualitative research techniques
Anselm Strauss and Juliet Corbin. 1998 · 1998
Earlier work this paper cites.
Analysis of multiple query reformulations on the web: The interactive information retrieval context
Soo Young Rieh and Hong (Iris) Xie. 2006 · 2005
Earlier work this paper cites.
Comments on context and conversation
Teun A Van Dijk. 2007 · 2007
Earlier work this paper cites.
Language, meaning, and social cognition
Thomas M Holtgraves and Yoshihisa Kashima. 2008 · 2008
Earlier work this paper cites.
Understanding User’s Query Intent with Wikipedia. In Proceedings of the 18th International Conference on World Wide Web (Madrid, Spain) (WWW ’09) . Association for Computing Machinery, New York, NY, USA, 471–480
Jian Hu, Gang Wang, Fred Lochovsky, Jian-tao Sun, and Zheng Chen. 2009 · 2009
Earlier work this paper cites.
Context-Sensitive Query Auto-Completion. In Proceedings of the 20th International Conference on World Wide Web (Hyderabad, India) (WWW ’11) . Association for Computing Machinery, New York, NY, USA, 107–116
Ziv Bar-Yossef and Naama Kraus. 2011 · 2011
Earlier work this paper cites.
People as contexts in conversation
Sarah Brown-Schmidt, Si On Yoon, and Rachel Anna Ryskin. 2015 · 2015
Earlier work this paper cites.
Like Having a Really Bad PA" The Gulf between User Expectation and Experience of Conversational Agents. In Proceedings of the 2016 CHI conference on human factors in computing systems . 5286–5297
Ewa Luger and Abigail Sellen. 2016 · 2016
Earlier work this paper cites.
Deep reinforcement learning from human preferences
Paul F Christiano, Jan Leike, Tom Brown, Miljan Martic, Shane Legg, and Dario Amodei. 2017 · 2017
Earlier work this paper cites.
Patterns for how users overcome obstacles in voice user interfaces. In Proceedings of the 2018 CHI conference on human factors in computing systems . 1–7
Chelsea Myers, Anushay Furqan, Jessica Nebolsky, Karina Caro, and Jichen Zhu. 2018 · 2018
Earlier work this paper cites.
Voice interfaces in everyday life. In proceedings of the 2018 CHI conference on human factors in computing systems . 1–12
Martin Porcheron, Joel E Fischer, Stuart Reeves, and Sarah Sharples. 2018 · 2018
Earlier work this paper cites.
Help! Is my chatbot falling into the uncanny valley? An empirical study of user experience in human–chatbot interaction
Marita Skjuve, Ida Maria Haugstveit, Asbjørn Følstad, and Petter Brandtzaeg. 2019 · 2019
Earlier work this paper cites.
Fine-tuning language models from human preferences
Daniel M Ziegler, Nisan Stiennon, Jeffrey Wu, Tom B Brown, Alec Radford, Dario Amodei, Paul Christiano, and Geoffrey Irving. 2019 · 2019
Earlier work this paper cites.
Realtoxicityprompts: Evaluating neural toxic degeneration in language models
Samuel Gehman, Suchin Gururangan, Maarten Sap, Yejin Choi, and Noah A Smith. 2020 · 2020
Earlier work this paper cites.
Persistent anti-muslim bias in large language models. In Proceedings of the 2021 AAAI/ACM Conference on AI, Ethics, and Society . 298–306
Abubakar Abid, Maheen Farooqi, and James Zou. 2021 · 2021
Earlier work this paper cites.
On the dangers of stochastic parrots: Can language models be too big?. In Proceedings of the 2021 ACM conference on fairness, accountability, and transparency . 610–623
Emily M Bender, Timnit Gebru, Angelina McMillan-Major, and Shmargaret Shmitchell. 2021 · 2021
Earlier work this paper cites.
Measuring and improving consistency in pretrained language models
Yanai Elazar, Nora Kassner, Shauli Ravfogel, Abhilasha Ravichander, Eduard Hovy, Hinrich Sch"̈utze, and Yoav Goldberg. 2021 · 2021
Earlier work this paper cites.
Gender and representation bias in GPT-3 generated stories. In Proceedings of the Third Workshop on Narrative Understanding . 48–55
Li Lucy and David Bamman. 2021 · 2021
Earlier work this paper cites.
Prompt programming for large language models: Beyond the few-shot paradigm. In Extended Abstracts of the 2021 CHI Conference on Human Factors in Computing Systems . 1–7
Laria Reynolds and Kyle McDonell. 2021 · 2021
Earlier work this paper cites.
Ethical and social risks of harm from language models
Laura Weidinger, John Mellor, Maribeth Rauh, Conor Griffin, Jonathan Uesato, Po-Sen Huang, Myra Cheng, Mia Glaese, Borja Balle, Atoosa Kasirzadeh, et al · 2021
Earlier work this paper cites.
ChatGPT is a new AI chatbot that can answer questions and write essays
Accessed on 10/08/2023a · 2022
Earlier work this paper cites.
A review on language models as knowledge bases
Badr AlKhamissi, Millicent Li, Asli Celikyilmaz, Mona Diab, and Marjan Ghazvininejad. 2022 · 2022
Earlier work this paper cites.
ChatGPT Usage and Limitations. (Dec. 2022)
Amos Azaria. 2022 · 2022
Earlier work this paper cites.
What does it mean for a language model to preserve privacy?. In Proceedings of the 2022 ACM Conference on Fairness, Accountability, and Transparency . 2280–2292
Hannah Brown, Katherine Lee, Fatemehsadat Mireshghallah, Reza Shokri, and Florian Tramèr. 2022 · 2022
Earlier work this paper cites.
Hai Dang, Lukas Mecke, Florian Lehmann, Sven Goller, and Daniel Buschek. 2022 · 2022
Earlier work this paper cites.
A Survey of Natural Language Generation
Chenhe Dong, Yinghui Li, Haifan Gong, Miaoxin Chen, Junxin Li, Ying Shen, and Min Yang. 2022 · 2022
Earlier work this paper cites.
Towards reasoning in large language models: A survey
Jie Huang and Kevin Chen-Chuan Chang. 2022 · 2022
Earlier work this paper cites.
BECEL: Benchmark for Consistency Evaluation of Language Models. In International Conference on Computational Linguistics
Myeongjun Jang, Deuk Sin Kwon, and Thomas Lukasiewicz. 2022 · 2022
Earlier work this paper cites.
Large language models are zero-shot reasoners
Takeshi Kojima, Shixiang Shane Gu, Machel Reid, Yutaka Matsuo, and Yusuke Iwasawa. 2022 · 2022
Earlier work this paper cites.
Language models of code are few-shot commonsense learners
Aman Madaan, Shuyan Zhou, Uri Alon, Yiming Yang, and Graham Neubig. 2022 · 2022
Earlier work this paper cites.
Law informs code: A legal informatics approach to aligning artificial intelligence with humans
John J Nay. 2022 · 2022
Earlier work this paper cites.
Training language models to follow instructions with human feedback
Long Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida, Carroll Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, et al · 2022
Earlier work this paper cites.
Discovering Language Model Behaviors with Model-Written Evaluations
Ethan Perez, Sam Ringer, Kamilė Lukošiūtė, Karina Nguyen, Edwin Chen, Scott Heiner, Craig Pettit, Catherine Olsson, Sandipan Kundu, Saurav Kadavath, Andy Jones, Anna Chen, Ben Mann, Brian Israel, Bryan Seethor, Cameron McKinnon, Christopher Olah, Da Yan, Daniela Amodei, Dario Amodei, Dawn Drain, Dustin Li, Eli Tran-Johnson, Guro Khundadze, Jackson Kernion, James Landis, Jamie Kerr, Jared Mueller, Jeeyoon Hyun, Joshua Landau, Kamal Ndousse, Landon Goldberg, Liane Lovitt, Martin Lucas, Michael Sellitto, Miranda Zhang, Neerav Kingsland, Nelson Elhage, Nicholas Joseph, Noemí Mercado, Nova DasSarma, Oliver Rausch, Robin Larson, Sam McCandlish, Scott Johnston, Shauna Kravec, Sheer El Showk, Tamera Lanham, Timothy Telleen-Lawton, Tom Brown, Tom Henighan, Tristan Hume, Yuntao Bai, Zac Hatfield-Dodds, Jack Clark, Samuel R. Bowman, Amanda Askell, Roger Grosse, Danny Hernandez, Deep Ganguli, Evan Hubinger, Nicholas Schiefer, and Jared Kaplan. 2022 · 2022
Earlier work this paper cites.
Limitations of language models in arithmetic and symbolic induction
Jing Qian, Hong Wang, Zekun Li, Shiyang Li, and Xifeng Yan. 2022 · 2022
Cited alongside, same era.
Self-consistency improves chain of thought reasoning in language models
Xuezhi Wang, Jason Wei, Dale Schuurmans, Quoc Le, Ed Chi, Sharan Narang, Aakanksha Chowdhery, and Denny Zhou. 2022 · 2022
Cited alongside, same era.
Chain-of-thought prompting elicits reasoning in large language models
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Fei Xia, Ed Chi, Quoc V Le, Denny Zhou, et al · 2022
Cited alongside, same era.
Taxonomy of risks posed by language models. In Proceedings of the 2022 ACM Conference on Fairness, Accountability, and Transparency . 214–229
Laura Weidinger, Jonathan Uesato, Maribeth Rauh, Conor Griffin, Po-Sen Huang, John Mellor, Amelia Glaese, Myra Cheng, Borja Balle, Atoosa Kasirzadeh, et al · 2022
Cited alongside, same era.
ChatGPT-Reshaping medical education and clinical management
Rehan Ahmed Khan, Masood Jawaid, Aymen Rehan Khan, and Madiha Sajjad. 2023 · 2023
Closest in time.
ChatGPT is shaping the future of medical writing but still requires human judgment
Felipe C Kitamura. 2023 · 2023
Closest in time.
Analysis of ChatGPT tool to assess the potential of its utility for academic writing in biomedical domain
Arun HS Kumar. 2023 · 2023
Closest in time.
Revolutionizing radiology with GPT-based models: Current applications, future possibilities and limitations of ChatGPT
Augustin Lecler, Loïc Duron, and Philippe Soyer. 2023 · 2023
Closest in time.
Evaluating the logical reasoning ability of chatgpt and gpt-4
Hanmeng Liu, Ruoxi Ning, Zhiyang Teng, Jian Liu, Qiji Zhou, and Yue Zhang. 2023b · 2023
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
American== white in multimodal language-and-image ai. In Proceedings of the 2022 AAAI/ACM Conference on AI, Ethics, and Society . 800–812
Robert Wolfe and Aylin Caliskan. 2022 · 2022
Cited alongside, same era.
Automatic chain of thought prompting in large language models
Zhuosheng Zhang, Aston Zhang, Mu Li, and Alex Smola. 2022 · 2022
Cited alongside, same era.
Large language models are human-level prompt engineers
Yongchao Zhou, Andrei Ioan Muresanu, Ziwen Han, Keiran Paster, Silviu Pitis, Harris Chan, and Jimmy Ba. 2022 · 2022
Cited alongside, same era.
LLM Jailbreak Study
Accessed on 10/06/2023 · 2023
Cited alongside, same era.
gpt-4-system-card.pdf
Accessed on 10/08/2023 · 2023
Cited alongside, same era.
Evaluating the performance of chatgpt in ophthalmology: An analysis of its successes and shortcomings
Fares Antaki, Samir Touma, Daniel Milad, Jonathan El-Khoury, and Renaud Duval. 2023 · 2023
Cited alongside, same era.
Yejin Bang, Samuel Cahyawijaya, Nayeon Lee, Wenliang Dai, Dan Su, Bryan Wilie, Holy Lovenia, Ziwei Ji, Tiezheng Yu, Willy Chung, Quyet V. Do, Yan Xu, and Pascale Fung. 2023 · 2023
Cited alongside, same era.
Morteza Behrooz, William Ngan, Joshua Lane, Giuliano Morse, Benjamin Babcock, Kurt Shuster, Mojtaba Komeili, Moya Chen, Melanie Kambadur, Y-Lan Boureau, et al · 2023
Cited alongside, same era.
Aman Madaan, Niket Tandon, Prakhar Gupta, Skyler Hallinan, Luyu Gao, Sarah Wiegreffe, Uri Alon, Nouha Dziri, Shrimai Prabhumoye, Yiming Yang, et al · 2023
Closest in time.
Artificial intelligence discusses the role of artificial intelligence in translational medicine: a JACC: basic to translational science interview with ChatGPT
Douglas L Mann. 2023 · 2023
Closest in time.
Recent advances in natural language processing via large pre-trained language models: A survey
Bonan Min, Hayley Ross, Elior Sulem, Amir Pouran Ben Veyseh, Thien Huu Nguyen, Oscar Sainz, Eneko Agirre, Ilana Heintz, and Dan Roth. 2023 · 2023
Closest in time.
Is ChatGPT a Good Tool for T&CM Students in Studying Pharmacology?
Saima Nisar and Muhammad Shahzad Aslam. 2023 · 2023
Closest in time.
OpenAI. 2023 · 2023
Closest in time.
Engineering education in the era of ChatGPT: Promise and pitfalls of generative AI for education. In 2023 IEEE Global Engineering Education Conference (EDUCON) . IEEE, 1–9
Junaid Qadir. 2023 · 2023
Closest in time.
Is ChatGPT a general-purpose natural language processing task solver?
Chengwei Qin, Aston Zhang, Zhuosheng Zhang, Jiaao Chen, Michihiro Yasunaga, and Diyi Yang. 2023 · 2023
Closest in time.
ChatGPT for Education and Research: Opportunities, Threats, and Strategies
Md. Mostafizer Rahman and Yutaka Watanobe. 2023 · 2023
Closest in time.
Evaluating ChatGPT as an adjunct for radiologic decision-making. medRxiv, 2023-02
A Rao, J Kim, M Kamineni, M Pang, W Lie, and MD Succi. 2023 · 2023
Closest in time.
ChatGPT: A comprehensive review on background, applications, key challenges, bias, ethics, limitations and future scope
Partha Pratim Ray. 2023 · 2023
Closest in time.
ChatGPT utility in healthcare education, research, and practice: systematic review on the promising perspectives and valid concerns. In Healthcare , Vol. 11. MDPI, 887
Malik Sallam. 2023 · 2023
Closest in time.
ChatGPT in drug discovery
Gaurav Sharma and Abhishek Thakur. 2023 · 2023
Closest in time.
Measuring Inductive Biases of In-Context Learning with Underspecified Demonstrations
Chenglei Si, Dan Friedman, Nitish Joshi, Shi Feng, Danqi Chen, and He He. 2023 · 2023
Closest in time.
Large language models in medicine
Arun James Thirunavukarasu, Darren Shu Jeng Ting, Kabilan Elangovan, Laura Gutierrez, Ting Fang Tan, and Daniel Shu Wei Ting. 2023 · 2023
Closest in time.
ChatGPT is fun, but not an author
H Holden Thorp. 2023 · 2023
Closest in time.
Opportunities and Challenges for ChatGPT and Large Language Models in Biomedicine and Health
Shubo Tian, Qiao Jin, Lana Yeganova, Po-Ting Lai, Qingqing Zhu, Xiuying Chen, Yifan Yang, Qingyu Chen, Won Kim, Donald C. Comeau, Rezarta Islamaj, Aadit Kapoor, Xin Gao, and Zhiyong Lu. 2023 · 2023
Closest in time.
Can ChatGPT write a good boolean query for systematic review literature search?
Shuai Wang, Harrisen Scells, Bevan Koopman, and Guido Zuccon. 2023 · 2023
Closest in time.
A prompt pattern catalog to enhance prompt engineering with chatgpt
Jules White, Quchen Fu, Sam Hays, Michael Sandborn, Carlos Olea, Henry Gilbert, Ashraf Elnashar, Jesse Spencer-Smith, and Douglas C Schmidt. 2023 · 2023
Closest in time.
Exploring the Limits of ChatGPT for Query or Aspect-based Text Summarization
Xianjun Yang, Yan Li, Xinlu Zhang, Haifeng Chen, and Wei Cheng. 2023 · 2023
Closest in time.
Cognitive Mirage: A Review of Hallucinations in Large Language Models
Hongbin Ye, Tong Liu, Aijia Zhang, Wei Hua, and Weiqiang Jia. 2023 · 2023
Closest in time.
Assessing the performance of ChatGPT in answering questions regarding cirrhosis and hepatocellular carcinoma
Yee Hui Yeo, Jamil S Samaan, Wee Han Ng, Peng-Sheng Ting, Hirsh Trivedi, Aarshi Vipani, Walid Ayoub, Ju Dong Yang, Omer Liran, Brennan Spiegel, et al · 2023
Closest in time.
How well do Large Language Models perform in Arithmetic tasks?
Zheng Yuan, Hongyi Yuan, Chuanqi Tan, Wei Wang, and Songfang Huang. 2023 · 2023
Closest in time.
Why Johnny can’t prompt: how non-AI experts try (and fail) to design LLM prompts. In Proceedings of the 2023 CHI Conference on Human Factors in Computing Systems . 1–21
JD Zamfirescu-Pereira, Richmond Y Wong, Bjoern Hartmann, and Qian Yang. 2023 · 2023
Closest in time.
One Small Step for Generative AI, One Giant Leap for AGI: A Complete Survey on ChatGPT in AIGC Era
Chaoning Zhang, Chenshuang Zhang, Chenghao Li, Yu Qiao, Sheng Zheng, Sumit Kumar Dam, Mengchun Zhang, Jung Uk Kim, Seong Tae Kim, Jinwoo Choi, Gyeong-Moon Park, Sung-Ho Bae, Lik-Hang Lee, Pan Hui, In So Kweon, and Choong Seon Hong. 2023 · 2023
Closest in time.
A Survey of Large Language Models
Wayne Xin Zhao, Kun Zhou, Junyi Li, Tianyi Tang, Xiaolei Wang, Yupeng Hou, Yingqian Min, Beichen Zhang, Junjie Zhang, Zican Dong, Yifan Du, Chen Yang, Yushuo Chen, Zhipeng Chen, Jinhao Jiang, Ruiyang Ren, Yifan Li, Xinyu Tang, Zikang Liu, Peiyu Liu, Jian-Yun Nie, and Ji-Rong Wen. 2023 · 2023
Closest in time.
Why Does ChatGPT Fall Short in Providing Truthful Answers?
Shen Zheng, Jie Huang, and Kevin Chen-Chuan Chang. 2023 · 2023
Closest in time.
Navigating the grey area: Expressions of overconfidence and uncertainty in language models
Kaitlyn Zhou, Dan Jurafsky, and Tatsunori Hashimoto. 2023 · 2023
Closest in time.