Fetching the paper…
Reading the bibliography…
The emergence of ChatGPT has generated much speculation in the press about its potential to disrupt social and economic systems.
RoBERTa: A robustly optimized BERT pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019 · 1907
Earlier work this paper cites.
Diagnostic psychological testing: The theory, statistical evaluation, and diagnostic application of a battery of tests: Volume II
David Rapaport, Merton Gill, and Roy Schafer. 1946 · 1946
Earlier work this paper cites.
Logical versus analogical or symbolic versus connectionist or neat versus scruffy
Marvin L Minsky. 1991 · 1991
Earlier work this paper cites.
Thinking Fast and Slow
Daniel Kahneman. 2000 · 2000
Earlier work this paper cites.
Introduction to the CoNLL-2003 shared task: Language-independent named entity recognition
Erik Tjong Kim Sang and Fien De Meulder. 2003 · 2003
Earlier work this paper cites.
ROUGE: A package for automatic evaluation of summaries
Chin-Yew Lin. 2004 · 2004
Earlier work this paper cites.
The predicting brain: Unconscious repetition, conscious reflection and therapeutic change
Regina Pally. 2007 · 2007
Earlier work this paper cites.
Phrase-based statistical language generation using graphical models and active learning
François Mairesse, Milica Gasic, Filip Jurcicek, Simon Keizer, Blaise Thomson, Kai Yu, and Steve Young. 2010 · 2010
Earlier work this paper cites.
The buying brain: Secrets for selling to the subconscious mind
Anantha Krishnan Pradeep. 2010 · 2010
Earlier work this paper cites.
Representation learning: A review and new perspectives
Yoshua Bengio, Aaron Courville, and Pascal Vincent. 2013 · 2013
Earlier work this paper cites.
Jumping NLP curves: A review of natural language processing research
Erik Cambria and Bebo White. 2014 · 2014
Earlier work this paper cites.
Teaching with rewards and punishments: Reinforcement or communication?
Mark K Ho, Michael L Littman, Fiery Cushman, and Joseph L Austerweil. 2015 · 2015
Earlier work this paper cites.
chrF: character n-gram F-score for automatic MT evaluation
Maja Popović. 2015 · 2015
Earlier work this paper cites.
ImageNet large scale visual recognition challenge
Olga Russakovsky, Jia Deng, Hao Su, Jonathan Krause, Sanjeev Satheesh, Sean Ma, Zhiheng Huang, Andrej Karpathy, Aditya Khosla, Michael Bernstein, et al. 2015 · 2015
Earlier work this paper cites.
Learning through human feedback
Jan Leike, Miljan Martic, and Shane Legg. 2017 · 2017
Earlier work this paper cites.
NEWSROOM: A dataset of 1.3 million summaries with diverse extractive strategies
Max Grusky, Mor Naaman, and Yoav Artzi. 2018 · 2018
Earlier work this paper cites.
How people learn II: Learners, contexts, and cultures
NASEM, National Academies of Sciences, Engineering, and Medicine and others. 2018 · 2018
Earlier work this paper cites.
BERT: Pre-training of deep bidirectional Transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Earlier work this paper cites.
Language models are unsupervised multitask learners
Alec Radford, Jeff Wu, Rewon Child, David Luan, Dario Amodei, and Ilya Sutskever. 2019 · 2019
Earlier work this paper cites.
BERTScore: Evaluating text generation with BERT
Tianyi Zhang, Varsha Kishore, Felix Wu, Kilian Q Weinberger, and Yoav Artzi. 2019 · 2019
Earlier work this paper cites.
Re-evaluating evaluation in text summarization
Manik Bhandari, Pranav Narayan Gour, Atabak Ashfaq, Pengfei Liu, and Graham Neubig. 2020 · 2020
Earlier work this paper cites.
USR: An unsupervised and reference free evaluation metric for dialog generation
Shikib Mehri and Maxine Eskenazi. 2020 · 2020
Earlier work this paper cites.
Going in circles is the way forward: The role of recurrence in visual inference
Ruben S van Bergen and Nikolaus Kriegeskorte. 2020 · 2020
Earlier work this paper cites.
SummEval: Re-evaluating summarization evaluation
Alexander R Fabbri, Wojciech Kryściński, Bryan McCann, Caiming Xiong, Richard Socher, and Dragomir Radev. 2021 · 2021
Earlier work this paper cites.
OpenMEVA: A benchmark for evaluating open-ended story generation metrics
Jian Guan, Zhexin Zhang, Zhuoer Feng, Zitao Liu, Wenbiao Ding, Xiaoxi Mao, Changjie Fan, and Minlie Huang. 2021 · 2021
Earlier work this paper cites.
Multilingual translation from denoising pre-training
Yuqing Tang, Chau Tran, Xian Li, Peng-Jen Chen, Naman Goyal, Vishrav Chaudhary, Jiatao Gu, and Angela Fan. 2021 · 2021
Earlier work this paper cites.
BARTScore: Evaluating generated text as text generation
Weizhe Yuan, Graham Neubig, and Pengfei Liu. 2021 · 2021
Earlier work this paper cites.
SenticNet 7: A commonsense-based neurosymbolic AI framework for explainable sentiment analysis
Erik Cambria, Qian Liu, Sergio Decherchi, Frank Xing, and Kenneth Kwok. 2022 · 2022
Earlier work this paper cites.
Solving quantitative reasoning problems with language models
Aitor Lewkowycz, Anders Johan Andreassen, David Dohan, Ethan Dyer, Henryk Michalewski, Vinay Venkatesh Ramasesh, Ambrose Slone, Cem Anil, Imanol Schlag, Theo Gutman-Solo, et al. 2022 · 2022
Earlier work this paper cites.
Rethinking the role of demonstrations: What makes in-context learning work?
Sewon Min, Xinxi Lyu, Ari Holtzman, Mikel Artetxe, Mike Lewis, Hannaneh Hajishirzi, and Luke Zettlemoyer. 2022 · 2022
Earlier work this paper cites.
Can we trust the evaluation on ChatGPT?
Rachith Aiyappa, Jisun An, Haewoon Kwak, and Yong-Yeol Ahn. 2023 · 2023
Earlier work this paper cites.
A wide evaluation of ChatGPT on affective computing tasks
Mostafa M Amin, Rui Mao, Erik Cambria, and Björn W Schuller. 2023 · 2023
Earlier work this paper cites.
Evaluating the performance of ChatGPT in ophthalmology: An analysis of its successes and shortcomings
Fares Antaki, Samir Touma, Daniel Milad, Jonathan El-Khoury, and Renaud Duval. 2023 · 2023
Earlier work this paper cites.
Exploring the boundaries of reality: Investigating the phenomenon of artificial intelligence hallucination in scientific writing through ChatGPT references
Sai Anirudh Athaluri, Sandeep Varma Manthena, VSR Krishna Manoj Kesapragada, Vineel Yarlagadda, Tirth Dave, and Rama Tulasi Siri Duddumpudi. 2023 · 2023
Earlier work this paper cites.
Yejin Bang, Samuel Cahyawijaya, Nayeon Lee, Wenliang Dai, Dan Su, Bryan Wilie, Holy Lovenia, Ziwei Ji, Tiezheng Yu, Willy Chung, et al. 2023 · 2023
Earlier work this paper cites.
Missing information, unresponsive authors, experimental flaws: The impossibility of assessing the reproducibility of previous human evaluations in NLP
Anja Belz, Craig Thomson, and Ehud Reiter. 2023 · 2023
Earlier work this paper cites.
ChatGPT participates in a computer science exam
Sebastian Bordt and Ulrike von Luxburg. 2023 · 2023
Earlier work this paper cites.
A categorical archive of ChatGPT failures
Ali Borji. 2023 · 2023
Earlier work this paper cites.
Sparks of artificial general intelligence: Early experiments with GPT-4
Sébastien Bubeck, Varun Chandrasekaran, Ronen Eldan, Johannes Gehrke, Eric Horvitz, Ece Kamar, Peter Lee, Yin Tat Lee, Yuanzhi Li, Scott Lundberg, Harsha Nori, Hamid Palangi, Marco Tulio Ribeiro, and Yi Zhang. 2023 · 2023
Earlier work this paper cites.
Zeno chatbot report
Alex Cabrera and Graham Neubig. 2023 · 2023
Cited alongside, same era.
Seven pillars for the future of Artificial Intelligence
Erik Cambria, Rui Mao, Melvin Chen, Zhaoxia Wang, and Seng-Beng Ho. 2023 · 2023
Cited alongside, same era.
ChatGPT for tourism: Applications, benefits and risks
Inês Carvalho and Stanislav Ivanov. 2023 · 2023
Cited alongside, same era.
Vicuna: An open-source chatbot impressing GPT-4 with 90%* ChatGPT quality
Wei-Lin Chiang, Zhuohan Li, Zi Lin, Ying Sheng, Zhanghao Wu, Hao Zhang, Lianmin Zheng, Siyuan Zhuang, Yonghao Zhuang, Joseph E. Gonzalez, Ion Stoica, and Eric P. Xing. 2023 · 2023
Cited alongside, same era.
ChatGPT goes to law school
Jonathan H Choi, Kristin E Hickman, Amy Monahan, and Daniel Schwarcz. 2023 · 2023
Cited alongside, same era.
Can AI tell good stories? Narrative transportation and persuasion with ChatGPT
Haoran Chu and Sixiao Liu. 2023 · 2023
Jinyang Li, Binyuan Hui, Ge Qu, Binhua Li, Jiaxi Yang, Bowen Li, Bailin Wang, Bowen Qin, Rongyu Cao, Ruiying Geng, et al. 2023 · 2023
Closest in time.
Perspectives on the social impacts of reinforcement learning with human feedback
Gabrielle Kaili-May Liu. 2023 · 2023
Closest in time.
How not to test GPT
Gary Marcus and Ernest Davis. 2023 · 2023
Closest in time.
Re-evaluating GPT-4’s bar exam performance
Eric Martínez. 2023 · 2023
Closest in time.
ChatGPT as a medical doctor? A diagnostic accuracy study on common and rare diseases
Lars Mehnen, Stefanie Gruarin, Mina Vasileva, and Bernhard Knapp. 2023 · 2023
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Investigating the use of an artificial intelligence chatbot with general chemistry exam questions
Ted M Clark. 2023 · 2023
Cited alongside, same era.
Evaluating language models for mathematics through interactions
Katherine M. Collins, Albert Qiaochu Jiang, Simon Frieder, Li Siang Wong, Miri Zilka, Umang Bhatt, Thomas Lukasiewicz, Yuhuai Wu, Joshua B. Tenenbaum, William Hart, Timothy Gowers, Wen-Ding Li, Adrian Weller, and Mateja Jamnik. 2023 · 2023
Cited alongside, same era.
Merten Nikolay Dahlkemper, Simon Zacharias Lahme, and Pascal Klein. 2023 · 2023
Cited alongside, same era.
Benchmarks for automated commonsense reasoning: A survey
Ernest Davis. 2023 · 2023
Cited alongside, same era.
An evaluation on large language model outputs: Discourse and memorization
Adrian de Wynter, Xun Wang, Alex Sokolov, Qilong Gu, and Si-Qing Chen. 2023 · 2023
Cited alongside, same era.
Beyond the safeguards: Exploring the security risks of ChatGPT
Erik Derner and Kristina Batistič. 2023 · 2023
Cited alongside, same era.
Shima Rahimi Moghaddam and Christopher J. Honey. 2023 · 2023
Closest in time.
Introducing MPT-7B: A new standard for open-source, commercially usable LLMs
NLP Team MosaicML. 2023 · 2023
Closest in time.
More human than human: Measuring ChatGPT political bias
Fabio Motoki, Valdemar Pinho Neto, and Victor Rodrigues. 2023 · 2023
Closest in time.
Introducing ChatGPT
OpenAI. 2023b · 2023
Closest in time.
Privacy protection with AI: Survey of data-anonymization techniques
Brad Payne. 2020 · 2023
Closest in time.
On the security vulnerabilities of text-to-SQL models
Xutan Peng, Yipeng Zhang, Jingfeng Yang, and Mark Stevenson. 2023 · 2023
Closest in time.
Summarization is (almost) dead
Xiao Pu, Mingqi Gao, and Xiaojun Wan. 2023 · 2023
Closest in time.
Is ChatGPT a general-purpose natural language processing task solver?
Chengwei Qin, Aston Zhang, Zhuosheng Zhang, Jiaao Chen, Michihiro Yasunaga, and Diyi Yang. 2023 · 2023
Closest in time.
Assessing the utility of ChatGPT throughout the entire clinical workflow
Arya S Rao, Michael Pang, John Kim, Meghana Kamineni, Winston Lie, Anoop K Prasad, Adam Landman, Keith Dryer, and Marc D Succi. 2023 · 2023
Closest in time.
You can generate it again: Data-to-text generation with verification and correction prompting
Xuan Ren and Lingqiao Liu. 2023 · 2023
Closest in time.
Marketing with ChatGPT: Navigating the ethical terrain of GPT-based chatbot technology
Pablo Rivas and Liang Zhao. 2023 · 2023
Closest in time.
The political biases of ChatGPT
David Rozado. 2023 · 2023
Closest in time.
The self-perception and political biases of ChatGPT
Jérôme Rutinowski, Sven Franke, Jan Endendyk, Ina Dormuth, and Markus Pauly. 2023 · 2023
Closest in time.
ChatGPT utility in healthcare education, research, and practice: Systematic review on the promising perspectives and valid concerns
Malik Sallam. 2023 · 2023
Closest in time.
Do ChatGPT and other AI chatbots pose a cybersecurity risk?: An exploratory study
Glorin Sebastian. 2023 · 2023
Closest in time.
ChatGPT: Not all languages are equal
Mohamed L Seghier. 2023 · 2023
Closest in time.
The curse of recursion: Training on generated data makes models forget
Ilia Shumailov, Zakhar Shumaylov, Yiren Zhao, Yarin Gal, Nicolas Papernot, and Ross Anderson. 2023 · 2023
Closest in time.
Mayank Soni and Vincent Wade. 2023 · 2023
Closest in time.
Is ChatGPT good at search? Investigating large language models as re-ranking agent
Weiwei Sun, Lingyong Yan, Xinyu Ma, Pengjie Ren, Dawei Yin, and Zhaochun Ren. 2023 · 2023
Closest in time.
Alpaca: A strong, replicable instruction-following model
Rohan Taori, Ishaan Gulrajani, Tianyi Zhang, Yann Dubois, Xuechen Li, Carlos Guestrin, Percy Liang, and Tatsunori B Hashimoto. 2023 · 2023
Closest in time.
LLaMa: Open and efficient foundation language models
Hugo Touvron, Thibaut Lavril, Gautier Izacard, Xavier Martinet, Marie-Anne Lachaux, Timothée Lacroix, Baptiste Rozière, Naman Goyal, Eric Hambro, Faisal Azhar, et al. 2023 · 2023
Closest in time.
ChatLog: Recording and analyzing ChatGPT across time
Shangqing Tu, Chunyang Li, Jifan Yu, Xiaozhi Wang, Lei Hou, and Juanzi Li. 2023 · 2023
Closest in time.
Zero-shot information extraction via chatting with ChatGPT
Xiang Wei, Xingyu Cui, Ning Cheng, Xiaobin Wang, Xin Zhang, Shen Huang, Pengjun Xie, Jinan Xu, Yufeng Chen, Meishan Zhang, et al. 2023 · 2023
Closest in time.
Qianqian Xie, Weiguang Han, Yanzhao Lai, Min Peng, and Jimin Huang. 2023 · 2023
Closest in time.
SuperCLUE: A benchmark for foundation models in Chinese
Liang Xu and others from SuperCLUE team. 2023 · 2023
Closest in time.
The dawn of lmms: Preliminary explorations with gpt-4v(ision)
Zhengyuan Yang, Linjie Li, Kevin Lin, Jianfeng Wang, Chung-Ching Lin, Zicheng Liu, and Lijuan Wang. 2023 · 2023
Closest in time.
Assessing hidden risks of LLMs: An empirical study on robustness, consistency, and credibility
Wentao Ye, Mingfeng Ou, Tianyi Li, Yipeng chen, Xuetao Ma, Yifan Yanggong, Sai Wu, Jie Fu, Gang Chen, Haobo Wang, and Junbo Zhao. 2023 · 2023
Closest in time.
Evaluating generative models for graph-to-text generation
Shuzhou Yuan and Michael F"arber. 2023 · 2023
Closest in time.
Red teaming ChatGPT via jailbreaking: Bias, robustness, reliability and toxicity
Terry Yue Zhuo, Yujin Huang, Chunyang Chen, and Zhenchang Xing. 2023 · 2023
Closest in time.
A survey on semantic processing techniques
Rui Mao, Kai He, Xulang Zhang, Guanyi Chen, Jinjie Ni, Zonglin Yang, and Erik Cambria. 2024 · 2024
Closest in time.
MM1: Methods, analysis & insights from multimodal llm pre-training
Brandon McKinzie, Zhe Gan, Jean-Philippe Fauconnier, Sam Dodge, Bowen Zhang, Philipp Dufter, Dhruti Shah, Xianzhi Du, Futang Peng, Floris Weers, Anton Belyi, Haotian Zhang, Karanjeet Singh, Doug Kang, Hongyu Hè, Max Schwarzer, Tom Gunter, Xiang Kong, Aonan Zhang, Jianyu Wang, Chong Wang, Nan Du, Tao Lei, Sam Wiseman, Mark Lee, Zirui Wang, Ruoming Pang, Peter Grasch, Alexander Toshev, and Yinfei Yang. 2024 · 2024
Closest in time.
A comparative analysis of metaphorical cognition in ChatGPT and human minds
Rui Mao, Guanyi Chen, Xiao Li, Mengshi Ge, and Erik Cambria. 2025 · 2025
Closest in time.
Towards a unified multi-dimensional evaluator for text generation
Ming Zhong, Yang Liu, Da Yin, Yuning Mao, Yizhu Jiao, Pengfei Liu, Chenguang Zhu, Heng Ji, and Jiawei Han. 2022 · 2038
Closest in time.