Fetching the paper…
Reading the bibliography…
Large language models have gained considerable interest for their impressive performance on various tasks.
Bleu: A method for automatic evaluation of machine translation
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu · 2002
Earlier work this paper cites.
ROUGE: A package for automatic evaluation of summaries
Chin-Yew Lin · 2004
Earlier work this paper cites.
Interpreting bleu/nist scores: How much improvement do we need to have a better system?
Y. Zhang, Stephan Vogel, and Alexander H. Waibel · 2004
Earlier work this paper cites.
METEOR: An automatic metric for MT evaluation with improved correlation with human judgments
Satanjeev Banerjee and Alon Lavie · 2005
Earlier work this paper cites.
Meteor, M-BLEU and M-TER: Evaluation metrics for high-correlation with human rankings of machine translation output
Abhaya Agarwal and Alon Lavie · 2008
Earlier work this paper cites.
Coqa: A conversational question answering challenge, 2018
Siva Reddy, Danqi Chen, and Christopher D. Manning · 2018
Earlier work this paper cites.
Codah: An adversarially authored question-answer dataset for common sense, 2019
Michael Chen, Mike D’Arcy, Alisa Liu, Jared Fernandez, and Doug Downey · 2019
Earlier work this paper cites.
Dialfact: A benchmark for fact-checking in dialogue, 2021
Prakhar Gupta, Chien-Sheng Wu, Wenhao Liu, and Caiming Xiong · 2021
Earlier work this paper cites.
Faviq: Fact verification from information-seeking questions, 2021
Jungsoo Park, Sewon Min, Jaewoo Kang, Luke Zettlemoyer, and Hannaneh Hajishirzi · 2021
Earlier work this paper cites.
Bartscore: Evaluating generated text as text generation, 2021
Weizhe Yuan, Graham Neubig, and Pengfei Liu · 2021
Earlier work this paper cites.
"i think this is the most disruptive technology": Exploring sentiments of chatgpt early adopters using twitter data, 2022
Mubin Ul Haque, Isuru Dharmadasa, Zarrin Tasnim Sworna, Roshan Namal Rajapakse, and Hussain Ahmad · 2022
Earlier work this paper cites.
What is ai chatbot phenomenon chatgpt and could it replace humans?, Dec 2022
2022
Cited alongside, same era.
Chatgpt: The end of online exam integrity?, 2022
Teo Susnjak · 2022
Cited alongside, same era.
Chatgpt makes medicine easy to swallow: An exploratory case study on simplified radiology reports, 2022
Katharina Jeblick, Balthasar Schachtner, Jakob Dexl, Andreas Mittermeier, Anna Theresa Stüber, Johanna Topalis, Tobias Weber, Philipp Wesp, Bastian Sabel, Jens Ricke, and Michael Ingrisch · 2022
Cited alongside, same era.
Uzh_clyp at semeval-2023 task 9: Head-first fine-tuning and chatgpt data generation for cross-lingual learning in tweet intimacy prediction, 2023
Andrianos Michail, Stefanos Konstantinou, and Simon Clematide · 2023
Cited alongside, same era.
Can chatgpt assess human personalities? a general evaluation framework, 2023
Haocong Rao, Cyril Leung, and Chunyan Miao · 2023
Cited alongside, same era.
Is chatgpt a general-purpose natural language processing task solver?, 2023
Chengwei Qin, Aston Zhang, Zhuosheng Zhang, Jiaao Chen, Michihiro Yasunaga, and Diyi Yang · 2023
Closest in time.
Regulating chatgpt and other large generative ai models, 2023
Philipp Hacker, Andreas Engel, and Marco Mauer · 2023
Closest in time.
ChatGPT and other large language models are double-edged swords
Yiqiu Shen, Laura Heacock, Jonathan Elias, Keith D Hentel, Beatriu Reig, George Shih, and Linda Moy · 2023
Closest in time.
Exploring ai ethics of chatgpt: A diagnostic analysis, 2023
Terry Yue Zhuo, Yujin Huang, Chunyang Chen, and Zhenchang Xing · 2023
Closest in time.
Using chatgpt for human-computer interaction research: A primer, 03 2023
Wilbert Tabone and Joost de Winter · 2023
Closest in time.
Will chatgpt get you caught? rethinking of plagiarism detection, 2023
Mohammad Khalil and Erkan Er · 2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Towards human-bot collaborative software architecting with chatgpt, 2023
Aakash Ahmad, Muhammad Waseem, Peng Liang, Mahdi Fehmideh, Mst Shamima Aktar, and Tommi Mikkonen · 2023
Cited alongside, same era.
How will language modelers like chatgpt affect occupations and industries?, 2023
Ed Felten, Manav Raj, and Robert Seamans · 2023
Cited alongside, same era.
Jaccard metric losses: Optimizing the jaccard index with soft labels, 2023
Zifu Wang and Matthew B. Blaschko · 2023
Cited alongside, same era.
ChatGPT: five priorities for research
Eva A M van Dis, Johan Bollen, Willem Zuidema, Robert van Rooij, and Claudi L Bockting · 2023
Cited alongside, same era.
A multitask, multilingual, multimodal evaluation of chatgpt on reasoning, hallucination, and interactivity, 2023
Yejin Bang, Samuel Cahyawijaya, Nayeon Lee, Wenliang Dai, Dan Su, Bryan Wilie, Holy Lovenia, Ziwei Ji, Tiezheng Yu, Willy Chung, Quyet V. Do, Yan Xu, and Pascale Fung · 2023
Cited alongside, same era.
Chatgpt is not all you need. a state of the art review of large generative ai models, 2023
Roberto Gozalo-Brizuela and Eduardo C. Garrido-Merchan · 2023
Cited alongside, same era.
Closest in time.
Exploring the role of artificial intelligence in enhancing academic performance: A case study of chatgpt
https://papers.ssrn.com/sol3/papers.cfm?abstract_id=4312358/ · 2023
Closest in time.
How close is chatgpt to human experts? comparison corpus, evaluation, and detection, 2023
Biyang Guo, Xin Zhang, Ziyuan Wang, Minqi Jiang, Jinran Nie, Yuxuan Ding, Jianwei Yue, and Yupeng Wu · 2023
Closest in time.
How close is chatgpt to human experts? comparison corpus, evaluation, and detection, 2023
Biyang Guo, Xin Zhang, Ziyuan Wang, Minqi Jiang, Jinran Nie, Yuxuan Ding, Jianwei Yue, and Yupeng Wu · 2023
Closest in time.
Chatgpt and the future of medical writing
Som Biswas · 2023
Closest in time.