Fetching the paper…
Reading the bibliography…
This paper aims to quantitatively evaluate the performance of ChatGPT, an interactive large language model, on inter-sentential relations such as temporal relations, causal relations, and discourse relations.
Roberta: A robustly optimized BERT pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019 · 1907
Earlier work this paper cites.
Scaling laws for neural language models
Jared Kaplan, Sam McCandlish, Tom Henighan, Tom B. Brown, Benjamin Chess, Rewon Child, Scott Gray, Alec Radford, Jeffrey Wu, and Dario Amodei. 2020 · 2001
Earlier work this paper cites.
A logic for causal reasoning
Alexander Bochman. 2003 · 2003
Earlier work this paper cites.
Timeml: Robust specification of event and temporal expressions in text
James Pustejovsky, José M. Castaño, Robert Ingria, Roser Saurí, Robert J. Gaizauskas, Andrea Setzer, Graham Katz, and Dragomir R. Radev. 2003a · 2003
Earlier work this paper cites.
The timebank corpus
James Pustejovsky, Patrick Hanks, Roser Sauri, Andrew See, Robert Gaizauskas, Andrea Setzer, Dragomir Radev, Beth Sundheim, David Day, Lisa Ferro, et al. 2003b · 2003
Earlier work this paper cites.
The penn discourse treebank 2.0
Rashmi Prasad, Nikhil Dinesh, Alan Lee, Eleni Miltsakaki, Livio Robaldo, Aravind K. Joshi, and Bonnie L. Webber. 2008 · 2008
Earlier work this paper cites.
RSGT: relational structure guided temporal relation extraction
Jie Zhou, Shenpo Dong, Hongkui Tu, Xiaodong Wang, and Yong Dou. 2022b · 2010
Earlier work this paper cites.
Semeval-2012 task 7: Choice of plausible alternatives: An evaluation of commonsense causal reasoning
Andrew S. Gordon, Zornitsa Kozareva, and Melissa Roemmele. 2012 · 2012
Earlier work this paper cites.
Semeval-2013 task 1: Tempeval-3: Evaluating time expressions, events, and temporal relations
Naushad UzZaman, Hector Llorens, Leon Derczynski, James F. Allen, Marc Verhagen, and James Pustejovsky. 2013 · 2013
Earlier work this paper cites.
An annotation framework for dense event ordering
Taylor Cassidy, Bill McDowell, Nathanael Chambers, and Steven Bethard. 2014 · 2014
Earlier work this paper cites.
Discourse complements lexical semantics for non-factoid answer reranking
Peter Jansen, Mihai Surdeanu, and Peter Clark. 2014 · 2014
Earlier work this paper cites.
Discourse parsing for multi-party chat dialogues
Stergos D. Afantenos, Eric Kow, Nicholas Asher, and Jérémy Perret. 2015 · 2015
Earlier work this paper cites.
One vector is not enough: Entity-augmented distributed semantics for discourse relations
Yangfeng Ji and Jacob Eisenstein. 2015 · 2015
Earlier work this paper cites.
The ubuntu dialogue corpus: A large dataset for research in unstructured multi-turn dialogue systems
Ryan Lowe, Nissan Pow, Iulian Serban, and Joelle Pineau. 2015 · 2015
Earlier work this paper cites.
Discourse structure and dialogue acts in multiparty dialogue: the STAC corpus
Nicholas Asher, Julie Hunter, Mathieu Morey, Farah Benamara, and Stergos D. Afantenos. 2016 · 2016
Earlier work this paper cites.
Integer linear programming for discourse parsing
Jérémy Perret, Stergos D. Afantenos, Nicholas Asher, and Mathieu Morey. 2016 · 2016
Earlier work this paper cites.
Filling in the blanks in understanding discourse adverbials: Consistency, conflict, and context-dependence in a crowdsourced elicitation task
Hannah Rohde, Anna Dickinson, Nathan Schneider, Christopher N. L. Clark, Annie Louis, and Bonnie L. Webber. 2016 · 2016
Earlier work this paper cites.
Creating causal embeddings for question answering with minimal supervision
Rebecca Sharp, Mihai Surdeanu, Peter Jansen, Peter Clark, and Michael Hammond. 2016 · 2016
Earlier work this paper cites.
Examples and specifications that prove a point: Identifying elaborative and argumentative discourse relations
Merel C. J. Scholman and Vera Demberg. 2017 · 2017
Earlier work this paper cites.
Conceptnet 5.5: An open multilingual graph of general knowledge
Robyn Speer, Joshua Chin, and Catherine Havasi. 2017 · 2017
Earlier work this paper cites.
Discourse-aware neural rewards for coherent text generation
Antoine Bosselut, Asli Celikyilmaz, Xiaodong He, Jianfeng Gao, Po-Sen Huang, and Yejin Choi. 2018 · 2018
Earlier work this paper cites.
Joint reasoning for temporal and causal relations
Qiang Ning, Zhili Feng, Hao Wu, and Dan Roth. 2018a · 2018
Earlier work this paper cites.
A multi-axis annotation scheme for event temporal relations
Qiang Ning, Hao Wu, and Dan Roth. 2018b · 2018
Earlier work this paper cites.
A regularization approach for incorporating event knowledge and coreference relations into neural discourse parsing
Zeyu Dai and Ruihong Huang. 2019 · 2019
Earlier work this paper cites.
Discofuse: A large-scale dataset for discourse-based sentence fusion
Mor Geva, Eric Malmi, Idan Szpektor, and Jonathan Berant. 2019 · 2019
Earlier work this paper cites.
Tddiscourse: A dataset for discourse-level temporal ordering of events
Aakanksha Naik, Luke Breitfeller, and Carolyn P. Rosé. 2019 · 2019
Earlier work this paper cites.
ATOMIC: an atlas of machine commonsense for if-then reasoning
Maarten Sap, Ronan Le Bras, Emily Allaway, Chandra Bhagavatula, Nicholas Lourie, Hannah Rashkin, Brendan Roof, Noah A. Smith, and Yejin Choi. 2019 · 2019
Earlier work this paper cites.
A deep sequential model for discourse parsing on multi-party dialogues
Zhouxing Shi and Minlie Huang. 2019 · 2019
Earlier work this paper cites.
Mining discourse markers for unsupervised sentence representation learning
Damien Sileo, Tim Van de Cruys, Camille Pradel, and Philippe Muller. 2019 · 2019
Earlier work this paper cites.
Discourse relation prediction: Revisiting word pairs with convolutional networks
Siddharth Varia, Christopher Hidey, and Tuhin Chakrabarty. 2019 · 2019
Cited alongside, same era.
Language models are few-shot learners
Tom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, Sandhini Agarwal, Ariel Herbert-Voss, Gretchen Krueger, Tom Henighan, Rewon Child, Aditya Ramesh, Daniel M. Ziegler, Jeffrey Wu, Clemens Winter, Christopher Hesse, Mark Chen, Eric Sigler, Mateusz Litwin, Scott Gray, Benjamin Chess, Jack Clark, Christopher Berner, Sam McCandlish, Alec Radford, Ilya Sutskever, and Dario Amodei. 2020 · 2020
Cited alongside, same era.
ELECTRA: pre-training text encoders as discriminators rather than generators
Kevin Clark, Minh-Thang Luong, Quoc V. Le, and Christopher D. Manning. 2020 · 2020
Cited alongside, same era.
Molweni: A challenge multiparty dialogues-based machine reading comprehension dataset with discourse structure
Jiaqi Li, Ming Liu, Min-Yen Kan, Zihao Zheng, Zekun Wang, Wenqiang Lei, Ting Liu, and Bing Qin. 2020 · 2020
Cited alongside, same era.
On the importance of word and sentence representation learning in implicit discourse relation classification
Finetuned language models are zero-shot learners
Jason Wei, Maarten Bosma, Vincent Y. Zhao, Kelvin Guu, Adams Wei Yu, Brian Lester, Nan Du, Andrew M. Dai, and Quoc V. Le. 2022 · 2022
Later among the works it cites.
Label distributions help implicit discourse relation classification
Frances Yung, Kaveri Anuranjana, Merel Scholman, and Vera Demberg. 2022 · 2022
Later among the works it cites.
Crake: Causal-enhanced table-filler for question answering over large scale knowledge base
Minhao Zhang, Ruoyu Zhang, Yanzeng Li, and Lei Zou. 2022b · 2022
Later among the works it cites.
Active example selection for in-context learning
Yiming Zhang, Shi Feng, and Chenhao Tan. 2022c · 2022
Later among the works it cites.
Prompt-based connective prediction method for fine-grained implicit discourse relation recognition
Hao Zhou, Man Lan, Yuanbin Wu, Yuefeng Chen, and Meirong Ma. 2022a · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Xin Liu, Jiefu Ou, Yangqiu Song, and Xin Jiang. 2020 · 2020
Cited alongside, same era.
GLUCOSE: generalized and contextualized story explanations
Nasrin Mostafazadeh, Aditya Kalyanpur, Lori Moon, David W. Buchanan, Lauren Berkowitz, Or Biran, and Jennifer Chu-Carroll. 2020 · 2020
Cited alongside, same era.
XCOPA: A multilingual dataset for causal commonsense reasoning
Edoardo Maria Ponti, Goran Glavas, Olga Majewska, Qianchu Liu, Ivan Vulic, and Anna Korhonen. 2020 · 2020
Cited alongside, same era.
ASER: A large-scale eventuality knowledge graph
Hongming Zhang, Xin Liu, Haojie Pan, Yangqiu Song, and Cane Wing-Ki Leung. 2020 · 2020
Cited alongside, same era.
Knowledge-enriched event causality identification via latent structure induction networks
Pengfei Cao, Xinyu Zuo, Yubo Chen, Kang Liu, Jun Zhao, Yuguang Chen, and Weihua Peng. 2021 · 2021
Cited alongside, same era.
Benchmarking commonsense knowledge base population with an effective evaluation dataset
Tianqing Fang, Weiqi Wang, Sehyun Choi, Shibo Hao, Hongming Zhang, Yangqiu Song, and Bin He. 2021 · 2021
Cited alongside, same era.
PTR: prompt tuning with rules for text classification
Xu Han, Weilin Zhao, Ning Ding, Zhiyuan Liu, and Maosong Sun. 2021 · 2021
Cited alongside, same era.
Exploring discourse structures for argument impact classification
Xin Liu, Jiefu Ou, Yangqiu Song, and Xin Jiang. 2021b · 2021
Cited alongside, same era.
Yejin Bang, Samuel Cahyawijaya, Nayeon Lee, Wenliang Dai, Dan Su, Bryan Wilie, Holy Lovenia, Ziwei Ji, Tiezheng Yu, Willy Chung, Quyet V. Do, Yan Xu, and Pascale Fung. 2023 · 2023
Closest in time.
Sparks of artificial general intelligence: Early experiments with GPT-4
Sébastien Bubeck, Varun Chandrasekaran, Ronen Eldan, Johannes Gehrke, Eric Horvitz, Ece Kamar, Peter Lee, Yin Tat Lee, Yuanzhi Li, Scott M. Lundberg, Harsha Nori, Hamid Palangi, Marco Túlio Ribeiro, and Yi Zhang. 2023 · 2023
Closest in time.
Discourse-aware prompt for argument impact classification
Chunkit Chan and Tsz Ho Chan. 2023 · 2023
Closest in time.
Jiayang Cheng, Lin Qiu, Tsz Ho Chan, Tianqing Fang, Weiqi Wang, Chunkit Chan, Dongyu Ru, Qipeng Guo, Hongming Zhang, Yangqiu Song, Yue Zhang, and Zheng Zhang. 2023 · 2023
Closest in time.
Benchmarks for automated commonsense reasoning: A survey
Ernest Davis. 2023 · 2023
Closest in time.
Ckbp v2: An expert-annotated evaluation set for commonsense knowledge base population
Tianqing Fang, Quyet V. Do, Sehyun Choi, Weiqi Wang, and Yangqiu Song. 2023 · 2023
Closest in time.
Mathematical capabilities of chatgpt
Simon Frieder, Luca Pinchetti, Ryan-Rhys Griffiths, Tommaso Salvatori, Thomas Lukasiewicz, Philipp Christian Petersen, Alexis Chevalier, and Julius Berner. 2023 · 2023
Closest in time.
Chatgpt outperforms crowd-workers for text-annotation tasks
Fabrizio Gilardi, Meysam Alizadeh, and Maël Kubli. 2023 · 2023
Closest in time.
Lion: Adversarial distillation of closed-source large language model
Yuxin Jiang, Chunkit Chan, Mingyang Chen, and Wei Wang. 2023 · 2023
Closest in time.
Is chatgpt A good translator? A preliminary study
Wenxiang Jiao, Wenxuan Wang, Jen-tse Huang, Xing Wang, and Zhaopeng Tu. 2023 · 2023
Closest in time.
Chatgpt: Jack of all trades, master of none
Jan Kocon, Igor Cichecki, Oliwier Kaszyca, Mateusz Kochanek, Dominika Szydlo, Joanna Baran, Julita Bielaniewicz, Marcin Gruza, Arkadiusz Janz, Kamil Kanclerz, Anna Kocon, Bartlomiej Koptyra, Wiktoria Mieleszczenko-Kowszewicz, Piotr Milkowski, Marcin Oleksy, Maciej Piasecki, Lukasz Radlinski, Konrad Wojtasik, Stanislaw Wozniak, and Przemyslaw Kazienko. 2023 · 2023
Closest in time.
Multi-step jailbreaking privacy attacks on chatgpt
Haoran Li, Dadi Guo, Wei Fan, Mingshi Xu, Jie Huang, Fanpu Meng, and Yangqiu Song. 2023b · 2023
Closest in time.
Finding supporting examples for in-context learning
Xiaonan Li and Xipeng Qiu. 2023 · 2023
Closest in time.
Analyzing leakage of personally identifiable information in language models
Nils Lukas, Ahmed Salem, Robert Sim, Shruti Tople, Lukas Wutschitz, and Santiago Zanella Béguelin. 2023 · 2023
Closest in time.
OpenAI. 2023 · 2023
Closest in time.
What happens before and after: Multi-event commonsense in event coreference resolution
Sahithya Ravi, Chris Tanner, Raymond Ng, and Vered Shwartz. 2023 · 2023
Closest in time.
Are emergent abilities of large language models a mirage?
Rylan Schaeffer, Brando Miranda, and Sanmi Koyejo. 2023 · 2023
Closest in time.
Let’s have a chat! A conversation with chatgpt: Technology, applications, and limitations
Sakib Shahriar and Kadhim Hayawi. 2023 · 2023
Closest in time.
Language models are causal knowledge extractors for zero-shot video question answering
Hung-Ting Su, Yulei Niu, Xudong Lin, Winston H. Hsu, and Shih-Fu Chang. 2023 · 2023
Closest in time.
Petter Törnberg. 2023 · 2023
Closest in time.
Causal-discovery performance of chatgpt in the context of neuropathic pain diagnosis
Ruibo Tu, Chao Ma, and Cheng Zhang. 2023 · 2023
Closest in time.
Survey on factuality in large language models: Knowledge, retrieval and domain-specificity
Cunxiang Wang, Xiaoze Liu, Yuanhao Yue, Xiangru Tang, Tianhang Zhang, Cheng Jiayang, Yunzhi Yao, Wenyang Gao, Xuming Hu, Zehan Qi, et al. 2023 · 2023
Closest in time.
Zero-shot temporal relation extraction with chatgpt
Chenhan Yuan, Qianqian Xie, and Sophia Ananiadou. 2023 · 2023
Closest in time.
Weiqi Wang, Tianqing Fang, Chunyang Li, Haochen Shi, Wenxuan Ding, Baixuan Xu, Zhaowei Wang, Jiaxin Bai, Xin Liu, Jiayang Cheng, et al. 2024 · 2024
Closest in time.