Fetching the paper…
Reading the bibliography…
Although large language models (LLMs) have apparently acquired a certain level of grammatical knowledge and the ability to make generalizations, they fail to interpret negation, a crucial step in Natural Language Processing.
WordNet: An Electronic Lexical Database
C. Fellbaum, editor. 1998 · 1998
Earlier work this paper cites.
A Simple Algorithm for Identifying Negated Findings and Diseases in Discharge Summaries
Wendy W. Chapman, Will Bridewell, Paul Hanbury, Gregory F. Cooper, and Bruce G. Buchanan. 2001 · 2001
Earlier work this paper cites.
Negation , pages 785–850. Cambridge University Press
Geoffrey K. Pullum, Rodney Huddleston, Rodney Huddleston, and Geoffrey K. Pullum. 2002 · 2002
Earlier work this paper cites.
Revising the Wordnet Domains Hierarchy: semantics, coverage and balancing
Luisa Bentivogli, Pamela Forner, Bernardo Magnini, and Emanuele Pianta. 2004 · 2004
Earlier work this paper cites.
Adding dense, weighted connections to WordNet
Jordan Boyd-Graber, Christiane Fellbaum, Daniel Osherson, and Robert Schapire. 2006 · 2006
Earlier work this paper cites.
Exploring the Automatic Selection of Basic Level Concepts
Rubén Izquierdo, Armando Suárez, and German Rigau. 2007 · 2007
Earlier work this paper cites.
Putting semantics into WordNet’s “Morphosemantic” Links
Christiane Fellbaum, Anne Osherson, and Peter E. Clark. 2009 · 2009
Earlier work this paper cites.
ConanDoyle-neg: Annotation of negation cues and their scope in Conan Doyle stories
Roser Morante and Walter Daelemans. 2012 · 2012
Earlier work this paper cites.
On the usefulness of lexical and syntactic processing in polarity classification of Twitter messages
David Vilares, Miguel A. Alonso, and Carlos Gómez-Rodríguez. 2015 · 2015
Earlier work this paper cites.
Neural versus Phrase-Based Machine Translation Quality: a Case Study
Luisa Bentivogli, Arianna Bisazza, Mauro Cettolo, and Marcello Federico. 2016 · 2016
Earlier work this paper cites.
Do Language Models Understand Anything? On the Ability of LSTMs to Understand Negative Polarity Items
Jaap Jumelet and Dieuwke Hupkes. 2018 · 2018
Earlier work this paper cites.
Linguistic generalization and compositionality in modern artificial neural networks
Marco Baroni. 2020 · 2020
Earlier work this paper cites.
Not a cute stroke: Analysis of Rule- and Neural Network-based Information Extraction Systems for Brain Radiology Reports
Andreas Grivas, Beatrice Alex, Claire Grover, Richard Tobin, and William Whiteley. 2020 · 2020
Cited alongside, same era.
An Analysis of Natural Language Inference Benchmarks through the Lens of Negation
Md Mosharaf Hossain, Venelin Kovatchev, Pranoy Dutta, Tiffany Kao, Elizabeth Wei, and Eduardo Blanco. 2020 · 2020
Cited alongside, same era.
Negated and Misprimed Probes for Pretrained Language Models: Birds Can Talk, But Cannot Fly
Nora Kassner and Hinrich Schütze. 2020 · 2020
Cited alongside, same era.
Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer
Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, and Peter J. Liu. 2020 · 2020
Cited alongside, same era.
Improving sentiment analysis with multi-task learning of negation
Jeremy Barnes, Erik Velldal, and Lilja Øvrelid. 2021 · 2021
Cited alongside, same era.
LoRA: Low-Rank Adaptation of Large Language Models
Edward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen. 2022 · 2022
Later among the works it cites.
Not another Negation Benchmark: The NaN-NLI Test Suite for Sub-clausal Negation
Thinh Hung Truong, Yulia Otmakhova, Timothy Baldwin, Trevor Cohn, Jey Han Lau, and Karin Verspoor. 2022 · 2022
Later among the works it cites.
Falcon-40B: an open large language model with state-of-the-art performance
Ebtesam Almazrouei, Hamza Alobeidli, Abdulaziz Alshamsi, Alessandro Cappelli, Ruxandra Cojocaru, Merouane Debbah, Etienne Goffinet, Daniel Heslow, Julien Launay, Quentin Malartic, Badreddine Noune, Baptiste Pannier, and Guilherme Penedo. 2023 · 2023
Closest in time.
Pythia: A Suite for Analyzing Large Language Models Across Training and Scaling
Stella Biderman, Hailey Schoelkopf, Quentin Anthony, Herbie Bradley, Kyle O’Brien, Eric Hallahan, Mohammad Aflah Khan, Shivanshu Purohit, USVSN Sai Prashanth, Edward Raff, Aviya Skowron, Lintang Sutawika, and Oskar van der Wal. 2023 · 2023
Closest in time.
Say What You Mean! Large Language Models Speak Too Positively about Negative Commonsense Knowledge
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Compositional Generalization in Semantic Parsing: Pre-training vs. Specialized Architectures
Daniel Furrer, Marc van Zee, Nathan Scales, and Nathanael Schärli. 2021 · 2021
Cited alongside, same era.
Revisiting Negation in Neural Machine Translation
Gongbo Tang, Philipp Rönchen, Rico Sennrich, and Joakim Nivre. 2021 · 2021
Cited alongside, same era.
UnCommonSense: Informative Negative Knowledge about Everyday Concepts
Hiba Arnaout, Simon Razniewski, Gerhard Weikum, and Jeff Z. Pan. 2022 · 2022
Cited alongside, same era.
Scaling instruction-finetuned language models
Hyung Won Chung, Le Hou, Shayne Longpre, Barret Zoph, Yi Tay, William Fedus, Eric Li, Xuezhi Wang, Mostafa Dehghani, Siddhartha Brahma, Albert Webson, Shixiang Shane Gu, Zhuyun Dai, Mirac Suzgun, Xinyun Chen, Aakanksha Chowdhery, Sharan Narang, Gaurav Mishra, Adams Yu, Vincent Y. Zhao, Yanping Huang, Andrew M. Dai, Hongkun Yu, Slav Petrov, Ed H. Chi, Jeff Dean, Jacob Devlin, Adam Roberts, Denny Zhou, Quoc V. Le, and Jason Wei. 2022 · 2022
Cited alongside, same era.
LLM.int8(): 8-bit Matrix Multiplication for Transformers at Scale
Tim Dettmers, Mike Lewis, Younes Belkada, and Luke Zettlemoyer. 2022 · 2022
Cited alongside, same era.
An Analysis of Negation in Natural Language Understanding Corpora
Md Mosharaf Hossain, Dhivya Chinnappa, and Eduardo Blanco. 2022 · 2022
Cited alongside, same era.
Jiangjie Chen, Wei Shi, Ziquan Fu, Sijie Cheng, Lei Li, and Yanghua Xiao. 2023 · 2023
Closest in time.
Vicuna: An Open-Source Chatbot Impressing GPT-4 with 90%* ChatGPT Quality
Wei-Lin Chiang, Zhuohan Li, Zi Lin, Ying Sheng, Zhanghao Wu, Hao Zhang, Lianmin Zheng, Siyuan Zhuang, Yonghao Zhuang, Joseph E. Gonzalez, Ion Stoica, and Eric P. Xing. 2023 · 2023
Closest in time.
Free Dolly: Introducing the World’s First Truly Open Instruction-Tuned LLM
Mike Conover, Matt Hayes, Ankit Mathur, Xiangrui Meng, Jianwei Xie, Jun Wan, Sam Shah, Ali Ghodsi, Patrick Wendell, Matei Zaharia, and Reynold Xin. 2023 · 2023
Closest in time.
QLoRA: Efficient Finetuning of Quantized LLMs
Tim Dettmers, Artidoro Pagnoni, Ari Holtzman, and Luke Zettlemoyer. 2023 · 2023
Closest in time.
Training Language Models with Language Feedback at Scale
Jérémy Scheurer, Jon Ander Campos, Tomasz Korbak, Jun Shern Chan, Angelica Chen, Kyunghyun Cho, and Ethan Perez. 2023 · 2023
Closest in time.
LLaMA: Open and Efficient Foundation Language Models
Hugo Touvron, Thibaut Lavril, Gautier Izacard, Xavier Martinet, Marie-Anne Lachaux, Timothée Lacroix, Baptiste Rozière, Naman Goyal, Eric Hambro, Faisal Azhar, Aurélien Rodriguez, Armand Joulin, Edouard Grave, and Guillaume Lample. 2023 · 2023
Closest in time.
WizardLM: Empowering Large Language Models to Follow Complex Instructions
Can Xu, Qingfeng Sun, Kai Zheng, Xiubo Geng, Pu Zhao, Jiazhan Feng, Chongyang Tao, and Daxin Jiang. 2023 · 2023
Closest in time.