Fetching the paper…
Reading the bibliography…
The advent of pre-trained Language Models (LMs) has markedly advanced natural language processing, but their efficacy in out-of-distribution (OOD) scenarios remains a significant challenge.
Roberta: A robustly optimized bert pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019 · 1907
Earlier work this paper cites.
A new readability yardstick
Rudolph Flesch. 1948 · 1948
Earlier work this paper cites.
The uses of argument
Stephen E. Toulmin. 1960 · 1960
Earlier work this paper cites.
Classification and analysis of multivariate observations
J MacQueen. 1967 · 1967
Earlier work this paper cites.
Least squares quantization in PCM
Stuart P. Lloyd. 1982 · 1982
Earlier work this paper cites.
Comparing partitions
Lawrence Hubert and Phipps Arabie. 1985 · 1985
Earlier work this paper cites.
Convex optimization
Stephen P Boyd and Lieven Vandenberghe. 2004 · 2004
Earlier work this paper cites.
On Rhetoric: A Theory of Civic Discourse
Aristotle and George A. Kennedy. ca. 350 B.C.E., translated 2007 · 2007
Earlier work this paper cites.
Biographies, Bollywood, boom-boxes and blenders: Domain adaptation for sentiment classification
John Blitzer, Mark Dredze, and Fernando Pereira. 2007 · 2007
Earlier work this paper cites.
Cross-language text classification using structural correspondence learning
Peter Prettenhofer and Benno Stein. 2010 · 2010
Earlier work this paper cites.
A corpus for research on deliberation and debate
Marilyn Walker, Jean Fox Tree, Pranav Anand, Rob Abbott, and Joseph King. 2012 · 2012
Earlier work this paper cites.
Emergent: a novel data-set for stance classification
William Ferreira and Andreas Vlachos. 2016 · 2016
Earlier work this paper cites.
Argumentation mining: State of the art and emerging trends
Marco Lippi and Paolo Torroni. 2016 · 2016
Earlier work this paper cites.
SemEval-2016 task 6: Detecting stance in tweets
Saif Mohammad, Svetlana Kiritchenko, Parinaz Sobhani, Xiaodan Zhu, and Colin Cherry. 2016 · 2016
Earlier work this paper cites.
Argumentation mining in user-generated web discourse
Ivan Habernal and Iryna Gurevych. 2017 · 2017
Earlier work this paper cites.
Reporting score distributions makes a difference: Performance study of LSTM-networks for sequence tagging
Nils Reimers and Iryna Gurevych. 2017 · 2017
Earlier work this paper cites.
Cross-lingual argumentation mining: Machine translation (and a bit of projection) is all you need!
Steffen Eger, Johannes Daxenberger, Christian Stab, and Iryna Gurevych. 2018 · 2018
Earlier work this paper cites.
Scitail: A textual entailment dataset from science question answering
Tushar Khot, Ashish Sabharwal, and Peter Clark. 2018 · 2018
Earlier work this paper cites.
Cross-domain sentiment classification with target domain specific information
Minlong Peng, Qi Zhang, Yu-gang Jiang, and Xuanjing Huang. 2018 · 2018
Earlier work this paper cites.
Will it blend? blending weak and strong labeled data in a neural network for argumentation mining
Eyal Shnarch, Carlos Alzate, Lena Dankin, Martin Gleize, Yufang Hou, Leshem Choshen, Ranit Aharonov, and Noam Slonim. 2018 · 2018
Earlier work this paper cites.
Cross-topic argument mining from heterogeneous sources
Christian Stab, Tristan Miller, Benjamin Schiller, Pranav Rai, and Iryna Gurevych. 2018 · 2018
Earlier work this paper cites.
GLUE: A multi-task benchmark and analysis platform for natural language understanding
Alex Wang, Amanpreet Singh, Julian Michael, Felix Hill, Omer Levy, and Samuel Bowman. 2018 · 2018
Earlier work this paper cites.
Cross-target stance classification with self-attention networks
Chang Xu, Cécile Paris, Surya Nepal, and Ross Sparks. 2018 · 2018
Earlier work this paper cites.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Earlier work this paper cites.
Using pre-training can improve model robustness and uncertainty
Dan Hendrycks, Kimin Lee, and Mantas Mazeika. 2019 · 2019
Cited alongside, same era.
Decoupled weight decay regularization
Ilya Loshchilov and Frank Hutter. 2019 · 2019
Cited alongside, same era.
Right for the wrong reasons: Diagnosing syntactic heuristics in natural language inference
Tom McCoy, Ellie Pavlick, and Tal Linzen. 2019 · 2019
Cited alongside, same era.
Sentence-BERT: Sentence embeddings using Siamese BERT-networks
Nils Reimers and Iryna Gurevych. 2019 · 2019
Cited alongside, same era.
Classification and clustering of arguments with contextualized word embeddings
Nils Reimers, Benjamin Schiller, Tilman Beck, Johannes Daxenberger, Christian Stab, and Iryna Gurevych. 2019 · 2019
Cited alongside, same era.
Automatic argument quality assessment - new datasets and methods
Assaf Toledo, Shai Gretz, Edo Cohen-Karlik, Roni Friedman, Elad Venezian, Dan Lahav, Michal Jacovi, Ranit Aharonov, and Noam Slonim. 2019 · 2019
Deberta: decoding-enhanced bert with disentangled attention
Pengcheng He, Xiaodong Liu, Jianfeng Gao, and Weizhu Chen. 2021b · 2021
Later among the works it cites.
Oodformer: Out-of-distribution detection transformer
Rajat Koner, Poulami Sinhamahapatra, Karsten Roscher, Stephan Günnemann, and Volker Tresp. 2021 · 2021
Later among the works it cites.
The power of scale for parameter-efficient prompt tuning
Brian Lester, Rami Al-Rfou, and Noah Constant. 2021 · 2021
Later among the works it cites.
On the stability of fine-tuning BERT: misconceptions, explanations, and strong baselines
Marius Mosbach, Maksym Andriushchenko, and Dietrich Klakow. 2021 · 2021
Later among the works it cites.
It’s not just size that matters: Small language models are also few-shot learners
Timo Schick and Hinrich Schütze. 2021b · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Handbook of argumentation theory: A critical survey of classical backgrounds and modern studies , volume 7
Frans H Van Eemeren, Rob Grootendorst, and Tjark Kruiger. 2019 · 2019
Cited alongside, same era.
Zero-Shot Stance Detection: A Dataset and Model using Generalized Topic Representations
Emily Allaway and Kathleen McKeown. 2020 · 2020
Cited alongside, same era.
ELECTRA: pre-training text encoders as discriminators rather than generators
Kevin Clark, Minh-Thang Luong, Quoc V. Le, and Christopher D. Manning. 2020 · 2020
Cited alongside, same era.
Is BERT really robust? A strong baseline for natural language attack on text classification and entailment
Di Jin, Zhijing Jin, Joey Tianyi Zhou, and Peter Szolovits. 2020 · 2020
Cited alongside, same era.
Cross-lingual ability of multilingual BERT: an empirical study
Karthikeyan K, Zihan Wang, Stephen Mayhew, and Dan Roth. 2020 · 2020
Cited alongside, same era.
Attention is not only a weight: Analyzing transformers with vector norms
Goro Kobayashi, Tatsuki Kuribayashi, Sho Yokoi, and Kentaro Inui. 2020 · 2020
Cited alongside, same era.
Zheyan Shen, Jiashuo Liu, Yue He, Xingxuan Zhang, Renzhe Xu, Han Yu, and Peng Cui. 2021 · 2021
Later among the works it cites.
An autonomous debating system
Noam Slonim, Yonatan Bilu, Carlos Alzate, Roy Bar-Haim, Ben Bogin, Francesca Bonin, Leshem Choshen, Edo Cohen-Karlik, Lena Dankin, Lilach Edelstein, et al. 2021 · 2021
Later among the works it cites.
Infobert: Improving robustness of language models from an information theoretic perspective
Boxin Wang, Shuohang Wang, Yu Cheng, Zhe Gan, Ruoxi Jia, Bo Li, and Jingjing Liu. 2021 · 2021
Later among the works it cites.
Scaling instruction-finetuned language models
Hyung Won Chung, Le Hou, Shayne Longpre, Barret Zoph, Yi Tay, William Fedus, Eric Li, Xuezhi Wang, Mostafa Dehghani, Siddhartha Brahma, et al. 2022 · 2022
Later among the works it cites.
Lora: Low-rank adaptation of large language models
Edward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen. 2022 · 2022
Later among the works it cites.
Fine-tuning can distort pretrained features and underperform out-of-distribution
Ananya Kumar, Aditi Raghunathan, Robbie Matthew Jones, Tengyu Ma, and Percy Liang. 2022 · 2022
Later among the works it cites.
Scientia potentia est—on the role of knowledge in computational argumentation
Anne Lauscher, Henning Wachsmuth, Iryna Gurevych, and Goran Glavaš. 2022 · 2022
Later among the works it cites.
JointCL: A joint contrastive learning framework for zero-shot stance detection
Bin Liang, Qinglin Zhu, Xiang Li, Min Yang, Lin Gui, Yulan He, and Ruifeng Xu. 2022 · 2022
Later among the works it cites.
Training language models to follow instructions with human feedback
Long Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida, Carroll L. Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, John Schulman, Jacob Hilton, Fraser Kelton, Luke Miller, Maddie Simens, Amanda Askell, Peter Welinder, Paul F. Christiano, Jan Leike, and Ryan Lowe. 2022 · 2022
Later among the works it cites.
On the effect of sample and topic sizes for argument mining datasets
Benjamin Schiller, Johannes Daxenberger, and Iryna Gurevych. 2022 · 2022
Later among the works it cites.
Out-of-distribution detection with deep nearest neighbors
Yiyou Sun, Yifei Ming, Xiaojin Zhu, and Yixuan Li. 2022 · 2022
Later among the works it cites.
Probing out-of-distribution robustness of language models with parameter-efficient transfer learning
Hyunsoo Cho, Choonghyun Park, Junyeop Kim, Hyuhng Joon Kim, Kang Min Yoo, and Sang goo Lee. 2023 · 2023
Closest in time.
A taxonomy and review of generalization research in nlp
Dieuwke Hupkes, Mario Giulianelli, Verna Dankers, Mikel Artetxe, Yanai Elazar, Tiago Pimentel, Christos Christodoulopoulos, Karim Lasri, Naomi Saphra, Arabella Sinclair, et al. 2023 · 2023
Closest in time.
Orca 2: Teaching small language models how to reason
Arindam Mitra, Luciano Del Corro, Shweti Mahajan, Andres Codas, Clarisse Simoes, Sahaj Agarwal, Xuxi Chen, Anastasia Razdaibiedina, Erik Jones, Kriti Aggarwal, et al. 2023 · 2023
Closest in time.
Model-tuning via prompts makes NLP models adversarially robust
Mrigank Raman, Pratyush Maini, J. Zico Kolter, Zachary C. Lipton, and Danish Pruthi. 2023 · 2023
Closest in time.
Llama 2: Open foundation and fine-tuned chat models
Hugo Touvron, Louis Martin, Kevin Stone, Peter Albert, Amjad Almahairi, Yasmine Babaei, Nikolay Bashlykov, Soumya Batra, Prajjwal Bhargava, Shruti Bhosale, et al. 2023 · 2023
Closest in time.
GLUE-X: Evaluating natural language understanding models from an out-of-distribution generalization perspective
Linyi Yang, Shuibai Zhang, Libo Qin, Yafu Li, Yidong Wang, Hanmeng Liu, Jindong Wang, Xing Xie, and Yue Zhang. 2023 · 2023
Closest in time.
Revisiting out-of-distribution robustness in nlp: Benchmark, analysis, and llms evaluations
Lifan Yuan, Yangyi Chen, Ganqu Cui, Hongcheng Gao, Fangyuan Zou, Xingyi Cheng, Heng Ji, Zhiyuan Liu, and Maosong Sun. 2023 · 2023
Closest in time.
Dive into the chasm: Probing the gap between in- and cross-topic generalization
Andreas Waldis, Yufang Hou, and Iryna Gurevych. 2024 · 2024
Closest in time.