Fetching the paper…
Reading the bibliography…
Generative AI offers a simple, prompt-based alternative to fine-tuning smaller BERT-style LLMs for text classification tasks.
“Language models are few-shot learners”
Tom Brown et al · 1901
Earlier work this paper cites.
“How much is majority status in the US Congress worth?”
Gary. Cox and Eric Magar · 1999
Earlier work this paper cites.
“Jurisprudential regimes in Supreme Court decision making”
Mark. Richards and Herbert. Kritzer · 2002
Earlier work this paper cites.
“Feature hashing for large scale multitask learning”
Kilian Weinberger et al · 2009
Earlier work this paper cites.
“Distributed representations of words and phrases and their compositionality”
Tomas Mikolov et al · 2013
Earlier work this paper cites.
“Efficient estimation of word representations in vector space”
Tomas Mikolov, Kai Chen, Greg Corrado and Jeffrey Dean · 2013
Earlier work this paper cites.
“Glove: Global vectors for word representation”
Jeffrey Pennington, Richard Socher and Christopher Manning · 2014
Earlier work this paper cites.
“Improving neural machine translation models with monolingual data”
Rico Sennrich, Barry Haddow and Alexandra Birch · 2015
Earlier work this paper cites.
“Enriching word vectors with subword information”
Piotr Bojanowski, Edouard Grave, Armand Joulin and Tomas Mikolov · 2017
Earlier work this paper cites.
“Outrageously large neural networks: The sparsely-gated mixture-of-experts layer”
Noam Shazeer et al · 2017
Earlier work this paper cites.
“Attention is all you need”
Ashish Vaswani et al · 2017
Earlier work this paper cites.
“Bert: Pre-training of deep bidirectional transformers for language understanding”
Jacob Devlin, Ming-Wei Chang, Kenton Lee and Kristina Toutanova · 2018
Earlier work this paper cites.
“Improving language understanding by generative pre-training”
Alec Radford, Karthik Narasimhan, Tim Salimans and Ilya Sutskever · 2018
Earlier work this paper cites.
Mike Lewis et al · 2019
Earlier work this paper cites.
“Roberta: A robustly optimized bert pretraining approach”
Yinhan Liu et al · 2019
Earlier work this paper cites.
“Eda: Easy data augmentation techniques for boosting performance on text classification tasks”
Jason Wei and Kai Zou · 2019
Earlier work this paper cites.
“Xlnet: Generalized autoregressive pretraining for language understanding”
Zhilin Yang et al · 2019
Cited alongside, same era.
“Using Word Order in Political Text Classification with Long Short-term Memory Models”
Charles Chang and Michael Masterson · 2020
Cited alongside, same era.
“Electra: Pre-training text encoders as discriminators rather than generators”
Kevin Clark, Minh-Thang Luong, Quoc Le and Christopher Manning · 2020
Cited alongside, same era.
“Don’t stop pretraining: Adapt language models to domains and tasks”
Suchin Gururangan et al · 2020
Cited alongside, same era.
“Active learning approaches for labeling text: review and assessment of the performance of active learning approaches”
Blake Miller, Fridolin Linder and Walter. Mebane · 2020
Cited alongside, same era.
“Emotion and reason in political language”
Gloria Gennaro and Elliott Ash · 2022
Later among the works it cites.
“Training language models to follow instructions with human feedback”
Long Ouyang et al · 2022
Later among the works it cites.
“Tweet emotion dynamics: Emotion word usage in tweets from US and Canada”
Krishnapriya Vishnubhotla and Saif. Mohammad · 2022
Later among the works it cites.
“Creating and comparing dictionary, word embedding, and transformer-based models to measure discrete emotions in german political text”
Tobias Widmann and Maximilian Wich · 2022
Later among the works it cites.
“Do we still need BERT in the age of GPT? Comparing the benefits of domain-adaptation and in-context-learning approaches to using LLMs for Political Science Research”, 2023
Mitchell Bosley, Musashi Jacobs-Harukawa, Hauke Licht and Alexander Hoyle · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
“Transformers: State-of-the-art natural language processing”
Thomas Wolf et al · 2020
Cited alongside, same era.
“Automated text classification of news articles: A practical guide”
Pablo Barberá et al · 2021
Cited alongside, same era.
“Better few-shot text classification with pre-trained language model”
Zheng Chen and Yunchen Zhang · 2021
Cited alongside, same era.
“Whose news? Class-biased economic reporting in the United States”
Alan. Jacobs, J. Matthews, Timothy Hicks and Eric Merkley · 2021
Cited alongside, same era.
“Playing to the gallery: Emotive rhetoric in parliaments”
Moritz Osnabrügge, Sara. Hobolt and Toni Rodon · 2021
Cited alongside, same era.
“Fake it ‘til you make it: A natural experiment to identify European politicians’ benefit from Twitter bots”
Bruno Silva and Sven-Oliver Proksch · 2021
Cited alongside, same era.
“Finetuned language models are zero-shot learners”
Jason Wei et al · 2021
Cited alongside, same era.
“Large language models for text classification: From zero-shot learning to fine-tuning”
Youngjin Chae and Thomas Davidson · 2023
Later among the works it cites.
“Palm-e: An embodied multimodal language model”
Danny Driess et al · 2023
Later among the works it cites.
“ChatGPT outperforms crowd workers for text-annotation tasks”
Fabrizio Gilardi, Meysam Alizadeh and Maël Kubli · 2023
Later among the works it cites.
Suriya Gunasekar et al · 2023
Later among the works it cites.
“Less Annotating, More Classifying: Addressing the Data Scarcity Issue of Supervised Machine Learning with Deep Transfer Learning and BERT-NLI”
Moritz Laurer, Wouter van Atteveldt, Andreu Casas and Kasper Welbers · 2023
Later among the works it cites.
“Learning from precedent: how the British Brexit experience shapes nationalist rhetoric outside the UK”
Marco Martini and Stefanie Walter · 2023
Later among the works it cites.
“Are emergent abilities of Large Language Models a mirage?”
Rylan Schaeffer, Brando Miranda and Sanmi Koyejo · 2023
Later among the works it cites.
“Can chatgpt understand too? a comparative study on chatgpt and fine-tuned bert”
Qihuang Zhong et al · 2023
Later among the works it cites.
“The Claude 3 Model Family: Opus, Sonnet, Haiku”, 2024
Anthropic · 2024
Closest in time.
“Language Models for Text Classification: Is In-Context Learning Enough?”
Aleksandra Edwards and Jose Camacho-Collados · 2024
Closest in time.