Fetching the paper…
Reading the bibliography…
A key component of generating text from modern language models (LM) is the selection and tuning of decoding algorithms.
“Language models are few-shot learners”
Tom Brown et al · 1901
Earlier work this paper cites.
“On information and sufficiency”
Solomon Kullback and Richard Leibler · 1951
Earlier work this paper cites.
“The Kolmogorov-Smirnov test for goodness of fit”
Frank Massey · 1951
Earlier work this paper cites.
“A learning algorithm for Boltzmann machines”
David Ackley, Geoffrey Hinton and Terrence Sejnowski · 1985
Earlier work this paper cites.
“Generating sequences with recurrent neural networks”
Alex Graves · 2013
Earlier work this paper cites.
“Neural machine translation by jointly learning to align and translate”
Dzmitry Bahdanau, Kyunghyun Cho and Yoshua Bengio · 2014
Earlier work this paper cites.
“Model inversion attacks that exploit confidence information and basic countermeasures”
Matt Fredrikson, Somesh Jha and Thomas Ristenpart · 2015
Earlier work this paper cites.
“Abstractive text summarization using sequence-to-sequence rnns and beyond”
Ramesh Nallapati, Bowen Zhou, Caglar Gulcehre and Bing Xiang · 2016
Earlier work this paper cites.
“Stealing machine learning models via prediction { \{ APIs } \} ”
Florian Tramèr et al · 2016
Earlier work this paper cites.
“Diverse beam search: Decoding diverse solutions from neural sequence models”
Ashwin Vijayakumar et al · 2016
Earlier work this paper cites.
“Controlling linguistic style aspects in neural language generation”
Jessica Ficler and Yoav Goldberg · 2017
Earlier work this paper cites.
“Badnets: Identifying vulnerabilities in the machine learning model supply chain”
Tianyu Gu, Brendan Dolan-Gavitt and Siddharth Garg · 2017
Earlier work this paper cites.
Louis Shao et al · 2017
Earlier work this paper cites.
“Membership inference attacks against machine learning models”
Reza Shokri, Marco Stronati, Congzheng Song and Vitaly Shmatikov · 2017
Earlier work this paper cites.
“Attention is All You Need”, 2017
Ashish Vaswani et al · 2017
Earlier work this paper cites.
“Neural text generation in stories using entity representations as context”
Elizabeth Clark, Yangfeng Ji and Noah Smith · 2018
Earlier work this paper cites.
“Bert: Pre-training of deep bidirectional transformers for language understanding”
Jacob Devlin, Ming-Wei Chang, Kenton Lee and Kristina Toutanova · 2018
Earlier work this paper cites.
“Hierarchical neural story generation”
Angela Fan, Mike Lewis and Yann Dauphin · 2018
Earlier work this paper cites.
“Learning to write with cooperative discriminators”
Ari Holtzman et al · 2018
Earlier work this paper cites.
“Sharp Nearby, Fuzzy Far Away: How Neural Language Models Use Context”
Urvashi Khandelwal, He He, Peng Qi and Dan Jurafsky · 2018
Cited alongside, same era.
“Towards Reverse-Engineering Black-Box Neural Networks”
Seong Oh, Max Augustin, Mario Fritz and Bernt Schiele · 2018
Cited alongside, same era.
“Stealing hyperparameters in machine learning”
Binghui Wang and Neil Gong · 2018
Cited alongside, same era.
“Unifying human and statistical evaluation for natural language generation”
Tatsunori Hashimoto, Hugh Zhang and Percy Liang · 2019
Cited alongside, same era.
“The curious case of neural text degeneration”
Ari Holtzman et al · 2019
Cited alongside, same era.
“Model reconstruction from model explanations”
“Scarecrow: A framework for scrutinizing machine text”
Yao Dou et al · 2021
Later among the works it cites.
“Model extraction and adversarial transferability, your bert is vulnerable!”
Xuanli He, Lingjuan Lyu, Qiongkai Xu and Lichao Sun · 2021
Later among the works it cites.
“Backdoor attacks on pre-trained models by layerwise weight poisoning”
Linyang Li et al · 2021
Later among the works it cites.
“Killing two birds with one stone: Stealing model and inferring attribute from bert-based apis”
Lingjuan Lyu, Xuanli He, Fangzhao Wu and Lichao Sun · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Smitha Milli, Ludwig Schmidt, Anca Dragan and Moritz Hardt · 2019
Cited alongside, same era.
“Knockoff nets: Stealing functionality of black-box models”
Tribhuvanesh Orekondy, Bernt Schiele and Mario Fritz · 2019
Cited alongside, same era.
“Language models are unsupervised multitask learners”
Alec Radford et al · 2019
Cited alongside, same era.
“Do Massively Pretrained Language Models Make Better Storytellers?”
Abigail See et al · 2019
Cited alongside, same era.
“Adversarial neural network inversion via auxiliary knowledge alignment”
Ziqi Yang, Ee-Chien Chang and Zhenkai Liang · 2019
Cited alongside, same era.
“Language GANs Falling Short”
Massimo Caccia et al · 2020
Cited alongside, same era.
“Evaluation of text generation: A survey”
Asli Celikyilmaz, Elizabeth Clark and Jianfeng Gao · 2020
Cited alongside, same era.
Saeed Mahloujifar et al · 2021
Later among the works it cites.
“Mauve: Measuring the gap between neural text and human text using divergence frontiers”
Krishna Pillutla et al · 2021
Later among the works it cites.
“Membership inference attacks against nlp classification models”
Virat Shejwalkar, Huseyin Inan, Amir Houmansadr and Robert Sim · 2021
Later among the works it cites.
“Backdoor pre-trained models can transfer to all”
Lujia Shen et al · 2021
Later among the works it cites.
“Dawn: Dynamic adversarial watermarking of neural networks”
Sebastian Szyller, Buse Atli, Samuel Marchal and N Asokan · 2021
Later among the works it cites.
“Trojaning language models for fun and profit”
Xinyang Zhang, Zheng Zhang, Shouling Ji and Ting Wang · 2021
Later among the works it cites.
“GPT-Neo: Large Scale Autoregressive Language Modeling with Mesh-Tensorflow, 2021”
Sid Black et al · 2022
Later among the works it cites.
“Repairing the cracked foundation: A survey of obstacles in evaluation practices for generated text”
Sebastian Gehrmann, Elizabeth Clark and Thibault Sellam · 2022
Later among the works it cites.
“Truncation Sampling as Language Model Desmoothing”
John Hewitt, Christopher Manning and Percy Liang · 2022
Later among the works it cites.
“RankGen: Improving Text Generation with Large Ranking Models”
Kalpesh Krishna, Yapei Chang, John Wieting and Mohit Iyyer · 2022
Later among the works it cites.
“Contrastive decoding: Open-ended text generation as optimization”
Xiang Li et al · 2022
Later among the works it cites.
“A conversational paradigm for program synthesis”
Erik Nijkamp et al · 2022
Later among the works it cites.
“Contrastive search is what you need for neural text generation”
Yixuan Su and Nigel Collier · 2022
Later among the works it cites.
“Opt: Open pre-trained transformer language models”
Susan Zhang et al · 2022
Later among the works it cites.
“Reverse-Engineering Decoding Strategies Given Blackbox Access to a Language Generation System”
Daphne Ippolito et al · 2023
Closest in time.