Fetching the paper…
Reading the bibliography…
Despite the remarkable advances in language modeling, current mainstream decoding methods still struggle to generate texts that align with human texts across different aspects.
Way off-policy batch deep reinforcement learning of implicit human preferences in dialog
Natasha Jaques, Asma Ghandeharioun, Judy Hanwen Shen, Craig Ferguson, Àgata Lapedriza, Noah Jones, Shixiang Gu, and Rosalind W. Picard · 1907
Earlier work this paper cites.
Roberta: A robustly optimized BERT pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov · 1907
Earlier work this paper cites.
CTRL: A conditional transformer language model for controllable generation
Nitish Shirish Keskar, Bryan McCann, Lav R. Varshney, Caiming Xiong, and Richard Socher · 1909
Earlier work this paper cites.
Fine-tuning language models from human preferences
Daniel M. Ziegler, Nisan Stiennon, Jeffrey Wu, Tom B. Brown, Alec Radford, Dario Amodei, Paul F. Christiano, and Geoffrey Irving · 1909
Earlier work this paper cites.
Distributional reinforcement learning for energy-based sequential models
Tetiana Parshakova, Jean-Marc Andreoli, and Marc Dymetman · 1912
Earlier work this paper cites.
Prediction and entropy of printed english
Claude E. Shannon · 1951
Earlier work this paper cites.
Equation of state calculations by fast computing machines
Nicholas C. Metropolis, Arianna W. Rosenbluth, Marshall N. Rosenbluth, and A. H. Teller · 1953
Earlier work this paper cites.
Stochastic relaxation, gibbs distributions, and the bayesian restoration of images
Stuart Geman and Donald Geman · 1984
Earlier work this paper cites.
Using the sir algorithm to simulate posterior distributions
RUBIN DB · 1988
Earlier work this paper cites.
Bayesian inference in econometric models using monte carlo integration
John Geweke · 1989
Earlier work this paper cites.
Bayesian statistics without tears: a sampling–resampling perspective
Adrian FM Smith and Alan E Gelfand · 1992
Earlier work this paper cites.
Weighted average importance sampling and defensive mixture distributions
Tim Hesterberg · 1995
Earlier work this paper cites.
A maximum entropy approach to adaptive statistical language modelling
R ROSENFELD · 1996
Earlier work this paper cites.
Information projections revisited
Imre Csiszár and Frantisek Matús · 2000
Earlier work this paper cites.
Whole-sentence exponential language models: a vehicle for linguistic-statistical integration
Ronald Rosenfeld, Stanley F. Chen, and Xiaojin Zhu · 2001
Earlier work this paper cites.
Training products of experts by minimizing contrastive divergence
Geoffrey E Hinton · 2002
Earlier work this paper cites.
Improved sampling-importance resampling and reduced bias importance sampling
Øivind Skare, Erik Bølviken, and Lars Holden · 2003
Earlier work this paper cites.
Trading off diversity and quality in natural language generation
Hugh Zhang, Daniel Duckworth, Daphne Ippolito, and Arvind Neelakantan · 2004
Earlier work this paper cites.
Pattern recognition and machine learning , volume 4
Christopher M Bishop and Nasser M Nasrabadi · 2006
Earlier work this paper cites.
A tutorial on energy-based learning
Yann LeCun, Sumit Chopra, Raia Hadsell, M Ranzato, and Fujie Huang · 2006
Earlier work this paper cites.
Posterior regularization for structured latent variable models
Kuzman Ganchev, João Graça, Jennifer Gillenwater, and Ben Taskar · 2010
Earlier work this paper cites.
Noise-contrastive estimation of unnormalized statistical models, with applications to natural image statistics
Michael Gutmann and Aapo Hyvärinen · 2012
Earlier work this paper cites.
A survey of forecast error measures
MV Shcherbakov, A Brebels, NL Shcherbakova, AP Tyukov, TA Janovsky, and VAe Kamaev · 2013
Cited alongside, same era.
How (not) to train your generative model: Scheduled sampling, likelihood, adversary?
Ferenc Huszar · 2015
Cited alongside, same era.
Sequence level training with recurrent neural networks
Marc’Aurelio Ranzato, Sumit Chopra, Michael Auli, and Wojciech Zaremba · 2016
Cited alongside, same era.
Minimum risk training for neural machine translation
Shiqi Shen, Yong Cheng, Zhongjun He, Wei He, Hua Wu, Maosong Sun, and Yang Liu · 2016
Cited alongside, same era.
Pointer sentinel mixture models
Stephen Merity, Caiming Xiong, James Bradbury, and Richard Socher · 2017
Cited alongside, same era.
Energy-based reranking: Improving neural machine translation using energy-based models
Sumanta Bhattacharyya, Amirmohammad Rooshenas, Subhajit Naskar, Simeng Sun, Mohit Iyyer, and Andrew McCallum · 2021
Later among the works it cites.
Relating neural text degeneration to exposure bias
Ting-Rui Chiang and Yun-Nung Chen · 2021
Later among the works it cites.
A theoretical analysis of the repetition problem in text generation
Zihao Fu, Wai Lam, Anthony Man-Cho So, and Bei Shi · 2021
Later among the works it cites.
Simcse: Simple contrastive learning of sentence embeddings
Tianyu Gao, Xingcheng Yao, and Danqi Chen · 2021
Later among the works it cites.
Discodvt: Generating long text with discourse-aware discrete variational transformer
Haozhe Ji and Minlie Huang · 2021
Later among the works it cites.
A distributional approach to controlled text generation
Muhammad Khalifa, Hady Elsahar, and Marc Dymetman · 2021
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Lantao Yu, Weinan Zhang, Jun Wang, and Yong Yu · 2017
Cited alongside, same era.
Hierarchical neural story generation
Angela Fan, Mike Lewis, and Yann N. Dauphin · 2018
Cited alongside, same era.
Long text generation via adversarial training with leaked information
Jiaxian Guo, Sidi Lu, Han Cai, Weinan Zhang, Yong Yu, and Jun Wang · 2018
Cited alongside, same era.
Noise contrastive estimation and negative sampling for conditional models: Consistency and statistical efficiency
Zhuang Ma and Michael Collins · 2018
Cited alongside, same era.
Toward diverse text generation with inverse reinforcement learning
Zhan Shi, Xinchi Chen, Xipeng Qiu, and Xuanjing Huang · 2018
Cited alongside, same era.
Reverse kl-divergence training of prior networks: Improved uncertainty and adversarial robustness
Andrey Malinin and Mark Gales · 2019
Cited alongside, same era.
Global autoregressive models for data-efficient sequence learning
Tetiana Parshakova, Jean-Marc Andreoli, and Marc Dymetman · 2019
Cited alongside, same era.
Later among the works it cites.
Straight to the gradient: Learning to use novel tokens for neural text generation
Xiang Lin, Simeng Han, and Shafiq R. Joty · 2021
Later among the works it cites.
Text generation by learning from demonstrations
Richard Yuanzhe Pang and He He · 2021
Later among the works it cites.
An information divergence measure between neural text and human text
Krishna Pillutla, Swabha Swayamdipta, Rowan Zellers, John Thickstun, Sean Welleck, Yejin Choi, and Zaid Harchaoui · 2021
Later among the works it cites.
Why exposure bias matters: An imitation learning perspective of error accumulation in language generation
Kushal Arora, Layla El Asri, Hareesh Bahuleyan, and Jackie Chi Kit Cheung · 2022
Later among the works it cites.
Greedification operators for policy optimization: Investigating forward and reverse kl divergences
Alan Chan, Hugo Silva, Sungsu Lim, Tadashi Kozuno, A Rupam Mahmood, and Martha White · 2022
Later among the works it cites.
Is GPT-3 text indistinguishable from human text? scarecrow: A framework for scrutinizing machine text
Yao Dou, Maxwell Forbes, Rik Koncel-Kedziorski, Noah A. Smith, and Yejin Choi · 2022
Later among the works it cites.
Quality-aware decoding for neural machine translation
Patrick Fernandes, António Farinhas, Ricardo Rei, José GC de Souza, Perez Ogayo, Graham Neubig, and André FT Martins · 2022
Later among the works it cites.
High quality rather than high model probability: Minimum Bayes risk decoding with neural metrics
Markus Freitag, David Grangier, Qijun Tan, and Bowen Liang · 2022
Later among the works it cites.
Contrastive decoding: Open-ended text generation as optimization
Xiang Lisa Li, Ari Holtzman, Daniel Fried, Percy Liang, Jason Eisner, Tatsunori Hashimoto, Luke Zettlemoyer, and Mike Lewis · 2022
Later among the works it cites.
Typical decoding for natural language generation
Clara Meister, Tiago Pimentel, Gian Wiher, and Ryan Cotterell · 2022
Later among the works it cites.
Training language models to follow instructions with human feedback
Long Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida, Carroll Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Gray, John Schulman, Jacob Hilton, Fraser Kelton, Luke Miller, Maddie Simens, Amanda Askell, Peter Welinder, Paul Christiano, Jan Leike, and Ryan Lowe · 2022
Later among the works it cites.
An empirical study on contrastive search and contrastive decoding for open-ended text generation
Yixuan Su and Jialu Xu · 2022
Later among the works it cites.
A contrastive framework for neural text generation
Yixuan Su, Tian Lan, Yan Wang, Dani Yogatama, Lingpeng Kong, and Nigel Collier · 2022
Later among the works it cites.
Learning to break the loop: Analyzing and mitigating repetitions for neural text generation
Jin Xu, Xiaojiang Liu, Jianhao Yan, Deng Cai, Huayang Li, and Jian Li · 2022
Later among the works it cites.
OPT: open pre-trained transformer language models
Susan Zhang, Stephen Roller, Naman Goyal, Mikel Artetxe, Moya Chen, Shuohui Chen, Christopher Dewan, Mona T. Diab, Xian Li, Xi Victoria Lin, Todor Mihaylov, Myle Ott, Sam Shleifer, Kurt Shuster, Daniel Simig, Punit Singh Koura, Anjali Sridhar, Tianlu Wang, and Luke Zettlemoyer · 2022
Later among the works it cites.
Tailoring language generation models under total variation distance
Haozhe Ji, Pei Ke, Zhipeng Hu, Rongsheng Zhang, and Minlie Huang · 2023
Closest in time.