Fetching the paper…
Reading the bibliography…
Numerous recent techniques for text style transfer characterize their approaches as variants of reinforcement learning and preference optimization.
Transfertransfo: A transfer learning approach for neural network based conversational agents
Thomas Wolf, Victor Sanh, Julien Chaumond, and Clement Delangue. 2019 · 1901
Earlier work this paper cites.
Formality style transfer with hybrid textual annotations
Ruochen Xu, Tao Ge, and Furu Wei. 2019 · 1903
Earlier work this paper cites.
Roberta: A robustly optimized bert pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019b · 1907
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu. 2002 · 2002
Earlier work this paper cites.
Minimum error rate training in statistical machine translation
Franz Josef Och. 2003 · 2003
Earlier work this paper cites.
Online large-margin training of syntactic and structural translation features
David Chiang, Yuval Marton, and Philip Resnik. 2008 · 2008
Earlier work this paper cites.
A monolingual tree-based translation model for sentence simplification
Zhemin Zhu, Delphine Bernhard, and Iryna Gurevych. 2010 · 2010
Earlier work this paper cites.
Tuning as ranking
Mark Hopkins and Jonathan May. 2011 · 2011
Earlier work this paper cites.
Batch tuning strategies for statistical machine translation
Colin Cherry and George Foster. 2012 · 2012
Earlier work this paper cites.
Hope and fear for discriminative training of statistical translation models
David Chiang. 2012 · 2012
Earlier work this paper cites.
Proximal policy optimization algorithms
John Schulman, Filip Wolski, Prafulla Dhariwal, Alec Radford, and Oleg Klimov. 2017 · 2017
Earlier work this paper cites.
Dear sir or madam, may I introduce the GYAFC dataset: Corpus, benchmarks and metrics for formality style transfer
Sudha Rao and Joel Tetreault. 2018 · 2018
Earlier work this paper cites.
A4NT: Author attribute anonymity by adversarial training of neural machine translation
Rakshith Shetty, Bernt Schiele, and Mario Fritz. 2018 · 2018
Earlier work this paper cites.
ParaNMT-50M: Pushing the limits of paraphrastic sentence embeddings with millions of machine translations
John Wieting and Kevin Gimpel. 2018 · 2018
Earlier work this paper cites.
Reinforcement learning based text style transfer without parallel training corpus
Hongyu Gong, Suma Bhat, Lingfei Wu, JinJun Xiong, and Wen-mei Hwu. 2019 · 2019
Earlier work this paper cites.
Disentangled representation learning for non-parallel text style transfer
Vineet John, Lili Mou, Hareesh Bahuleyan, and Olga Vechtomova. 2019 · 2019
Earlier work this paper cites.
Multiple-attribute text rewriting
Guillaume Lample, Sandeep Subramanian, Eric Smith, Ludovic Denoyer, Marc’Aurelio Ranzato, and Y-Lan Boureau. 2019 · 2019
Earlier work this paper cites.
Sentence-BERT: Sentence embeddings using Siamese BERT-networks
Nils Reimers and Iryna Gurevych. 2019 · 2019
Earlier work this paper cites.
Semi-supervised text style transfer: Cross projection in latent space
Mingyue Shang, Piji Li, Zhenxin Fu, Lidong Bing, Dongyan Zhao, Shuming Shi, and Rui Yan. 2019 · 2019
Cited alongside, same era.
Harnessing pre-trained neural networks with rules for formality style transfer
Yunli Wang, Yu Wu, Lili Mou, Zhoujun Li, and Wenhan Chao. 2019 · 2019
Cited alongside, same era.
Neural network acceptability judgments
Alex Warstadt, Amanpreet Singh, and Samuel R. Bowman. 2019 · 2019
Cited alongside, same era.
On the weaknesses of reinforcement learning for neural machine translation
Leshem Choshen, Lior Fox, Zohar Aizenbud, and Omri Abend. 2020 · 2020
Cited alongside, same era.
Hooks in the headline: Learning to generate headlines with controlled styles
Di Jin, Zhijing Jin, Joey Tianyi Zhou, Lisa Orii, and Peter Szolovits. 2020 · 2020
Cited alongside, same era.
Reformulating unsupervised style transfer as paraphrase generation
Improving iterative text revision by learning where to edit from other revision tasks
Zae Myung Kim, Wanyu Du, Vipul Raheja, Dhruv Kumar, and Dongyeop Kang. 2022 · 2022
Later among the works it cites.
Semi-supervised formality style transfer with consistency training
Ao Liu, An Wang, and Naoaki Okazaki. 2022 · 2022
Later among the works it cites.
QUARK: Controllable text generation with reinforced unlearning
Ximing Lu, Sean Welleck, Jack Hessel, Liwei Jiang, Lianhui Qin, Peter West, Prithviraj Ammanabrolu, and Yejin Choi. 2022 · 2022
Later among the works it cites.
Training language models to follow instructions with human feedback
Long Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida, Carroll Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, et al. 2022 · 2022
Later among the works it cites.
A recipe for arbitrary text style transfer with large language models
Emily Reif, Daphne Ippolito, Ann Yuan, Andy Coenen, Chris Callison-Burch, and Jason Wei. 2022 · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Kalpesh Krishna, John Wieting, and Mohit Iyyer. 2020 · 2020
Cited alongside, same era.
Politeness transfer: A tag and generate approach
Aman Madaan, Amrith Setlur, Tanmay Parekh, Barnabas Poczos, Graham Neubig, Yiming Yang, Ruslan Salakhutdinov, Alan W Black, and Shrimai Prabhumoye. 2020 · 2020
Cited alongside, same era.
Unsupervised text style transfer with padded masked language models
Eric Malmi, Aliaksei Severyn, and Sascha Rothe. 2020 · 2020
Cited alongside, same era.
Parallel data augmentation for formality style transfer
Yi Zhang, Tao Ge, and Xu Sun. 2020 · 2020
Cited alongside, same era.
ER-AE: Differentially private text generation for authorship anonymization
Haohan Bo, Steven H. H. Ding, Benjamin C. M. Fung, and Farkhund Iqbal. 2021 · 2021
Cited alongside, same era.
Text detoxification using large pre-trained neural models
David Dale, Anton Voronov, Daryna Dementieva, Varvara Logacheva, Olga Kozlova, Nikita Semenov, and Alexander Panchenko. 2021 · 2021
Cited alongside, same era.
Revisiting the weaknesses of reinforcement learning for neural machine translation
Samuel Kiegeland and Julia Kreutzer. 2021 · 2021
Cited alongside, same era.
Prompt-and-rerank: A method for zero-shot and few-shot arbitrary textual style transfer with small language models
Mirac Suzgun, Luke Melas-Kyriazi, and Dan Jurafsky. 2022 · 2022
Later among the works it cites.
STEER: Unified style transfer with expert reinforcement
Skyler Hallinan, Faeze Brahman, Ximing Lu, Jaehun Jung, Sean Welleck, and Yejin Choi. 2023a · 2023
Later among the works it cites.
Prompt-based editing for text style transfer
Guoqing Luo, Yu Han, Lili Mou, and Mauajama Firdaus. 2023 · 2023
Later among the works it cites.
Low-resource authorship style transfer: Can non-famous authors be imitated?
Ajay Patel, Nicholas Andrews, and Chris Callison-Burch. 2023 · 2023
Later among the works it cites.
Direct preference optimization: Your language model is secretly a reward model
Rafael Rafailov, Archit Sharma, Eric Mitchell, Christopher D Manning, Stefano Ermon, and Chelsea Finn. 2023 · 2023
Later among the works it cites.
CoEdIT: Text editing by task-specific instruction tuning
Vipul Raheja, Dhruv Kumar, Ryan Koo, and Dongyeop Kang. 2023 · 2023
Later among the works it cites.
Llama 2: Open foundation and fine-tuned chat models
Hugo Touvron, Louis Martin, Kevin Stone, Peter Albert, Amjad Almahairi, Yasmine Babaei, Nikolay Bashlykov, Soumya Batra, Prajjwal Bhargava, Shruti Bhosale, Dan Bikel, Lukas Blecher, Cristian Canton Ferrer, Moya Chen, Guillem Cucurull, David Esiobu, Jude Fernandes, Jeremy Fu, Wenyin Fu, Brian Fuller, Cynthia Gao, Vedanuj Goswami, Naman Goyal, Anthony Hartshorn, Saghar Hosseini, Rui Hou, Hakan Inan, Marcin Kardas, Viktor Kerkez, Madian Khabsa, Isabel Kloumann, Artem Korenev, Punit Singh Koura, Marie-Anne Lachaux, Thibaut Lavril, Jenya Lee, Diana Liskovich, Yinghai Lu, Yuning Mao, Xavier Martinet, Todor Mihaylov, Pushkar Mishra, Igor Molybog, Yixin Nie, Andrew Poulton, Jeremy Reizenstein, Rashi Rungta, Kalyan Saladi, Alan Schelten, Ruan Silva, Eric Michael Smith, Ranjan Subramanian, Xiaoqing Ellen Tan, Binh Tang, Ross Taylor, Adina Williams, Jian Xiang Kuan, Puxin Xu, Zheng Yan, Iliyan Zarov, Yuchen Zhang, Angela Fan, Melanie Kambadur, Sharan Narang, Aurelien Rodriguez, Robert Stojnic, Sergey Edunov, and Thomas Scialom. 2023 · 2023
Later among the works it cites.
Self-play fine-tuning converts weak language models to strong language models
Zixiang Chen, Yihe Deng, Huizhuo Yuan, Kaixuan Ji, and Quanquan Gu. 2024 · 2024
Closest in time.
Authorship style transfer with policy optimization
Shuai Liu, Shantanu Agarwal, and Jonathan May. 2024 · 2024
Closest in time.
Iterative reasoning preference optimization
Richard Yuanzhe Pang, Weizhe Yuan, Kyunghyun Cho, He He, Sainbayar Sukhbaatar, and Jason Weston. 2024 · 2024
Closest in time.
Iterative preference learning from human feedback: Bridging theory and practice for rlhf under kl-constraint
Wei Xiong, Hanze Dong, Chenlu Ye, Ziqi Wang, Han Zhong, Heng Ji, Nan Jiang, and Tong Zhang. 2023 · 2024
Closest in time.
Self-rewarding language models
Weizhe Yuan, Richard Yuanzhe Pang, Kyunghyun Cho, Xian Li, Sainbayar Sukhbaatar, Jing Xu, and Jason Weston. 2024 · 2024
Closest in time.