Fetching the paper…
Reading the bibliography…
Learning from human feedback has been shown to improve text-to-image models.
Reinforcement learning by reward-weighted regression for operational space control
Peters, Jan and Schaal, Stefan · 2007
Earlier work this paper cites.
Learning multiple layers of features from tiny images
Krizhevsky, Alex, Hinton, Geoffrey, et al · 2009
Earlier work this paper cites.
Deep learning face attributes in the wild
Liu, Ziwei, Luo, Ping, Wang, Xiaogang, and Tang, Xiaoou · 2015
Earlier work this paper cites.
U-net: Convolutional networks for biomedical image segmentation
Ronneberger, Olaf, Fischer, Philipp, and Brox, Thomas · 2015
Earlier work this paper cites.
High-dimensional continuous control using generalized advantage estimation
Schulman, John, Moritz, Philipp, Levine, Sergey, Jordan, Michael, and Abbeel, Pieter · 2015
Earlier work this paper cites.
Deep unsupervised learning using nonequilibrium thermodynamics
Sohl-Dickstein, Jascha, Weiss, Eric, Maheswaranathan, Niru, and Ganguli, Surya · 2015
Earlier work this paper cites.
Proximal policy optimization algorithms
Schulman, John, Wolski, Filip, Dhariwal, Prafulla, Radford, Alec, and Klimov, Oleg · 2017
Earlier work this paper cites.
Generative adversarial networks
Goodfellow, Ian, Pouget-Abadie, Jean, Mirza, Mehdi, Xu, Bing, Warde-Farley, David, Ozair, Sherjil, Courville, Aaron, and Bengio, Yoshua · 2020
Earlier work this paper cites.
Denoising diffusion probabilistic models
Ho, Jonathan, Jain, Ajay, and Abbeel, Pieter · 2020
Earlier work this paper cites.
Diffwave: A versatile diffusion model for audio synthesis
Kong, Zhifeng, Ping, Wei, Huang, Jiaji, Zhao, Kexin, and Catanzaro, Bryan · 2020
Earlier work this paper cites.
Exploring the limits of transfer learning with a unified text-to-text transformer
Raffel, Colin, Shazeer, Noam, Roberts, Adam, Lee, Katherine, Narang, Sharan, Matena, Michael, Zhou, Yanqi, Li, Wei, and Liu, Peter J · 2020
Earlier work this paper cites.
Denoising diffusion implicit models
Song, Jiaming, Meng, Chenlin, and Ermon, Stefano · 2020
Earlier work this paper cites.
Improved techniques for training score-based generative models
Song, Yang and Ermon, Stefano · 2020
Earlier work this paper cites.
Learning to summarize with human feedback
Stiennon, Nisan, Ouyang, Long, Wu, Jeffrey, Ziegler, Daniel, Lowe, Ryan, Voss, Chelsea, Radford, Alec, Amodei, Dario, and Christiano, Paul F · 2020
Earlier work this paper cites.
Diffusion models beat gans on image synthesis
Dhariwal, Prafulla and Nichol, Alexander · 2021
Cited alongside, same era.
Lora: Low-rank adaptation of large language models
Hu, Edward J, Shen, Yelong, Wallis, Phillip, Allen-Zhu, Zeyuan, Li, Yuanzhi, Wang, Shean, Wang, Lu, and Chen, Weizhu · 2021
Cited alongside, same era.
Lee, Kimin, Smith, Laura, and Abbeel, Pieter · 2021
Cited alongside, same era.
Learning transferable visual models from natural language supervision
Radford, Alec, Kim, Jong Wook, Hallacy, Chris, Ramesh, Aditya, Goh, Gabriel, Agarwal, Sandhini, Sastry, Girish, Askell, Amanda, Mishkin, Pamela, Clark, Jack, et al · 2021
Cited alongside, same era.
Laion-400m: Open dataset of clip-filtered 400 million image-text pairs
Hierarchical text-conditional image generation with clip latents
Ramesh, Aditya, Dhariwal, Prafulla, Nichol, Alex, Chu, Casey, and Chen, Mark · 2022
Later among the works it cites.
High-resolution image synthesis with latent diffusion models
Rombach, Robin, Blattmann, Andreas, Lorenz, Dominik, Esser, Patrick, and Ommer, Björn · 2022
Later among the works it cites.
Photorealistic text-to-image diffusion models with deep language understanding
Saharia, Chitwan, Chan, William, Saxena, Saurabh, Li, Lala, Whang, Jay, Denton, Emily, Ghasemipour, Seyed Kamyar Seyed, Ayan, Burcu Karagol, Mahdavi, S Sara, Lopes, Rapha Gontijo, et al · 2022
Later among the works it cites.
Laion-5b: An open large-scale dataset for training next generation image-text models
Schuhmann, Christoph, Beaumont, Romain, Vencu, Richard, Gordon, Cade, Wightman, Ross, Cherti, Mehdi, Coombes, Theo, Katta, Aarush, Mullis, Clayton, Wortsman, Mitchell, et al · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Schuhmann, Christoph, Vencu, Richard, Beaumont, Romain, Kaczmarczyk, Robert, Mullis, Clayton, Katta, Aarush, Coombes, Theo, Jitsev, Jenia, and Komatsuzaki, Aran · 2021
Cited alongside, same era.
Maximum likelihood training of score-based diffusion models
Song, Yang, Durkan, Conor, Murray, Iain, and Ermon, Stefano · 2021
Cited alongside, same era.
Training a helpful and harmless assistant with reinforcement learning from human feedback
Bai, Yuntao, Jones, Andy, Ndousse, Kamal, Askell, Amanda, Chen, Anna, DasSarma, Nova, Drain, Dawn, Fort, Stanislav, Ganguli, Deep, Henighan, Tom, et al · 2022
Cited alongside, same era.
Training-free structured diffusion guidance for compositional text-to-image synthesis
Feng, Weixi, He, Xuehai, Fu, Tsu-Jui, Jampani, Varun, Akula, Arjun, Narayana, Pradyumna, Basu, Sugato, Wang, Xin Eric, and Wang, William Yang · 2022
Cited alongside, same era.
Benchmarking spatial relationships in text-to-image generation
Gokhale, Tejas, Palangi, Hamid, Nushi, Besmira, Vineet, Vibhav, Horvitz, Eric, Kamar, Ece, Baral, Chitta, and Yang, Yezhou · 2022
Cited alongside, same era.
Classifier-free diffusion guidance
Ho, Jonathan and Salimans, Tim · 2022
Cited alongside, same era.
Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation
Li, Junnan, Li, Dongxu, Xiong, Caiming, and Hoi, Steven · 2022
Cited alongside, same era.
Training language models to follow instructions with human feedback
Ouyang, Long, Wu, Jeffrey, Jiang, Xu, Almeida, Diogo, Wainwright, Carroll, Mishkin, Pamela, Zhang, Chong, Agarwal, Sandhini, Slama, Katarina, Ray, Alex, et al · 2022
Cited alongside, same era.
Black, Kevin, Janner, Michael, Du, Yilun, Kostrikov, Ilya, and Levine, Sergey · 2023
Closest in time.
Diffusion policy: Visuomotor policy learning via action diffusion
Chi, Cheng, Feng, Siyuan, Du, Yilun, Xu, Zhenjia, Cousineau, Eric, Burchfiel, Benjamin, and Song, Shuran · 2023
Closest in time.
Optimizing ddpm sampling with shortcut fine-tuning
Fan, Ying and Lee, Kangwook · 2023
Closest in time.
Tifa: Accurate and interpretable text-to-image faithfulness evaluation with question answering
Hu, Yushi, Liu, Benlin, Kasai, Jungo, Wang, Yizhong, Ostendorf, Mari, Krishna, Ranjay, and Smith, Noah A · 2023
Closest in time.
Pick-a-pic: An open dataset of user preferences for text-to-image generation
Kirstain, Yuval, Polyak, Adam, Singer, Uriel, Matiana, Shahbuland, Penna, Joe, and Levy, Omer · 2023
Closest in time.
Aligning text-to-image models using human feedback
Lee, Kimin, Liu, Hao, Ryu, Moonkyung, Watkins, Olivia, Du, Yuqing, Boutilier, Craig, Abbeel, Pieter, Ghavamzadeh, Mohammad, and Gu, Shixiang Shane · 2023
Closest in time.
Chain of hindsight aligns language models with feedback
Liu, Hao, Sferrazza, Carmelo, and Abbeel, Pieter · 2023
Closest in time.
Better aligning text-to-image models with human preference
Wu, Xiaoshi, Sun, Keqiang, Zhu, Feng, Zhao, Rui, and Li, Hongsheng · 2023
Closest in time.
Imagereward: Learning and evaluating human preferences for text-to-image generation
Xu, Jiazheng, Liu, Xiao, Wu, Yuchen, Tong, Yuxuan, Li, Qinkai, Ding, Ming, Tang, Jie, and Dong, Yuxiao · 2023
Closest in time.