Simple diffusion: End-to-end diffusion for high resolution images, 2023
Emiel Hoogeboom, Jonathan Heek, and Tim Salimans · 2023
Later among the works it cites.
Language is not all you need: Aligning perception with language models
Original
Shaohan Huang, Li Dong, Wenhui Wang, Yaru Hao, Saksham Singhal, Shuming Ma, Tengchao Lv, Lei Cui, Owais Khan Mohammed, Qiang Liu, et al · 2023
Later among the works it cites.
Text2video-zero: Text-to-image diffusion models are zero-shot video generators, 2023
Levon Khachatryan, Andranik Movsisyan, Vahram Tadevosyan, Roberto Henschel, Zhangyang Wang, Shant Navasardyan, and Humphrey Shi · 2023
Later among the works it cites.
Obelics: An open web-scale filtered dataset of interleaved image-text documents, 2023
Hugo Laurençon, Lucile Saulnier, Léo Tronchon, Stas Bekman, Amanpreet Singh, Anton Lozhkov, Thomas Wang, Siddharth Karamcheti, Alexander M. Rush, Douwe Kiela, Matthieu Cord, and Victor Sanh · 2023
Later among the works it cites.
Scalable diffusion models with transformers, 2023
William Peebles and Saining Xie · 2023
Later among the works it cites.
Sdxl: improving latent diffusion models for high-resolution image synthesis
Original
Dustin Podell, Zion English, Kyle Lacey, Andreas Blattmann, Tim Dockhorn, Jonas Müller, Joe Penna, and Robin Rombach · 2023
Later among the works it cites.
Exploring the limits of transfer learning with a unified text-to-text transformer, 2023
Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, and Peter J. Liu · 2023
Later among the works it cites.
Dreambooth: Fine tuning text-to-image diffusion models for subject-driven generation, 2023
Nataniel Ruiz, Yuanzhen Li, Varun Jampani, Yael Pritch, Michael Rubinstein, and Kfir Aberman · 2023
Later among the works it cites.
Adversarial diffusion distillation, 2023
Axel Sauer, Dominik Lorenz, Andreas Blattmann, and Robin Rombach · 2023
Later among the works it cites.
Styledrop: Text-to-image generation in any style, 2023
Kihyuk Sohn, Nataniel Ruiz, Kimin Lee, Daniel Castro Chin, Irina Blok, Huiwen Chang, Jarred Barber, Lu Jiang, Glenn Entis, Yuanzhen Li, Yuan Hao, Irfan Essa, Michael Rubinstein, and Dilip Krishnan · 2023
Later among the works it cites.
Journeydb: A benchmark for generative image understanding, 2023
Keqiang Sun, Junting Pan, Yuying Ge, Hao Li, Haodong Duan, Xiaoshi Wu, Renrui Zhang, Aojun Zhou, Zipeng Qin, Yi Wang, Jifeng Dai, Yu Qiao, Limin Wang, and Hongsheng Li · 2023
Later among the works it cites.
Attention is all you need, 2023
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N. Gomez, Lukasz Kaiser, and Illia Polosukhin · 2023
Later among the works it cites.
On the de-duplication of laion-2b, 2023
Ryan Webster, Julien Rabin, Loic Simon, and Frederic Jurie · 2023
Later among the works it cites.
Scaling autoregressive multi-modal models: Pretraining and instruction tuning, 2023
Lili Yu, Bowen Shi, Ramakanth Pasunuru, Benjamin Muller, Olga Golovneva, Tianlu Wang, Arun Babu, Binh Tang, Brian Karrer, Shelly Sheynin, Candace Ross, Adam Polyak, Russell Howes, Vasu Sharma, Puxin Xu, Hovhannes Tamoyan, Oron Ashual, Uriel Singer, Shang-Wen Li, Susan Zhang, Richard James, Gargi Ghosh, Yaniv Taigman, Maryam Fazel-Zarandi, Asli Celikyilmaz, Luke Zettlemoyer, and Armen Aghajanyan · 2023
Later among the works it cites.
Fast sampling of diffusion models with exponential integrator, 2023
Qinsheng Zhang and Yongxin Chen · 2023
Later among the works it cites.
Unipc: A unified predictor-corrector framework for fast sampling of diffusion models
Original
Wenliang Zhao, Lujia Bai, Yongming Rao, Jie Zhou, and Jiwen Lu · 2023
Later among the works it cites.
Dpm-solver-v3: Improved diffusion ode solver with empirical model statistics, 2023
Kaiwen Zheng, Cheng Lu, Jianfei Chen, and Jun Zhu · 2023
Later among the works it cites.