Fetching the paper…
Reading the bibliography…
State-of-the-art T2I models are capable of generating high-resolution images given textual prompts.
Benchmark for compositional text-to-image synthesis
D. H. Park, S. Azadi, X. Liu, T. Darrell, and A. Rohrbach · 2021
Earlier work this paper cites.
Training-free structured diffusion guidance for compositional text-to-image synthesis
W. Feng, X. He, T.-J. Fu, V. Jampani, A. Akula, P. Narayana, S. Basu, X. E. Wang, and W. Y. Wang · 2022
Earlier work this paper cites.
High-resolution image synthesis with latent diffusion models
R. Rombach, A. Blattmann, D. Lorenz, P. Esser, and B. Ommer · 2022
Earlier work this paper cites.
Attend-and-excite: Attention-based semantic guidance for text-to-image diffusion models
H. Chefer, Y. Alaluf, Y. Vinker, L. Wolf, and D. Cohen-Or · 2023
Earlier work this paper cites.
Optimizing ddpm sampling with shortcut fine-tuning
Y. Fan and K. Lee · 2023
Earlier work this paper cites.
Dpok: Reinforcement learning for fine-tuning text-to-image diffusion models
Y. Fan, O. Watkins, Y. Du, H. Liu, M. Ryu, C. Boutilier, P. Abbeel, M. Ghavamzadeh, K. Lee, and K. Lee · 2023
Earlier work this paper cites.
Optimizing prompts for text-to-image generation
Y. Hao, Z. Chi, L. Dong, and F. Wei · 2023
Earlier work this paper cites.
Tifa: Accurate and interpretable text-to-image faithfulness evaluation with question answering
Y. Hu, B. Liu, J. Kasai, Y. Wang, M. Ostendorf, R. Krishna, and N. A. Smith · 2023
Cited alongside, same era.
T2i-compbench: A comprehensive benchmark for open-world compositional text-to-image generation
K. Huang, K. Sun, E. Xie, Z. Li, and X. Liu · 2023
Cited alongside, same era.
Aligning text-to-image models using human feedback
K. Lee, H. Liu, M. Ryu, O. Watkins, Y. Du, C. Boutilier, P. Abbeel, M. Ghavamzadeh, and S. S. Gu · 2023
Cited alongside, same era.
L. Lian, B. Li, A. Yala, and T. Darrell · 2023
Cited alongside, same era.
Dall·e 3 system card, Oct 2023
OpenAI · 2023
Cited alongside, same era.
Harnessing the spatial-temporal attention of diffusion models for high-fidelity text-to-image synthesis
Q. Wu, Y. Liu, H. Zhao, T. Bui, Z. Lin, Y. Zhang, and S. Chang · 2023
Later among the works it cites.
Imagereward: Learning and evaluating human preferences for text-to-image generation
J. Xu, X. Liu, Y. Wu, Y. Tong, Q. Li, M. Ding, J. Tang, and Y. Dong · 2023
Later among the works it cites.
Training-free layout control with cross-attention guidance
M. Chen, I. Laina, and A. Vedaldi · 2024
Later among the works it cites.
Scaling rectified flow transformers for high-resolution image synthesis
P. Esser, S. Kulal, A. Blattmann, R. Entezari, J. Müller, H. Saini, Y. Levi, D. Lorenz, A. Sauer, F. Boesel, et al · 2024
Later among the works it cites.
Kto: Model alignment as prospect theoretic optimization
K. Ethayarajh, W. Xu, N. Muennighoff, D. Jurafsky, and D. Kiela · 2024
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Dreamsync: Aligning text-to-image generation with image understanding feedback
J. Sun, D. Fu, Y. Hu, S. Wang, R. Rassin, D.-C. Juan, D. Alon, C. Herrmann, S. van Steenkiste, R. Krishna, et al · 2023
Cited alongside, same era.
Gemini: a family of highly capable multimodal models
G. Team, R. Anil, S. Borgeaud, J.-B. Alayrac, J. Yu, R. Soricut, J. Schalkwyk, A. M. Dai, A. Hauth, K. Millican, et al · 2023
Cited alongside, same era.
Training diffusion models with reinforcement learning
K. Black, M. Janner, Y. Du, I. Kostrikov, and S. Levine
Cited in the paper.
Directly fine-tuning diffusion models on differentiable rewards
K. Clark, P. Vicol, K. Swersky, and D. J. Fleet
Cited in the paper.
Aligning diffusion models by optimizing human utility
S. Li, K. Kallidromitis, A. Gokul, Y. Kato, and K. Kozuka
Cited in the paper.
Aligning text-to-image diffusion models with reward backpropagation
M. Prabhudesai, A. Goyal, D. Pathak, and K. Fragkiadaki
Cited in the paper.
T2i-compbench++: An enhanced and comprehensive benchmark for compositional text-to-image generation
K. Huang, C. Duan, K. Sun, E. Xie, Z. Li, and X. Liu · 2025
Closest in time.