Fetching the paper…
Reading the bibliography…
In this paper, we address the limitations of existing text-to-image diffusion models in generating demographically fair results when given human-related descriptions.
A simple framework for contrastive learning of visual representations
Chen, T., Kornblith, S., Norouzi, M., and Hinton, G. E · 2002
Earlier work this paper cites.
Efficient estimation of word representations in vector space
Mikolov, T., Chen, K., Corrado, G., and Dean, J · 2013
Earlier work this paper cites.
Auto-encoding variational bayes
Kingma, D. P. and Welling, M · 2014
Earlier work this paper cites.
Euclidean distance matrices: Essential theory, algorithms, and applications
Dokmanic, I., Parhizkar, R., Ranieri, J., and Vetterli, M · 2015
Earlier work this paper cites.
Equality of opportunity in supervised learning
Hardt, M., Price, E., and Srebro, N · 2016
Earlier work this paper cites.
What is real-world data? a review of definitions based on literature and stakeholder interviews
Makady, A., de Boer, A., Hillege, H., Klungel, O., and Goettsch, W · 2017
Earlier work this paper cites.
Keep drawing it: Iterative language-based image generation and editing
El-Nouby, A., Sharma, S., Schulz, H., Hjelm, R. D., Asri, L. E., Kahou, S. E., Bengio, Y., and Taylor, G. W · 2018
Earlier work this paper cites.
A style-based generator architecture for generative adversarial networks
Karras, T., Laine, S., and Aila, T · 2018
Earlier work this paper cites.
Robust and discriminative speaker embedding via intra-class distance variance regularization
Le, N. and Odobez, J · 2018
Earlier work this paper cites.
Polarity and intensity: the two aspects of sentiment analysis
Tian, L., Lai, C., and Moore, J. D · 2018
Earlier work this paper cites.
Realtoxicityprompts: Evaluating neural toxic degeneration in language models
Gehman, S., Gururangan, S., Sap, M., Choi, Y., and Smith, N. A · 2020
Earlier work this paper cites.
Denoising diffusion probabilistic models
Ho, J., Jain, A., and Abbeel, P · 2020
Earlier work this paper cites.
Persistent anti-muslim bias in large language models
Abid, A., Farooqi, M., and Zou, J · 2021
Earlier work this paper cites.
On the dangers of stochastic parrots: Can language models be too big?
Bender, E. M., Gebru, T., McMillan-Major, A., and Shmitchell, S · 2021
Earlier work this paper cites.
Multimodal datasets: misogyny, pornography, and malignant stereotypes
Birhane, A., Prabhu, V. U., and Kahembwe, E · 2021
Earlier work this paper cites.
Implicit stereotypes in pre-trained classifiers
Dehouche, N · 2021
Earlier work this paper cites.
Diffusion models beat gans on image synthesis
Dhariwal, P. and Nichol, A. Q · 2021
Earlier work this paper cites.
Clipscore: A reference-free evaluation metric for image captioning
Hessel, J., Holtzman, A., Forbes, M., Bras, R. L., and Choi, Y · 2021
Earlier work this paper cites.
Fairface: Face attribute dataset for balanced race, gender, and age for bias measurement and mitigation
Kärkkäinen, K. and Joo, J · 2021
Earlier work this paper cites.
Learning transferable visual models from natural language supervision
Radford, A., Kim, J. W., Hallacy, C., Ramesh, A., Goh, G., Agarwal, S., Sastry, G., Askell, A., Mishkin, P., Clark, J., Krueger, G., and Sutskever, I · 2021
Earlier work this paper cites.
LAION-400M: open dataset of clip-filtered 400 million image-text pairs
Schuhmann, C., Vencu, R., Beaumont, R., Kaczmarczyk, R., Mullis, C., Katta, A., Coombes, T., Jitsev, J., and Komatsuzaki, A · 2021
Cited alongside, same era.
Denoising diffusion implicit models
Song, J., Meng, C., and Ermon, S · 2021
Cited alongside, same era.
ediff-i: Text-to-image diffusion models with an ensemble of expert denoisers
Balaji, Y., Nah, S., Huang, X., Vahdat, A., Song, J., Kreis, K., Aittala, M., Aila, T., Laine, S., Catanzaro, B., Karras, T., and Liu, M · 2022
Cited alongside, same era.
How well can text-to-image generative models understand ethical natural language interventions?
Bansal, H., Yin, D., Monajatipoor, M., and Chang, K · 2022
Cited alongside, same era.
A prompt array keeps the bias away: Debiasing vision-language models with adversarial learning
Berg, H., Hall, S., Bhalgat, Y., Kirk, H., Shtedritski, A., and Bain, M · 2022
Align your latents: High-resolution video synthesis with latent diffusion models, 2023
Blattmann, A., Rombach, R., Ling, H., Dockhorn, T., Kim, S. W., Fidler, S., and Kreis, K · 2023
Closest in time.
Pix2video: Video editing using image diffusion
Ceylan, D., Huang, C. P., and Mitra, N. J · 2023
Closest in time.
Debiasing vision-language models via biased prompts
Chuang, C.-Y., Jampani, V., Li, Y., Torralba, A., and Jegelka, S · 2023
Closest in time.
Training-free structured diffusion guidance for compositional text-to-image synthesis
Feng, W., He, X., Fu, T., Jampani, V., Akula, A. R., Narayana, P., Basu, S., Wang, X. E., and Wang, W. Y · 2023
Closest in time.
A friendly face: Do text-to-image systems rely on stereotypes when the input is under-specified?
Fraser, K. C., Kiritchenko, S., and Nejadgholi, I · 2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Word embeddings via causal inference: Gender bias reducing and semantic information preserving
Ding, L., Yu, D., Xie, J., Guo, W., Hu, S., Liu, M., Kong, L., Dai, H., Bao, Y., and Jiang, B · 2022
Cited alongside, same era.
Classifier-free diffusion guidance
Ho, J. and Salimans, T · 2022
Cited alongside, same era.
Mass-editing memory in a transformer
Meng, K., Sharma, A. S., Andonian, A., Belinkov, Y., and Bau, D · 2022
Cited alongside, same era.
Improving adversarial robustness by contrastive guided diffusion process
Ouyang, Y., Xie, L., and Cheng, G · 2022
Cited alongside, same era.
Hierarchical text-conditional image generation with CLIP latents
Ramesh, A., Dhariwal, P., Nichol, A., Chu, C., and Chen, M · 2022
Cited alongside, same era.
High-resolution image synthesis with latent diffusion models
Rombach, R., Blattmann, A., Lorenz, D., Esser, P., and Ommer, B · 2022
Cited alongside, same era.
Photorealistic text-to-image diffusion models with deep language understanding
Saharia, C., Chan, W., Saxena, S., Li, L., Whang, J., Denton, E. L., Ghasemipour, S. K. S., Lopes, R. G., Ayan, B. K., Salimans, T., Ho, J., Fleet, D. J., and Norouzi, M · 2022
Cited alongside, same era.
Closest in time.
Fair diffusion: Instructing text-to-image generation models on fairness
Friedrich, F., Schramowski, P., Brack, M., Struppek, L., Hintersdorf, D., Luccioni, S., and Kersting, K · 2023
Closest in time.
Localized text-to-image generation for free via cross attention control
He, Y., Salakhutdinov, R., and Kolter, J. Z · 2023
Closest in time.
Rs-corrector: Correcting the racial stereotypes in latent diffusion models
Jiang, Y., Lyu, Y., Ma, T., Peng, B., and Dong, J · 2023
Closest in time.
Dense text-to-image generation with attention modulation
Kim, Y., Lee, J., Kim, J., Ha, J., and Zhu, J · 2023
Closest in time.
Elucidating the exposure bias in diffusion models
Ning, M., Li, M., Su, J., Salah, A. A., and Ertugrul, I. Ö · 2023
Closest in time.
Editing implicit assumptions in text-to-image diffusion models
Orgad, H., Kawar, B., and Belinkov, Y · 2023
Closest in time.
Toward verifiable and reproducible human evaluation for text-to-image generation
Otani, M., Togashi, R., Sawai, Y., Ishigami, R., Nakashima, Y., Rahtu, E., Heikkilä, J., and Satoh, S · 2023
Closest in time.
Dreamfusion: Text-to-3d using 2d diffusion
Poole, B., Jain, A., Barron, J. T., and Mildenhall, B · 2023
Closest in time.
Safe latent diffusion: Mitigating inappropriate degeneration in diffusion models
Schramowski, P., Brack, M., Deiseroth, B., and Kersting, K · 2023
Closest in time.
Finetuning text-to-image diffusion models for fairness
Shen, X., Du, C., Pang, T., Lin, M., Wong, Y., and Kankanhalli, M · 2023
Closest in time.
Any-to-any generation via composable diffusion
Tang, Z., Yang, Z., Zhu, C., Zeng, M., and Bansal, M · 2023
Closest in time.
Human motion diffusion model
Tevet, G., Raab, S., Gordon, B., Shafir, Y., Cohen-Or, D., and Bermano, A. H · 2023
Closest in time.
Mixfairface: Towards ultimate fairness via mixfair adapter in face recognition
Wang, F., Wang, C., Sun, M., and Lai, S · 2023
Closest in time.
Unified concept editing in diffusion models
Gandikota, R., Orgad, H., Belinkov, Y., Materzyńska, J., and Bau, D · 2024
Closest in time.