Fetching the paper…
Reading the bibliography…
Warning: this paper includes model outputs showing offensive content.
The information bottleneck method
Tishby, N., Pereira, F. C., and Bialek, W. (2000) · 2000
Earlier work this paper cites.
The im algorithm: a variational approach to information maximization
Barber, D. and Agakov, F. (2003) · 2003
Earlier work this paper cites.
Rouge: A package for automatic evaluation of summaries
Lin, C.-Y. (2004) · 2004
Earlier work this paper cites.
Mutual information estimation reveals global associations between stimuli and biological processes
Suzuki, T., Sugiyama, M., Kanamori, T., and Sese, J. (2009) · 2009
Earlier work this paper cites.
Squared-loss mutual information regularization: A novel information-theoretic approach to semi-supervised learning
Niu, G., Jitkrittum, W., Dai, B., Hachiya, H., and Sugiyama, M. (2013) · 2013
Earlier work this paper cites.
Microsoft coco: Common objects in context
Lin, T.-Y., Maire, M., Belongie, S., Hays, J., Perona, P., Ramanan, D., Dollár, P., and Zitnick, C. L. (2014) · 2014
Earlier work this paper cites.
From image descriptions to visual denotations: New similarity metrics for semantic inference over event descriptions
Young, P., Lai, A., Hodosh, M., and Hockenmaier, J. (2014) · 2014
Earlier work this paper cites.
Vqa: Visual question answering
Antol, S., Agrawal, A., Lu, J., Mitchell, M., Batra, D., Zitnick, C. L., and Parikh, D. (2015) · 2015
Earlier work this paper cites.
Deep variational information bottleneck
Alemi, A. A., Fischer, I., Dillon, J. V., and Murphy, K. (2016) · 2016
Earlier work this paper cites.
Spice: Semantic propositional image caption evaluation
Anderson, P., Fernando, B., Johnson, M., and Gould, S. (2016) · 2016
Earlier work this paper cites.
Improved techniques for training gans
Salimans, T., Goodfellow, I., Zaremba, W., Cheung, V., Radford, A., and Chen, X. (2016) · 2016
Earlier work this paper cites.
Gans trained by a two time-scale update rule converge to a local nash equilibrium
Heusel, M., Ramsauer, H., Unterthiner, T., Nessler, B., and Hochreiter, S. (2017) · 2017
Earlier work this paper cites.
Attention is all you need
Vaswani, A., Shazeer, N., Parmar, N., Uszkoreit, J., Jones, L., Gomez, A. N., Kaiser, Ł., and Polosukhin, I. (2017) · 2017
Earlier work this paper cites.
Protest activity detection and perceived violence estimation from social media images
Won, D., Steinert-Threlkeld, Z. C., and Joo, J. (2017) · 2017
Earlier work this paper cites.
BERT: Pre-training of deep bidirectional transformers for language understanding
Devlin, J., Chang, M.-W., Lee, K., and Toutanova, K. (2019) · 2019
Earlier work this paper cites.
Decoupled weight decay regularization
Loshchilov, I. and Hutter, F. (2019) · 2019
Earlier work this paper cites.
Language models are unsupervised multitask learners
Radford, A., Wu, J., Child, R., Luan, D., Amodei, D., Sutskever, I., et al. (2019) · 2019
Earlier work this paper cites.
Universal adversarial triggers for attacking and analyzing NLP
Wallace, E., Feng, S., Kandpal, N., Gardner, M., and Singh, S. (2019) · 2019
Earlier work this paper cites.
Plug and play language models: A simple approach to controlled text generation
Dathathri, S., Madotto, A., Lan, J., Hung, J., Frank, E., Molino, P., Yosinski, J., and Liu, R. (2020) · 2020
Earlier work this paper cites.
Realtoxicityprompts: Evaluating neural toxic degeneration in language models
Gehman, S., Gururangan, S., Sap, M., Choi, Y., and Smith, N. A. (2020a) · 2020
Earlier work this paper cites.
RealToxicityPrompts: Evaluating neural toxic degeneration in language models
Gehman, S., Gururangan, S., Sap, M., Choi, Y., and Smith, N. A. (2020b) · 2020
Earlier work this paper cites.
Denoising diffusion probabilistic models
Ho, J., Jain, A., and Abbeel, P. (2020) · 2020
Earlier work this paper cites.
Oscar: Object-semantics aligned pre-training for vision-language tasks
Li, X., Yin, X., Li, C., Zhang, P., Hu, X., Zhang, L., Wang, L., Hu, H., Dong, L., Wei, F., et al. (2020) · 2020
Earlier work this paper cites.
On relations between the relative entropy and χ \chi 2-divergence, generalizations and applications
Nishiyama, T. and Sason, I. (2020) · 2020
Earlier work this paper cites.
Multimodal datasets: misogyny, pornography, and malignant stereotypes
Birhane, A., Prabhu, V. U., and Kahembwe, E. (2021) · 2021
Cited alongside, same era.
On the opportunities and risks of foundation models
Bommasani, R., Hudson, D. A., Adeli, E., Altman, R., Arora, S., von Arx, S., Bernstein, M. S., Bohg, J., Bosselut, A., Brunskill, E., et al. (2021) · 2021
Cited alongside, same era.
Conceptual 12m: Pushing web-scale image-text pre-training to recognize long-tail visual concepts
Changpinyo, S., Sharma, P., Ding, N., and Soricut, R. (2021) · 2021
Cited alongside, same era.
Fairfil: Contrastive neural debiasing method for pretrained text encoders
Cheng, P., Hao, W., Yuan, S., Si, S., and Carin, L. (2021) · 2021
Cited alongside, same era.
Text detoxification using large pre-trained neural models
Dall-eval: Probing the reasoning skills and social biases of text-to-image generative transformers
Cho, J., Zala, A., and Bansal, M. (2022) · 2022
Later among the works it cites.
Write and paint: Generative vision-language models are unified modal learners
Diao, S., Zhou, W., Zhang, X., and Wang, J. (2022) · 2022
Later among the works it cites.
Cogview2: Faster and better text-to-image generation via hierarchical transformers
Ding, M., Zheng, W., Hong, W., and Tang, J. (2022) · 2022
Later among the works it cites.
Transformer feed-forward layers build predictions by promoting concepts in the vocabulary space
Geva, M., Caciularu, A., Wang, K. R., and Goldberg, Y. (2022) · 2022
Later among the works it cites.
ToxiGen: A large-scale machine-generated dataset for adversarial and implicit hate speech detection
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Dale, D., Voronov, A., Dementieva, D., Logacheva, V., Kozlova, O., Semenov, N., and Panchenko, A. (2021) · 2021
Cited alongside, same era.
Dall·e mini
Dayma, B., Patil, S., Cuenca, P., Saifullah, K., Abraham, T., LeKhac, P., Melas, L., and Ghosh, R. (2021) · 2021
Cited alongside, same era.
Cogview: Mastering text-to-image generation via transformers
Ding, M., Yang, Z., Hong, W., Zheng, W., Zhou, C., Yin, D., Lin, J., Zou, X., Shao, Z., Yang, H., et al. (2021) · 2021
Cited alongside, same era.
Latent hatred: A benchmark for understanding implicit hate speech
ElSherief, M., Ziems, C., Muchlinski, D., Anupindi, V., Seybolt, J., De Choudhury, M., and Yang, D. (2021) · 2021
Cited alongside, same era.
Clipscore: A reference-free evaluation metric for image captioning
Hessel, J., Holtzman, A., Forbes, M., Le Bras, R., and Choi, Y. (2021) · 2021
Cited alongside, same era.
Unifying multimodal transformer for bi-directional image and text generation
Huang, Y., Xue, H., Liu, B., and Lu, Y. (2021) · 2021
Cited alongside, same era.
Gedi: Generative discriminator guided sequence generation
Krause, B., Gotmare, A. D., McCann, B., Keskar, N. S., Joty, S., Socher, R., and Rajani, N. F. (2021) · 2021
Cited alongside, same era.
DExperts: Decoding-time controlled text generation with experts and anti-experts
Liu, A., Sap, M., Lu, X., Swayamdipta, S., Bhagavatula, C., Smith, N. A., and Choi, Y. (2021) · 2021
Cited alongside, same era.
Hartvigsen, T., Gabriel, S., Palangi, H., Sap, M., Ray, D., and Kamar, E. (2022) · 2022
Later among the works it cites.
Quantifying societal bias amplification in image captioning
Hirota, Y., Nakashima, Y., and Garcia, N. (2022) · 2022
Later among the works it cites.
Du-vlg: Unifying vision-and-language generation via dual sequence-to-sequence pre-training
Huang, L., Niu, G., Liu, J., Xiao, X., and Wu, H. (2022) · 2022
Later among the works it cites.
L-verse: Bidirectional generation between image and text
Kim, T., Song, G., Lee, S., Kim, S., Seo, Y., Lee, S., Kim, S. H., Lee, H., and Bae, K. (2022) · 2022
Later among the works it cites.
A new generation of perspective api: Efficient multilingual character-level transformers
Lees, A., Tran, V. Q., Tay, Y., Sorensen, J., Gupta, J., Metzler, D., and Vasserman, L. (2022) · 2022
Later among the works it cites.
Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation
Li, J., Li, D., Xiong, C., and Hoi, S. (2022) · 2022
Later among the works it cites.
Grit: Faster and better image captioning transformer using dual visual features
Nguyen, V.-Q., Suganuma, M., and Okatani, T. (2022) · 2022
Later among the works it cites.
Hierarchical text-conditional image generation with clip latents
Ramesh, A., Dhariwal, P., Nichol, A., Chu, C., and Chen, M. (2022) · 2022
Later among the works it cites.
High-resolution image synthesis with latent diffusion models
Rombach, R., Blattmann, A., Lorenz, D., Esser, P., and Ommer, B. (2022) · 2022
Later among the works it cites.
How much can CLIP benefit vision-and-language tasks?
Shen, S., Li, L. H., Tan, H., Bansal, M., Rohrbach, A., Chang, K.-W., Yao, Z., and Keutzer, K. (2022) · 2022
Later among the works it cites.
Measuring representational harms in image captioning
Wang, A., Barocas, S., Laird, K., and Wallach, H. (2022a) · 2022
Later among the works it cites.
Opt: Open pre-trained transformer language models
Zhang, S., Roller, S., Goyal, N., Artetxe, M., Chen, M., Chen, S., Dewan, C., Diab, M., Li, X., Lin, X. V., et al. (2022) · 2022
Later among the works it cites.
Towards language-free training for text-to-image generation
Zhou, Y., Zhang, R., Chen, C., Li, C., Tensmeyer, C., Yu, T., Gu, J., Xu, J., and Sun, T. (2022) · 2022
Later among the works it cites.
Instructpix2pix: Learning to follow image editing instructions
Brooks, T., Holynski, A., and Efros, A. A. (2023) · 2023
Closest in time.
Palm-e: An embodied multimodal language model
Driess, D., Xia, F., Sajjadi, M. S., Lynch, C., Chowdhery, A., Ichter, B., Wahid, A., Tompson, J., Vuong, Q., Yu, T., et al. (2023) · 2023
Closest in time.
Li, J., Li, D., Savarese, S., and Hoi, S. (2023) · 2023
Closest in time.
Liu, H., Li, C., Wu, Q., and Lee, Y. J. (2023) · 2023
Closest in time.
Llama: Open and efficient foundation language models
Touvron, H., Lavril, T., Izacard, G., Martinet, X., Lachaux, M.-A., Lacroix, T., Rozière, B., Goyal, N., Hambro, E., Azhar, F., et al. (2023) · 2023
Closest in time.
Unified detoxifying and debiasing in language generation via inference-time adaptive optimization
Yang, Z., Yi, X., Li, P., Liu, Y., and Xie, X. (2023) · 2023
Closest in time.