Scaling up vision-language pre-training for image captioning
Hu, X., Gan, Z., Wang, J., Yang, Z., Liu, Z., Lu, Y., and Wang, L · 2022
Later among the works it cites.
Prompt waywardness: The curious case of discretized interpretation of continuous prompts
Khashabi, D., Lyu, X., Min, S., Qin, L., Richardson, K., Welleck, S., Hajishirzi, H., Khot, T., Sabharwal, A., Singh, S., and Choi, Y · 2022
Later among the works it cites.
Gradient-based constrained sampling from language models
Original
Kumar, S., Paria, B., and Tsvetkov, Y · 2022
Later among the works it cites.
Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation
Original
Li, J., Li, D., Xiong, C., and Hoi, S · 2022
Later among the works it cites.
Grips: Gradient-free, edit-based instruction search for prompting large language models
Original
Prasad, A., Hase, P., Zhou, X., and Bansal, M · 2022
Later among the works it cites.
Red-teaming the stable diffusion safety filter
Original
Rando, J., Paleka, D., Lindner, D., Heim, L., and Tramèr, F · 2022
Later among the works it cites.
High-resolution image synthesis with latent diffusion models
Rombach, R., Blattmann, A., Lorenz, D., Esser, P., and Ommer, B · 2022
Later among the works it cites.
Multitask prompted training enables zero-shot task generalization
Sanh, V., Webson, A., Raffel, C., Bach, S., Sutawika, L., Alyafeai, Z., Chaffin, A., Stiegler, A., Raja, A., Dey, M., Bari, M. S., Xu, C., Thakker, U., Sharma, S. S., Szczechla, E., Kim, T., Chhablani, G., Nayak, N., Datta, D., Chang, J., Jiang, M. T.-J., Wang, H., Manica, M., Shen, S., Yong, Z. X., Pandey, H., Bawden, R., Wang, T., Neeraj, T., Rozen, J., Sharma, A., Santilli, A., Fevry, T., Fries, J. A., Teehan, R., Scao, T. L., Biderman, S., Gao, L., Wolf, T., and Rush, A. M · 2022
Later among the works it cites.
Gustavosta/Stable-Diffusion-Prompts ⋅ \cdot Datasets at Hugging Face, December 2022
Santana, G · 2022
Later among the works it cites.
Laion-5b: An open large-scale dataset for training next generation image-text models
Original
Schuhmann, C., Beaumont, R., Vencu, R., Gordon, C., Wightman, R., Cherti, M., Coombes, T., Katta, A., Mullis, C., Wortsman, M., et al · 2022
Later among the works it cites.
Toward human readable prompt tuning: Kubrick’s the shining is a good movie, and a good prompt too?
Original
Shi, W., Han, X., Gonen, H., Holtzman, A., Tsvetkov, Y., and Zettlemoyer, L · 2022
Later among the works it cites.
Opt: Open pre-trained transformer language models
Original
Zhang, S., Roller, S., Goyal, N., Artetxe, M., Chen, M., Chen, S., Dewan, C., Diab, M., Li, X., Lin, X. V., et al · 2022
Later among the works it cites.