Fetching the paper…
Reading the bibliography…
The growing proliferation of customized and pretrained generative models has made it infeasible for a user to be fully cognizant of every model in existence.
Plug-and-play diffusion features for text-driven image-to-image translation. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition . 1921–1930
Narek Tumanyan, Michal Geyer, Shai Bagon, and Tali Dekel. 2023 · 1930
Earlier work this paper cites.
Multi-concept customization of text-to-image diffusion
Nupur Kumari, Bingliang Zhang, Richard Zhang, Eli Shechtman, and Jun-Yan Zhu. 2023 · 1941
Earlier work this paper cites.
Content based image retrieval systems
Venkat N Gudivada and Vijay V Raghavan. 1995 · 1995
Earlier work this paper cites.
Modern information retrieval . Vol. 463
Ricardo Baeza-Yates, Berthier Ribeiro-Neto, et al · 1999
Earlier work this paper cites.
Content-based image retrieval at the end of the early years
Arnold WM Smeulders, Marcel Worring, Simone Santini, Amarnath Gupta, and Ramesh Jain. 2000 · 2000
Earlier work this paper cites.
Modeling the shape of the scene: A holistic representation of the spatial envelope
Aude Oliva and Antonio Torralba. 2001 · 2001
Earlier work this paper cites.
Training products of experts by minimizing contrastive divergence
Geoffrey E Hinton. 2002 · 2002
Earlier work this paper cites.
Video Google: A text retrieval approach to object matching in videos. In ICCV
Josef Sivic and Andrew Zisserman. 2003 · 2003
Earlier work this paper cites.
Distinctive image features from scale-invariant keypoints
David G Lowe. 2004 · 2004
Earlier work this paper cites.
Histograms of oriented gradients for human detection. In CVPR
Navneet Dalal and Bill Triggs. 2005 · 2005
Earlier work this paper cites.
Image retrieval: Ideas, influences, and trends of the new age
Ritendra Datta, Dhiraj Joshi, Jia Li, and James Z Wang. 2008 · 2008
Earlier work this paper cites.
Small codes and large image databases for recognition. In CVPR
Antonio Torralba, Rob Fergus, and Yair Weiss. 2008 · 2008
Earlier work this paper cites.
Spectral hashing. In NeurIPS
Yair Weiss, Antonio Torralba, and Rob Fergus. 2008 · 2008
Earlier work this paper cites.
ImageNet: A Large-Scale Hierarchical Image Database. In CVPR
J. Deng, W. Dong, R. Socher, L.-J. Li, K. Li, and L. Fei-Fei. 2009 · 2009
Earlier work this paper cites.
Sketch-based image retrieval: Benchmark and bag-of-features descriptors
Mathias Eitz, Kristian Hildebrand, Tamy Boubekeur, and Marc Alexa. 2010 · 2010
Earlier work this paper cites.
Aggregating local descriptors into a compact image representation. In 2010 IEEE computer society conference on computer vision and pattern recognition . IEEE, 3304–3311
Hervé Jégou, Matthijs Douze, Cordelia Schmid, and Patrick Pérez. 2010 · 2010
Earlier work this paper cites.
Introduction to information retrieval
Christopher Manning, Prabhakar Raghavan, and Hinrich Schütze. 2010 · 2010
Earlier work this paper cites.
A survey on visual content-based video indexing and retrieval
Weiming Hu, Nianhua Xie, Li Li, Xianglin Zeng, and Stephen Maybank. 2011 · 2011
Earlier work this paper cites.
Using very deep autoencoders for content-based image retrieval.. In ESANN , Vol. 1. Citeseer, 2
Alex Krizhevsky and Geoffrey E Hinton. 2011 · 2011
Earlier work this paper cites.
Three things everyone should know to improve object retrieval. In CVPR
Relja Arandjelović and Andrew Zisserman. 2012 · 2012
Earlier work this paper cites.
Iterative quantization: A procrustean approach to learning binary codes for large-scale image retrieval
Yunchao Gong, Svetlana Lazebnik, Albert Gordo, and Florent Perronnin. 2012 · 2012
Earlier work this paper cites.
Devise: A deep visual-semantic embedding model. In NeurIPS
Andrea Frome, Greg S Corrado, Jon Shlens, Samy Bengio, Jeff Dean, Marc’Aurelio Ranzato, and Tomas Mikolov. 2013 · 2013
Earlier work this paper cites.
3D sub-query expansion for improving sketch-based multi-view image retrieval. In ICCV
Yen-Liang Lin, Cheng-Yu Huang, Hao-Jeng Wang, and Winston Hsu. 2013 · 2013
Earlier work this paper cites.
Neural codes for image retrieval. In ECCV
Artem Babenko, Anton Slesarev, Alexandr Chigorin, and Victor Lempitsky. 2014 · 2014
Earlier work this paper cites.
Generative adversarial nets. In NeurIPS
Ian Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron Courville, and Yoshua Bengio. 2014 · 2014
Earlier work this paper cites.
Deep fragment embeddings for bidirectional image sentence mapping. In NeurIPS
Andrej Karpathy, Armand Joulin, and Li F Fei-Fei. 2014 · 2014
Earlier work this paper cites.
Auto-encoding variational bayes. In ICLR
Diederik P Kingma and Max Welling. 2014 · 2014
Earlier work this paper cites.
Grounded compositional semantics for finding and describing images with sentences
Richard Socher, Andrej Karpathy, Quoc V Le, Christopher D Manning, and Andrew Y Ng. 2014 · 2014
Earlier work this paper cites.
Lsun: Construction of a large-scale image dataset using deep learning with humans in the loop
Fisher Yu, Ari Seff, Yinda Zhang, Shuran Song, Thomas Funkhouser, and Jianxiong Xiao. 2015 · 2015
Earlier work this paper cites.
Conditional Image Generation with PixelCNN Decoders. In NeurIPS
Aaron van den Oord, Nal Kalchbrenner, Oriol Vinyals, Lasse Espeholt, Alex Graves, and Koray Kavukcuoglu. 2016 · 2016
Earlier work this paper cites.
The sketchy database: learning to retrieve badly drawn bunnies
Patsorn Sangkloy, Nathan Burnell, Cusuh Ham, and James Hays. 2016 · 2016
Earlier work this paper cites.
Sketch me that shoe. In CVPR
Qian Yu, Feng Liu, Yi-Zhe Song, Tao Xiang, Timothy M Hospedales, and Chen-Change Loy. 2016 · 2016
Earlier work this paper cites.
Vse++: Improving visual-semantic embeddings with hard negatives. In BMVC
Fartash Faghri, David J Fleet, Jamie Ryan Kiros, and Sanja Fidler. 2017 · 2017
Earlier work this paper cites.
GANs trained by a two time-scale update rule converge to a local Nash equilibrium. In NeurIPS
Martin Heusel, Hubert Ramsauer, Thomas Unterthiner, Bernhard Nessler, and Sepp Hochreiter. 2017 · 2017
Earlier work this paper cites.
Deep sketch hashing: Fast free-hand sketch-based image retrieval. In CVPR
Li Liu, Fumin Shen, Yuming Shen, Xianglong Liu, and Ling Shao. 2017 · 2017
Earlier work this paper cites.
SIFT meets CNN: A decade survey of instance retrieval
Liang Zheng, Yi Yang, and Qi Tian. 2017 · 2017
Earlier work this paper cites.
David Ha and Jürgen Schmidhuber. 2018 · 2018
Earlier work this paper cites.
Progressive growing of gans for improved quality, stability, and variation. In ICLR
Tero Karras, Timo Aila, Samuli Laine, and Jaakko Lehtinen. 2018 · 2018
Earlier work this paper cites.
Representation learning with contrastive predictive coding
Aaron van den Oord, Yazhe Li, and Oriol Vinyals. 2018 · 2018
Cited alongside, same era.
Deep shape matching. In ECCV
Filip Radenovic, Giorgos Tolias, and Ondrej Chum. 2018 · 2018
Cited alongside, same era.
Transferring gans: generating images from limited data. In ECCV
Yaxing Wang, Chenshen Wu, Luis Herranz, Joost van de Weijer, Abel Gonzalez-Garcia, and Bogdan Raducanu. 2018 · 2018
Cited alongside, same era.
This vessel does not exist
Derek Philip Au. 2019 · 2019
Cited alongside, same era.
Large scale gan training for high fidelity natural image synthesis. In ICLR
Andrew Brock, Jeff Donahue, and Karen Simonyan. 2019 · 2019
Cited alongside, same era.
AI is blurring the definition of artist: Advanced algorithms are using machine learning to create art autonomously
Datasetgan: Efficient labeled data factory with minimal human effort. In CVPR
Yuxuan Zhang, Huan Ling, Jun Gao, Kangxue Yin, Jean-Francois Lafleche, Adela Barriuso, Antonio Torralba, and Sanja Fidler. 2021 · 2021
Later among the works it cites.
Barbershop: Gan-based image compositing using segmentation masks
Peihao Zhu, Rameen Abdal, John Femiani, and Peter Wonka. 2021 · 2021
Later among the works it cites.
Civit AI
2022 · 2022
Closest in time.
Stable Diffusion Dreambooth Concepts Library
2022 · 2022
Closest in time.
Blended diffusion for text-driven editing of natural images. In CVPR
Omri Avrahami, Dani Lischinski, and Ohad Fried. 2022 · 2022
Closest in time.
State-of-the-Art in the Architecture, Methods and Applications of StyleGAN
Amit H Bermano, Rinon Gal, Yuval Alaluf, Ron Mokady, Yotam Nitzan, Omer Tov, Or Patashnik, and Daniel Cohen-Or. 2022 · 2022
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Ahmed Elgammal. 2019 · 2019
Cited alongside, same era.
A style-based generator architecture for generative adversarial networks. In CVPR
Tero Karras, Samuli Laine, and Timo Aila. 2019 · 2019
Cited alongside, same era.
Image generation from small datasets via batch statistics adaptation. In ICCV
Atsuhiro Noguchi and Tatsuya Harada. 2019 · 2019
Cited alongside, same era.
Generating diverse high-fidelity images with vq-vae-2. In NeurIPS
Ali Razavi, Aaron van den Oord, and Oriol Vinyals. 2019 · 2019
Cited alongside, same era.
Rewriting a deep generative model. In ECCV
David Bau, Steven Liu, Tongzhou Wang, Jun-Yan Zhu, and Antonio Torralba. 2020 · 2020
Cited alongside, same era.
DeepFaceDrawing: Deep generation of face images from sketches
Shu-Yu Chen, Wanchao Su, Lin Gao, Shihong Xia, and Hongbo Fu. 2020 · 2020
Cited alongside, same era.
StarGAN v2: Diverse Image Synthesis for Multiple Domains. In CVPR
Yunjey Choi, Youngjung Uh, Jaejun Yoo, and Jung-Woo Ha. 2020 · 2020
Cited alongside, same era.
Closest in time.
Retrieval-augmented diffusion models
Andreas Blattmann, Robin Rombach, Kaan Oktay, Jonas Müller, and Björn Ommer. 2022 · 2022
Closest in time.
Learning to generate line drawings that convey geometry and semantics. In CVPR
Caroline Chan, Fredo Durand, and Phillip Isola. 2022 · 2022
Closest in time.
An image is worth one word: Personalizing text-to-image generation using textual inversion
Rinon Gal, Yuval Alaluf, Yuval Atzmon, Or Patashnik, Amit H Bermano, Gal Chechik, and Daniel Cohen-Or. 2022a · 2022
Closest in time.
StyleGAN-NADA: CLIP-Guided Domain Adaptation of Image Generators
Rinon Gal, Or Patashnik, Haggai Maron, Gal Chechik, and Daniel Cohen-Or. 2022b · 2022
Closest in time.
When, Why, and Which Pretrained GANs Are Useful?. In ICLR
Timofey Grigoryev, Andrey Voynov, and Artem Babenko. 2022 · 2022
Closest in time.
Prompt-to-prompt image editing with cross attention control
Amir Hertz, Ron Mokady, Jay Tenenbaum, Kfir Aberman, Yael Pritch, and Daniel Cohen-Or. 2022 · 2022
Closest in time.
Multimodal Conditional Image Synthesis with Product-of-Experts GANs. In ECCV
Xun Huang, Arun Mallya, Ting-Chun Wang, and Ming-Yu Liu. 2022 · 2022
Closest in time.
Ensembling Off-the-shelf Models for GAN Training. In CVPR
Nupur Kumari, Richard Zhang, Eli Shechtman, and Jun-Yan Zhu. 2022 · 2022
Closest in time.
Junnan Li, Dongxu Li, Caiming Xiong, and Steven Hoi. 2022 · 2022
Closest in time.
A convnet for the 2020s. In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition . 11976–11986
Zhuang Liu, Hanzi Mao, Chao-Yuan Wu, Christoph Feichtenhofer, Trevor Darrell, and Saining Xie. 2022 · 2022
Closest in time.
Datasets and pretrained Models for StyleGAN3
lucid layers. 2022 · 2022
Closest in time.
Self-Distilled StyleGAN: Towards Generation from Internet Photos. In ACM SIGGRAPH
Ron Mokady, Michal Yarom, Omer Tov, Oran Lang, Daniel Cohen-Or, Tali Dekel, Michal Irani, and Inbar Mosseri. 2022 · 2022
Closest in time.
MyStyle: A Personalized Generative Prior
Yotam Nitzan, Kfir Aberman, Qiurui He, Orly Liba, Michal Yarom, Yossi Gandelsman, Inbar Mosseri, Yael Pritch, and Daniel Cohen-Or. 2022 · 2022
Closest in time.
On Buggy Resizing Libraries and Surprising Subtleties in FID Calculation. In CVPR
Gaurav Parmar, Richard Zhang, and Jun-Yan Zhu. 2022 · 2022
Closest in time.
Dreamfusion: Text-to-3d using 2d diffusion
Ben Poole, Ajay Jain, Jonathan T Barron, and Ben Mildenhall. 2022 · 2022
Closest in time.
Hierarchical text-conditional image generation with clip latents
Aditya Ramesh, Prafulla Dhariwal, Alex Nichol, Casey Chu, and Mark Chen. 2022 · 2022
Closest in time.
High-Resolution Image Synthesis with Latent Diffusion Models. In CVPR
Robin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser, and Björn Ommer. 2022 · 2022
Closest in time.
Nataniel Ruiz, Yuanzhen Li, Varun Jampani, Yael Pritch, Michael Rubinstein, and Kfir Aberman. 2022 · 2022
Closest in time.
Photorealistic text-to-image diffusion models with deep language understanding
Chitwan Saharia, William Chan, Saurabh Saxena, Lala Li, Jay Whang, Emily L Denton, Kamyar Ghasemipour, Raphael Gontijo Lopes, Burcu Karagol Ayan, Tim Salimans, et al · 2022
Closest in time.
Stylegan-xl: Scaling stylegan to large diverse datasets. In ACM SIGGRAPH
Axel Sauer, Katja Schwarz, and Andreas Geiger. 2022 · 2022
Closest in time.
Unsplash
Unsplash. 2022 · 2022
Closest in time.
Rewriting Geometric Rules of a GAN
Sheng-Yu Wang, David Bau, and Jun-Yan Zhu. 2022 · 2022
Closest in time.
Instructpix2pix: Learning to follow image editing instructions. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition . 18392–18402
Tim Brooks, Aleksander Holynski, and Alexei A Efros. 2023 · 2023
Closest in time.
Re-imagen: Retrieval-augmented text-to-image generator. In ICLR
Wenhu Chen, Hexiang Hu, Chitwan Saharia, and William W Cohen. 2023 · 2023
Closest in time.
Encoder-based domain tuning for fast personalization of text-to-image models
Rinon Gal, Moab Arar, Yuval Atzmon, Amit H Bermano, Gal Chechik, and Daniel Cohen-Or. 2023 · 2023
Closest in time.
Svdiff: Compact parameter space for diffusion fine-tuning
Ligong Han, Yinxiao Li, Han Zhang, Peyman Milanfar, Dimitris Metaxas, and Feng Yang. 2023 · 2023
Closest in time.
Imagic: Text-based real image editing with diffusion models
Bahjat Kawar, Shiran Zada, Oran Lang, Omer Tov, Huiwen Chang, Tali Dekel, Inbar Mosseri, and Michal Irani. 2023 · 2023
Closest in time.
Unified multi-modal latent diffusion for joint subject and text conditional image generation
Yiyang Ma, Huan Yang, Wenjing Wang, Jianlong Fu, and Jiaying Liu. 2023 · 2023
Closest in time.
Hugginggpt: Solving ai tasks with chatgpt and its friends in huggingface
Yongliang Shen, Kaitao Song, Xu Tan, Dongsheng Li, Weiming Lu, and Yueting Zhuang. 2023 · 2023
Closest in time.
Adding conditional control to text-to-image diffusion models. In Proceedings of the IEEE/CVF International Conference on Computer Vision . 3836–3847
Lvmin Zhang, Anyi Rao, and Maneesh Agrawala. 2023 · 2023
Closest in time.
Styleclip: Text-driven manipulation of stylegan imagery. In ICCV . 2085–2094
Or Patashnik, Zongze Wu, Eli Shechtman, Daniel Cohen-Or, and Dani Lischinski. 2021 · 2094
Closest in time.