Fetching the paper…
Reading the bibliography…
This paper presents a novel generative model, Collaborative Competitive Agents (CCA), which leverages the capabilities of multiple Large Language Models (LLMs) based agents to execute complex tasks.
Image inpainting
Bertalmio M, Sapiro G, Caselles V, Ballester C · 2000
Earlier work this paper cites.
Object removal by exemplar-based inpainting
Criminisi A, Perez P, Toyama K · 2003
Earlier work this paper cites.
Image completion with structure propagation
Sun J, Yuan L, Jia J, Shum H Y · 2005
Earlier work this paper cites.
Generative adversarial nets
Goodfellow I, Pouget-Abadie J, Mirza M, Xu B, Warde-Farley D, Ozair S, Courville A, Bengio Y · 2014
Earlier work this paper cites.
Perceptual losses for real-time style transfer and super-resolution
Johnson J, Alahi A, Fei-Fei L · 2016
Earlier work this paper cites.
Image style transfer using convolutional neural networks
Gatys L A, Ecker A S, Bethge M · 2016
Earlier work this paper cites.
Unpaired image-to-image translation using cycle-consistent adversarial networks
Zhu J Y, Park T, Isola P, Efros A A · 2017
Earlier work this paper cites.
Image-to-image translation with conditional adversarial networks
Isola P, Zhu J Y, Zhou T, Efros A A · 2017
Earlier work this paper cites.
Large scale gan training for high fidelity natural image synthesis
Brock A, Donahue J, Simonyan K · 2018
Earlier work this paper cites.
Reinforcement learning: An introduction
Sutton R S, Barto A G · 2018
Earlier work this paper cites.
Arbitrary style transfer with deep feature reshuffle
Gu S, Chen C, Liao J, Yuan L · 2018
Earlier work this paper cites.
A style-based generator architecture for generative adversarial networks
Karras T, Laine S, Aila T · 2019
Earlier work this paper cites.
Analyzing and improving the image quality of StyleGAN
Karras T, Laine S, Aittala M, Hellsten J, Lehtinen J, Aila T · 2020
Earlier work this paper cites.
Denoising diffusion probabilistic models
Ho J, Jain A, Abbeel P · 2020
Earlier work this paper cites.
Interpreting the latent space of gans for semantic face editing
Shen Y, Gu J, Tang X, Zhou B · 2020
Earlier work this paper cites.
In-domain gan inversion for real image editing
Zhu J, Shen Y, Zhao D, Zhou B · 2020
Earlier work this paper cites.
Score-based generative modeling through stochastic differential equations
Song Y, Sohl-Dickstein J, Kingma D P, Kumar A, Ermon S, Poole B · 2021
Earlier work this paper cites.
Diffusion models beat gans on image synthesis
Dhariwal P, Nichol A · 2021
Earlier work this paper cites.
Improved denoising diffusion probabilistic models
Nichol A Q, Dhariwal P · 2021
Earlier work this paper cites.
Sdedit: Guided image synthesis and editing with stochastic differential equations
Meng C, He Y, Song Y, Song J, Wu J, Zhu J Y, Ermon S · 2021
Earlier work this paper cites.
Learning transferable visual models from natural language supervision
Radford A, Kim J W, Hallacy C, Ramesh A, Goh G, Agarwal S, Sastry G, Askell A, Mishkin P, Clark J, others · 2021
Earlier work this paper cites.
Training language models to follow instructions with human feedback
Ouyang L, Wu J, Jiang X, Almeida D, Wainwright C, Mishkin P, Zhang C, Agarwal S, Slama K, Ray A, others · 2022
Earlier work this paper cites.
Webshop: Towards scalable real-world web interaction with grounded language agents
Yao S, Chen H, Yang J, Narasimhan K · 2022
Earlier work this paper cites.
Elucidating the design space of diffusion-based generative models
Karras T, Aittala M, Aila T, Laine S · 2022
Earlier work this paper cites.
Prompt-to-prompt image editing with cross-attention control
Hertz A, Mokady R, Tenenbaum J, Aberman K, Pritch Y, Cohen-or D · 2022
Cited alongside, same era.
Chain-of-thought prompting elicits reasoning in large language models
Wei J, Wang X, Schuurmans D, Bosma M, Xia F, Chi E, Le Q V, Zhou D, others · 2022
Cited alongside, same era.
React: Synergizing reasoning and acting in language models
Yao S, Zhao J, Yu D, Du N, Shafran I, Narasimhan K, Cao Y · 2022
Cited alongside, same era.
Gan inversion: A survey
Xia W, Zhang Y, Yang Y, Xue J H, Zhou B, Yang M H · 2022
Cited alongside, same era.
Vector quantized diffusion model for text-to-image synthesis
Gu S, Chen D, Bao J, Wen F, Zhang B, Chen D, Yuan L, Guo B · 2022
Cited alongside, same era.
High-resolution image synthesis with latent diffusion models
Self-refine: Iterative refinement with self-feedback
Madaan A, Tandon N, Gupta P, Hallinan S, Gao L, Wiegreffe S, Alon U, Dziri N, Prabhumoye S, Yang Y, others · 2023
Later among the works it cites.
Idea2img: Iterative self-refinement with gpt-4v (ision) for automatic image design and generation
Yang Z, Wang J, Li L, Lin K, Lin C C, Liu Z, Wang L · 2023
Later among the works it cites.
Hugginggpt: Solving ai tasks with chatgpt and its friends in huggingface
Shen Y, Song K, Tan X, Li D, Lu W, Zhuang Y · 2023
Later among the works it cites.
Palm-e: An embodied multimodal language model
Driess D, Xia F, Sajjadi M S, Lynch C, Chowdhery A, Ichter B, Wahid A, Tompson J, Vuong Q, Yu T, others · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Rombach R, Blattmann A, Lorenz D, Esser P, Ommer B · 2022
Cited alongside, same era.
Hierarchical text-conditional image generation with clip latents
Ramesh A, Dhariwal P, Nichol A, Chu C, Chen M · 2022
Cited alongside, same era.
Photorealistic text-to-image diffusion models with deep language understanding
Saharia C, Chan W, Saxena S, Li L, Whang J, Denton E L, Ghasemipour K, Gontijo Lopes R, Karagol Ayan B, Salimans T, others · 2022
Cited alongside, same era.
ediffi: Text-to-image diffusion models with an ensemble of expert denoisers
Balaji Y, Nah S, Huang X, Vahdat A, Song J, Kreis K, Aittala M, Aila T, Laine S, Catanzaro B, others · 2022
Cited alongside, same era.
Improved aesthetic predictor, 2022
Schuhmann C · 2022
Cited alongside, same era.
Gpt-4v(ision) system card, 9 2023
OpenAI · 2023
Cited alongside, same era.
Gpt-4 technical report, 2023
OpenAI · 2023
Cited alongside, same era.
Li G, Hammoud H A A K, Itani H, Khizbullin D, Ghanem B · 2023
Later among the works it cites.
Agentverse: Facilitating multi-agent collaboration and exploring emergent behaviors in agents
Chen W, Su Y, Zuo J, Yang C, Yuan C, Qian C, Chan C M, Qin Y, Lu Y, Xie R, others · 2023
Later among the works it cites.
Chateval: Towards better llm-based evaluators through multi-agent debate
Chan C M, Chen W, Su Y, Yu J, Xue W, Zhang S, Fu J, Liu Z · 2023
Later among the works it cites.
Language-guided face animation by recurrent stylegan-based generator
Hang T, Yang H, Liu B, Fu J, Geng X, Guo B · 2023
Later among the works it cites.
Null-text inversion for editing real images using guided diffusion models
Mokady R, Hertz A, Aberman K, Pritch Y, Cohen-Or D · 2023
Later among the works it cites.
Paint by example: Exemplar-based image editing with diffusion models
Yang B, Gu S, Zhang B, Zhang T, Chen X, Sun X, Chen D, Wen F · 2023
Later among the works it cites.
Magicbrush: A manually annotated dataset for instruction-guided image editing, 2023
Zhang K, Mo L, Chen W, Sun H, Su Y · 2023
Later among the works it cites.
Edict: Exact diffusion inversion via coupled transformations
Wallace B, Gokul A, Naik N · 2023
Later among the works it cites.
Dreambooth: Fine tuning text-to-image diffusion models for subject-driven generation
Ruiz N, Li Y, Jampani V, Pritch Y, Rubinstein M, Aberman K · 2023
Later among the works it cites.
Visual programming: Compositional visual reasoning without training
Gupta T, Kembhavi A · 2023
Later among the works it cites.
Improved baselines with visual instruction tuning
Liu H, Li C, Li Y, Lee Y J · 2023
Later among the works it cites.
Grounding dino: Marrying dino with grounded pre-training for open-set object detection
Liu S, Zeng Z, Ren T, Li F, Zhang H, Yang J, Li C, Yang J, Su H, Zhu J, others · 2023
Later among the works it cites.
Improved noise schedule for diffusion training, 2024
Hang T, Gu S, Geng X, Guo B · 2024
Closest in time.
Fine-grained control of generative data augmentation in iot sensing
Wang T, Yang Q, Wang R, Sun D, Li J, Chen Y, Hu Y, Yang C, Kimura T, Kara D, others · 2024
Closest in time.
Composerx: Multi-agent symbolic music composition with llms
Deng Q, Yang Q, Yuan R, Huang Y, Wang Y, Liu X, Tian Z, Pan J, Zhang G, Lin H, others · 2024
Closest in time.
A glance at in-context learning
Yongliang WU X Y · 2024
Closest in time.
Instructdiffusion: A generalist modeling interface for vision tasks
Geng Z, Yang B, Hang T, Li C, Gu S, Zhang T, Bao J, Zhang Z, Hu H, Chen D, Guo B · 2024
Closest in time.
Regional style and color transfer
Ding Z, Li P, Yang Q, Li S, Gong Q · 2024
Closest in time.
Styleclip: Text-driven manipulation of stylegan imagery
Patashnik O, Wu Z, Shechtman E, Cohen-Or D, Lischinski D · 2094
Closest in time.