Fetching the paper…
Reading the bibliography…
In this paper, we introduce a novel generative model, Diffusion Layout Transformers without Autoencoder (Dolfin), which significantly improves the modeling capability with reduced complexity compared to existing methods.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber · 1997
Earlier work this paper cites.
Unsupervised learning of models for recognition
Markus Weber, Max Welling, and Pietro Perona · 2000
Earlier work this paper cites.
Object class recognition by unsupervised scale-invariant learning
Robert Fergus, Pietro Perona, and Andrew Zisserman · 2003
Earlier work this paper cites.
Learning hierarchical models of scenes, objects, and parts
Erik B Sudderth, Antonio Torralba, William T Freeman, and Alan S Willsky · 2005
Earlier work this paper cites.
Learning generative models via discriminative approaches
Zhuowen Tu · 2007
Earlier work this paper cites.
Auto-encoding variational bayes, 2013
Diederik P Kingma and Max Welling · 2013
Earlier work this paper cites.
Generative adversarial nets
Ian Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron Courville, and Yoshua Bengio · 2014
Earlier work this paper cites.
Microsoft coco: Common objects in context
Tsung-Yi Lin, Michael Maire, Serge Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Dollár, and C Lawrence Zitnick · 2014
Earlier work this paper cites.
Unsupervised representation learning with deep convolutional generative adversarial networks
Alec Radford, Luke Metz, and Soumith Chintala · 2015
Earlier work this paper cites.
Automatic generation of visual-textual presentation layout
Xuyong Yang, Tao Mei, Ying-Qing Xu, Yong Rui, and Shipeng Li · 2016
Earlier work this paper cites.
Wasserstein generative adversarial networks
Martin Arjovsky, Soumith Chintala, and Léon Bottou · 2017
Earlier work this paper cites.
Rico: A mobile app dataset for building data-driven design applications
Biplab Deka, Zifeng Huang, Chad Franzen, Joshua Hibschman, Daniel Afergan, Yang Li, Jeffrey Nichols, and Ranjitha Kumar · 2017
Earlier work this paper cites.
Introspective classification with convolutional nets
Long Jin, Justin Lazarow, and Zhuowen Tu · 2017
Earlier work this paper cites.
Introspective neural networks for generative modeling
Justin Lazarow, Long Jin, and Zhuowen Tu · 2017
Earlier work this paper cites.
Inception-v4, inception-resnet and the impact of residual connections on learning
Christian Szegedy, Sergey Ioffe, Vincent Vanhoucke, and Alexander Alemi · 2017
Earlier work this paper cites.
Attention is all you need, 2017
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N. Gomez, Lukasz Kaiser, and Illia Polosukhin · 2017
Earlier work this paper cites.
Gans trained by a two time-scale update rule converge to a local nash equilibrium, 2018
Martin Heusel, Hubert Ramsauer, Thomas Unterthiner, Bernhard Nessler, and Sepp Hochreiter · 2018
Earlier work this paper cites.
Learning to parse wireframes in images of man-made environments
Kun Huang, Yifan Wang, Zihan Zhou, Tianjiao Ding, Shenghua Gao, and Yi Ma · 2018
Earlier work this paper cites.
Progressive growing of gans for improved quality, stability, and variation
Tero Karras, Timo Aila, Samuli Laine, and Jaakko Lehtinen · 2018
Cited alongside, same era.
Wasserstein introspective neural networks
Kwonjoon Lee, Weijian Xu, Fan Fan, and Zhuowen Tu · 2018
Cited alongside, same era.
Layoutvae: Stochastic scene layout generation from a label set
Akash Abdu Jyothi, Thibaut Durand, Jiawei He, Leonid Sigal, and Greg Mori · 2019
Cited alongside, same era.
Bart: Denoising sequence-to-sequence pre-training for natural language generation, translation, and comprehension, 2019
Mike Lewis, Yinhan Liu, Naman Goyal, Marjan Ghazvininejad, Abdelrahman Mohamed, Omer Levy, Ves Stoyanov, and Luke Zettlemoyer · 2019
Cited alongside, same era.
Layoutgan: Generating graphic layouts with wireframe discriminators, 2019
Jianan Li, Jimei Yang, Aaron Hertzmann, Jianming Zhang, and Tingfa Xu · 2019
Cited alongside, same era.
Maskgit: Masked generative image transformer, 2022
Huiwen Chang, Han Zhang, Lu Jiang, Ce Liu, and William T. Freeman · 2022
Later among the works it cites.
Doc2ppt: Automatic presentation slides generation from scientific documents
Tsu-Jui Fu, William Yang Wang, Daniel McDuff, and Yale Song · 2022
Later among the works it cites.
Vector quantized diffusion model for text-to-image synthesis, 2022
Shuyang Gu, Dong Chen, Jianmin Bao, Fang Wen, Bo Zhang, Dongdong Chen, Lu Yuan, and Baining Guo · 2022
Later among the works it cites.
Blt: bidirectional layout transformer for controllable layout generation
Xiang Kong, Lu Jiang, Huiwen Chang, Han Zhang, Yuan Hao, Haifeng Gong, and Irfan Essa · 2022
Later among the works it cites.
Glide: Towards photorealistic image generation and editing with text-guided diffusion models
Alexander Quinn Nichol, Prafulla Dhariwal, Aditya Ramesh, Pranav Shyam, Pamela Mishkin, Bob Mcgrew, Ilya Sutskever, and Mark Chen · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Ying Cao Xinru Zheng, Xiaotian Qiao and Rynson W.H. Lau · 2019
Cited alongside, same era.
Content-aware generative modeling of graphic design layouts
Xinru Zheng, Xiaotian Qiao, Ying Cao, and Rynson W. H. Lau · 2019
Cited alongside, same era.
Publaynet: largest dataset ever for document layout analysis, 2019
Xu Zhong, Jianbin Tang, and Antonio Jimeno Yepes · 2019
Cited alongside, same era.
Denoising diffusion probabilistic models
Jonathan Ho, Ajay Jain, and Pieter Abbeel · 2020
Cited alongside, same era.
Neural design network: Graphic layout generation with constraints, 2020
Hsin-Ying Lee, Lu Jiang, Irfan Essa, Phuong B Le, Haifeng Gong, Ming-Hsuan Yang, and Weilong Yang · 2020
Cited alongside, same era.
House-gan: Relational generative adversarial networks for graph-constrained house layout generation
Nelson Nauata, Kai-Hung Chang, Chin-Yi Cheng, Greg Mori, and Yasutaka Furukawa · 2020
Cited alongside, same era.
Variational transformer networks for layout generation
Diego Martin Arroyo, Janis Postels, and Federico Tombari · 2021
Cited alongside, same era.
William Peebles and Saining Xie · 2022
Later among the works it cites.
High-resolution image synthesis with latent diffusion models
Robin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser, and Björn Ommer · 2022
Later among the works it cites.
Housediffusion: Vector floorplan generation via a diffusion model with discrete and continuous denoising, 2022
Mohammad Amin Shabani, Sepidehsadat Hosseini, and Yasutaka Furukawa · 2022
Later among the works it cites.
3d neural field generation using triplane diffusion
J Ryan Shue, Eric Ryan Chan, Ryan Po, Zachary Ankner, Jiajun Wu, and Gordon Wetzstein · 2022
Later among the works it cites.
Paint2pix: Interactive painting based progressive image synthesis and editing, 2022
Jaskirat Singh, Liang Zheng, Cameron Smith, and Jose Echevarria · 2022
Later among the works it cites.
Denoising diffusion implicit models, 2022
Jiaming Song, Chenlin Meng, and Stefano Ermon · 2022
Later among the works it cites.
Layoutdm: Transformer-based diffusion model for layout generation, 2023
Shang Chai, Liansheng Zhuang, and Fengying Yan · 2023
Closest in time.
Single-stage diffusion nerf: A unified approach to 3d generation and reconstruction
Hansheng Chen, Jiatao Gu, Anpei Chen, Wei Tian, Zhuowen Tu, Lingjie Liu, and Hao Su · 2023
Closest in time.
Play: Parametrically conditioned layout generation using latent diffusion, 2023
Chin-Yi Cheng, Forrest Huang, Gang Li, and Yang Li · 2023
Closest in time.
Layoutdm: Discrete diffusion model for controllable layout generation
Naoto Inoue, Kotaro Kikuchi, Edgar Simo-Serra, Mayu Otani, and Kota Yamaguchi · 2023
Closest in time.
Dlt: Conditioned layout generation with joint discrete-continuous diffusion layout transformer
Elad Levi, Eli Brosh, Mykola Mykhailych, and Meir Perez · 2023
Closest in time.
Adding conditional control to text-to-image diffusion models
Lvmin Zhang and Maneesh Agrawala · 2023
Closest in time.