Fetching the paper…
Reading the bibliography…
The visual world around us can be described as a structured set of objects and their associated relations.
Layoutgan: Generating graphic layouts with wireframe discriminators
Jianan Li, Jimei Yang, Aaron Hertzmann, Jianming Zhang, and Tingfa Xu · 1901
Earlier work this paper cites.
Training products of experts by minimizing contrastive divergence
Geoffrey E. Hinton · 2002
Earlier work this paper cites.
Compositional visual generation and inference with energy based models
Yilun Du, Shuang Li, and Igor Mordatch · 2004
Earlier work this paper cites.
Navigation among movable obstacles: Real-time reasoning in complex environments
Mike Stilman and James J Kuffner · 2005
Earlier work this paper cites.
A tutorial on energy-based learning
Yann LeCun, Sumit Chopra, Raia Hadsell, Marc’Aurelio Ranzato, and Fu-Jie Huang · 2006
Earlier work this paper cites.
Improved contrastive divergence training of energy based models
Yilun Du, Shuang Li, Joshua Tenenbaum, and Igor Mordatch · 2012
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba · 2014
Earlier work this paper cites.
Draw: A recurrent neural network for image generation
Karol Gregor, Ivo Danihelka, Alex Graves, Danilo Rezende, and Daan Wierstra · 2015
Earlier work this paper cites.
Human-level concept learning through probabilistic program induction
B. Lake, R. Salakhutdinov, and J. Tenenbaum · 2015
Earlier work this paper cites.
Generating images from captions with attention
Elman Mansimov, Emilio Parisotto, Jimmy Lei Ba, and Ruslan Salakhutdinov · 2015
Earlier work this paper cites.
Interaction networks for learning about objects, relations and physics
Peter W Battaglia, Razvan Pascanu, Matthew Lai, Danilo Rezende, and Koray Kavukcuoglu · 2016
Earlier work this paper cites.
Deep directed generative models with energy-based probability estimation
Taesup Kim and Yoshua Bengio · 2016
Earlier work this paper cites.
Learning physical intuition of block towers by example
Adam Lerer, Sam Gross, and Rob Fergus · 2016
Earlier work this paper cites.
Conditional image generation with pixelcnn decoders
Aaron van den Oord, Nal Kalchbrenner, Oriol Vinyals, Lasse Espeholt, Alex Graves, and Koray Kavukcuoglu · 2016
Earlier work this paper cites.
A theory of generative convnet
Jianwen Xie, Yang Lu, Song-Chun Zhu, and Yingnian Wu · 2016
Earlier work this paper cites.
Clevr: A diagnostic dataset for compositional language and elementary visual reasoning
Justin Johnson, Bharath Hariharan, Laurens Van Der Maaten, Li Fei-Fei, C Lawrence Zitnick, and Ross Girshick · 2017
Earlier work this paper cites.
Visual genome: Connecting language and vision using crowdsourced dense image annotations
Ranjay Krishna, Yuke Zhu, Oliver Groth, Justin Johnson, Kenji Hata, Joshua Kravitz, Stephanie Chen, Yannis Kalantidis, Li-Jia Li, David A Shamma, et al · 2017
Cited alongside, same era.
Discovering objects and their relations from entangled scene representations
David Raposo, Adam Santoro, David Barrett, Razvan Pascanu, Timothy Lillicrap, and Peter Battaglia · 2017
Cited alongside, same era.
Parallel multiscale autoregressive density estimation
Scott Reed, Aäron Oord, Nal Kalchbrenner, Sergio Gómez Colmenarejo, Ziyu Wang, Yutian Chen, Dan Belov, and Nando Freitas · 2017
Cited alongside, same era.
Stackgan: Text to photo-realistic image synthesis with stacked generative adversarial networks
Han Zhang, Tao Xu, Hongsheng Li, Shaoting Zhang, Xiaogang Wang, Xiaolei Huang, and Dimitris Metaxas · 2017
Cited alongside, same era.
Blender - a 3D modelling and rendering package
Blender Online Community · 2018
Cited alongside, same era.
Using scene graph context to improve image generation
Subarna Tripathi, Anahita Bhiwandiwalla, Alexei Bastidas, and Hanlin Tang · 2019
Later among the works it cites.
Stackgan++: Realistic image synthesis with stacked generative adversarial networks
Han Zhang, Tao Xu, Hongsheng Li, Shaoting Zhang, Xiaogang Wang, Xiaolei Huang, and Dimitris N. Metaxas · 2019
Later among the works it cites.
Compositional video synthesis with action graphs
Amir Bar, Roei Herzig, Xiaolong Wang, Gal Chechik, Trevor Darrell, and Amir Globerson · 2020
Later among the works it cites.
Exponential family estimation via adversarial dynamics embedding, 2020
Bo Dai, Zhen Liu, Hanjun Dai, Niao He, Arthur Gretton, Le Song, and Dale Schuurmans · 2020
Later among the works it cites.
Flow contrastive estimation of energy-based models
Ruiqi Gao, Erik Nijkamp, Diederik P Kingma, Zhen Xu, Andrew M Dai, and Ying Nian Wu · 2020
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
An implicit generative model for small molecular graphs. arxiv preprint 2018
N De Cao and Kipf TMGAN · 2018
Cited alongside, same era.
Keep drawing it: Iterative language-based image generation and editing
Alaaeldin El-Nouby, Shikhar Sharma, Hannes Schulz, Devon Hjelm, Layla El Asri, Samira Ebrahimi Kahou, Yoshua Bengio, and Graham W Taylor · 2018
Cited alongside, same era.
Inferring semantic layout for hierarchical text-to-image synthesis
Seunghoon Hong, Dingdong Yang, Jongwook Choi, and Honglak Lee · 2018
Cited alongside, same era.
Image generation from scene graphs
Justin Johnson, Agrim Gupta, and Li Fei-Fei · 2018
Cited alongside, same era.
Learning neural random fields with inclusive auxiliary generators
Yunfu Song and Zhijian Ou · 2018
Cited alongside, same era.
Specifying object attributes and relations in interactive scene generation
Oron Ashual and Lior Wolf · 2019
Cited alongside, same era.
Implicit generation and generalization in energy-based models
Yilun Du and Igor Mordatch · 2019
Cited alongside, same era.
Later among the works it cites.
Learning canonical representations for scene graph to image generation
Roei Herzig, Amir Bar, Huijuan Xu, Gal Chechik, Trevor Darrell, and Amir Globerson · 2020
Later among the works it cites.
Training generative adversarial networks with limited data
Tero Karras, M. Aittala, Janne Hellsten, S. Laine, J. Lehtinen, and Timo Aila · 2020
Later among the works it cites.
House-gan: Relational generative adversarial networks for graph-constrained house layout generation
Nelson Nauata, Kai-Hung Chang, Chin-Yi Cheng, Greg Mori, and Yasutaka Furukawa · 2020
Later among the works it cites.
On the anatomy of mcmc-based maximum likelihood learning of energy-based models
Erik Nijkamp, Mitch Hill, Tian Han, Song-Chun Zhu, and Y. Wu · 2020
Later among the works it cites.
igibson, a simulation environment for interactive tasks in large realistic scenes
Bokui Shen, Fei Xia, Chengshu Li, Roberto Martín-Martín, Linxi Fan, Guanzhi Wang, Shyamal Buch, Claudia D’Arpino, Sanjana Srivastava, Lyne P Tchapmi, et al · 2020
Later among the works it cites.
Concept grounding with modular action-capsules in semantic video prediction
Wei Yu, Wenxin Chen, Songhenh Yin, Steve Easterbrook, and Animesh Garg · 2020
Later among the works it cites.
Integrated task and motion planning
Caelan Reed Garrett, Rohan Chitnis, Rachel Holladay, Beomjoon Kim, Tom Silver, Leslie Pack Kaelbling, and Tomás Lozano-Pérez · 2021
Closest in time.
Exploiting relationship for complex-scene image generation
Tianyu Hua, Hongdong Zheng, Yalong Bai, Wei Zhang, X. Zhang, and Tao Mei · 2021
Closest in time.
Learning transferable visual models from natural language supervision
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, et al · 2021
Closest in time.
Zero-shot text-to-image generation
Aditya Ramesh, Mikhail Pavlov, Gabriel Goh, Scott Gray, Chelsea Voss, Alec Radford, Mark Chen, and Ilya Sutskever · 2021
Closest in time.