Fetching the paper…
Reading the bibliography…
AI-generated content (AIGC) methods aim to produce text, images, videos, 3D assets, and other media using AI algorithms.
Geometric modeling using octree encoding
Donald Meagher · 1982
Earlier work this paper cites.
Signal estimation from modified short-time fourier transform
Daniel Griffin and Jae Lim · 1984
Earlier work this paper cites.
Animated conversation: rule-based generation of facial expression, gesture & spoken intonation for multiple conversational agents
Justine Cassell, Catherine Pelachaud, Norman Badler, et al · 1994
Earlier work this paper cites.
A morphable model for the synthesis of 3d faces
Volker Blanz and Thomas Vetter · 1999
Earlier work this paper cites.
Pose space deformation: a unified approach to shape interpolation and skeleton-driven deformation
John P Lewis et al · 2000
Earlier work this paper cites.
Multi-weight enveloping: least-squares approximation techniques for skin animation
Xiaohuan Corina Wang and Cary Phillips · 2002
Earlier work this paper cites.
A first look at music composition using lstm recurrent neural networks
Douglas Eck and Juergen Schmidhuber · 2002
Earlier work this paper cites.
On visual similarity based 3d model retrieval
Ding-Yun Chen, Xiao-Pei Tian, Yu-Te Shen, et al · 2003
Earlier work this paper cites.
Rhythmic-motion synthesis based on motion-beat analysis
Tae-hoon Kim, Sang Il Park, and Sung Yong Shin · 2003
Earlier work this paper cites.
Deformation transfer for triangle meshes
Robert W Sumner and Jovan Popović · 2004
Earlier work this paper cites.
Recognizing human actions: a local svm approach
Christian Schuldt, Ivan Laptev, and Barbara Caputo · 2004
Earlier work this paper cites.
Estimating 3d shape and texture using pixel intensity, edges, specular highlights, texture constraints and a prior
Sami Romdhani and Thomas Vetter · 2005
Earlier work this paper cites.
Face transfer with multilinear models
Daniel Vlasic, Matthew Brand, Hanspeter Pfister, et al · 2006
Earlier work this paper cites.
A 3d facial expression database for facial behavior research
Lijun Yin, Xiaozhou Wei, Yi Sun, et al · 2006
Earlier work this paper cites.
An audio-visual corpus for speech perception and automatic speech recognition
Martin Cooke, Jon Barker, Stuart Cunningham, et al · 2006
Earlier work this paper cites.
Cat head detection-how to effectively exploit shape and texture features
Weiwei Zhang, Jian Sun, and Xiaoou Tang · 2008
Earlier work this paper cites.
Geometric skinning with approximate dual quaternion blending
Ladislav Kavan, Steven Collins, Jiří Žára, et al · 2008
Earlier work this paper cites.
Automated flower classification over a large number of classes
Maria-Elena Nilsback and Andrew Zisserman · 2008
Earlier work this paper cites.
Learning multiple layers of features from tiny images
Alex Krizhevsky · 2009
Earlier work this paper cites.
Robust blind watermarking of point-sampled geometry
Parag Agarwal and Balakrishnan Prabhakaran · 2009
Earlier work this paper cites.
A 3d face model for pose and illumination invariant face recognition
Pascal Paysan, Reinhard Knothe, Brian Amberg, et al · 2009
Earlier work this paper cites.
The caltech-ucsd birds-200-2011 dataset
Catherine Wah, Steve Branson, Peter Welinder, et al · 2011
Earlier work this paper cites.
Collecting highly parallel data for paraphrase evaluation
David Chen and William B Dolan · 2011
Earlier work this paper cites.
Ucf101: A dataset of 101 human actions classes from videos in the wild
Khurram Soomro, Amir Roshan Zamir, and Mubarak Shah · 2012
Earlier work this paper cites.
Reconstructing detailed dynamic face geometry from monocular video
Pablo Garrido, Levi Valgaerts, Chenglei Wu, et al · 2013
Earlier work this paper cites.
Facewarehouse: A 3d facial expression database for visual computing
Chen Cao, Yanlin Weng, Shun Zhou, et al · 2013
Earlier work this paper cites.
Human3. 6m: Large scale datasets and predictive methods for 3d human sensing in natural environments
Catalin Ionescu, Dragos Papava, Vlad Olaru, et al · 2013
Earlier work this paper cites.
Generative adversarial nets
Ian Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, et al · 2014
Earlier work this paper cites.
Auto-encoding variational bayes
Diederik P Kingma and Max Welling · 2014
Earlier work this paper cites.
Stochastic backpropagation and approximate inference in deep generative models
Danilo Jimenez Rezende, Shakir Mohamed, and Daan Wierstra · 2014
Earlier work this paper cites.
Conditional generative adversarial nets
Mehdi Mirza and Simon Osindero · 2014
Earlier work this paper cites.
Microsoft coco: Common objects in context
Tsung-Yi Lin, Michael Maire, Serge Belongie, et al · 2014
Earlier work this paper cites.
Beyond pascal: A benchmark for 3d object detection in the wild
Yu Xiang, Roozbeh Mottaghi, and Silvio Savarese · 2014
Earlier work this paper cites.
3d shapenets: A deep representation for volumetric shapes
Zhirong Wu, Shuran Song, Aditya Khosla, et al · 2015
Earlier work this paper cites.
Nice: Non-linear independent components estimation
Laurent Dinh, David Krueger, and Yoshua Bengio · 2015
Earlier work this paper cites.
Importance weighted autoencoders
Yuri Burda, Roger Grosse, and Ruslan Salakhutdinov · 2015
Earlier work this paper cites.
Variational inference with normalizing flows
Danilo Rezende and Shakir Mohamed · 2015
Earlier work this paper cites.
Deep unsupervised learning using nonequilibrium thermodynamics
Jascha Sohl-Dickstein, Eric Weiss, Niru Maheswaranathan, et al · 2015
Earlier work this paper cites.
Unsupervised learning of video representations using lstms
Nitish Srivastava, Elman Mansimov, and Ruslan Salakhudinov · 2015
Earlier work this paper cites.
Shapenet: An information-rich 3d model repository
Angel X Chang, Thomas Funkhouser, Leonidas Guibas, et al · 2015
Earlier work this paper cites.
Smpl: A skinned multi-person linear model
Matthew Loper, Naureen Mahmood, Javier Romero, et al · 2015
Earlier work this paper cites.
Recurrent network models for human dynamics
Katerina Fragkiadaki, Sergey Levine, Panna Felsen, et al · 2015
Earlier work this paper cites.
Dynamic 3d avatar creation from hand-held video input
Alexandru Eugen Ichim, Sofien Bouaziz, and Mark Pauly · 2015
Earlier work this paper cites.
Lsun: Construction of a large-scale image dataset using deep learning with humans in the loop
Fisher Yu, Ari Seff, Yinda Zhang, et al · 2015
Earlier work this paper cites.
A large annotated corpus for learning natural language inference
Samuel Bowman, Gabor Angeli, Christopher Potts, et al · 2015
Earlier work this paper cites.
Deep learning face attributes in the wild
Ziwei Liu, Ping Luo, Xiaogang Wang, et al · 2015
Earlier work this paper cites.
A large-scale car dataset for fine-grained categorization and verification
Linjie Yang, Ping Luo, Chen Change Loy, et al · 2015
Earlier work this paper cites.
Microsoft coco captions: Data collection and evaluation server
Xinlei Chen, Hao Fang, Tsung-Yi Lin, et al · 2015
Earlier work this paper cites.
Generating videos with scene dynamics
Carl Vondrick, Hamed Pirsiavash, and Antonio Torralba · 2016
Earlier work this paper cites.
A deep learning framework for character motion synthesis and editing
Daniel Holden, Jun Saito, and Taku Komura · 2016
Earlier work this paper cites.
Wavenet: A generative model for raw audio
Aaron van den Oord, Sander Dieleman, Heiga Zen, et al · 2016
Earlier work this paper cites.
Improved techniques for training gans
Tim Salimans, Ian Goodfellow, Wojciech Zaremba, et al · 2016
Earlier work this paper cites.
Variational lossy autoencoder
Xi Chen, Diederik P Kingma, Tim Salimans, et al · 2016
Earlier work this paper cites.
Improved variational inference with inverse autoregressive flow
Durk P Kingma, Tim Salimans, Rafal Jozefowicz, et al · 2016
Earlier work this paper cites.
Towards conceptual compression
Karol Gregor, Frederic Besse, Danilo Jimenez Rezende, et al · 2016
Earlier work this paper cites.
Context encoders: Feature learning by inpainting
Deepak Pathak, Philipp Krahenbuhl, Jeff Donahue, et al · 2016
Earlier work this paper cites.
Generative visual manipulation on the natural image manifold
Jun-Yan Zhu, Philipp Krähenbühl, Eli Shechtman, et al · 2016
Earlier work this paper cites.
An uncertain future: Forecasting from static images using variational autoencoders
Jacob Walker, Carl Doersch, Abhinav Gupta, et al · 2016
Earlier work this paper cites.
Unsupervised learning for physical interaction through video prediction
Chelsea Finn, Ian Goodfellow, and Sergey Levine · 2016
Earlier work this paper cites.
Generating sentences from a continuous space
Samuel Bowman, Luke Vilnis, Oriol Vinyals, et al · 2016
Earlier work this paper cites.
Deep reinforcement learning for dialogue generation
Jiwei Li, Will Monroe, Alan Ritter, et al · 2016
Earlier work this paper cites.
Learning a probabilistic latent space of object shapes via 3d generative-adversarial modeling
Jiajun Wu, Chengkai Zhang, et al · 2016
Earlier work this paper cites.
Song from pi: A musically plausible network for pop music generation
Hang Chu, Raquel Urtasun, and Sanja Fidler · 2016
Earlier work this paper cites.
Deep learning for music
Allen Huang and Raymond Wu · 2016
Earlier work this paper cites.
Generating images from captions with attention
Elman Mansimov, Emilio Parisotto, Jimmy Lei Ba, et al · 2016
Earlier work this paper cites.
Generative adversarial text to image synthesis
Scott Reed, Zeynep Akata, Xinchen Yan, et al · 2016
Earlier work this paper cites.
Learning what and where to draw
Scott E Reed, Zeynep Akata, Santosh Mohan, et al · 2016
Earlier work this paper cites.
Learning deep representations of fine-grained visual descriptions
Scott Reed, Zeynep Akata, Honglak Lee, et al · 2016
Earlier work this paper cites.
Msr-vtt: A large video description dataset for bridging video and language
Jun Xu, Tao Mei, Ting Yao, et al · 2016
Earlier work this paper cites.
3d-r2n2: A unified approach for single and multi-view 3d object reconstruction
Christopher B Choy, Danfei Xu, JunYoung Gwak, et al · 2016
Earlier work this paper cites.
Learning a predictable and generative vector representation for objects
Rohit Girdhar, David F Fouhey, Mikel Rodriguez, et al · 2016
Earlier work this paper cites.
Unsupervised learning of 3d structure from images
Danilo Jimenez Rezende, SM Eslami, Shakir Mohamed, et al · 2016
Earlier work this paper cites.
Perspective transformer nets: Learning single-view 3d object reconstruction without 3d supervision
Xinchen Yan, Jimei Yang, et al · 2016
Earlier work this paper cites.
Adaptive 3d face reconstruction from unconstrained photo collections
Joseph Roth, Yiying Tong, and Xiaoming Liu · 2016
Earlier work this paper cites.
Real-time facial animation with image-based dynamic avatars
Chen Cao, Hongzhi Wu, Yanlin Weng, et al · 2016
Earlier work this paper cites.
Face2face: Real-time face capture and reenactment of rgb videos
Justus Thies, Michael Zollhofer, Marc Stamminger, et al · 2016
Earlier work this paper cites.
The cityscapes dataset for semantic urban scene understanding
Marius Cordts, Mohamed Omran, Sebastian Ramos, et al · 2016
Earlier work this paper cites.
Youtube-8m: A large-scale video classification benchmark
Sami Abu-El-Haija, Nisarg Kothari, Joonseok Lee, et al · 2016
Earlier work this paper cites.
The lambada dataset: Word prediction requiring a broad discourse context
Denis Paperno, Germán Kruszewski, Angeliki Lazaridou, et al · 2016
Earlier work this paper cites.
Squad: 100,000+ questions for machine comprehension of text
Pranav Rajpurkar, Jian Zhang, Konstantin Lopyrev, et al · 2016
Earlier work this paper cites.
Deepfashion: Powering robust clothes recognition and retrieval with rich annotations
Ziwei Liu, Ping Luo, Shi Qiu, et al · 2016
Earlier work this paper cites.
The kit motion-language dataset
Matthias Plappert, Christian Mandery, and Tamim Asfour · 2016
Earlier work this paper cites.
Ntu rgb+ d: A large scale dataset for 3d human activity analysis
Amir Shahroudy, Jun Liu, Tian-Tsong Ng, et al · 2016
Earlier work this paper cites.
Octnet: Learning deep 3d representations at high resolutions
Gernot Riegler, Ali Osman Ulusoy, and Andreas Geiger · 2017
Earlier work this paper cites.
Surfnet: Generating 3d shape surfaces using deep residual networks
Ayan Sinha, Asim Unmesh, Qixing Huang, et al · 2017
Earlier work this paper cites.
Gans trained by a two time-scale update rule converge to a local nash equilibrium
Martin Heusel, Hubert Ramsauer, et al · 2017
Earlier work this paper cites.
Improved training of wasserstein gans
Ishaan Gulrajani, Faruk Ahmed, Martin Arjovsky, et al · 2017
Earlier work this paper cites.
Neural discrete representation learning
Aaron Van Den Oord, Oriol Vinyals, et al · 2017
Earlier work this paper cites.
Density estimation using real nvp
Laurent Dinh, Jascha Sohl-Dickstein, and Samy Bengio · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, et al · 2017
Earlier work this paper cites.
Wasserstein generative adversarial networks
Martin Arjovsky, Soumith Chintala, and Léon Bottou · 2017
Earlier work this paper cites.
beta-vae: Learning basic visual concepts with a constrained variational framework
Irina Higgins, Loic Matthey, Arka Pal, et al · 2017
Earlier work this paper cites.
Conditional image synthesis with auxiliary classifier gans
Augustus Odena, Christopher Olah, and Jonathon Shlens · 2017
Earlier work this paper cites.
Plug & play generative networks: Conditional iterative generation of images in latent space
Anh Nguyen, Jeff Clune, Yoshua Bengio, et al · 2017
Earlier work this paper cites.
A learned representation for artistic style
Vincent Dumoulin, Jonathon Shlens, and Manjunath Kudlur · 2017
Earlier work this paper cites.
Globally and locally consistent image completion
Satoshi Iizuka, Edgar Simo-Serra, and Hiroshi Ishikawa · 2017
Earlier work this paper cites.
Image-to-image translation with conditional adversarial networks
Phillip Isola, Jun-Yan Zhu, Tinghui Zhou, et al · 2017
Earlier work this paper cites.
Unpaired image-to-image translation using cycle-consistent adversarial networks
Jun-Yan Zhu, Taesung Park, Phillip Isola, et al · 2017
Earlier work this paper cites.
Unsupervised image-to-image translation networks
Ming-Yu Liu, Thomas Breuel, and Jan Kautz · 2017
Earlier work this paper cites.
Learning to discover cross-domain relations with generative adversarial networks
Taeksoo Kim, Moonsu Cha, Hyunsoo Kim, et al · 2017
Earlier work this paper cites.
Temporal generative adversarial nets with singular value clipping
Masaki Saito, Eiichi Matsumoto, and Shunta Saito · 2017
Earlier work this paper cites.
Deep predictive coding networks for video prediction and unsupervised learning
William Lotter, Gabriel Kreiman, and David Cox · 2017
Earlier work this paper cites.
Triviaqa: A large scale distantly supervised challenge dataset for reading comprehension
Mandar Joshi, Eunsol Choi, et al · 2017
Earlier work this paper cites.
Improved variational autoencoders for text modeling using dilated convolutions
Zichao Yang, Zhiting Hu, Ruslan Salakhutdinov, et al · 2017
Earlier work this paper cites.
Seqgan: Sequence generative adversarial nets with policy gradient
Lantao Yu, Weinan Zhang, Jun Wang, et al · 2017
Earlier work this paper cites.
Adversarial feature matching for text generation
Yizhe Zhang, Zhe Gan, Kai Fan, et al · 2017
Earlier work this paper cites.
Shape completion using 3d-encoder-predictor cnns and shape synthesis
Angela Dai, Charles Ruizhongtai Qi, and Matthias Nießner · 2017
Earlier work this paper cites.
Grass: Generative recursive autoencoders for shape structures
Jun Li, Kai Xu, Siddhartha Chaudhuri, et al · 2017
Earlier work this paper cites.
Octree generating networks: Efficient convolutional architectures for high-resolution 3d outputs
Maxim Tatarchenko et al · 2017
Earlier work this paper cites.
Octnetfusion: Learning depth fusion from data
Gernot Riegler, Ali Osman Ulusoy, Horst Bischof, et al · 2017
Earlier work this paper cites.
Revisiting classifier two-sample tests
David Lopez-Paz and Maxime Oquab · 2017
Earlier work this paper cites.
Carla: An open urban driving simulator
Alexey Dosovitskiy, German Ros, Felipe Codevilla, et al · 2017
Earlier work this paper cites.
Learning a model of facial shape and expression from 4d scans
Tianye Li, Timo Bolkart, Michael J Black, et al · 2017
Earlier work this paper cites.
The lj speech dataset
Keith Ito and Linda Johnson · 2017
Earlier work this paper cites.
SampleRNN: An unconditional end-to-end neural audio generation model
Soroush Mehri, Kundan Kumar, Ishaan Gulrajani, et al · 2017
Earlier work this paper cites.
Neural audio synthesis of musical notes with wavenet autoencoders
Jesse Engel, Cinjon Resnick, Adam Roberts, et al · 2017
Earlier work this paper cites.
Midinet: A convolutional generative adversarial network for symbolic-domain music generation
Li-Chia Yang et al · 2017
Earlier work this paper cites.
Deepbach: a steerable model for bach chorales generation
Gaëtan Hadjeres, François Pachet, and Frank Nielsen · 2017
Earlier work this paper cites.
Stackgan: Text to photo-realistic image synthesis with stacked generative adversarial networks
Han Zhang, Tao Xu, Hongsheng Li, et al · 2017
Earlier work this paper cites.
Attentive semantic video generation using captions
Tanya Marwah, Gaurav Mittal, and Vineeth N Balasubramanian · 2017
Earlier work this paper cites.
A point set generation network for 3d object reconstruction from a single image
Haoqiang Fan, Hao Su, and Leonidas J Guibas · 2017
Earlier work this paper cites.
Learning detailed face reconstruction from a single image
Elad Richardson, Matan Sela, Roy Or-El, et al · 2017
Earlier work this paper cites.
Mofa: Model-based deep convolutional face autoencoder for unsupervised monocular reconstruction
Ayush Tewari, Michael Zollhofer, Hyeongwoo Kim, et al · 2017
Earlier work this paper cites.
Lip reading sentences in the wild
Joon Son Chung, Andrew Senior, Oriol Vinyals, et al · 2017
Earlier work this paper cites.
Audio-driven facial animation by joint end-to-end learning of pose and emotion
Tero Karras, Timo Aila, Samuli Laine, et al · 2017
Earlier work this paper cites.
Synthesizing obama: learning lip sync from audio
Supasorn Suwajanakorn, Steven M Seitz, and Ira Kemelmacher-Shlizerman · 2017
Earlier work this paper cites.
Tacotron: Towards end-to-end speech synthesis
Yuxuan Wang, RJ Skerry-Ryan, Daisy Stanton, et al · 2017
Earlier work this paper cites.
Revisiting unreasonable effectiveness of data in deep learning era
Chen Sun, Abhinav Shrivastava, Saurabh Singh, et al · 2017
Earlier work this paper cites.
Self-supervised visual planning with temporal skip connections
Frederik Ebert, Chelsea Finn, Alex X Lee, et al · 2017
Earlier work this paper cites.
Dynamic faust: Registering human bodies in motion
Federica Bogo, Javier Romero, Gerard Pons-Moll, et al · 2017
Earlier work this paper cites.
Dense-captioning events in videos
Ranjay Krishna, Kenji Hata, Frederic Ren, et al · 2017
Earlier work this paper cites.
Matterport3d: Learning from rgb-d data in indoor environments
Angel Chang, Angela Dai, Thomas Funkhouser, et al · 2017
Earlier work this paper cites.
Scannet: Richly-annotated 3d reconstructions of indoor scenes
Angela Dai, Angel X Chang, Manolis Savva, et al · 2017
Earlier work this paper cites.
Voxceleb: A large-scale speaker identification dataset
Arsha Nagrani, Joon Son Chung, and Andrew Zisserman · 2017
Earlier work this paper cites.
Monocular 3d human pose estimation in the wild using improved cnn supervision
Dushyant Mehta, Helge Rhodin, Dan Casas, et al · 2017
Earlier work this paper cites.
Audio set: An ontology and human-labeled dataset for audio events
Jort F Gemmeke, Daniel PW Ellis, Dylan Freedman, et al · 2017
Earlier work this paper cites.
Lip reading in the wild
Joon Son Chung and Andrew Zisserman · 2017
Earlier work this paper cites.
Mocogan: Decomposing motion and content for video generation
Sergey Tulyakov, Ming-Yu Liu, Xiaodong Yang, et al · 2018
Earlier work this paper cites.
Improving language understanding by generative pre-training
Alec Radford, Karthik Narasimhan, Tim Salimans, et al · 2018
Earlier work this paper cites.
Learning representations and generative models for 3d point clouds
Panos Achlioptas, Olga Diamanti, Ioannis Mitliagkas, et al · 2018
Earlier work this paper cites.
Efficient neural audio synthesis
Nal Kalchbrenner, Erich Elsen, Karen Simonyan, et al · 2018
Earlier work this paper cites.
Progressive growing of GANs for improved quality, stability, and variation
Tero Karras, Timo Aila, Samuli Laine, et al · 2018
Earlier work this paper cites.
Unsupervised representation learning with deep convolutional generative adversarial networks
Alec Radford, Luke Metz, et al · 2018
Earlier work this paper cites.
Spectral normalization for generative adversarial networks
Takeru Miyato, Toshiki Kataoka, Masanori Koyama, et al · 2018
Earlier work this paper cites.
Glow: Generative flow with invertible 1x1 convolutions
Durk P Kingma and Prafulla Dhariwal · 2018
Earlier work this paper cites.
Understanding disentangling in beta-vae
Christopher P Burgess, Irina Higgins, Arka Pal, et al · 2018
Earlier work this paper cites.
Neural ordinary differential equations
Ricky TQ Chen, Yulia Rubanova, Jesse Bettencourt, et al · 2018
Earlier work this paper cites.
Generative image inpainting with contextual attention
Jiahui Yu, Zhe Lin, Jimei Yang, et al · 2018
Earlier work this paper cites.
Shift-net: Image inpainting via deep feature rearrangement
Zhaoyi Yan, Xiaoming Li, Mu Li, et al · 2018
Earlier work this paper cites.
Image inpainting for irregular holes using partial convolutions
Guilin Liu, Fitsum A Reda, Kevin J Shih, et al · 2018
Earlier work this paper cites.
Diverse image-to-image translation via disentangled representations
Hsin-Ying Lee, Hung-Yu Tseng, Jia-Bin Huang, et al · 2018
Earlier work this paper cites.
Stargan: Unified generative adversarial networks for multi-domain image-to-image translation
Yunjey Choi, Minje Choi, et al · 2018
Earlier work this paper cites.
Sketchygan: Towards diverse and realistic sketch to image synthesis
Wengling Chen and James Hays · 2018
Earlier work this paper cites.
Learning to generate time-lapse videos using multi-stage dynamic generative adversarial networks
Wei Xiong, Wenhan Luo, Lin Ma, et al · 2018
Earlier work this paper cites.
Video-to-video synthesis
Ting-Chun Wang, Ming-Yu Liu, Jun-Yan Zhu, et al · 2018
Earlier work this paper cites.
Stochastic variational video prediction
Mohammad Babaeizadeh, Chelsea Finn, Dumitru Erhan, et al · 2018
Earlier work this paper cites.
Pose guided human video generation
Ceyuan Yang, Zhe Wang, Xinge Zhu, et al · 2018
Earlier work this paper cites.
Controllable video generation with sparse trajectories
Zekun Hao, Xun Huang, and Serge Belongie · 2018
Earlier work this paper cites.
Think you have solved question answering? try arc, the ai2 reasoning challenge
Peter Clark, Isaac Cowhey, Oren Etzioni, et al · 2018
Earlier work this paper cites.
Long text generation via adversarial training with leaked information
Jiaxian Guo, Sidi Lu, Han Cai, et al · 2018
Earlier work this paper cites.
Maskgan: Better text generation via filling in the _
William Fedus, Ian Goodfellow, and Andrew M Dai · 2018
Earlier work this paper cites.
Global-to-local generative model for 3d shapes
Hao Wang, Nadav Schor, Ruizhen Hu, et al · 2018
Earlier work this paper cites.
Convolutional sequence to sequence model for human dynamics
Chen Li, Zhen Zhang, Wee Sun Lee, et al · 2018
Earlier work this paper cites.
Hp-gan: Probabilistic 3d human motion prediction via gan
Emad Barsoum, John Kender, and Zicheng Liu · 2018
Earlier work this paper cites.
Speech commands: A dataset for limited-vocabulary speech recognition
Pete Warden · 2018
Earlier work this paper cites.
Parallel wavenet: Fast high-fidelity speech synthesis
Aaron Oord, Yazhe Li, Igor Babuschkin, et al · 2018
Earlier work this paper cites.
The challenge of realistic music generation: modelling raw audio at scale
Sander Dieleman, Aaron van den Oord, et al · 2018
Earlier work this paper cites.
Stackgan++: Realistic image synthesis with stacked generative adversarial networks
Han Zhang, Tao Xu, Hongsheng Li, et al · 2018
Earlier work this paper cites.
Attngan: Fine-grained text to image generation with attentional generative adversarial networks
Tao Xu, Pengchuan Zhang, et al · 2018
Earlier work this paper cites.
Video generation from text
Yitong Li, Martin Min, Dinghan Shen, et al · 2018
Earlier work this paper cites.
Neural 3d mesh renderer
Hiroharu Kato, Yoshitaka Ushiku, and Tatsuya Harada · 2018
Earlier work this paper cites.
Learning category-specific mesh reconstruction from image collections
Angjoo Kanazawa, Shubham Tulsiani, Alexei A Efros, et al · 2018
Earlier work this paper cites.
Pixel2mesh: Generating 3d mesh models from single rgb images
Nanyang Wang, Yinda Zhang, Zhuwen Li, et al · 2018
Earlier work this paper cites.
Deep marching cubes: Learning explicit surface representations
Yiyi Liao, Simon Donne, and Andreas Geiger · 2018
Earlier work this paper cites.
Video based reconstruction of 3d people models
Thiemo Alldieck, Marcus Magnor, Weipeng Xu, et al · 2018
Earlier work this paper cites.
End-to-end recovery of human shape and pose
Angjoo Kanazawa, Michael J Black, David W Jacobs, et al · 2018
Earlier work this paper cites.
Bodynet: Volumetric inference of 3d human body shapes
Gul Varol, Duygu Ceylan, Bryan Russell, et al · 2018
Earlier work this paper cites.
Self-supervised multi-level face model learning for monocular reconstruction at over 250 hz
Ayush Tewari, Michael Zollhöfer, et al · 2018
Earlier work this paper cites.
Text2action: Generative adversarial synthesis from language to action
Hyemin Ahn, Timothy Ha, Yunho Choi, et al · 2018
Earlier work this paper cites.
Dance with melody: An lstm-autoencoder approach to music-oriented dance synthesis
Taoran Tang, Jia Jia, and Hanyang Mao · 2018
Earlier work this paper cites.
Evaluation of speech-to-gesture generation using bi-directional lstm network
Dai Hasegawa, Naoshi Kaneko, Shinichi Shirakawa, et al · 2018
Earlier work this paper cites.
X2face: A network for controlling face generation using images, audio, and pose codes
Olivia Wiles, A Koepke, and Andrew Zisserman · 2018
Earlier work this paper cites.
Deep voice 3: Scaling text-to-speech with convolutional sequence learning
Wei Ping, Kainan Peng, Andrew Gibiansky, et al · 2018
Earlier work this paper cites.
Natural tts synthesis by conditioning wavenet on mel spectrogram predictions
Jonathan Shen, Ruoming Pang, Ron J Weiss, et al · 2018
Earlier work this paper cites.
A short note about kinetics-600
Joao Carreira, Eric Noland, Andras Banki-Horvath, et al · 2018
Earlier work this paper cites.
Faceforensics: A large-scale video dataset for forgery detection in human faces
Andreas Rössler, Davide Cozzolino, et al · 2018
Earlier work this paper cites.
Can a suit of armor conduct electricity? a new dataset for open book question answering
Todor Mihaylov, Peter Clark, et al · 2018
Earlier work this paper cites.
Conceptual captions: A cleaned, hypernymed, image alt-text dataset for automatic image captioning
Piyush Sharma, Nan Ding, et al · 2018
Earlier work this paper cites.
Pix3d: Dataset and methods for single-image 3d shape modeling
Xingyuan Sun, Jiajun Wu, Xiuming Zhang, et al · 2018
Earlier work this paper cites.
Voxceleb2: Deep speaker recognition
J Chung, A Nagrani, and A Zisserman · 2018
Earlier work this paper cites.
Recovering accurate 3d human pose in the wild using imus and a moving camera
Timo Von Marcard, Roberto Henschel, et al · 2018
Earlier work this paper cites.
Generating educational game levels with multistep deep convolutional generative adversarial networks
Kyungjin Park et al · 2019
Earlier work this paper cites.
Language models are unsupervised multitask learners
Alec Radford, Jeffrey Wu, Rewon Child, et al · 2019
Earlier work this paper cites.
Pointflow: 3d point cloud generation with continuous normalizing flows
Guandao Yang, Xun Huang, Zekun Hao, et al · 2019
Earlier work this paper cites.
Learning implicit fields for generative shape modeling
Zhiqin Chen and Hao Zhang · 2019
Earlier work this paper cites.
Deepsdf: Learning continuous signed distance functions for shape representation
Jeong Joon Park, Peter Florence, Julian Straub, et al · 2019
Earlier work this paper cites.
Hologan: Unsupervised learning of 3d representations from natural images
Thu Nguyen-Phuoc, Chuan Li, Lucas Theis, et al · 2019
Earlier work this paper cites.
A style-based generator architecture for generative adversarial networks
Tero Karras, Samuli Laine, and Timo Aila · 2019
Cited alongside, same era.
Self-attention generative adversarial networks
Han Zhang, Ian Goodfellow, Dimitris Metaxas, et al · 2019
Cited alongside, same era.
Large scale GAN training for high fidelity natural image synthesis
Andrew Brock, Jeff Donahue, Karen Simonyan, et al · 2019
Cited alongside, same era.
Preventing posterior collapse with delta-vaes
Ali Razavi, Aaron van den Oord, Ben Poole, et al · 2019
Cited alongside, same era.
Invertible residual networks
Jens Behrmann, Will Grathwohl, Ricky TQ Chen, et al · 2019
Cited alongside, same era.
Ffjord: Free-form continuous dynamics for scalable reversible generative models
Grathwohl Will, T. Q. Chen Ricky, et al · 2019
Cited alongside, same era.
Lion: Latent point diffusion models for 3d shape generation
Arash Vahdat, Francis Williams, Zan Gojcic, et al · 2022
Later among the works it cites.
gdna: Towards generative detailed neural avatars
Xu Chen, Tianjian Jiang, Jie Song, et al · 2022
Later among the works it cites.
Avatargen: a 3d generative model for animatable human avatars
Jianfeng Zhang, Zihang Jiang, Dingdong Yang, et al · 2022
Later among the works it cites.
Headnerf: A real-time nerf-based parametric head model
Yang Hong, Bo Peng, Haiyao Xiao, et al · 2022
Later among the works it cites.
High-resolution image synthesis with latent diffusion models
Robin Rombach, Andreas Blattmann, Dominik Lorenz, et al · 2022
Later among the works it cites.
ViTGAN: Training GANs with vision transformers
Kwonjoon Lee, Huiwen Chang, Lu Jiang, et al · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Residual flows for invertible generative modeling
Ricky TQ Chen, Jens Behrmann, David K Duvenaud, et al · 2019
Cited alongside, same era.
Flow++: Improving flow-based generative models with variational dequantization and architecture design
Jonathan Ho, Xi Chen, et al · 2019
Cited alongside, same era.
Generative modeling by estimating gradients of the data distribution
Yang Song and Stefano Ermon · 2019
Cited alongside, same era.
Generating diverse high-fidelity images with vq-vae-2
Ali Razavi, Aaron Van den Oord, and Oriol Vinyals · 2019
Cited alongside, same era.
Free-form image inpainting with gated convolution
Jiahui Yu, Zhe Lin, Jimei Yang, et al · 2019
Cited alongside, same era.
Few-shot unsupervised image-to-image translation
Ming-Yu Liu, Xun Huang, Arun Mallya, et al · 2019
Cited alongside, same era.
Styleswin: Transformer-based gan for high-resolution image generation
Bowen Zhang, Shuyang Gu, Bo Zhang, et al · 2022
Later among the works it cites.
Elucidating the design space of diffusion-based generative models
Tero Karras, Miika Aittala, Timo Aila, et al · 2022
Later among the works it cites.
Progressive distillation for fast sampling of diffusion models
Tim Salimans and Jonathan Ho · 2022
Later among the works it cites.
Cascaded diffusion models for high fidelity image generation
Jonathan Ho, Chitwan Saharia, William Chan, et al · 2022
Later among the works it cites.
Tackling the generative learning trilemma with denoising diffusion GANs
Zhisheng Xiao, Karsten Kreis, and Arash Vahdat · 2022
Later among the works it cites.
Palette: Image-to-image diffusion models
Chitwan Saharia, William Chan, Huiwen Chang, et al · 2022
Later among the works it cites.
Repaint: Inpainting using denoising diffusion probabilistic models
Andreas Lugmayr, Martin Danelljan, Andres Romero, et al · 2022
Later among the works it cites.
Paint2pix: interactive painting based progressive image synthesis and editing
Jaskirat Singh, Liang Zheng, Cameron Smith, et al · 2022
Later among the works it cites.
Stylegan-v: A continuous video generator with the price, image quality and perks of stylegan2
Ivan Skorokhodov, Sergey Tulyakov, et al · 2022
Later among the works it cites.
Generating videos with dynamics-aware implicit generative adversarial networks
Sihyun Yu, Jihoon Tack, Sangwoo Mo, et al · 2022
Later among the works it cites.
Video diffusion models
Jonathan Ho, Tim Salimans, Alexey Gritsenko, et al · 2022
Later among the works it cites.
Long video generation with time-agnostic vqgan and time-sensitive transformer
Songwei Ge, Thomas Hayes, Harry Yang, et al · 2022
Later among the works it cites.
Flexible diffusion modeling of long videos
William Harvey, Saeid Naderiparizi, Vaden Masrani, et al · 2022
Later among the works it cites.
Latent image animator: Learning to animate images via latent space navigation
Yaohui Wang, Di Yang, Francois Bremond, et al · 2022
Later among the works it cites.
Controllable animation of fluid elements in still images
Aniruddha Mahapatra and Kuldeep Kulkarni · 2022
Later among the works it cites.
Switch transformers: Scaling to trillion parameter models with simple and efficient sparsity
William Fedus, Barret Zoph, et al · 2022
Later among the works it cites.
Training compute-optimal large language models
Jordan Hoffmann, Sebastian Borgeaud, Arthur Mensch, et al · 2022
Later among the works it cites.
Palm: Scaling language modeling with pathways
Aakanksha Chowdhery, Sharan Narang, Jacob Devlin, et al · 2022
Later among the works it cites.
Training language models to follow instructions with human feedback
Long Ouyang, Jeffrey Wu, Xu Jiang, et al · 2022
Later among the works it cites.
Diffusion-lm improves controllable text generation
Xiang Li, John Thickstun, and Ishaan Gulrajani · 2022
Later among the works it cites.
Recent advancements in learning algorithms for point clouds: An updated overview
Elena Camuffo, Daniele Mari, et al · 2022
Later among the works it cites.
Autosdf: Shape priors for 3d completion, reconstruction and generation
Paritosh Mittal, Yen-Chi Cheng, Maneesh Singh, et al · 2022
Later among the works it cites.
Auv-net: Learning aligned uv maps for texture transfer and synthesis
Zhiqin Chen, Kangxue Yin, and Sanja Fidler · 2022
Later among the works it cites.
Texturify: Generating textures on 3d shape surfaces
Yawar Siddiqui, Justus Thies, Fangchang Ma, et al · 2022
Later among the works it cites.
Get3d: A generative model of high quality 3d textured shapes learned from images
Jun Gao, Tianchang Shen, Zian Wang, et al · 2022
Later among the works it cites.
Neumesh: Learning disentangled neural mesh-based implicit field for geometry and texture editing
Bangbang Yang, Chong Bao, et al · 2022
Later among the works it cites.
Voxgraf: Fast 3d-aware image synthesis with sparse voxel grids
Katja Schwarz, Axel Sauer, Michael Niemeyer, et al · 2022
Later among the works it cites.
Stylesdf: High-resolution 3d-consistent image and geometry generation
Roy Or-El, Xuan Luo, Mengyi Shan, et al · 2022
Later among the works it cites.
Gram: Generative radiance manifolds for 3d-aware image generation
Yu Deng, Jiaolong Yang, Jianfeng Xiang, et al · 2022
Later among the works it cites.
Stylenerf: A style-based 3d aware generator for high-resolution image synthesis
Jiatao Gu, Lingjie Liu, Peng Wang, et al · 2022
Later among the works it cites.
Efficient geometry-aware 3d generative adversarial networks
Eric R Chan, Connor Z Lin, Matthew A Chan, et al · 2022
Later among the works it cites.
3d-aware image synthesis via learning structural and textural representations
Yinghao Xu, Sida Peng, Ceyuan Yang, et al · 2022
Later among the works it cites.
Multi-view consistent generative adversarial networks for 3d-aware image synthesis
Xuanmeng Zhang, Zhedong Zheng, et al · 2022
Later among the works it cites.
Giraffe hd: A high-resolution 3d-aware generative model
Yang Xue, Yuheng Li, Krishna Kumar Singh, et al · 2022
Later among the works it cites.
Mip-nerf 360: Unbounded anti-aliased neural radiance fields
Jonathan T Barron, Ben Mildenhall, Dor Verbin, et al · 2022
Later among the works it cites.
Tensorf: Tensorial radiance fields
Anpei Chen, Zexiang Xu, Andreas Geiger, et al · 2022
Later among the works it cites.
Nerf-editing: geometry editing of neural radiance fields
Yu-Jie Yuan, Yang-Tian Sun, Yu-Kun Lai, et al · 2022
Later among the works it cites.
Deforming radiance fields with cages
Tianhan Xu and Tatsuya Harada · 2022
Later among the works it cites.
Back to mlp: A simple baseline for human motion prediction
Wen Guo, Yuming Du, Xi Shen, et al · 2022
Later among the works it cites.
Generative adversarial graph convolutional networks for human action synthesis
Bruno Degardin, Joao Neves, Vasco Lopes, et al · 2022
Later among the works it cites.
Mugl: Large scale multi person conditional action generation with locomotion
Shubh Maheshwari, Debtanu Gupta, et al · 2022
Later among the works it cites.
Photorealistic text-to-image diffusion models with deep language understanding
Chitwan Saharia, William Chan, et al · 2022
Later among the works it cites.
Df-gan: A simple and effective baseline for text-to-image synthesis
Ming Tao, Hao Tang, Fei Wu, et al · 2022
Later among the works it cites.
Cogview2: Faster and better text-to-image generation via hierarchical transformers
Ming Ding, Wendi Zheng, Wenyi Hong, et al · 2022
Later among the works it cites.
Nüwa: Visual synthesis pre-training for neural visual world creation
Chenfei Wu, Jian Liang, Lei Ji, et al · 2022
Later among the works it cites.
Make-a-scene: Scene-based text-to-image generation with human priors
Oran Gafni, Adam Polyak, Oron Ashual, et al · 2022
Later among the works it cites.
Scaling autoregressive models for content-rich text-to-image generation
Jiahui Yu, Yuanzhong Xu, Jing Yu Koh, et al · 2022
Later among the works it cites.
Glide: Towards photorealistic image generation and editing with text-guided diffusion models
Alexander Quinn Nichol et al · 2022
Later among the works it cites.
Vector quantized diffusion model for text-to-image synthesis
Shuyang Gu, Dong Chen, Jianmin Bao, et al · 2022
Later among the works it cites.
Hierarchical text-conditional image generation with clip latents
Aditya Ramesh, Prafulla Dhariwal, Alex Nichol, et al · 2022
Later among the works it cites.
ediffi: Text-to-image diffusion models with an ensemble of expert denoisers
Yogesh Balaji, Seungjun Nah, Xun Huang, et al · 2022
Later among the works it cites.
Retrieval-augmented diffusion models
Andreas Blattmann, Robin Rombach, Kaan Oktay, et al · 2022
Later among the works it cites.
Diffusionclip: Text-guided diffusion models for robust image manipulation
Gwanghyun Kim, Taesung Kwon, and Jong Chul Ye · 2022
Later among the works it cites.
Blended diffusion for text-driven editing of natural images
Omri Avrahami, Dani Lischinski, and Ohad Fried · 2022
Later among the works it cites.
Sketch-guided text-to-image diffusion models
Andrey Voynov, Kfir Aberman, and Daniel Cohen-Or · 2022
Later among the works it cites.
Pivotal tuning for latent-based editing of real images
Daniel Roich, Ron Mokady, Amit H Bermano, et al · 2022
Later among the works it cites.
On aliased resizing and surprising subtleties in gan evaluation
Gaurav Parmar, Richard Zhang, and Jun-Yan Zhu · 2022
Later among the works it cites.
Nuwa-infinity: Autoregressive over autoregressive generation for infinite visual synthesis
Jian Liang, Chenfei Wu, et al · 2022
Later among the works it cites.
Imagen video: High definition video generation with diffusion models
Jonathan Ho, William Chan, Chitwan Saharia, et al · 2022
Later among the works it cites.
Text2live: Text-driven layered image and video editing
Omer Bar-Tal, Dolev Ofri-Amar, Rafail Fridman, et al · 2022
Later among the works it cites.
Tune-a-video: One-shot tuning of image diffusion models for text-to-video generation
Jay Zhangjie Wu, Yixiao Ge, Xintao Wang, et al · 2022
Later among the works it cites.
Shapecrafter: A recursive text-conditioned 3d shape generation model
Rao Fu, Xiao Zhan, Yiwen Chen, et al · 2022
Later among the works it cites.
Clip-forge: Towards zero-shot text-to-shape generation
Aditya Sanghi, Hang Chu, Joseph G Lambourne, et al · 2022
Later among the works it cites.
Text2mesh: Text-driven neural stylization for meshes
Oscar Michel, Roi Bar-On, Richard Liu, et al · 2022
Later among the works it cites.
Zero-shot text-guided object generation with dream fields
Ajay Jain, Ben Mildenhall, Jonathan T Barron, et al · 2022
Later among the works it cites.
Pix2nerf: Unsupervised conditional p-gan for single image to neural radiance fields translation
Shengqu Cai, Anton Obukhov, et al · 2022
Later among the works it cites.
D^2neRF: Self-supervised decoupling of dynamic and static objects from a monocular video
Tianhao Walter Wu et al · 2022
Later among the works it cites.
Selfrecon: Self reconstruction your digital avatar from monocular video
Boyi Jiang, Yang Hong, Hujun Bao, et al · 2022
Later among the works it cites.
Clipface: Text-guided editing of textured 3d morphable models
Shivangi Aneja, Justus Thies, Angela Dai, et al · 2022
Later among the works it cites.
Photorealistic monocular 3d reconstruction of humans wearing clothing
Thiemo Alldieck, Mihai Zanfir, and Cristian Sminchisescu · 2022
Later among the works it cites.
Realistic one-shot mesh-based head avatars
Taras Khakhulin, Vanessa Sklyarova, Victor Lempitsky, et al · 2022
Later among the works it cites.
Mofanerf: Morphable facial neural radiance field
Yiyu Zhuang, Hao Zhu, Xusen Sun, et al · 2022
Later among the works it cites.
Icon: Implicit clothed humans obtained from normals
Yuliang Xiu, Jinlong Yang, Dimitrios Tzionas, et al · 2022
Later among the works it cites.
Capturing and animation of body and clothing from monocular video
Yao Feng, Jinlong Yang, Marc Pollefeys, et al · 2022
Later among the works it cites.
Generating diverse and natural 3d human motions from text
Chuan Guo, Shihao Zou, Xinxin Zuo, et al · 2022
Later among the works it cites.
Motionclip: Exposing human motion generation to clip space
Guy Tevet, Brian Gordon, Amir Hertz, et al · 2022
Later among the works it cites.
Avatarclip: zero-shot text-driven generation and animation of 3d avatars
Fangzhou Hong, Mingyuan Zhang, Liang Pan, et al · 2022
Later among the works it cites.
Rhythm is a dancer: Music-driven motion synthesis with global structure
Andreas Aristidou, Anastasios Yiannakidis, et al · 2022
Later among the works it cites.
Learning hierarchical cross-modal association for co-speech gesture generation
Xian Liu, Qianyi Wu, Hang Zhou, et al · 2022
Later among the works it cites.
Diffsinger: Singing voice synthesis via shallow diffusion mechanism
Jinglin Liu, Chengxi Li, Yi Ren, et al · 2022
Later among the works it cites.
Mulan: A joint embedding of music audio and natural language
Qingqing Huang, Aren Jansen, Joonseok Lee, et al · 2022
Later among the works it cites.
Stylegan-human: A data-centric odyssey of human generation
Jianglin Fu, Shikai Li, Yuming Jiang, et al · 2022
Later among the works it cites.
Text2human: Text-driven controllable human image generation
Yuming Jiang, Shuai Yang, Haonan Qiu, et al · 2022
Later among the works it cites.
Discovering transferable forensic features for cnn-generated images detection
Keshigeyan Chandrasegaran, Ngoc-Trung Tran, et al · 2022
Later among the works it cites.
Dall-e 2 pre-training mitigations
Alex Nichol · 2022
Later among the works it cites.
Large language models are few-shot clinical information extractors
Monica Agrawal, Stefan Hegselmann, Hunter Lang, et al · 2022
Later among the works it cites.
Factuality enhanced language models for open-ended text generation
Nayeon Lee, Wei Ping, Peng Xu, et al · 2022
Later among the works it cites.
Fast model editing at scale
Eric Mitchell, Charles Lin, Antoine Bosselut, et al · 2022
Later among the works it cites.
Locating and editing factual associations in gpt
Kevin Meng, David Bau, Alex Andonian, et al · 2022
Later among the works it cites.
Education in the era of generative artificial intelligence (ai): Understanding the potential benefits of chatgpt in promoting teaching and learning
David Baidoo-Anu and Leticia Owusu Ansah · 2023
Closest in time.
Chatgpt and other large language models as evolutionary engines for online interactive collaborative game design
Pier Luca Lanzi and Daniele Loiacono · 2023
Closest in time.
Artificial intelligence as a facilitator for film production process
Hardeep Singh, Kamaljeet Kaur, and Preet Pinder Singh · 2023
Closest in time.
A proposed meta-reality immersive development pipeline: Generative ai models and extended reality (xr) content for the metaverse
Jeremiah Ratican, James Hutson, and Andrew Wright · 2023
Closest in time.
Scenehgn: Hierarchical graph networks for 3d indoor scene generation with fine-grained geometry
Lin Gao, Jia-Mu Sun, Kaichun Mo, et al · 2023
Closest in time.
Modi: Unconditional motion synthesis from diverse data
Sigal Raab, Inbal Leibovitch, Peizhuo Li, et al · 2023
Closest in time.
Make-a-video: Text-to-video generation without text-video data
Uriel Singer, Adam Polyak, Thomas Hayes, et al · 2023
Closest in time.
Dreamfusion: Text-to-3d using 2d diffusion
Ben Poole, Ajay Jain, Jonathan T. Barron, et al · 2023
Closest in time.
Ai-generated content (aigc): A survey
Jiayang Wu, Wensheng Gan, Zefeng Chen, et al · 2023
Closest in time.
Chatgpt: A comprehensive review on background, applications, key challenges, bias, ethics, limitations and future scope
Partha Pratim Ray · 2023
Closest in time.
Harnessing the power of llms in practice: A survey on chatgpt and beyond
Jingfeng Yang, Hongye Jin, Ruixiang Tang, et al · 2023
Closest in time.
One small step for generative ai, one giant leap for agi: A complete survey on chatgpt in aigc era
Chaoning Zhang et al · 2023
Closest in time.
Chatgpt is not all you need. a state of the art review of large generative ai models
Roberto Gozalo-Brizuela and Eduardo C Garrido-Merchan · 2023
Closest in time.
Unleashing the power of edge-cloud generative ai in mobile networks: A survey of aigc services
Minrui Xu, Hongyang Du, et al · 2023
Closest in time.
A survey on generative modeling with limited data, few shots, and zero shot
Milad Abdollahzadeh, Touba Malekzadeh, et al · 2023
Closest in time.
A comprehensive survey on pretrained foundation models: A history from bert to chatgpt
Ce Zhou, Qian Li, Chen Li, et al · 2023
Closest in time.
A complete survey on generative ai (aigc): Is chatgpt from gpt-4 to gpt-5 all you need?
Chaoning Zhang, Chenshuang Zhang, et al · 2023
Closest in time.
A comprehensive survey of ai-generated content (aigc): A history of generative ai from gan to chatgpt
Yihan Cao, Siyu Li, et al · 2023
Closest in time.
Consistency models
Yang Song, Prafulla Dhariwal, Mark Chen, et al · 2023
Closest in time.
Drag your gan: Interactive point-based manipulation on the generative image manifold
Xingang Pan, Ayush Tewari, et al · 2023
Closest in time.
Self-guided diffusion models
Vincent Tao Hu, David W. Zhang, Yuki M. Asano, et al · 2023
Closest in time.
Dragdiffusion: Harnessing diffusion models for interactive point-based image editing
Yujun Shi, Chuhui Xue, Jiachun Pan, et al · 2023
Closest in time.
High-resolution image reconstruction with latent diffusion models from human brain activity
Yu Takagi and Shinji Nishimoto · 2023
Closest in time.
Layoutdm: Transformer-based diffusion model for layout generation
Shang Chai, Liansheng Zhuang, and Fengying Yan · 2023
Closest in time.
Videofusion: Decomposed diffusion models for high-quality video generation
Zhengxiong Luo, Dayou Chen, Yingya Zhang, et al · 2023
Closest in time.
Video probabilistic diffusion models in projected latent space
Sihyun Yu, Kihyuk Sohn, Subin Kim, et al · 2023
Closest in time.
Align your latents: High-resolution video synthesis with latent diffusion models
Andreas Blattmann, Robin Rombach, Huan Ling, et al · 2023
Closest in time.
Nuwa-xl: Diffusion over diffusion for extremely long video generation
Shengming Yin, Chenfei Wu, Huan Yang, et al · 2023
Closest in time.
Conditional image-to-video generation with latent flow diffusion models
Haomiao Ni, Changhao Shi, Kai Li, et al · 2023
Closest in time.
Llama: Open and efficient foundation language models
Hugo Touvron, Thibaut Lavril, Gautier Izacard, et al · 2023
Closest in time.
Diffuseq: Sequence to sequence text generation with diffusion models
Shansan Gong, Mukai Li, Jiangtao Feng, et al · 2023
Closest in time.
Octree transformer: Autoregressive 3d shape generation on hierarchically structured sequences
Moritz Ibing et al · 2023
Closest in time.
Diffusion-based signed distance fields for 3d shape generation
Jaehyeok Shim, Changwoo Kang, and Kyungdon Joo · 2023
Closest in time.
Controllable mesh generation through sparse latent point diffusion models
Zhaoyang Lyu, Jinyi Wang, Yuwei An, et al · 2023
Closest in time.
Holodiffusion: Training a 3d diffusion model using 2d images
Animesh Karnewar, Andrea Vedaldi, David Novotny, et al · 2023
Closest in time.
Diffrf: Rendering-guided 3d radiance field diffusion
Norman Müller, Yawar Siddiqui, Lorenzo Porzi, et al · 2023
Closest in time.
3d neural field generation using triplane diffusion
J Ryan Shue, Eric Ryan Chan, Ryan Po, et al · 2023
Closest in time.
Eva3d: Compositional 3d human generation from 2d image collections
Fangzhou Hong, Zhaoxi Chen, LAN Yushi, et al · 2023
Closest in time.
Human motion diffusion model
Guy Tevet, Sigal Raab, Brian Gordon, et al · 2023
Closest in time.
Diffpose: Toward more reliable 3d pose estimation
Jia Gong, Lin Geng Foo, Zhipeng Fan, Qiuhong Ke, Hossein Rahmani, and Jun Liu · 2023
Closest in time.
Unified pose sequence modeling
Lin Geng Foo, Tianjiao Li, Hossein Rahmani, Qiuhong Ke, and Jun Liu · 2023
Closest in time.
Dreambooth: Fine tuning text-to-image diffusion models for subject-driven generation
Nataniel Ruiz, Yuanzhen Li, Varun Jampani, et al · 2023
Closest in time.
Imagic: Text-based real image editing with diffusion models
Bahjat Kawar, Shiran Zada, Oran Lang, et al · 2023
Closest in time.
Gligen: Open-set grounded text-to-image generation
Yuheng Li, Haotian Liu, Qingyang Wu, et al · 2023
Closest in time.
knn-diffusion: Image generation via large-scale retrieval
Shelly Sheynin, Oron Ashual, Adam Polyak, et al · 2023
Closest in time.
Adding conditional control to text-to-image diffusion models
Lvmin Zhang and Maneesh Agrawala · 2023
Closest in time.
An image is worth one word: Personalizing text-to-image generation using textual inversion
Rinon Gal, Yuval Alaluf, et al · 2023
Closest in time.
Multi-concept customization of text-to-image diffusion
Nupur Kumari, Bingliang Zhang, Richard Zhang, et al · 2023
Closest in time.
Hyperdreambooth: Hypernetworks for fast personalization of text-to-image models
Nataniel Ruiz, Yuanzhen Li, et al · 2023
Closest in time.
Cogvideo: Large-scale pretraining for text-to-video generation via transformers
Wenyi Hong, Ming Ding, Wendi Zheng, et al · 2023
Closest in time.
Make-a-story: Visual memory conditioned consistent story generation
Tanzila Rahman, Hsin-Ying Lee, Jian Ren, et al · 2023
Closest in time.
Video-p2p: Video editing with cross-attention control
Shaoteng Liu, Yuechen Zhang, Wenbo Li, et al · 2023
Closest in time.
Structure and content-guided video synthesis with diffusion models
Patrick Esser, Johnathan Chiu, Parmida Atighehchian, et al · 2023
Closest in time.
Dream3d: Zero-shot text-to-3d synthesis using 3d shape prior and text-to-image diffusion models
Jiale Xu, Xintao Wang, et al · 2023
Closest in time.
Magic3d: High-resolution text-to-3d content creation
Chen-Hsuan Lin, Jun Gao, Luming Tang, et al · 2023
Closest in time.
Latent-nerf for shape-guided generation of 3d shapes and textures
Gal Metzer, Elad Richardson, Or Patashnik, et al · 2023
Closest in time.
Sdfusion: Multimodal 3d shape completion, reconstruction, and generation
Yen-Chi Cheng, Hsin-Ying Lee, Sergey Tulyakov, et al · 2023
Closest in time.
Texture: Text-guided texturing of 3d shapes
Elad Richardson, Gal Metzer, Yuval Alaluf, et al · 2023
Closest in time.
Pc2: Projection-conditioned point cloud diffusion for single-image 3d reconstruction
Luke Melas-Kyriazi, Christian Rupprecht, et al · 2023
Closest in time.
Text2scene: Text-driven indoor scene stylization with part-aware details
Inwoo Hwang, Hyeonwoo Kim, and Young Min Kim · 2023
Closest in time.
Text2nerf: Text-driven 3d scene generation with neural radiance fields
Jingbo Zhang, Xiaoyu Li, Ziyu Wan, et al · 2023
Closest in time.
Text2room: Extracting textured 3d meshes from 2d text-to-image models
Lukas Höllein, Ang Cao, Andrew Owens, et al · 2023
Closest in time.
Compositional 3d scene generation using locally conditioned diffusion
Ryan Po and Gordon Wetzstein · 2023
Closest in time.
Text-to-4D dynamic scene generation
Uriel Singer, Shelly Sheynin, Adam Polyak, et al · 2023
Closest in time.
Generative novel view synthesis with 3d-aware diffusion models
Eric R Chan, Koki Nagano, Matthew A Chan, et al · 2023
Closest in time.
Dreamface: Progressive generation of animatable 3d faces under text guidance
Longwen Zhang, Qiwei Qiu, Hongyang Lin, et al · 2023
Closest in time.
High-fidelity 3d face generation from natural language descriptions
Menghua Wu, Hao Zhu, Linjia Huang, et al · 2023
Closest in time.
Distribution-aligned diffusion for human mesh recovery
Lin Geng Foo, Jia Gong, Hossein Rahmani, and Jun Liu · 2023
Closest in time.
High-fidelity clothed avatar reconstruction from a single image
Tingting Liao, Xiaomei Zhang, Yuliang Xiu, et al · 2023
Closest in time.
Vid2avatar: 3d avatar reconstruction from videos in the wild via self-supervised scene decomposition
Chen Guo, Tianjian Jiang, et al · 2023
Closest in time.
Pointavatar: Deformable point-based head avatars from videos
Yufeng Zheng, Wang Yifan, Gordon Wetzstein, et al · 2023
Closest in time.
T2m-gpt: Generating human motion from textual descriptions with discrete representations
Jianrong Zhang, Yangsong Zhang, et al · 2023
Closest in time.
Mofusion: A framework for denoising-diffusion-based motion synthesis
Rishabh Dabral, Muhammad Hamza Mughal, et al · 2023
Closest in time.
Text and image guided 3d avatar generation and manipulation
Zehranaz Canfes, M Furkan Atasoy, Alara Dirik, et al · 2023
Closest in time.
Edge: Editable dance generation from music
Jonathan Tseng, Rodrigo Castellon, and Karen Liu · 2023
Closest in time.
Taming diffusion models for audio-driven co-speech gesture generation
Lingting Zhu, Xian Liu, Xuanyu Liu, et al · 2023
Closest in time.
Geneface: Generalized and high-fidelity audio-driven 3d talking face synthesis
Zhenhui Ye, Ziyue Jiang, Yi Ren, et al · 2023
Closest in time.
One-shot high-fidelity talking-head synthesis with deformable neural radiance field
Weichuang Li, Longhao Zhang, Don Wang, et al · 2023
Closest in time.
Diffsound: Discrete diffusion model for text-to-sound generation
Dongchao Yang, Jianwei Yu, Helin Wang, et al · 2023
Closest in time.
Musiclm: Generating music from text
Andrea Agostinelli, Timo I Denk, Zalán Borsos, et al · 2023
Closest in time.
Noise2music: Text-conditioned music generation with diffusion models
Qingqing Huang, Daniel S Park, Tao Wang, et al · 2023
Closest in time.
Mo \ \backslash ˆ usai: Text-to-music generation with long-context latent diffusion
Flavio Schneider, Zhijing Jin, and Bernhard Schölkopf · 2023
Closest in time.
Study and analysis of chat gpt and its impact on different fields of study
Dinesh Kalla, Nathan Smith, Fnu Samaah, and Sivaraju Kuraku · 2023
Closest in time.
Audiolm: a language modeling approach to audio generation
Zalán Borsos, Raphaël Marinier, Damien Vincent, et al · 2023
Closest in time.
How to backdoor diffusion models?
Sheng-Yen Chou, Pin-Yu Chen, and Tsung-Yi Ho · 2023
Closest in time.
Quantifying memorization across neural language models
Nicholas Carlini, Daphne Ippolito, Matthew Jagielski, et al · 2023
Closest in time.
Extracting training data from diffusion models
Nicholas Carlini, Jamie Hayes, Milad Nasr, et al · 2023
Closest in time.
Diffusion art or digital forgery? investigating data replication in diffusion models
Gowthami Somepalli, Vasu Singla, Micah Goldblum, et al · 2023
Closest in time.
Capabilities of gpt-4 on medical challenge problems
Harsha Nori, Nicholas King, Scott Mayer McKinney, et al · 2023
Closest in time.
Survey of hallucination in natural language generation
Ziwei Ji, Nayeon Lee, Rita Frieske, et al · 2023
Closest in time.
Palm-e: An embodied multimodal language model
Danny Driess, Fei Xia, Mehdi S. M. Sajjadi, et al · 2023
Closest in time.
Gpt-4 technical report
OpenAI · 2023
Closest in time.
Plug-and-play diffusion features for text-driven image-to-image translation
Narek Tumanyan, Michal Geyer, Shai Bagon, et al · 2023
Closest in time.
Unite and conquer: Plug & play multi-modal synthesis using diffusion models
Nithin Gopalakrishnan Nair, Wele Gedara Chaminda Bandara, and Vishal M Patel · 2023
Closest in time.
Laga: Layered 3d avatar generation and customization via gaussian splatting
Jia Gong, Shenyu Ji, Lin Geng Foo, et al · 2024
Closest in time.
Humangaussian: Text-driven 3d human generation with gaussian splatting
Xian Liu, Xiaohang Zhan, Jiaxiang Tang, et al · 2024
Closest in time.
Avatar concept slider: Manipulate concepts in your human avatar with fine-grained control
Yixuan He, Lin Geng Foo, Ajmal Saeed Mian, et al · 2024
Closest in time.
Llms are good sign language translators
Jia Gong, Lin Geng Foo, Yixuan He, et al · 2024
Closest in time.
Visual anagrams: Generating multi-view optical illusions with diffusion models
Daniel Geng, Inbum Park, and Andrew Owens · 2024
Closest in time.
Llafs: When large language models meet few-shot segmentation
Lanyun Zhu, Tianrun Chen, Deyi Ji, et al · 2024
Closest in time.
Llms are good action recognizers
Haoxuan Qu, Yujun Cai, and Jun Liu · 2024
Closest in time.
Ltgc: Long-tail recognition via leveraging llms-driven generated content
Qihao Zhao, Yalun Dai, Hao Li, et al · 2024
Closest in time.