Fetching the paper…
Reading the bibliography…
Current deep networks are very data-hungry and benefit from training on largescale datasets, which are often time-consuming to collect and annotate.
The pascal visual object classes (voc) challenge
M. Everingham, L. Van Gool, C. K. Williams, J. Winn, and A. Zisserman · 2010
Earlier work this paper cites.
A naturalistic open source movie for optical flow evaluation
D. J. Butler, J. Wulff, G. B. Stanley, and M. J. Black · 2012
Earlier work this paper cites.
Teaching 3d geometry to deformable part models
B. Pepik, M. Stark, P. Gehler, and B. Schiele · 2012
Earlier work this paper cites.
Data-driven scene understanding from 3d models
S. Satkin, J. Lin, and M. Hebert · 2012
Earlier work this paper cites.
Indoor segmentation and support inference from rgbd images
N. Silberman, D. Hoiem, P. Kohli, and R. Fergus · 2012
Earlier work this paper cites.
Auto-encoding variational bayes
D. P. Kingma and M. Welling · 2013
Earlier work this paper cites.
Nice: Non-linear independent components estimation
L. Dinh, D. Krueger, and Y. Bengio · 2014
Earlier work this paper cites.
Depth map prediction from a single image using a multi-scale deep network
D. Eigen, C. Puhrsch, and R. Fergus · 2014
Earlier work this paper cites.
Microsoft coco: Common objects in context
T.-Y. Lin, M. Maire, S. Belongie, J. Hays, P. Perona, D. Ramanan, P. Dollár, and C. L. Zitnick · 2014
Earlier work this paper cites.
U-net: Convolutional networks for biomedical image segmentation
O. Ronneberger, P. Fischer, and T. Brox · 2015
Earlier work this paper cites.
Deep unsupervised learning using nonequilibrium thermodynamics
J. Sohl-Dickstein, E. Weiss, N. Maheswaranathan, and S. Ganguli · 2015
Earlier work this paper cites.
The cityscapes dataset for semantic urban scene understanding
M. Cordts, M. Omran, S. Ramos, T. Rehfeld, M. Enzweiler, R. Benenson, U. Franke, S. Roth, and B. Schiele · 2016
Earlier work this paper cites.
The cityscapes dataset for semantic urban scene understanding
M. Cordts, M. Omran, S. Ramos, T. Rehfeld, M. Enzweiler, R. Benenson, U. Franke, S. Roth, and B. Schiele · 2016
Earlier work this paper cites.
Virtual worlds as proxy for multi-object tracking analysis
A. Gaidon, Q. Wang, Y. Cabon, and E. Vig · 2016
Earlier work this paper cites.
Scribblesup: Scribble-supervised convolutional networks for semantic segmentation
D. Lin, J. Dai, J. Jia, K. He, and J. Sun · 2016
Earlier work this paper cites.
V-net: Fully convolutional neural networks for volumetric medical image segmentation
F. Milletari, N. Navab, and S.-A. Ahmadi · 2016
Earlier work this paper cites.
Decoupled weight decay regularization
I. Loshchilov and F. Hutter · 2017
Earlier work this paper cites.
Scene parsing through ade20k dataset
B. Zhou, H. Zhao, X. Puig, S. Fidler, A. Barriuso, and A. Torralba · 2017
Earlier work this paper cites.
Learning pixel-level semantic affinity with image-level supervision for weakly supervised semantic segmentation
J. Ahn and S. Kwak · 2018
Earlier work this paper cites.
Encoder-decoder with atrous separable convolution for semantic image segmentation
L.-C. Chen, Y. Zhu, G. Papandreou, F. Schroff, and H. Adam · 2018
Earlier work this paper cites.
Zero-shot semantic segmentation
M. Bucher, T.-H. Vu, M. Cord, and P. Pérez · 2019
Earlier work this paper cites.
Deep high-resolution representation learning for human pose estimation
K. Sun, B. Xiao, D. Liu, and J. Wang · 2019
Earlier work this paper cites.
Y. Cabon, N. Murray, and M. Humenberger · 2020
Earlier work this paper cites.
Gpt-3: Its nature, scope, limits, and consequences
L. Floridi and M. Chiriatti · 2020
Earlier work this paper cites.
Pose augmentation: Class-agnostic object pose transformation for object recognition
Y. Ge, J. Zhao, and L. Itti · 2020
Cited alongside, same era.
Generative adversarial networks
I. Goodfellow, J. Pouget-Abadie, M. Mirza, B. Xu, D. Warde-Farley, S. Ozair, A. Courville, and Y. Bengio · 2020
Cited alongside, same era.
Context-aware feature generation for zero-shot semantic segmentation
Z. Gu, S. Zhou, L. Niu, Z. Zhao, and L. Zhang · 2020
Cited alongside, same era.
Denoising diffusion probabilistic models
J. Ho, A. Jain, and P. Abbeel · 2020
Cited alongside, same era.
A survey of unsupervised deep domain adaptation
G. Wilson and D. J. Cook · 2020
Cited alongside, same era.
A survey of semi-and weakly supervised semantic segmentation of images
M. Zhang, Y. Zhou, J. Zhao, Y. Man, B. Liu, and R. Yao · 2020
Cited alongside, same era.
Z. Li, Z. Chen, X. Liu, and J. Jiang · 2022
Later among the works it cites.
Raregan: Generating samples for rare classes
Z. Lin, H. Liang, G. Fanti, and V. Sekar · 2022
Later among the works it cites.
Hierarchical text-conditional image generation with clip latents
A. Ramesh, P. Dhariwal, A. Nichol, C. Chu, and M. Chen · 2022
Later among the works it cites.
High-resolution image synthesis with latent diffusion models
R. Rombach, A. Blattmann, D. Lorenz, P. Esser, and B. Ommer · 2022
Later among the works it cites.
Photorealistic text-to-image diffusion models with deep language understanding
C. Saharia, W. Chan, S. Saxena, L. Li, J. Whang, E. L. Denton, K. Ghasemipour, R. Gontijo Lopes, B. Karagol Ayan, T. Salimans, et al · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Towards single stage weakly supervised semantic segmentation
P. Akiva and K. Dana · 2021
Cited alongside, same era.
Exploiting a joint embedding space for generalized zero-shot semantic segmentation
D. Baek, Y. Oh, and B. Ham · 2021
Cited alongside, same era.
Label-efficient semantic segmentation with diffusion models
D. Baranchuk, I. Rubachev, A. Voynov, V. Khrulkov, and A. Babenko · 2021
Cited alongside, same era.
Per-pixel classification is not all you need for semantic segmentation
B. Cheng, A. Schwing, and A. Kirillov · 2021
Cited alongside, same era.
Sign: Spatial-information incorporated generative network for generalized zero-shot semantic segmentation
J. Cheng, S. Nandi, P. Natarajan, and W. Abd-Almageed · 2021
Cited alongside, same era.
Taming transformers for high-resolution image synthesis
P. Esser, R. Rombach, and B. Ommer · 2021
Cited alongside, same era.
Laion-5b: An open large-scale dataset for training next generation image-text models
C. Schuhmann, R. Beaumont, R. Vencu, C. Gordon, R. Wightman, M. Cherti, T. Coombes, A. Katta, C. Mullis, M. Wortsman, et al · 2022
Later among the works it cites.
Deep generative mixture model for robust imbalance classification
X. Wang, L. Jing, Y. Lyu, M. Guo, J. Wang, H. Liu, J. Yu, and T. Zeng · 2022
Later among the works it cites.
Self-instruct: Aligning language model with self generated instructions
Y. Wang, Y. Kordi, S. Mishra, A. Liu, N. A. Smith, D. Khashabi, and H. Hajishirzi · 2022
Later among the works it cites.
Medsegdiff: Medical image segmentation with diffusion probabilistic model
J. Wu, H. Fang, Y. Zhang, Y. Yang, and Y. Xu · 2022
Later among the works it cites.
Synthetic data supervised salient object detection
Z. Wu, L. Wang, W. Wang, T. Shi, C. Chen, A. Hao, and S. Li · 2022
Later among the works it cites.
Handsoff: Labeled dataset generation with no additional human annotations
A. Xu, M. I. Vasileva, A. Dave, and A. Seshadri · 2022
Later among the works it cites.
Generalized decoding for pixel, image, and language
X. Zou, Z.-Y. Dou, J. Yang, Z. Gan, L. Li, C. Li, X. Dai, H. Behl, J. Wang, L. Yuan, et al · 2022
Later among the works it cites.
Sparks of artificial general intelligence: Early experiments with gpt-4
S. Bubeck, V. Chandrasekaran, R. Eldan, J. Gehrke, E. Horvitz, E. Kamar, P. Lee, Y. T. Lee, Y. Li, S. Lundberg, et al · 2023
Closest in time.
Imagebind: One embedding space to bind them all
R. Girdhar, A. El-Nouby, Z. Liu, M. Singh, K. V. Alwala, A. Joulin, and I. Misra · 2023
Closest in time.
Mix-of-show: Decentralized low-rank adaptation for multi-concept customization of diffusion models
Y. Gu, X. Wang, J. Z. Wu, Y. Shi, Y. Chen, Z. Fan, W. Xiao, R. Zhao, S. Chang, W. Wu, et al · 2023
Closest in time.
Medgen3d: A deep generative framework for paired 3d image and mask generation
K. Han, Y. Xiong, C. You, P. Khosravi, S. Sun, X. Yan, J. Duncan, and X. Xie · 2023
Closest in time.
How good are gpt models at machine translation? a comprehensive evaluation
A. Hendy, M. Abdelrehim, A. Sharaf, V. Raunak, M. Gabr, H. Matsushita, Y. J. Kim, M. Afify, and H. H. Awadalla · 2023
Closest in time.
Ablating concepts in text-to-image diffusion models
N. Kumari, B. Zhang, S.-Y. Wang, E. Shechtman, R. Zhang, and J.-Y. Zhu · 2023
Closest in time.
Guiding text-to-image diffusion model towards grounded generation
Z. Li, Q. Zhou, X. Zhang, Y. Zhang, Y. Wang, and W. Xie · 2023
Closest in time.
W. Wu, Y. Zhao, M. Z. Shou, H. Zhou, and C. Shen · 2023
Closest in time.
Open-vocabulary panoptic segmentation with text-to-image diffusion models
J. Xu, S. Liu, A. Vahdat, W. Byeon, X. Wang, and S. De Mello · 2023
Closest in time.
Unleashing text-to-image diffusion models for visual perception
W. Zhao, Y. Rao, Z. Liu, B. Liu, J. Zhou, and J. Lu · 2023
Closest in time.
Generative prompt model for weakly supervised object localization
Y. Zhao, Q. Ye, W. Wu, C. Shen, and F. Wan · 2023
Closest in time.