Fetching the paper…
Reading the bibliography…
In an effort to further advance semi-supervised generative and classification tasks, we propose a simple yet effective training strategy called dual pseudo training (DPT), built upon strong semi-supervised learners and diffusion models.
R. A. Fisher, “On the mathematical foundations of theoretical statistics,” Philosophical transactions of the Royal Society of London. Series A, containing papers of a mathematical or physical character , vol. 222, no. 594-604, pp. 309–368, 1922
1922
Earlier work this paper cites.
T. Joachims et al. , “Transductive inference for text classification using support vector machines,” in ICML , vol. 99, 1999, pp. 200–209
1999
Earlier work this paper cites.
Y. Grandvalet and Y. Bengio, “Semi-supervised learning by entropy minimization,” in Advances in Neural Information Processing Systems , vol. 17, 01 2004
2004
Earlier work this paper cites.
A. Krizhevsky and G. Hinton, “Learning multiple layers of features from tiny images,” Citeseer , 2009
2009
Earlier work this paper cites.
J. Deng, W. Dong, R. Socher, L. Li, K. Li, and L. Fei-Fei, “Imagenet: A large-scale hierarchical image database,” in IEEE Computer Society Conference on Computer Vision and Pattern Recognition , 2009, pp. 248–255
2009
Earlier work this paper cites.
D.-H. Lee et al. , “Pseudo-label: The simple and efficient semi-supervised learning method for deep neural networks,” in Workshop on challenges in representation learning, ICML , vol. 3, 2013, p. 896
2013
Earlier work this paper cites.
D. P. Kingma, S. Mohamed, D. J. Rezende, and M. Welling, “Semi-supervised learning with deep generative models,” in Advances in Neural Information Processing Systems , 2014
2014
Earlier work this paper cites.
J. Sohl-Dickstein, E. A. Weiss, N. Maheswaranathan, and S. Ganguli, “Deep unsupervised learning using nonequilibrium thermodynamics,” in Proceedings of the 32nd International Conference on Machine Learning , 2015
2015
Earlier work this paper cites.
A. Rasmus, M. Berglund, M. Honkala, H. Valpola, and T. Raiko, “Semi-supervised learning with ladder networks,” in Advances in Neural Information Processing Systems , 2015
2015
Earlier work this paper cites.
T. Salimans, I. Goodfellow, W. Zaremba, V. Cheung, A. Radford, and X. Chen, “Improved techniques for training GANs,” in Advances in Neural Information Processing Systems , 2016
2016
Earlier work this paper cites.
L. Maaløe, C. K. Sønderby, S. K. Sønderby, and O. Winther, “Auxiliary deep generative models,” in International conference on machine learning . PMLR, 2016, pp. 1445–1453
2016
Earlier work this paper cites.
J. T. Springenberg, “Unsupervised and semi-supervised learning with categorical generative adversarial networks,” in 4th International Conference on Learning Representations, ICLR 2016, San Juan, Puerto Rico, May 2-4, 2016, Conference Track Proceedings , Y. Bengio and Y. LeCun, Eds., 2016
2016
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2016, pp. 770–778
2016
Earlier work this paper cites.
C. Szegedy, V. Vanhoucke, S. Ioffe, J. Shlens, and Z. Wojna, “Rethinking the inception architecture for computer vision,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2016, pp. 2818–2826
2016
Earlier work this paper cites.
C. Li, K. Xu, J. Zhu, and B. Zhang, “Triple generative adversarial nets,” in Advances in Neural Information Processing Systems , 2017, pp. 4088–4098
2017
Earlier work this paper cites.
M. Heusel, H. Ramsauer, T. Unterthiner, B. Nessler, and S. Hochreiter, “Gans trained by a two time-scale update rule converge to a local nash equilibrium,” in Advances in Neural Information Processing Systems , 2017, pp. 6626–6637
2017
Earlier work this paper cites.
C. Li, J. Zhu, and B. Zhang, “Max-margin deep generative models for (semi-) supervised learning,” IEEE transactions on pattern analysis and machine intelligence , vol. 40, no. 11, pp. 2762–2775, 2017
2017
Earlier work this paper cites.
Z. Dai, Z. Yang, F. Yang, W. W. Cohen, and R. R. Salakhutdinov, “Good semi-supervised learning that requires a bad gan,” in Advances in Neural Information Processing Systems , 2017, pp. 6510–6520
2017
Earlier work this paper cites.
Z. Gan, L. Chen, W. Wang, Y. Pu, Y. Zhang, H. Liu, C. Li, and L. Carin, “Triangle generative adversarial networks,” in Advances in Neural Information Processing Systems , 2017, pp. 5247–5256
2017
Earlier work this paper cites.
A. Tarvainen and H. Valpola, “Mean teachers are better role models: Weight-averaged consistency targets improve semi-supervised deep learning results,” in Advances in Neural Information Processing Systems , 2017, pp. 1195–1204
2017
Earlier work this paper cites.
S. Laine and T. Aila, “Temporal ensembling for semi-supervised learning,” in International Conference on Learning Representations , 2017
2017
Earlier work this paper cites.
C. Szegedy, S. Ioffe, V. Vanhoucke, and A. Alemi, “Inception-v4, inception-resnet and the impact of residual connections on learning,” in Proceedings of the AAAI conference on artificial intelligence , vol. 31, no. 1, 2017
2017
Earlier work this paper cites.
T. Miyato, S.-i. Maeda, M. Koyama, and S. Ishii, “Virtual adversarial training: a regularization method for supervised and semi-supervised learning,” IEEE transactions on pattern analysis and machine intelligence , vol. 41, no. 8, pp. 1979–1993, 2018
2018
Earlier work this paper cites.
Y. Luo, J. Zhu, M. Li, Y. Ren, and B. Zhang, “Smooth neighbors on teacher graphs for semi-supervised learning,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2018, pp. 8896–8905
2018
Earlier work this paper cites.
A. Oliver, A. Odena, C. A. Raffel, E. D. Cubuk, and I. Goodfellow, “Realistic evaluation of deep semi-supervised learning algorithms,” in Advances in Neural Information Processing Systems , 2018, pp. 3235–3246
2018
Earlier work this paper cites.
A. Brock, J. Donahue, and K. Simonyan, “Large scale gan training for high fidelity natural image synthesis,” in International Conference on Learning Representations , 2018
2018
Earlier work this paper cites.
T. Miyato, S.-i. Maeda, M. Koyama, and S. Ishii, “Virtual adversarial training: a regularization method for supervised and semi-supervised learning,” IEEE transactions on pattern analysis and machine intelligence , vol. 41, no. 8, pp. 1979–1993, 2018
2018
Earlier work this paper cites.
J. Hu, L. Shen, and G. Sun, “Squeeze-and-excitation networks,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2018, pp. 7132–7141
2018
Earlier work this paper cites.
M. Lučić, M. Tschannen, M. Ritter, X. Zhai, O. Bachem, and S. Gelly, “High-fidelity image generation with fewer labels,” in International conference on machine learning . PMLR, 2019, pp. 4183–4192
2019
Earlier work this paper cites.
A. Iscen, G. Tolias, Y. Avrithis, and O. Chum, “Label propagation for deep semi-supervised learning,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2019, pp. 5070–5079
2019
Earlier work this paper cites.
B. Athiwaratkun, M. Finzi, P. Izmailov, and A. G. Wilson, “There are many consistent explanations of unlabeled data: Why you should average,” in 7th International Conference on Learning Representations, ICLR 2019, New Orleans, LA, USA, May 6-9, 2019 , 2019
2019
Earlier work this paper cites.
D. Berthelot, N. Carlini, I. Goodfellow, N. Papernot, A. Oliver, and C. A. Raffel, “Mixmatch: A holistic approach to semi-supervised learning,” in Advances in Neural Information Processing Systems , 2019, pp. 5049–5059
2019
Earlier work this paper cites.
D. Berthelot, N. Carlini, E. D. Cubuk, A. Kurakin, K. Sohn, H. Zhang, and C. Raffel, “Remixmatch: Semi-supervised learning with distribution matching and augmentation anchoring,” in International Conference on Learning Representations , 2019
2019
Earlier work this paper cites.
Y. Song and S. Ermon, “Generative modeling by estimating gradients of the data distribution,” Advances in Neural Information Processing Systems , vol. 32, 2019
2019
Earlier work this paper cites.
T. Kynkäänniemi, T. Karras, S. Laine, J. Lehtinen, and T. Aila, “Improved precision and recall metric for assessing generative models,” Advances in Neural Information Processing Systems , vol. 32, 2019
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
M. Tan and Q. Le, “Efficientnet: Rethinking model scaling for convolutional neural networks,” in International conference on machine learning . PMLR, 2019, pp. 6105–6114
2019
Earlier work this paper cites.
J. Ho, A. Jain, and P. Abbeel, “Denoising diffusion probabilistic models,” in Advances in Neural Information Processing Systems , 2020
2020
Earlier work this paper cites.
K. Sohn, D. Berthelot, N. Carlini, Z. Zhang, H. Zhang, C. A. Raffel, E. D. Cubuk, A. Kurakin, and C.-L. Li, “Fixmatch: Simplifying semi-supervised learning with consistency and confidence,” Advances in neural information processing systems , vol. 33, pp. 596–608, 2020
2020
Cited alongside, same era.
T. Chen, S. Kornblith, K. Swersky, M. Norouzi, and G. E. Hinton, “Big self-supervised models are strong semi-supervised learners,” Advances in neural information processing systems , vol. 33, pp. 22 243–22 255, 2020
2020
Cited alongside, same era.
J.-B. Grill, F. Strub, F. Altché, C. Tallec, P. Richemond, E. Buchatskaya, C. Doersch, B. Avila Pires, Z. Guo, M. Gheshlaghi Azar et al. , “Bootstrap your own latent-a new approach to self-supervised learning,” Advances in neural information processing systems , vol. 33, pp. 21 271–21 284, 2020
2020
Cited alongside, same era.
V. Besnier, H. Jain, A. Bursuc, M. Cord, and P. Pérez, “This dataset does not exist: training models from generated images,” in ICASSP 2020-2020 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) . IEEE, 2020, pp. 1–5
M. Assran, M. Caron, I. Misra, P. Bojanowski, F. Bordes, P. Vincent, A. Joulin, M. Rabbat, and N. Ballas, “Masked siamese networks for label-efficient learning,” in European Conference on Computer Vision . Springer, 2022, pp. 456–473
2022
Later among the works it cites.
J. Ho, C. Saharia, W. Chan, D. J. Fleet, M. Norouzi, and T. Salimans, “Cascaded diffusion models for high fidelity image generation.” J. Mach. Learn. Res. , vol. 23, pp. 47–1, 2022
2022
Later among the works it cites.
R. Rombach, A. Blattmann, D. Lorenz, P. Esser, and B. Ommer, “High-resolution image synthesis with latent diffusion models,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2022, pp. 10 684–10 695
2022
Later among the works it cites.
J. Ho and T. Salimans, “Classifier-free diffusion guidance,” arXiv preprint arXiv:2207.12598 , 2022
2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2020
Cited alongside, same era.
M. Noroozi, “Self-labeled conditional gans,” arXiv preprint arXiv:2012.02162 , 2020
2020
Cited alongside, same era.
S. Liu, T. Wang, D. Bau, J.-Y. Zhu, and A. Torralba, “Diverse image generation via self-conditioned gans,” in Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2020, pp. 14 286–14 295
2020
Cited alongside, same era.
T. Karras, M. Aittala, J. Hellsten, S. Laine, J. Lehtinen, and T. Aila, “Training generative adversarial networks with limited data,” in Advances in Neural Information Processing Systems , 2020
2020
Cited alongside, same era.
Q. Xie, Z. Dai, E. Hovy, T. Luong, and Q. Le, “Unsupervised data augmentation for consistency training,” Advances in Neural Information Processing Systems , vol. 33, pp. 6256–6268, 2020
2020
Cited alongside, same era.
K. He, H. Fan, Y. Wu, S. Xie, and R. B. Girshick, “Momentum contrast for unsupervised visual representation learning,” in IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2020, pp. 9726–9735
2020
Cited alongside, same era.
T. Chen, S. Kornblith, M. Norouzi, and G. E. Hinton, “A simple framework for contrastive learning of visual representations,” in International Conference on Machine Learning , vol. 119, 2020, pp. 1597–1607
2020
Cited alongside, same era.
Y. Song, J. Sohl-Dickstein, D. P. Kingma, A. Kumar, S. Ermon, and B. Poole, “Score-based generative modeling through stochastic differential equations,” in 9th International Conference on Learning Representations , 2021
2021
Cited alongside, same era.
P. Dhariwal and A. Nichol, “Diffusion models beat gans on image synthesis,” Advances in Neural Information Processing Systems , vol. 34, pp. 8780–8794, 2021
2021
Cited alongside, same era.
K. He, X. Chen, S. Xie, Y. Li, P. Dollár, and R. B. Girshick, “Masked autoencoders are scalable vision learners,” in IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2022, pp. 15 979–15 988
2022
Later among the works it cites.
M. Zheng, S. You, L. Huang, F. Wang, C. Qian, and C. Xu, “Simmatch: Semi-supervised learning with similarity matching,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2022, pp. 14 471–14 481
2022
Later among the works it cites.
H. Tang, L. Sun, and K. Jia, “Stochastic consensus: Enhancing semi-supervised learning with consistency of stochastic classifiers,” in European Conference on Computer Vision . Springer, 2022, pp. 330–346
2022
Later among the works it cites.
X. Wang, L. Lian, and S. X. Yu, “Unsupervised selective labeling for more effective semi-supervised learning,” in European Conference on Computer Vision . Springer, 2022, pp. 427–445
2022
Later among the works it cites.
B. Chen, J. Jiang, X. Wang, P. Wan, J. Wang, and M. Long, “Debiased self-training for semi-supervised learning,” Advances in Neural Information Processing Systems , vol. 35, pp. 32 424–32 437, 2022
2022
Later among the works it cites.
H. Tang and K. Jia, “Towards discovering the effectiveness of moderately confident samples for semi-supervised learning,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2022, pp. 14 658–14 667
2022
Later among the works it cites.
A. Jahanian, X. Puig, Y. Tian, and P. Isola, “Generative models as a data source for multiview representation learning,” in The Tenth International Conference on Learning Representations, ICLR 2022, Virtual Event, April 25-29, 2022 , 2022
2022
Later among the works it cites.
2022
Later among the works it cites.
A. Q. Nichol, P. Dhariwal, A. Ramesh, P. Shyam, P. Mishkin, B. McGrew, I. Sutskever, and M. Chen, “GLIDE: towards photorealistic image generation and editing with text-guided diffusion models,” in International Conference on Machine Learning, ICML 2022, 17-23 July 2022, Baltimore, Maryland, USA , vol. 162, 2022, pp. 16 784–16 804
2022
Later among the works it cites.
2022
Later among the works it cites.
S. Gu, D. Chen, J. Bao, F. Wen, B. Zhang, D. Chen, L. Yuan, and B. Guo, “Vector quantized diffusion model for text-to-image synthesis,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2022, pp. 10 696–10 706
2022
Later among the works it cites.
2022
Later among the works it cites.
2022
Later among the works it cites.
C. Meng, Y. He, Y. Song, J. Song, J. Wu, J. Zhu, and S. Ermon, “Sdedit: Guided image synthesis and editing with stochastic differential equations,” in The Tenth International Conference on Learning Representations, ICLR 2022, Virtual Event, April 25-29, 2022 , 2022
2022
Later among the works it cites.
M. Zhao, F. Bao, C. Li, and J. Zhu, “Egsde: Unpaired image-to-image translation via energy-guided stochastic differential equations,” Advances in Neural Information Processing Systems , 2022
2022
Later among the works it cites.
E. Hoogeboom, V. G. Satorras, C. Vignac, and M. Welling, “Equivariant diffusion for molecule generation in 3d,” in International Conference on Machine Learning . PMLR, 2022, pp. 8867–8887
2022
Later among the works it cites.
2022
Later among the works it cites.
F. Bao, C. Li, J. Zhu, and B. Zhang, “Analytic-dpm: an analytic estimate of the optimal reverse variance in diffusion probabilistic models,” in The Tenth International Conference on Learning Representations, ICLR 2022, Virtual Event, April 25-29, 2022 , 2022
2022
Later among the works it cites.
2022
Later among the works it cites.
Q. Zhang and Y. Chen, “Fast sampling of diffusion models with exponential integrator,” in NeurIPS 2022 Workshop on Score-Based Methods , 2022
2022
Later among the works it cites.
2022
Later among the works it cites.
F. Bao, C. Li, J. Sun, J. Zhu, and B. Zhang, “Estimating the optimal covariance with imperfect mean in diffusion probabilistic models,” in Proceedings of the 39th International Conference on Machine Learning , 2022, pp. 1555–1584
2022
Later among the works it cites.
T. Salimans and J. Ho, “Progressive distillation for fast sampling of diffusion models,” in The Tenth International Conference on Learning Representations, ICLR 2022, Virtual Event, April 25-29, 2022 , 2022
2022
Later among the works it cites.
2022
Later among the works it cites.
A. Sauer, K. Schwarz, and A. Geiger, “Stylegan-xl: Scaling stylegan to large diverse datasets,” in ACM SIGGRAPH 2022 Conference Proceedings , 2022, pp. 1–10
2022
Later among the works it cites.
J. Zhou, C. Wei, H. Wang, W. Shen, C. Xie, A. Yuille, and T. Kong, “ibot: Image bert pre-training with online tokenizer,” International Conference on Learning Representations (ICLR) , 2022
2022
Later among the works it cites.
Z. Xie, Z. Zhang, Y. Cao, Y. Lin, J. Bao, Z. Yao, Q. Dai, and H. Hu, “Simmim: A simple framework for masked image modeling,” in International Conference on Computer Vision and Pattern Recognition (CVPR) , 2022
2022
Later among the works it cites.
V. T. Hu, D. W. Zhang, Y. M. Asano, G. J. Burghouts, and C. G. Snoek, “Self-guided diffusion models,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2023, pp. 18 413–18 422
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
A. Alshenoudy, B. Sabrowsky-Hirsch, S. Thumfart, M. Giretzlehner, and E. Kobler, “Semi-supervised brain tumor segmentation using diffusion models,” in IFIP International Conference on Artificial Intelligence Applications and Innovations . Springer, 2023, pp. 314–325
2023
Closest in time.
S. Gong, C. Chen, Y. Gong, N. Y. Chan, W. Ma, C. H.-K. Mak, J. Abrigo, and Q. Dou, “Diffusion model based semi-supervised learning on brain hemorrhage images for efficient midline shift quantification,” in International Conference on Information Processing in Medical Imaging . Springer, 2023, pp. 69–81
2023
Closest in time.
Y. Chen, X. Tan, B. Zhao, Z. Chen, R. Song, J. Liang, and X. Lu, “Boosting semi-supervised learning by exploiting all unlabeled data,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2023, pp. 7548–7557
2023
Closest in time.