Fetching the paper…
Reading the bibliography…
Deep learning technology has developed unprecedentedly in the last decade and has become the primary choice in many application domains.
P. J. Werbos, “Backpropagation through time: what it does and how to do it,” Proceedings of the IEEE , vol. 78, no. 10, pp. 1550–1560, 1990
1990
Earlier work this paper cites.
Y. LeCun, L. Bottou, Y. Bengio, and P. Haffner, “Gradient-based learning applied to document recognition,” Proceedings of the IEEE , vol. 86, no. 11, pp. 2278–2324, 1998
1998
Earlier work this paper cites.
J. Platt et al. , “Probabilistic outputs for support vector machines and comparisons to regularized likelihood methods,” Advances in large margin classifiers , vol. 10, no. 3, pp. 61–74, 1999
1999
Earlier work this paper cites.
K. B. Petersen, M. S. Pedersen et al. , “The matrix cookbook,” Technical University of Denmark , vol. 7, no. 15, p. 510, 2008
2008
Earlier work this paper cites.
L. Van der Maaten and G. Hinton, “Visualizing data using t-sne.” Journal of machine learning research , vol. 9, no. 11, 2008
2008
Earlier work this paper cites.
A. Krizhevsky and G. Hinton, “Learning multiple layers of features from tiny images,” Citeseer, Tech. Rep., 2009
2009
Earlier work this paper cites.
S.-A. Rebuffi, A. Kolesnikov, G. Sperl, and C. H. Lampert, “icarl: Incremental classifier and representation learning,” in Proceedings of the IEEE conference on Computer Vision and Pattern Recognition , 2017, pp. 2001–2010
2010
Earlier work this paper cites.
Y. Netzer, T. Wang, A. Coates, A. Bissacco, B. Wu, and A. Y. Ng, “Reading digits in natural images with unsupervised feature learning,” 2011
2011
Earlier work this paper cites.
S. Sagiroglu and D. Sinanc, “Big data: A review,” in 2013 international conference on collaboration technologies and systems (CTS) . IEEE, 2013, pp. 42–47
2013
Earlier work this paper cites.
Y. Bengio, A. Courville, and P. Vincent, “Representation learning: A review and new perspectives,” IEEE transactions on pattern analysis and machine intelligence , vol. 35, no. 8, pp. 1798–1828, 2013
2013
Earlier work this paper cites.
D. Maclaurin, D. Duvenaud, and R. Adams, “Gradient-based hyperparameter optimization through reversible learning,” in International conference on machine learning . PMLR, 2015, pp. 2113–2122
2015
Earlier work this paper cites.
O. Russakovsky, J. Deng, H. Su, J. Krause, S. Satheesh, S. Ma, Z. Huang, A. Karpathy, A. Khosla, M. Bernstein et al. , “Imagenet large scale visual recognition challenge,” International journal of computer vision , vol. 115, pp. 211–252, 2015
2015
Earlier work this paper cites.
Y. Le and X. Yang, “Tiny imagenet visual recognition challenge,” CS 231N , vol. 7, p. 7, 2015
2015
Earlier work this paper cites.
S. Ren, K. He, R. Girshick, and J. Sun, “Faster r-cnn: Towards real-time object detection with region proposal networks,” Advances in neural information processing systems , vol. 28, 2015
2015
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2016, pp. 770–778
2016
Earlier work this paper cites.
C. Szegedy, V. Vanhoucke, S. Ioffe, J. Shlens, and Z. Wojna, “Rethinking the inception architecture for computer vision,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2016, pp. 2818–2826
2016
Earlier work this paper cites.
A. Krizhevsky, I. Sutskever, and G. E. Hinton, “Imagenet classification with deep convolutional neural networks,” Communications of the ACM , vol. 60, no. 6, pp. 84–90, 2017
2017
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin, “Attention is all you need,” Advances in neural information processing systems , vol. 30, 2017
2017
Earlier work this paper cites.
C. Szegedy, S. Ioffe, V. Vanhoucke, and A. Alemi, “Inception-v4, inception-resnet and the impact of residual connections on learning,” in Proceedings of the AAAI conference on artificial intelligence , vol. 31, no. 1, 2017
2017
Earlier work this paper cites.
H. Xiao, K. Rasul, and R. Vollgraf. (2017) Fashion-mnist: a novel image dataset for benchmarking machine learning algorithms
2017
Earlier work this paper cites.
G. Huang, Z. Liu, L. Van Der Maaten, and K. Q. Weinberger, “Densely connected convolutional networks,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2017, pp. 4700–4708
2017
Earlier work this paper cites.
B. McMahan, E. Moore, D. Ramage, S. Hampson, and B. A. y Arcas, “Communication-efficient learning of deep networks from decentralized data,” in Artificial intelligence and statistics . PMLR, 2017, pp. 1273–1282
2017
Earlier work this paper cites.
S. Herath, M. Harandi, and F. Porikli, “Going deeper into action recognition: A survey,” Image and vision computing , vol. 60, pp. 4–21, 2017
2017
Earlier work this paper cites.
2018
Earlier work this paper cites.
A. Jacot, F. Gabriel, and C. Hongler, “Neural tangent kernel: Convergence and generalization in neural networks,” Advances in neural information processing systems , vol. 31, 2018
2018
Earlier work this paper cites.
H. Li, Z. Xu, G. Taylor, C. Studer, and T. Goldstein, “Visualizing the loss landscape of neural nets,” Advances in neural information processing systems , vol. 31, 2018
2018
Earlier work this paper cites.
D. Zhang, J. Yin, X. Zhu, and C. Zhang, “Network representation learning: A survey,” IEEE transactions on Big Data , vol. 6, no. 1, pp. 3–28, 2018
2018
Earlier work this paper cites.
S. Gidaris and N. Komodakis, “Dynamic few-shot visual learning without forgetting,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2018, pp. 4367–4375
2018
Earlier work this paper cites.
F. M. Castro, M. J. Marín-Jiménez, N. Guil, C. Schmid, and K. Alahari, “End-to-end incremental learning,” in Proceedings of the European conference on computer vision (ECCV) , 2018, pp. 233–248
2018
Earlier work this paper cites.
E. Strubell, A. Ganesh, and A. McCallum, “Energy and policy considerations for deep learning in NLP,” in Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics . Florence, Italy: Association for Computational Linguistics, Jul. 2019, pp. 3645–3650. [Online]. Available: https://aclanthology.org/P19-1355
2019
Earlier work this paper cites.
J. Nalepa and M. Kawulok, “Selecting training sets for support vector machines: a review,” Artificial Intelligence Review , vol. 52, no. 2, pp. 857–900, 2019
2019
Earlier work this paper cites.
A. Bietti and J. Mairal, “On the inductive bias of neural tangent kernels,” Advances in Neural Information Processing Systems , vol. 32, 2019
2019
Earlier work this paper cites.
J. Devlin, M.-W. Chang, K. Lee, and K. Toutanova, “BERT: Pre-training of deep bidirectional transformers for language understanding,” in Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long and Short Papers) . Minneapolis, Minnesota: Association for Computational Linguistics, Jun. 2019, pp. 4171–4186. [Online]. Available: https://aclanthology.org/N19-1423
2019
Earlier work this paper cites.
Q. Yang, Y. Liu, Y. Cheng, Y. Kang, T. Chen, and H. Yu, “Federated learning,” Synthesis Lectures on Artificial Intelligence and Machine Learning , vol. 13, no. 3, pp. 1–207, 2019
2019
Earlier work this paper cites.
T. Chen, S. Kornblith, M. Norouzi, and G. Hinton, “A simple framework for contrastive learning of visual representations,” in International conference on machine learning . PMLR, 2020, pp. 1597–1607
2020
Earlier work this paper cites.
K. He, H. Fan, Y. Wu, S. Xie, and R. Girshick, “Momentum contrast for unsupervised visual representation learning,” in Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2020, pp. 9729–9738
2020
Earlier work this paper cites.
T. Brown, B. Mann, N. Ryder, M. Subbiah, J. D. Kaplan, P. Dhariwal, A. Neelakantan, P. Shyam, G. Sastry, A. Askell et al. , “Language models are few-shot learners,” Advances in neural information processing systems , vol. 33, pp. 1877–1901, 2020
2020
Earlier work this paper cites.
B. Mirzasoleiman, J. Bilmes, and J. Leskovec, “Coresets for data-efficient training of machine learning models,” in International Conference on Machine Learning . PMLR, 2020, pp. 6950–6960
2020
Cited alongside, same era.
C. Coleman, C. Yeh, S. Mussmann, B. Mirzasoleiman, P. Bailis, P. Liang, J. Leskovec, and M. Zaharia, “Selection via proxy: Efficient data selection for deep learning,” in International Conference on Learning Representations , 2020. [Online]. Available: https://openreview.net/forum?id=HJg2b0VYDr
2020
Cited alongside, same era.
O. Bohdal, Y. Yang, and T. Hospedales, “Flexible dataset distillation: Learn labels instead of images,” in Proceedings of the Advances in Neural Information Processing Systems (NeurIPS), Workshop , 2020
2020
Cited alongside, same era.
F. P. Such, A. Rawal, J. Lehman, K. Stanley, and J. Clune, “Generative teaching networks: Accelerating neural architecture search by learning to generate synthetic training data,” in International Conference on Machine Learning . PMLR, 2020, pp. 9206–9216
A. Carta, A. Cossu, V. Lomonaco, and D. Bacciu, “Distilled replay: Overcoming forgetting through synthetic samples,” in Continual Semi-Supervised Learning: First International Workshop, CSSL 2021, Virtual Event, August 19-20, 2021, Revised Selected Papers , vol. 13418. Springer Nature, 2022, p. 104
2022
Later among the works it cites.
M. Sangermano, A. Carta, A. Cossu, and D. Bacciu, “Sample condensation in online continual learning,” in 2022 International Joint Conference on Neural Networks (IJCNN) . IEEE, 2022, pp. 01–08
2022
Later among the works it cites.
V. P. C, C. White, P. Jain, S. Nayak, R. K. Iyer, and G. Ramakrishnan, “Speeding up NAS with adaptive subset selection,” in First Conference on Automated Machine Learning (Late-Breaking Workshop) , 2022. [Online]. Available: https://openreview.net/forum?id=52OEvDa5xrj
2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2020
Cited alongside, same era.
W. Masarczyk and I. Tautkute, “Reducing catastrophic forgetting with learning on synthetic data,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition Workshops , 2020, pp. 252–253
2020
Cited alongside, same era.
G. Li, G. Qian, I. C. Delgadillo, M. Muller, A. Thabet, and B. Ghanem, “Sgas: Sequential greedy architecture search,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2020, pp. 1620–1630
2020
Cited alongside, same era.
2020
Cited alongside, same era.
2020
Cited alongside, same era.
——, “Soft-label anonymous gastric x-ray image distillation,” in Proceedings of the IEEE International Conference on Image Processing (ICIP) , 2020, pp. 305–309
2020
Cited alongside, same era.
D. Medvedev and A. D’yakonov, “New properties of the data distillation method when working with tabular data,” in Proceedings of the International Conference on Analysis of Images, Social Networks and Texts (AIST) , 2020, pp. 379–390
2020
Cited alongside, same era.
I. Sucholutsky and M. Schonlau, “Soft-label dataset distillation and text dataset distillation,” in 2021 International Joint Conference on Neural Networks (IJCNN) . IEEE, 2021, pp. 1–8
2021
Cited alongside, same era.
T. Nguyen, Z. Chen, and J. Lee, “Dataset meta-learning from kernel ridge-regression,” in Proceedings of the International Conference on Learning Representations (ICLR) , 2021
2021
Cited alongside, same era.
2022
Later among the works it cites.
2022
Later among the works it cites.
2022
Later among the works it cites.
2022
Later among the works it cites.
G. Cazenavette, T. Wang, A. Torralba, A. A. Efros, and J.-Y. Zhu, “Wearable ImageNet: Synthesizing tileable textures via dataset distillation,” pp. 2278–2282, 2022
2022
Later among the works it cites.
Y. Chen, Z. Wu, Z. Shen, and J. Jia, “Learning from designers: Fashion compatibility analysis via dataset distillation,” in Proceedings of the IEEE International Conference on Image Processing (ICIP) , 2022, pp. 856–860
2022
Later among the works it cites.
——, “Compressed gastric image generation based on soft-label dataset distillation for medical data sharing,” Computer Methods and Programs in Biomedicine , vol. 227, p. 107189, 2022
2022
Later among the works it cites.
N. Sachdeva, M. Preet Dhaliwal, C.-J. Wu, and J. McAuley, “Infinite recommendation networks: A data-centric approach,” in Proceedings of the Advances in Neural Information Processing Systems (NeurIPS) , 2022
2022
Later among the works it cites.
2022
Later among the works it cites.
T. Dong, B. Zhao, and L. Liu, “Privacy for free: How does dataset condensation help privacy?” in Proceedings of the International Conference on Machine Learning (ICML) , 2022, pp. 5378–5396
2022
Later among the works it cites.
2022
Later among the works it cites.
D. Chen, R. Kerkouche, and M. Fritz, “Private set generation with discriminative information,” in Proceedings of the Advances in Neural Information Processing Systems (NeurIPS) , 2022
2022
Later among the works it cites.
2022
Later among the works it cites.
2022
Later among the works it cites.
N. Loo, R. Hasani, M. Lechner, and D. Rus, “Dataset distillation with convexified implicit gradients,” in Proceedings of the International Conference on Machine Learning (ICML) , 2023
2023
Closest in time.
J. Cui, R. Wang, S. Si, and C.-J. Hsieh, “Scaling up dataset distillation to imagenet-1k with constant memory,” in Proceedings of the International Conference on Machine Learning (ICML) , 2023
2023
Closest in time.
B. Zhao and H. Bilen, “Dataset condensation with distribution matching,” in Proceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision (WACV) , 2023
2023
Closest in time.
G. Zhao, G. Li, Y. Qin, and Y. Yu, “Improved distribution matching for dataset condensation,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2023, pp. 7856–7865
2023
Closest in time.
N. Sachdeva and J. McAuley, “Data distillation: A survey,” Transactions on Machine Learning Research , 2023, survey Certification. [Online]. Available: https://openreview.net/forum?id=lmXMXP74TO
2023
Closest in time.
G. Cazenavette, T. Wang, A. Torralba, A. A. Efros, and J.-Y. Zhu, “Generalizing dataset distillation via deep generative prior,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , 2023
2023
Closest in time.
2023
Closest in time.
Y. Liu, J. Gu, K. Wang, Z. Zhu, W. Jiang, and Y. You, “DREAM: Efficient dataset distillation by representative matching,” in Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV) , 2023
2023
Closest in time.
A. Maekawa, N. Kobayashi, K. Funakoshi, and M. Okumura, “Dataset distillation with attention labels for fine-tuning bert,” in Proceedings of the Annual Meeting of the Association for Computational Linguistics (ACL) , 2023, pp. 119–127
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
Y. Xiong, R. Wang, M. Cheng, F. Yu, and C.-J. Hsieh, “FedDM: Iterative distribution matching for communication-efficient federated learning,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , 2023
2023
Closest in time.
G. Li, R. Togo, T. Ogawa, and M. Haseyama, “Dataset distillation for medical dataset sharing,” pp. 1–6, 2023
2023
Closest in time.
2023
Closest in time.
T. Zheng and B. Li, “Differentially private dataset condensation,” 2023. [Online]. Available: https://openreview.net/forum?id=H8XpqEkbua_
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
D. Zhu, B. Lei, J. Zhang, Y. Fang, R. Zhang, Y. Xie, and D. Xu, “Rethinking data distillation: Do not overlook calibration,” in Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV) , 2023
2023
Closest in time.