Fetching the paper…
Reading the bibliography…
Dataset distillation compresses large datasets into smaller synthetic coresets which retain performance with the aim of reducing the storage and computational burden of processing the entire dataset.
Introduction to coresets: Accurate coresets
I. Jubran, A. Maalouf, and D. Feldman · 1910
Earlier work this paper cites.
The influence curve and its role in robust estimation
F. R. Hampel · 1974
Earlier work this paper cites.
Bayesian Learning for Neural Networks
R. M. Neal · 1996
Earlier work this paper cites.
Bayesian classification with gaussian processes
C. K. I. Williams and D. Barber · 1998
Earlier work this paper cites.
Input space versus feature space in kernel-based methods
B. Scholkopf, S. Mika, C. Burges, P. Knirsch, K.-R. Muller, G. Ratsch, and A. Smola · 1999
Earlier work this paper cites.
Probabilistic outputs for support vector machines and comparisons to regularized likelihood methods
J. Platt · 2000
Earlier work this paper cites.
On coresets for k-means and k-median clustering
S. Har-Peled and S. Mazumdar · 2004
Earlier work this paper cites.
Gaussian processes for machine learning
C. E. Rasmussen and C. K. I. Williams · 2006
Earlier work this paper cites.
Sparse gaussian processes using pseudo-inputs
E. Snelson and Z. Ghahramani · 2006
Earlier work this paper cites.
Random features for large-scale kernel machines
A. Rahimi and B. Recht · 2007
Earlier work this paper cites.
Learning multiple layers of features from tiny images
A. Krizhevsky, G. Hinton, et al · 2009
Earlier work this paper cites.
Variational learning of inducing variables in sparse gaussian processes
M. Titsias · 2009
Earlier work this paper cites.
Super-samples from kernel herding
Y. Chen, M. Welling, and A. Smola · 2010
Earlier work this paper cites.
Mnist handwritten digit database
Y. LeCun, C. Cortes, and C. Burges · 2010
Earlier work this paper cites.
Prototype selection for interpretable classification
J. Bien and R. Tibshirani · 2011
Earlier work this paper cites.
Reading digits in natural images with unsupervised feature learning
Y. Netzer, T. Wang, A. Coates, A. Bissacco, B. Wu, and A. Y. Ng · 2011
Earlier work this paper cites.
Understanding classifier errors by examining influential neighbors
M. Kabra, A. Robie, and K. Branson · 2015
Earlier work this paper cites.
Gradient-based hyperparameter optimization through reversible learning
D. Maclaurin, D. Duvenaud, and R. Adams · 2015
Earlier work this paper cites.
Deep learning with differential privacy
M. Abadi, A. Chu, I. Goodfellow, H. B. McMahan, I. Mironov, K. Talwar, and L. Zhang · 2016
Earlier work this paper cites.
Toward deeper understanding of neural networks: The power of initialization and a dual view on expressivity
A. Daniely, R. Frostig, and Y. Singer · 2016
Earlier work this paper cites.
Examples are not enough, learn to criticize! criticism for interpretability
B. Kim, R. Khanna, and O. Koyejo · 2016
Earlier work this paper cites.
J. M. Phillips · 2016
Earlier work this paper cites.
Matching networks for one shot learning
O. Vinyals, C. Blundell, T. Lillicrap, D. Wierstra, et al · 2016
Earlier work this paper cites.
Input sparsity time low-rank approximation via ridge leverage score sampling
M. B. Cohen, C. Musco, and C. Musco · 2017
Cited alongside, same era.
Understanding black-box predictions via influence functions
P. W. Koh and P. Liang · 2017
Cited alongside, same era.
Training gaussian mixture models at scale via coresets
M. Lucic, M. Faulkner, A. Krause, and D. Feldman · 2017
Cited alongside, same era.
icarl: Incremental classifier and representation learning
S.-A. Rebuffi, A. Kolesnikov, G. Sperl, and C. H. Lampert · 2017
Cited alongside, same era.
Prototypical networks for few-shot learning
J. Snell, K. Swersky, and R. Zemel · 2017
Cited alongside, same era.
Fashion-mnist: a novel image dataset for benchmarking machine learning algorithms
H. Xiao, K. Rasul, and R. Vollgraf · 2017
Tensor programs i: Wide feedforward or recurrent neural networks of any architecture are gaussian processes
G. Yang · 2019
Later among the works it cites.
Scail: Classifier weights scaling for class incremental learning
E. Belouadah and A. Popescu · 2020
Later among the works it cites.
Flexible dataset distillation: Learn labels instead of images
O. Bohdal, Y. Yang, and T. Hospedales · 2020
Later among the works it cites.
Coresets via bilevel optimization for continual learning and streaming
Z. Borsos, M. Mutnỳ, and A. Krause · 2020
Later among the works it cites.
When do neural networks outperform kernel methods?
B. Ghorbani, S. Mei, T. Misiakiewicz, and A. Montanari · 2020
Later among the works it cites.
Infinite attention: Nngp and ntk for deep attention networks
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
End-to-end incremental learning
F. M. Castro, M. J. Marín-Jiménez, N. Guil, C. Schmid, and K. Alahari · 2018
Cited alongside, same era.
Gaussian process behaviour in wide deep neural networks
A. G. de G. Matthews, J. Hron, M. Rowland, R. E. Turner, and Z. Ghahramani · 2018
Cited alongside, same era.
Neural tangent kernel: Convergence and generalization in neural networks
A. Jacot, F. Gabriel, and C. Hongler · 2018
Cited alongside, same era.
Deep neural networks as gaussian processes
J. Lee, J. Sohl-dickstein, J. Pennington, R. Novak, S. Schoenholz, and Y. Bahri · 2018
Cited alongside, same era.
Dirichlet-based gaussian processes for large-scale calibrated classification
D. Milios, R. Camoriano, P. Michiardi, L. Rosasco, and M. Filippone · 2018
Cited alongside, same era.
T. Wang, J.-Y. Zhu, A. Torralba, and A. A. Efros · 2018
Cited alongside, same era.
J. Hron, Y. Bahri, J. Sohl-Dickstein, and R. Novak · 2020
Later among the works it cites.
Simple and effective regularization methods for training on noisily labeled data with generalization guarantee
W. Hu, Z. Li, and D. Yu · 2020
Later among the works it cites.
Neural circuit policies enabling auditable autonomy
M. Lechner, R. Hasani, A. Amini, T. A. Henzinger, D. Rus, and R. Grosu · 2020
Later among the works it cites.
Finite versus infinite neural networks: an empirical study
J. Lee, S. Schoenholz, J. Pennington, B. Adlam, L. Xiao, R. Novak, and J. Sohl-Dickstein · 2020
Later among the works it cites.
Optimizing millions of hyperparameters by implicit differentiation
J. Lorraine, P. Vicol, and D. Duvenaud · 2020
Later among the works it cites.
Coresets for data-efficient training of machine learning models
B. Mirzasoleiman, J. A. Bilmes, and J. Leskovec · 2020
Later among the works it cites.
Neural tangents: Fast and easy infinite neural networks in python
R. Novak, L. Xiao, J. Hron, J. Lee, A. A. Alemi, J. Sohl-Dickstein, and S. S. Schoenholz · 2020
Later among the works it cites.
Towards nngp-guided neural architecture search
D. Peng, D. S. Park, J. Lee, J. Sohl-dickstein, and Y. Cao · 2020
Later among the works it cites.
Neural kernels without tangents
V. Shankar, A. Fang, W. Guo, S. Fridovich-Keil, J. Ragan-Kelley, L. Schmidt, and B. Recht · 2020
Later among the works it cites.
Non-Gaussian processes and neural networks at finite widths
S. Yaida · 2020
Later among the works it cites.
Differentiable augmentation for data-efficient gan training
S. Zhao, Z. Liu, J. Lin, J.-Y. Zhu, and S. Han · 2020
Later among the works it cites.
Adabelief optimizer: Adapting stepsizes by the belief in observed gradients
J. Zhuang, T. Tang, Y. Ding, S. Tatikonda, N. Dvornek, X. Papademetris, and J. S. Duncan · 2020
Later among the works it cites.
M. Refinetti, S. Goldt, F. Krzakala, and L. Zdeborová · 2021
Later among the works it cites.
Scaling neural tangent kernels via sketching and random features
A. Zandieh, I. Han, H. Avron, N. Shoham, C. Kim, and J. Shin · 2021
Later among the works it cites.
Dataset condensation with gradient matching
B. Zhao, K. R. Mopuri, and H. Bilen · 2021
Later among the works it cites.
Evolution of neural tangent kernels under benign and adversarial training
N. Loo, R. Hasani, A. Amini, and D. Rus · 2022
Closest in time.
Fast finite width neural tangent kernel
R. Novak, J. Sohl-Dickstein, and S. S. Schoenholz · 2022
Closest in time.
Interpreting neural policies with disentangled tree representations
T.-H. Wang, W. Xiao, T. Seyde, R. Hasani, and D. Rus · 2022
Closest in time.