Fetching the paper…
Reading the bibliography…
We propose a new dataset distillation algorithm using reparameterization and convexification of implicit gradients (RCIG), that substantially improves the state-of-the-art.
MacKay, M., Vicol, P., Lorraine, J., Duvenaud, D., and Grosse, R. B · 1903
Earlier work this paper cites.
Meta-learning with implicit gradients
Rajeswaran, A., Finn, C., Kakade, S. M., and Levine, S · 1909
Earlier work this paper cites.
Such, F. P., Rawal, A., Lehman, J., Stanley, K. O., and Clune, J · 1912
Earlier work this paper cites.
The influence curve and its role in robust estimation
Hampel, F. R · 1974
Earlier work this paper cites.
Learning Internal Representations by Error Propagation , pp. 318–362
Rumelhart, D. E. and McClelland, J. L · 1987
Earlier work this paper cites.
Generalization of backpropagation with application to a recurrent gas market model
Werbos, P. J · 1988
Earlier work this paper cites.
Backpropagation through time: what it does and how to do it
Werbos, P · 1990
Earlier work this paper cites.
Fast Exact Multiplication by the Hessian
Pearlmutter, B. A · 1994
Earlier work this paper cites.
Gradient-based learning applied to document recognition
Lecun, Y., Bottou, L., Bengio, Y., and Haffner, P · 1998
Earlier work this paper cites.
Optimal use of regularization and cross-validation in neural network modeling
Chen, D. and Hagan, M · 1999
Earlier work this paper cites.
Gradient-based optimization of hyperparameters
Bengio, Y · 2000
Earlier work this paper cites.
Influence functions in deep learning are fragile, 2020
Basu, S., Pope, P., and Feizi, S · 2006
Earlier work this paper cites.
Learning multiple layers of features from tiny images
Krizhevsky, A · 2009
Earlier work this paper cites.
Deep learning via hessian-free optimization
Martens, J · 2010
Earlier work this paper cites.
Caltech-UCSD Birds 200
Welinder, P., Branson, S., Mita, T., Wah, C., Schroff, F., Belongie, S., and Perona, P · 2010
Earlier work this paper cites.
A unified framework for approximating and clustering data
Feldman, D. and Langberg, M · 2011
Earlier work this paper cites.
Generic methods for optimization-based modeling
Domke, J · 2012
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks
Krizhevsky, A., Sutskever, I., and Hinton, G. E · 2012
Earlier work this paper cites.
Intriguing properties of neural networks
Szegedy, C., Zaremba, W., Sutskever, I., Bruna, J., Erhan, D., Goodfellow, I., and Fergus, R · 2013
Earlier work this paper cites.
Imagenet large scale visual recognition challenge
Russakovsky, O., Deng, J., Su, H., Krause, J., Satheesh, S., Ma, S., Huang, Z., Karpathy, A., Khosla, A., Bernstein, M. S., Berg, A. C., and Fei-Fei, L · 2014
Earlier work this paper cites.
Very deep convolutional networks for large-scale image recognition
Simonyan, K. and Zisserman, A · 2014
Earlier work this paper cites.
Deep residual learning for image recognition
He, K., Zhang, X., Ren, S., and Sun, J · 2015
Earlier work this paper cites.
Batch normalization: Accelerating deep network training by reducing internal covariate shift
Ioffe, S. and Szegedy, C · 2015
Earlier work this paper cites.
Adam: A method for stochastic optimization
Kingma, D. P. and Ba, J · 2015
Cited alongside, same era.
Tiny imagenet visual recognition challenge
Le, Y. and Yang, X. S · 2015
Cited alongside, same era.
Gradient-based hyperparameter optimization through reversible learning
Maclaurin, D., Duvenaud, D., and Adams, R. P · 2015
Cited alongside, same era.
Deep learning with differential privacy
Abadi, M., Chu, A., Goodfellow, I., McMahan, H. B., Mironov, I., Talwar, K., and Zhang, L · 2016
Cited alongside, same era.
Second-order stochastic optimization for machine learning in linear time
Agarwal, N., Bullins, B., and Hazan, E · 2016
Cited alongside, same era.
Approximate k-means++ in sublinear time
The DeepMind JAX Ecosystem, 2020
Babuschkin, I., Baumli, K., Bell, A., Bhupatiraju, S., Bruce, J., Buchlovsky, P., Budden, D., Cai, T., Clark, A., Danihelka, I., Fantacci, C., Godwin, J., Jones, C., Hemsley, R., Hennigan, T., Hessel, M., Hou, S., Kapturowski, S., Keck, T., Kemaev, I., King, M., Kunesch, M., Martens, L., Merzic, H., Mikulik, V., Norman, T., Quan, J., Papamakarios, G., Ring, R., Ruiz, F., Sanchez, A., Schneider, R., Sezener, E., Spencer, S., Srinivasan, S., Wang, L., Stokowiec, W., and Viola, F · 2020
Later among the works it cites.
Coresets via bilevel optimization for continual learning and streaming
Borsos, Z., Mutny, M., and Krause, A · 2020
Later among the works it cites.
Deep learning versus kernel learning: an empirical study of loss landscape geometry and the time evolution of the neural tangent kernel
Fort, S., Dziugaite, G. K., Paul, M., Kharaghani, S., Roy, D. M., and Ganguli, S · 2020
Later among the works it cites.
Finite depth and width corrections to the neural tangent kernel
Hanin, B. and Nica, M · 2020
Later among the works it cites.
Flax: A neural network library and ecosystem for JAX, 2020
Heek, J., Levskaya, A., Oliver, A., Ritter, M., Rondepierre, B., Steiner, A., and van Zee, M · 2020
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Bachem, O., Lucic, M., Hassani, S. H., and Krause, A · 2016
Cited alongside, same era.
Coresets for scalable bayesian logistic regression
Huggins, J. H., Campbell, T., and Broderick, T · 2016
Cited alongside, same era.
Strong coresets for hard and soft bregman clustering with applications to exponential family mixtures
Lucic, M., Bachem, O., and Krause, A · 2016
Cited alongside, same era.
Equilibrium propagation: Bridging the gap between energy-based models and backpropagation, 2016
Scellier, B. and Bengio, Y · 2016
Cited alongside, same era.
Membership inference attacks against machine learning models
Shokri, R., Stronati, M., and Shmatikov, V · 2016
Cited alongside, same era.
Model-agnostic meta-learning for fast adaptation of deep networks
Finn, C., Abbeel, P., and Levine, S · 2017
Cited alongside, same era.
Understanding black-box predictions via influence functions
Koh, P. W. and Liang, P · 2017
Cited alongside, same era.
Later among the works it cites.
A smaller subset of 10 easily classified classes from imagenet, and a little more french, 2020
Howard, J · 2020
Later among the works it cites.
Finite versus infinite neural networks: an empirical study
Lee, J., Schoenholz, S., Pennington, J., Adlam, B., Xiao, L., Novak, R., and Sohl-Dickstein, J · 2020
Later among the works it cites.
Coresets for data-efficient training of machine learning models
Mirzasoleiman, B., Bilmes, J. A., and Leskovec, J · 2020
Later among the works it cites.
Coresets for near-convex functions
Tukan, M., Maalouf, A., and Feldman, D · 2020
Later among the works it cites.
Adabelief optimizer: Adapting stepsizes by the belief in observed gradients
Zhuang, J., Tang, T., Ding, Y., Tatikonda, S., Dvornek, N., Papademetris, X., and Duncan, J. S · 2020
Later among the works it cites.
Randomized automatic differentiation
Oktay, D., McGreivy, N., Aduol, J., Beatson, A., and Adams, R. P · 2021
Later among the works it cites.
Unbiased gradient estimation in unrolled computation graphs with persistent evolution strategies
Vicol, P., Metz, L., and Sohl-Dickstein, J · 2021
Later among the works it cites.
Dataset condensation with gradient matching
Zhao, B., Mopuri, K. R., and Bilen, H · 2021
Later among the works it cites.
A contrastive rule for meta-learning
Zucchet, N., Schug, S., von Oswald, J., Zhao, D., and Sacramento, J · 2021
Later among the works it cites.
If influence functions are the answer, then what is the question?
Bae, J., Ng, N. H., Lo, A., Ghassemi, M., and Grosse, R. B · 2022
Later among the works it cites.
No free lunch in ”privacy for free: How does dataset condensation help privacy”, 2022
Carlini, N., Feldman, V., and Nasr, M · 2022
Later among the works it cites.
Privacy for free: How does dataset condensation help privacy?
Dong, T., Zhao, B., and Lyu, L · 2022
Later among the works it cites.
Adaptive second order coresets for data-efficient machine learning
Pooladzandi, O., Davini, D., and Mirzasoleiman, B · 2022
Later among the works it cites.
Sample condensation in online continual learning, 2022
Sangermano, M., Carta, A., Cossu, A., and Bacciu, D · 2022
Later among the works it cites.
On implicit bias in overparameterized bilevel optimization
Vicol, P., Lorraine, J. P., Pedregosa, F., Duvenaud, D., and Grosse, R. B · 2022
Later among the works it cites.
Cafe: Learning to condense dataset by aligning features
Wang, K., Zhao, B., Peng, X., Zhu, Z., Yang, S., Wang, S., Huang, G., Bilen, H., Wang, X., and You, Y · 2022
Later among the works it cites.
Tct: Convexifying federated learning using bootstrapped neural tangent kernels, 2022
Yu, Y., Wei, A., Karimireddy, S. P., Ma, Y., and Jordan, M. I · 2022
Later among the works it cites.
Dataset distillation using neural feature regression
Zhou, Y., Nezhadarya, E., and Ba, J · 2022
Later among the works it cites.