Fetching the paper…
Reading the bibliography…
Collecting large quantities of high-quality data can be prohibitively expensive or impractical, and a bottleneck in machine learning.
Bradley Efron and Carl Morris, Stein’s paradox in statistics , Scientific American 236
1977
Earlier work this paper cites.
Charles M Stein, Estimation of the mean of a multivariate normal distribution , The annals of Statistics (1981), 1135–1151
1981
Earlier work this paper cites.
Markus Reiß, Asymptotic equivalence for nonparametric regression with multivariate and random design , The Annals of Statistics (2008), 1957–1982
1982
Earlier work this paper cites.
Yehoram Gordon, Some inequalities for gaussian processes and applications , Israel Journal of Mathematics 50
1985
Earlier work this paper cites.
Lawrence D Brown and Mark G Low, Asymptotic equivalence of nonparametric regression and white noise , The Annals of Statistics 24
1996
Earlier work this paper cites.
Aaad W van der Vaart, Asymptotic statistics , Cambridge University Press, 2000
2000
Earlier work this paper cites.
Steven Bird, Nltk: the natural language toolkit , Proceedings of the COLING/ACL 2006 Interactive Presentation Sessions, 2006, pp. 69–72
2006
Earlier work this paper cites.
Alexandre B. Tsybakov, Introduction to nonparametric estimation , Springer, 2009
2009
Earlier work this paper cites.
Shai Ben-David, John Blitzer, Koby Crammer, Alex Kulesza, Fernando Pereira, and Jennifer Wortman Vaughan, A theory of learning from different domains , Machine learning 79
2010
Earlier work this paper cites.
Alex Krizhevsky, Ilya Sutskever, and Geoffrey E Hinton, Imagenet classification with deep convolutional neural networks , Advances in neural information processing systems 25
2012
Earlier work this paper cites.
Christos Thrampoulidis, Samet Oymak, and Babak Hassibi, Regularized linear regression: A precise analysis of the estimation error , Proceedings of Machine Learning Research 40
2015
Earlier work this paper cites.
Andreas Maurer, Massimiliano Pontil, and Bernardino Romera-Paredes, The benefit of multitask representation learning , Journal of Machine Learning Research 17
2016
Earlier work this paper cites.
German Ros, Laura Sellart, Joanna Materzynska, David Vazquez, and Antonio M Lopez, The synthia dataset: A large collection of synthetic images for semantic segmentation of urban scenes , Proceedings of the IEEE conference on computer vision and pattern recognition, 2016, pp. 3234–3243
2016
Earlier work this paper cites.
2017
Earlier work this paper cites.
Matthew Johnson-Roberson, Charles Barto, Rounak Mehta, Sharath Nittur Sridhar, Karl Rosaen, and Ram Vasudevan, Driving in the matrix: Can virtual worlds replace human-generated annotations for real world tasks? , 2017 IEEE International Conference on Robotics and Automation (ICRA), IEEE, 2017, pp. 746–753
2017
Earlier work this paper cites.
Hassan Abu Alhaija, Siva Karthik Mustikovela, Lars Mescheder, Andreas Geiger, and Carsten Rother, Augmented reality meets computer vision: Efficient data generation for urban driving scenes , International Journal of Computer Vision 126
2018
Cited alongside, same era.
Christos Thrampoulidis, Ehsan Abbasi, and Babak Hassibi, Precise error analysis of regularized m m -estimators in high dimensions , IEEE Transactions on Information Theory 64
2018
Cited alongside, same era.
Jonathan Tremblay, Aayush Prakash, David Acuna, Mark Brophy, Varun Jampani, Cem Anil, Thang To, Eric Cameracci, Shaad Boochoon, and Stan Birchfield, Training deep networks with synthetic data: Bridging the reality gap by domain randomization , Proceedings of the IEEE conference on computer vision and pattern recognition workshops, 2018, pp. 969–977
2018
Cited alongside, same era.
Roman Vershynin, High-dimensional probability: An introduction with applications in data science , vol. 47, Cambridge university press, 2018
2018
Léo Miolane and Andrea Montanari, The distribution of the lasso: Uniform control over sparse balls and adaptive parameter tuning , The Annals of Statistics 49
2021
Later among the works it cites.
Yi Tay, Mostafa Dehghani, Jinfeng Rao, William Fedus, Samira Abnar, Hyung Won Chung, Sharan Narang, Dani Yogatama, Ashish Vaswani, and Donald Metzler, Scale efficiently: Insights from pretraining and finetuning transformers , International Conference on Learning Representations, 2021
2021
Later among the works it cites.
Ibrahim M Alabdulmohsin, Behnam Neyshabur, and Xiaohua Zhai, Revisiting neural scaling laws in language and vision , Advances in Neural Information Processing Systems 35
2022
Later among the works it cites.
Chen Cheng and Andrea Montanari, Dimension free ridge regression , arXiv:2210.08571 (2022)
2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Yuhua Chen, Wen Li, Xiaoran Chen, and Luc Van Gool, Learning semantic segmentation from synthetic data: A geometrically guided input-output adaptation approach , Proceedings of the IEEE/CVF conference on computer vision and pattern recognition, 2019, pp. 1841–1850
2019
Cited alongside, same era.
Jonathan S Rosenfeld, Amir Rosenfeld, Yonatan Belinkov, and Nir Shavit, A constructive prediction of the generalization error across scales , International Conference on Learning Representations, 2019
2019
Cited alongside, same era.
Connor Shorten and Taghi M Khoshgoftaar, A survey on image data augmentation for deep learning , Journal of big data 6
2019
Cited alongside, same era.
Mary J Goldman, Brian Craft, Mim Hastie, Kristupas Repečka, Fran McDade, Akhil Kamath, Ayan Banerjee, Yunhai Luo, Dave Rogers, Angela N Brooks, et al., Visualizing and interpreting cancer genomics data via the xena platform , Nature biotechnology 38
2020
Cited alongside, same era.
2020
Cited alongside, same era.
2020
Cited alongside, same era.
2020
Cited alongside, same era.
Nilesh Tripuraneni, Michael Jordan, and Chi Jin, On the theory of transfer learning: The importance of task diversity , Advances in neural information processing systems 33
2020
Cited alongside, same era.
2022
Later among the works it cites.
Xuanli He, Islam Nassar, Jamie Kiros, Gholamreza Haffari, and Mohammad Norouzi, Generate, annotate, and learn: Nlp with synthetic text , Transactions of the Association for Computational Linguistics 10
2022
Later among the works it cites.
2022
Later among the works it cites.
Yu Meng, Jiaxin Huang, Yu Zhang, and Jiawei Han, Generating training data with language models: Towards zero-shot language understanding , Advances in Neural Information Processing Systems 35
2022
Later among the works it cites.
Arthur Moreau, Nathan Piasco, Dzmitry Tsishkou, Bogdan Stanciulescu, and Arnaud de La Fortelle, Lens: Localization enhanced by nerf synthesis , Conference on Robot Learning, PMLR, 2022, pp. 1347–1356
2022
Later among the works it cites.
Lin Yen-Chen, Pete Florence, Jonathan T Barron, Tsung-Yi Lin, Alberto Rodriguez, and Phillip Isola, Nerf-supervision: Learning dense object descriptors from neural radiance fields , 2022 International Conference on Robotics and Automation (ICRA), IEEE, 2022, pp. 6496–6503
2022
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
Feiyang Kang, Hoang Anh Just, Anit Kumar Sahu, and Ruoxi Jia, Performance scaling via optimal transport: Enabling data selection from partially revealed sources , Advances in Neural Information Processing Systems 36
2024
Closest in time.