Fetching the paper…
Reading the bibliography…
A principled understanding of generalization in deep learning may require unifying disparate observations under a single conceptual framework.
Early Stopping in Deep Networks: Double Descent and How to Eliminate it, September 2020
Reinhard Heckel and Fatih Furkan Yilmaz · 2007
Earlier work this paper cites.
The basic ai drives
Stephen Omohundro · 2008
Earlier work this paper cites.
Superintelligence: Paths, Dangers, Strategies
Nick Bostrom · 2014
Earlier work this paper cites.
A closer look at memorization in deep networks
Devansh Arpit, Stanisław Jastrzębski, Nicolas Ballas, David Krueger, Emmanuel Bengio, Maxinder S Kanwal, Tegan Maharaj, Asja Fischer, Aaron Courville, Yoshua Bengio, et al · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin · 2017
Earlier work this paper cites.
Reconciling modern machine learning practice and the bias-variance trade-off
Mikhail Belkin, Daniel Hsu, Siyuan Ma, and Soumik Mandal · 2018
Earlier work this paper cites.
The building blocks of interpretability
Chris Olah, Arvind Satyanarayan, Ian Johnson, Shan Carter, Ludwig Schubert, Katherine Ye, and Alexander Mordvintsev · 2018
Earlier work this paper cites.
Explaining neural scaling laws
Yasaman Bahri, Ethan Dyer, Jared Kaplan, Jaehoon Lee, and Utkarsh Sharma · 2021
Earlier work this paper cites.
Eliciting latent knowledge, 2021
Paul Christiano, Mark Xu, and Ajeya Cotra · 2021
Cited alongside, same era.
Truthfulqa: Measuring how models mimic human falsehoods
Stephanie Lin, Jacob Hilton, and Owain Evans · 2021
Cited alongside, same era.
Deep double descent: Where bigger models and more data hurt
Preetum Nakkiran, Gal Kaplun, Yamini Bansal, Tristan Yang, Boaz Barak, and Ilya Sutskever · 2021
Cited alongside, same era.
Multi-scale Feature Learning Dynamics: Insights for Double Descent, December 2021
Mohammad Pezeshki, Amartya Mitra, Yoshua Bengio, and Guillaume Lajoie · 2021
Cited alongside, same era.
Visible thoughts project, 2021
Nate Soares · 2021
Cited alongside, same era.
Is power-seeking ai an existential risk?
Joseph Carlsmith · 2022
Later among the works it cites.
Externalized reasoning oversight, 2022
Tamera Lanham · 2022
Later among the works it cites.
Towards understanding grokking: An effective theory of representation learning
Ziming Liu, Ouail Kitouni, Niklas Nolte, Eric J Michaud, Max Tegmark, and Mike Williams · 2022
Later among the works it cites.
A solvable model of neural scaling laws
Alexander Maloney, Daniel A Roberts, and James Sully · 2022
Later among the works it cites.
A Mechanistic Interpretability Analysis of Grokking, August 2022
Neel Nanda and Tom Lieberum · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cory Stephenson and Tyler Lee · 2021
Cited alongside, same era.
Optimal policies tend to seek power
Alex Turner, Logan Smith, Rohin Shah, Andrew Critch, and Prasad Tadepalli · 2021
Cited alongside, same era.
Tensor programs iv: Feature learning in infinite-width neural networks
Greg Yang and Edward J. Hu · 2021
Cited alongside, same era.
Alethea Power, Yuri Burda, Harri Edwards, Igor Babuschkin, and Vedant Misra · 2022
Later among the works it cites.
Scaling laws from the data manifold dimension
Utkarsh Sharma and Jared Kaplan · 2022
Later among the works it cites.
Metadata archaeology: Unearthing data subsets by leveraging training dynamics
Shoaib Ahmed Siddiqui, Nitarshan Rajkumar, Tegan Maharaj, David Krueger, and Sara Hooker · 2022
Later among the works it cites.