Fetching the paper…
Reading the bibliography…
Symmetry is a fundamental tool in the exploration of a broad range of complex systems.
Johanni Brea, Berfin Simsek, Bernd Illing and Wulfram Gerstner · 1907
Earlier work this paper cites.
“Invariante Variationsprobleme”
E Noether · 1918
Earlier work this paper cites.
“A Relationship Between Arbitrary Positive Matrices and Doubly Stochastic Matrices”
Richard Sinkhorn · 1964
Earlier work this paper cites.
“Principles of Mathematical Analysis”
Walter Rudin · 1976
Earlier work this paper cites.
“Linear representations of finite groups”
Jean-Pierre Serre · 1977
Earlier work this paper cites.
“What is a random matrix”
Persi Diaconis · 2005
Earlier work this paper cites.
“Measuring Statistical Dependence with Hilbert-Schmidt Norms”
Arthur Gretton, Olivier Bousquet, Alex Smola and Bernhard Schölkopf · 2005
Earlier work this paper cites.
“Universal Kernels”
Charles. Micchelli, Yuesheng Xu and Haizhang Zhang · 2006
Earlier work this paper cites.
“Kernel methods in machine learning”
Thomas Hofmann, Bernhard Schölkopf and Alexander. Smola · 2008
Earlier work this paper cites.
“Imagenet: A large-scale hierarchical image database”
Jia Deng et al · 2009
Earlier work this paper cites.
“Visualizing higher-layer features of a deep network”
Dumitru Erhan, Yoshua Bengio, Aaron Courville and Pascal Vincent · 2009
Earlier work this paper cites.
“Learning multiple layers of features from tiny images”, 2009
Alex Krizhevsky · 2009
Earlier work this paper cites.
“Optimizing Mode Connectivity via Neuron Alignment”, 2020
N. Tatro et al · 2009
Earlier work this paper cites.
“Torchvision the machine-vision package of torch”
Sébastien Marcel and Yann Rodriguez · 2010
Earlier work this paper cites.
“Rectified Linear Units Improve Restricted Boltzmann Machines”
Vinod Nair and Geoffrey. Hinton · 2010
Earlier work this paper cites.
“Algorithms for learning kernels based on centered alignment”
Corinna Cortes, Mehryar Mohri and Afshin Rostamizadeh · 2012
Earlier work this paper cites.
“Neural Mechanics: Symmetry and Broken Conservation Laws in Deep Learning Dynamics”, 2021
Daniel Kunin et al · 2012
Earlier work this paper cites.
“Convex Relaxations for Permutation Problems”
Fajwel Fogel, Rodolphe Jenatton, Francis. Bach and Alexandre d’Aspremont · 2013
Earlier work this paper cites.
“Beyond the Birkhoff Polytope: Convex Relaxations for Vector Permutation Problems”
Cong Lim and Stephen. Wright · 2014
Earlier work this paper cites.
“Intriguing properties of neural networks”
Christian Szegedy et al · 2014
Earlier work this paper cites.
“Visualizing and understanding convolutional networks”
Matthew Zeiler and Rob Fergus · 2014
Earlier work this paper cites.
“Understanding Symmetries in Deep Networks”
Vijay Badrinarayanan, Bamdev Mishra and R. Cipolla · 2015
Cited alongside, same era.
“Batch Normalization: Accelerating Deep Network Training by Reducing Internal Covariate Shift”
S. Ioffe and Christian Szegedy · 2015
Cited alongside, same era.
“Convergent Learning: Do Different Neural Networks Learn the Same Representations?”
Yixuan Li et al · 2015
Cited alongside, same era.
“Understanding Image Representations by Measuring Their Equivariance and Equivalence”
Karel Lenc and A. Vedaldi · 2015
Cited alongside, same era.
“Understanding Neural Networks Through Deep Visualization”
Jason Yosinski et al · 2015
Cited alongside, same era.
“Positively Scale-Invariant Flatness of ReLU Neural Networks”
Mingyang Yi et al · 2019
Later among the works it cites.
“Cutmix: Regularization strategy to train strong classifiers with localizable features”
Sangdoo Yun et al · 2019
Later among the works it cites.
“Zoom in: An introduction to circuits”
Chris Olah et al · 2020
Later among the works it cites.
“SoftSort: A Continuous Relaxation for the argsort Operator”
Sebastian Prillo and Julian Eisenschlos · 2020
Later among the works it cites.
“Reverse-Engineering Deep ReLU Networks”
D. Rolnick and Konrad Kording · 2020
Later among the works it cites.
“SciPy 1.0: Fundamental Algorithms for Scientific Computing in Python”
Pauli Virtanen et al · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Bolei Zhou et al · 2015
Cited alongside, same era.
“Deep Learning” http://www.deeplearningbook.org
Ian Goodfellow, Yoshua Bengio and Aaron Courville · 2016
Cited alongside, same era.
“Gaussian Error Linear Units (GELUs)”
Dan Hendrycks and Kevin Gimpel · 2016
Cited alongside, same era.
“Network Dissection: Quantifying Interpretability of Deep Visual Representations”
David Bau et al · 2017
Cited alongside, same era.
“Topology and Geometry of Half-Rectified Network Optimization”
C. Freeman and Joan Bruna · 2017
Cited alongside, same era.
“Neural tangent kernel: convergence and generalization in neural networks (invited paper)”
Arthur Jacot, Franck Gabriel and Clément Hongler · 2018
Cited alongside, same era.
“Learning Latent Permutations with Gumbel-Sinkhorn Networks”
Gonzalo Mena, David Belanger, Scott Linderman and Jasper Snoek · 2018
Cited alongside, same era.
“Revisiting Model Stitching to Compare Neural Representations”
Yamini Bansal, Preetum Nakkiran and Boaz Barak · 2021
Later among the works it cites.
“Similarity and Matching of Neural Network Representations”
Adrián Csiszárik et al · 2021
Later among the works it cites.
“Grounding Representation Similarity with Statistical Testing”
Frances Ding, Jean-Stanislas Denain and J. Steinhardt · 2021
Later among the works it cites.
“Grounding Representation Similarity Through Statistical Testing”
Frances Ding, Jean-Stanislas Denain and Jacob Steinhardt · 2021
Later among the works it cites.
“A mathematical framework for transformer circuits”, 2021
N Elhage et al · 2021
Later among the works it cites.
“Escaping the big data paradigm with compact transformers”
Ali Hassani et al · 2021
Later among the works it cites.
“Do Wide and Deep Networks Learn the Same Things? Uncovering How Neural Network Representations Vary with Width and Depth”
Thao Nguyen, Maithra Raghu and Simon Kornblith · 2021
Later among the works it cites.
“Do vision transformers see like convolutional neural networks?”
Maithra Raghu et al · 2021
Later among the works it cites.
“Generalized Shape Metrics on Neural Representations”
Alex. Williams, Erin’Mara Kunz, Simon Kornblith and Scott. Linderman · 2021
Later among the works it cites.
“Git Re-Basin: Merging Models modulo Permutation Symmetries”
Samuel. Ainsworth, Jonathan Hayase and Siddhartha Srinivasa · 2022
Closest in time.
“Softmax Linear Units”
Nelson Elhage et al · 2022
Closest in time.
“The Role of Permutation Invariance in Linear Mode Connectivity of Neural Networks”
Rahim Entezari, Hanie Sedghi, Olga Saukh and Behnam Neyshabur · 2022
Closest in time.
“ffcv” commit 849, https://github.com/libffcv/ffcv/ , 2022
Guillaume Leclerc et al · 2022
Closest in time.
Zhuang Liu et al · 2022
Closest in time.