Fetching the paper…
Reading the bibliography…
It has been observed that representations learned by distinct neural networks conceal structural similarities when the models are trained under similar inductive biases.
Roberta: A robustly optimized bert pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov · 1907
Earlier work this paper cites.
Relations between two sets of variates
Harold Hotelling · 1992
Earlier work this paper cites.
A global geometric framework for nonlinear dimensionality reduction
Joshua B Tenenbaum, Vin de Silva, and John C Langford · 2000
Earlier work this paper cites.
Toward semantics-based answer pinpointing
Eduard Hovy, Laurie Gerber, Ulf Hermjakob, Chin-Yew Lin, and Deepak Ravichandran · 2001
Earlier work this paper cites.
Rethinking channel dimensions for efficient model design
Dongyoon Han, Sangdoo Yun, Byeongho Heo, and YoungJoon Yoo · 2007
Earlier work this paper cites.
Collective Classification in Network Data
Prithviraj Sen, Galileo Namata, Mustafa Bilgic, Lise Getoor, Brian Galligher, and Tina Eliassi-Rad · 2008
Earlier work this paper cites.
Learning multiple layers of features from tiny images
Alex Krizhevsky, Geoffrey Hinton, et al · 2009
Earlier work this paper cites.
The mnist database of handwritten digit images for machine learning research
Li Deng · 2012
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P. Kingma and Jimmy Ba · 2015
Earlier work this paper cites.
Character-level convolutional networks for text classification
Xiang Zhang, Junbo Jake Zhao, and Yann LeCun · 2015
Earlier work this paper cites.
Group equivariant convolutional networks
Taco Cohen and Max Welling · 2016
Earlier work this paper cites.
Adaptive data augmentation for image classification
Alhussein Fawzi, Horst Samulowitz, Deepak Turaga, and Pascal Frossard · 2016
Earlier work this paper cites.
The heat method for distance computation
Keenan Crane, Clarisse Weischedel, and Max Wardetzky · 2017
Earlier work this paper cites.
Svcca: Singular vector canonical correlation analysis for deep learning dynamics and interpretability
Maithra Raghu, Justin Gilmer, Jason Yosinski, and Jascha Sohl-Dickstein · 2017
Earlier work this paper cites.
Deep convolutional neural networks and data augmentation for environmental sound classification
Justin Salamon and Juan Pablo Bello · 2017
Earlier work this paper cites.
The riemannian geometry of deep generative models, 2017
Hang Shao, Abhishek Kumar, and P. Thomas Fletcher · 2017
Earlier work this paper cites.
Harmonic networks: Deep translation and rotation equivariance
Daniel E Worrall, Stephan J Garbin, Daniyar Turmukhambetov, and Gabriel J Brostow · 2017
Earlier work this paper cites.
Fashion-mnist: a novel image dataset for benchmarking machine learning algorithms
Han Xiao, Kashif Rasul, and Roland Vollgraf · 2017
Earlier work this paper cites.
Oriented response networks
Yanzhao Zhou, Qixiang Ye, Qiang Qiu, and Jianbin Jiao · 2017
Earlier work this paper cites.
Insights on representational similarity in neural networks with canonical correlation
Ari Morcos, Maithra Raghu, and Samy Bengio · 2018
Cited alongside, same era.
Martin Arjovsky, Léon Bottou, Ishaan Gulrajani, and David Lopez-Paz · 2019
Cited alongside, same era.
A general theory of equivariant cnns on homogeneous spaces
Taco S Cohen, Mario Geiger, and Maurice Weiler · 2019
Cited alongside, same era.
Unsupervised cross-lingual representation learning at scale
Alexis Conneau, Kartikay Khandelwal, Naman Goyal, Vishrav Chaudhary, Guillaume Wenzek, Francisco Guzmán, Edouard Grave, Myle Ott, Luke Zettlemoyer, and Veselin Stoyanov · 2019
Cited alongside, same era.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova · 2019
Cited alongside, same era.
Invariance learning in deep neural networks with differentiable laplace approximations
Alexander Immer, Tycho van der Ouderaa, Gunnar Rätsch, Vincent Fortuin, and Mark van der Wilk · 2022
Later among the works it cites.
Dvc: Data version control - git for data & models, 2022
Ruslan Kuprieiev, skshetry, Dmitry Petrov, Paweł Redzyński, Peter Rowlands, Casper da Costa-Luis, Alexander Schepanovski, Ivan Shcheklein, Batuhan Taskaya, Gao, Jorge Orpinel, David de la Iglesia Castro, Fábio Santos, Aman Sharma, Dave Berenbaum, Zhanibek, Dani Hodovic, daniele, Nikita Kodenko, Andrew Grigorev, Earl, Nabanita Dash, George Vyshnya, Ronan Lamy, maykulkarni, Max Hora, Vera, and Sanidhya Mangal · 2022
Later among the works it cites.
Learning invariant weights in neural networks
Tycho FA van der Ouderaa and Mark van der Wilk · 2022
Later among the works it cites.
N24News: A new dataset for multimodal news classification
Zhen Wang, Xu Shan, Xiangxie Zhang, and Jie Yang · 2022
Later among the works it cites.
Bootstrapping parallel anchors for relative representations
Irene Cannistraci, Luca Moschella, Valentino Maiorca, Marco Fumero, Antonio Norelli, and Emanuele Rodolà · 2023
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Incorporating rotational invariance in convolutional neural network architecture
Haribabu Kandi, Ayushi Jain, Swetha Velluva Chathoth, Deepak Mishra, and Gorthi RK Sai Subrahmanyam · 2019
Cited alongside, same era.
Similarity of neural network representations revisited
Simon Kornblith, Mohammad Norouzi, Honglak Lee, and Geoffrey Hinton · 2019
Cited alongside, same era.
Albert: A lite bert for self-supervised learning of language representations
Zhenzhong Lan, Mingda Chen, Sebastian Goodman, Kevin Gimpel, Piyush Sharma, and Radu Soricut · 2019
Cited alongside, same era.
Deep set prediction networks
Yan Zhang, Jonathon Hare, and Adam Prugel-Bennett · 2019
Cited alongside, same era.
Learning invariances in neural networks from training data
Gregory Benton, Marc Finzi, Pavel Izmailov, and Andrew G Wilson · 2020
Cited alongside, same era.
ELECTRA: pre-training text encoders as discriminators rather than generators
Kevin Clark, Minh-Thang Luong, Quoc V. Le, and Christopher D. Manning · 2020
Cited alongside, same era.
On the benefits of invariance in neural networks
Clare Lyle, Mark van der Wilk, Marta Kwiatkowska, Yarin Gal, and Benjamin Bloem-Reddy · 2020
Cited alongside, same era.
Weakly supervised vision-and-language pre-training with relative representations
Chi Chen, Peng Li, Maosong Sun, and Yang Liu · 2023
Closest in time.
From charts to atlas: Merging latent spaces into one
Donato Crisostomi, Irene Cannistraci, Luca Moschella, Pietro Barbiero, Marco Ciccone, Pietro Lio, and Emanuele Rodolà · 2023
Closest in time.
Similarity of neural network models: A survey of functional and representational measures
Max Klabunde, Tobias Schumacher, Markus Strohmaier, and Florian Lemmerich · 2023
Closest in time.
On the direct alignment of latent spaces
Zorah Lähner and Michael Moeller · 2023
Closest in time.
Metric space magnitude for evaluating unsupervised representation learning
Katharina Limbeck, Rayna Andreeva, Rik Sarkar, and Bastian Rieck · 2023
Closest in time.
Latent space translation via semantic alignment
Valentino Maiorca, Luca Moschella, Antonio Norelli, Marco Fumero, Francesco Locatello, and Emanuele Rodolà · 2023
Closest in time.
Harmonics of learning: Universal fourier features emerge in invariant networks
Giovanni Luca Marchetti, Christopher Hillar, Danica Kragic, and Sophia Sanborn · 2023
Closest in time.
Relative representations enable zero-shot latent space communication
Luca Moschella, Valentino Maiorca, Marco Fumero, Antonio Norelli, Francesco Locatello, and Emanuele Rodolà · 2023
Closest in time.
Asif: Coupled data turns unimodal models to multimodal without training
Antonio Norelli, Marco Fumero, Valentino Maiorca, Luca Moschella, Emanuele Rodolà, and Francesco Locatello · 2023
Closest in time.
Deep neural networks with efficient guaranteed invariances
Matthias Rath and Alexandru Paul Condurache · 2023
Closest in time.
Zero-shot stitching in reinforcement learning using relative representations
Antonio Pio Ricciardi, Valentino Maiorca, Luca Moschella, and Emanuele Rodolà · 2023
Closest in time.
A general framework for robust g-invariance in g-equivariant networks
Sophia Sanborn and Nina Miolane · 2023
Closest in time.
Mapping the multiverse of latent representations
Jeremy Wayland, Corinna Coupette, and Bastian Rieck · 2024
Closest in time.