Ablation of a Robot’s Brain: Neural Networks Under a Knife
Original
Peter E. Lillian, Richard Meyes, and Tobias Meisen · 2018
Later among the works it cites.
The Building Blocks of Interpretability
Chris Olah, Arvind Satyanarayan, Ian Johnson, Shan Carter, Ludwig Schubert, Katherine Ye, and Alexander Mordvintsev · 2018
Later among the works it cites.
Causal importance of orientation selectivity for generalization in image recognition
Jumpei Ukita · 2018
Later among the works it cites.
Interpreting Deep Visual Representations via Network Dissection
B. Zhou, D. Bau, A. Oliva, and A. Torralba · 2018
Later among the works it cites.
Revisiting the Importance of Individual Units in CNNs via Ablation
Original
Bolei Zhou, Yiyou Sun, David Bau, and Antonio Torralba · 2018
Later among the works it cites.
The role of untuned neurons in sensory information coding
Joel Zylberberg · 2018
Later among the works it cites.
What Is One Grain of Sand in the Desert? Analyzing Individual Neurons in Deep NLP Models
Fahim Dalvi, Nadir Durrani, Hassan Sajjad, Yonatan Belinkov, Anthony Bau, and James Glass · 2019
Later among the works it cites.
How Important is a Neuron
Kedar Dhamdhere, Mukund Sundararajan, and Qiqi Yan · 2019
Later among the works it cites.
On Interpretability and Feature Representations: An Analysis of the Sentiment Neuron
Jonathan Donnelly and Adam Roegiest · 2019
Later among the works it cites.
Selectivity metrics provide misleading estimates of the selectivity of single units in neural networks
Ella Gale, Ryan Blything, Nicholas Martin, Jeffrey S. Bowers, and Anh Nguyen · 2019
Later among the works it cites.
Oscillatory recurrent gated neural integrator circuits (organics), a unifying theoretical framework for neural dynamics
David J. Heeger and Wayne E. Mackey · 2019
Later among the works it cites.
A Benchmark for Interpretability Methods in Deep Neural Networks
Sara Hooker, Dumitru Erhan, Pieter-Jan Kindermans, and Been Kim · 2019
Later among the works it cites.
Trivializations for gradient-based optimization on manifolds
Mario Lezcano-Casado · 2019
Later among the works it cites.
Discovery of Natural Language Concepts in Individual Units of CNNs
Seil Na, Yo Joong Choe, Dong-Hyun Lee, and Gunhee Kim · 2019
Later among the works it cites.
PyTorch: An Imperative Style, High-Performance Deep Learning Library
Adam Paszke, Sam Gross, Francisco Massa, Adam Lerer, James Bradbury, Gregory Chanan, Trevor Killeen, Zeming Lin, Natalia Gimelshein, Luca Antiga, Alban Desmaison, Andreas Kopf, Edward Yang, Zachary DeVito, Martin Raison, Alykhan Tejani, Sasank Chilamkurthy, Benoit Steiner, Lu Fang, Junjie Bai, and Soumith Chintala · 2019
Later among the works it cites.
The language of the brain: real-world neural population codes
J Andrew Pruszynski and Joel Zylberberg · 2019
Later among the works it cites.
Understanding trained CNNs by indexing neuron selectivity
Original
Ivet Rafegas, Maria Vanrell, Luis A. Alexandre, and Guillem Arias · 2019
Later among the works it cites.
Towards the neural population doctrine
Shreya Saxena and John P Cunningham · 2019
Later among the works it cites.
Understanding the role of individual units in a deep neural network
David Bau, Jun-Yan Zhu, Hendrik Strobelt, Agata Lapedriza, Bolei Zhou, and Antonio Torralba · 2020
Closest in time.
Zoom In: An Introduction to Circuits
Chris Olah, Nick Cammarata, Ludwig Schubert, Gabriel Goh, Michael Petrov, and Shan Carter · 2020
Closest in time.
Cortical population activity within a preserved neural manifold underlies multiple motor behaviors
Juan A. Gallego, Matthew G. Perich, Stephanie N. Naufel, Christian Ethier, Sara A. Solla, and Lee E. Miller · 2041
Closest in time.
Spike-timing-dependent ensemble encoding by non-classically responsive cortical neurons
Michele N Insanally, Ioana Carcea, Rachel E Field, Chris C Rodgers, Brian DePasquale, Kanaka Rajan, Michael R DeWeese, Badr F Albanna, and Robert C Froemke · 2050
Closest in time.