Fetching the paper…
Reading the bibliography…
Advances in model editing through neuron pruning hold promise for removing undesirable concepts from large language models.
Distilbert, a distilled version of bert: smaller, faster, cheaper and lighter
Victor Sanh, Lysandre Debut, Julien Chaumond, and Thomas Wolf. 2019 · 1910
Earlier work this paper cites.
Optimal brain damage
Yann LeCun, John Denker, and Sara Solla. 1989 · 1989
Earlier work this paper cites.
An iterative pruning algorithm for feedforward neural networks
G. Castellano, A.M. Fanelli, and M. Pelillo. 1997 · 1997
Earlier work this paper cites.
Introduction to the CoNLL-2002 shared task: Language-independent named entity recognition
Erik F. Tjong Kim Sang. 2002 · 2002
Earlier work this paper cites.
Distributed representations of words and phrases and their compositionality
Tomás Mikolov, Ilya Sutskever, Kai Chen, Gregory S. Corrado, and Jeffrey Dean. 2013 · 2013
Earlier work this paper cites.
Investigating language universal and specific properties in word embeddings
Peng Qian, Xipeng Qiu, and Xuanjing Huang. 2016 · 2016
Earlier work this paper cites.
seqeval: A python framework for sequence labeling evaluation
Hiroki Nakayama. 2018 · 2018
Earlier work this paper cites.
Identifying and controlling important neurons in neural machine translation
Anthony Bau, Yonatan Belinkov, Hassan Sajjad, Nadir Durrani, Fahim Dalvi, and James Glass. 2019 · 2019
Cited alongside, same era.
Neurox: A toolkit for analyzing individual neurons in neural networks
Fahim Dalvi, Avery Nortonsmith, Anthony Bau, Yonatan Belinkov, Hassan Sajjad, Nadir Durrani, and James Glass. 2019 · 2019
Cited alongside, same era.
Decoupled weight decay regularization
Ilya Loshchilov and Frank Hutter. 2019 · 2019
Cited alongside, same era.
Studying the plasticity in deep convolutional neural networks using random pruning
Deepak Mittal, Shweta Bhardwaj, Mitesh M. Khapra, and Balaraman Ravindran. 2019 · 2019
Cited alongside, same era.
Language models are unsupervised multitask learners
Alec Radford, Jeffrey Wu, Rewon Child, David Luan, Dario Amodei, Ilya Sutskever, et al. 2019 · 2019
Cited alongside, same era.
What part of the neural network does this? understanding LSTMs by measuring and dissecting neurons
Analyzing individual neurons in pre-trained language models
Nadir Durrani, Hassan Sajjad, Fahim Dalvi, and Yonatan Belinkov. 2020 · 2020
Later among the works it cites.
Compositional explanations of neurons
Jesse Mu and Jacob Andreas. 2020 · 2020
Later among the works it cites.
Similarity analysis of contextual word representation models
John Wu, Yonatan Belinkov, Hassan Sajjad, Nadir Durrani, Fahim Dalvi, and James Glass. 2020 · 2020
Later among the works it cites.
On the pitfalls of analyzing individual neurons in language models
Omer Antverg and Yonatan Belinkov. 2022 · 2022
Later among the works it cites.
Discovering latent concepts learned in BERT
Fahim Dalvi, Abdul Khan, Firoj Alam, Nadir Durrani, Jia Xu, and Hassan Sajjad. 2022 · 2022
Later among the works it cites.
The unreasonable effectiveness of random pruning: Return of the most naive baseline for sparse training
Shiwei Liu, Tianlong Chen, Xiaohan Chen, Li Shen, Decebal Constantin Mocanu, Zhangyang Wang, and Mykola Pechenizkiy. 2022 · 2022
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Ji Xin, Jimmy Lin, and Yaoliang Yu. 2019 · 2019
Cited alongside, same era.
Analyzing redundancy in pretrained transformer models
Fahim Dalvi, Hassan Sajjad, Nadir Durrani, and Yonatan Belinkov. 2020 · 2020
Cited alongside, same era.
Later among the works it cites.