Generative pretraining from pixels
M. Chen, A. Radford, R. Child, J. Wu, H. Jun, D. Luan, and I. Sutskever · 2020
Later among the works it cites.
interpreting gpt: the logit lens, 2020
nostalgebraist · 2020
Later among the works it cites.
Improving neural language generation with spectrum control
L. Wang, J. Huang, K. Huang, Z. Hu, G. Wang, and Q. Gu · 2020
Later among the works it cites.
Transformers: State-of-the-art natural language processing
T. Wolf, L. Debut, V. Sanh, J. Chaumond, C. Delangue, A. Moi, P. Cistac, T. Rault, R. Louf, M. Funtowicz, J. Davison, S. Shleifer, P. von Platen, C. Ma, Y. Jernite, J. Plu, C. Xu, T. L. Scao, S. Gugger, M. Drame, Q. Lhoest, and A. M. Rush · 2020
Later among the works it cites.
Revisiting model stitching to compare neural representations
Y. Bansal, P. Nakkiran, and B. Barak · 2021
Later among the works it cites.
Similarity and matching of neural network representations
A. Csiszárik, P. Korösi-Szabó, Á. K. Matszangosz, G. Papp, and D. Varga · 2021
Later among the works it cites.
Knowledge neurons in pretrained transformers, 2021
Original
D. Dai, L. Dong, Y. Hao, Z. Sui, B. Chang, and F. Wei · 2021
Later among the works it cites.
A mathematical framework for transformer circuits, 2021
N. Elhage, N. Nanda, C. Olsson, T. Henighan, N. Joseph, B. Mann, A. Askell, Y. Bai, A. Chen, T. Conerly, N. DasSarma, D. Drain, D. Ganguli, Z. Hatfield-Dodds, D. Hernandez, A. Jones, J. Kernion, L. Lovitt, K. Ndousse, D. Amodei, T. Brown, J. Clark, J. Kaplan, S. McCandlish, and C. Olah · 2021
Later among the works it cites.
Isoscore: Measuring the uniformity of vector space utilization
Original
W. Rudman, N. Gillman, T. Rayne, and C. Eickhoff · 2021
Later among the works it cites.
Softmax linear units
N. Elhage, T. Hume, C. Olsson, N. Nanda, T. Henighan, S. Johnston, S. ElShowk, N. Joseph, N. DasSarma, B. Mann, D. Hernandez, A. Askell, K. Ndousse, A. Jones, D. Drain, A. Chen, Y. Bai, D. Ganguli, L. Lovitt, Z. Hatfield-Dodds, J. Kernion, T. Conerly, S. Kravec, S. Fort, S. Kadavath, J. Jacobson, E. Tran-Johnson, J. Kaplan, J. Clark, T. Brown, S. McCandlish, D. Amodei, and C. Olah · 2022
Closest in time.
How to dissect a muppet: The structure of transformer embedding spaces
Original
T. Mickus, D. Paperno, and M. Constant · 2022
Closest in time.
The multiBERTs: BERT reproductions for robustness analysis
T. Sellam, S. Yadlowsky, I. Tenney, J. Wei, N. Saphra, A. D’Amour, T. Linzen, J. Bastings, I. R. Turc, J. Eisenstein, D. Das, and E. Pavlick · 2022
Closest in time.
Opt: Open pre-trained transformer language models, 2022
Original
S. Zhang, S. Roller, N. Goyal, M. Artetxe, M. Chen, S. Chen, C. Dewan, M. Diab, X. Li, X. V. Lin, T. Mihaylov, M. Ott, S. Shleifer, K. Shuster, D. Simig, P. S. Koura, A. Sridhar, T. Wang, and L. Zettlemoyer · 2022
Closest in time.