Fetching the paper…
Reading the bibliography…
During the last decade, Natural Language Processing has become, after Computer Vision, the second field of Artificial Intelligence that was massively changed by the advent of Deep Learning.
Fast graph representation learning with pytorch geometric
Matthias Fey and Jan Eric Lenssen · 1903
Earlier work this paper cites.
Visualizing attention in transformer-based language representation models
Jesse Vig · 1904
Earlier work this paper cites.
What does BERT look at? an analysis of bert’s attention
Kevin Clark, Urvashi Khandelwal, Omer Levy, and Christopher D. Manning · 1906
Earlier work this paper cites.
Roberta: A robustly optimized BERT pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov · 1907
Earlier work this paper cites.
A closer look at data bias in neural extractive summarization models
Ming Zhong, Danqing Wang, Pengfei Liu, Xipeng Qiu, and Xuanjing Huang · 1909
Earlier work this paper cites.
Distilbert, a distilled version of BERT: smaller, faster, cheaper and lighter
Victor Sanh, Lysandre Debut, Julien Chaumond, and Thomas Wolf · 1910
Earlier work this paper cites.
Huggingface transformers: State-of-the-art natural language processing
Thomas Wolf, Lysandre Debut, Victor Sanh, Julien Chaumond, Clement Delangue, Anthony Moi, Pierric Cistac, Tim Rault, Rémi Louf, Morgan Funtowicz, and Jamie Brew · 1910
Earlier work this paper cites.
How can we fool LIME and shap? adversarial attacks on post hoc explanation methods
Dylan Slack, Sophie Hilgard, Emily Jia, Sameer Singh, and Himabindu Lakkaraju · 1911
Earlier work this paper cites.
Fractal Geometry of Nature
Benoit Mandelbrot · 1977
Earlier work this paper cites.
Self-Organizing Maps , volume 30 of Springer Series in Information Sciences
Teuvo Kohonen · 1995
Earlier work this paper cites.
A language modeling approach to information retrieval
Jay M. Ponte and W. Bruce Croft · 1998
Earlier work this paper cites.
Unifying graph convolutional neural networks and label propagation
Hongwei Wang and Jure Leskovec · 2002
Earlier work this paper cites.
Youwei Song, Jiahai Wang, Zhiwei Liang, Zhiyue Liu, and Tao Jiang · 2002
Earlier work this paper cites.
On the impressive performance of randomly weighted encoders in summarization tasks
Jonathan Pilault, Jaehong Park, and Christopher J. Pal · 2002
Earlier work this paper cites.
A survey of deep learning for scientific discovery
Maithra Raghu and Eric Schmidt · 2003
Earlier work this paper cites.
An introduction to variable and feature selection
Isabelle Guyon and André Elisseeff · 2003
Earlier work this paper cites.
Adv-bert: BERT is not robust on misspellings! generating nature adversarial samples on BERT
Lichao Sun, Kazuma Hashimoto, Wenpeng Yin, Akari Asai, Jia Li, Philip S. Yu, and Caiming Xiong · 2003
Earlier work this paper cites.
Information-theoretic probing with minimum description length
Elena Voita and Ivan Titov · 2003
Earlier work this paper cites.
Self-attention attribution: Interpreting information interactions inside transformer
Yaru Hao, Li Dong, Furu Wei, and Ke Xu · 2004
Earlier work this paper cites.
Neural additive models: Interpretable machine learning with neural nets
Rishabh Agarwal, Nicholas Frosst, Xuezhou Zhang, Rich Caruana, and Geoffrey E. Hinton · 2004
Earlier work this paper cites.
Causal mediation analysis for interpreting neural NLP: the case of gender bias
Jesse Vig, Sebastian Gehrmann, Yonatan Belinkov, Sharon Qian, Daniel Nevo, Yaron Singer, and Stuart M. Shieber · 2004
Earlier work this paper cites.
Attention module is not only a weight: Analyzing transformers with vector norms
Goro Kobayashi, Tatsuki Kuribayashi, Sho Yokoi, and Kentaro Inui · 2004
Earlier work this paper cites.
The Grammar of Graphics, Second Edition
Leland Wilkinson · 2005
Earlier work this paper cites.
Identifying necessary elements for bert’s multilinguality
Philipp Dufter and Hinrich Schütze · 2005
Earlier work this paper cites.
A tale of a probe and a parser
Rowan Hall Maudslay, Josef Valvoda, Tiago Pimentel, Adina Williams, and Ryan Cotterell · 2005
Earlier work this paper cites.
Attviz: Online exploration of self-attention for transparent neural language modeling
Blaz Skrlj, Nika Erzen, Shane Sheehan, Saturnino Luz, Marko Robnik-Sikonja, and Senja Pollak · 2005
Earlier work this paper cites.
Hotflip: White-box adversarial examples for text classification
Javid Ebrahimi, Anyi Rao, Daniel Lowd, and Dejing Dou · 2006
Earlier work this paper cites.
Badnl: Backdoor attacks against NLP models
Xiaoyi Chen, Ahmed Salem, Michael Backes, Shiqing Ma, and Yang Zhang · 2006
Earlier work this paper cites.
Bertology meets biology: Interpreting attention in protein language models
Jesse Vig, Ali Madani, Lav R. Varshney, Caiming Xiong, Richard Socher, and Nazneen Fatema Rajani · 2006
Earlier work this paper cites.
Openattack: An open-source textual adversarial attack toolkit
Guoyang Zeng, Fanchao Qi, Qianrui Zhou, Tingji Zhang, Bairu Hou, Yuan Zang, Zhiyuan Liu, and Maosong Sun · 2009
Earlier work this paper cites.
D 3 data-driven documents
Michael Bostock, Vadim Ogievetsky, and Jeffrey Heer · 2011
Earlier work this paper cites.
Extracting training data from large language models
Nicholas Carlini, Florian Tramèr, Eric Wallace, Matthew Jagielski, Ariel Herbert-Voss, Katherine Lee, Adam Roberts, Tom B. Brown, Dawn Song, Úlfar Erlingsson, Alina Oprea, and Colin Raffel · 2012
Earlier work this paper cites.
Exploring and visualizing variation in language resources
Peter Fankhauser, Jörg Knappen, and Elke Teich · 2014
Earlier work this paper cites.
Parallel coordinates for multidimensional data visualization: Basic concepts
Julian Heinrich and Daniel Weiskopf · 2015
Earlier work this paper cites.
Visualizing and understanding recurrent networks
Andrej Karpathy, Justin Johnson, and Fei-Fei Li · 2015
Earlier work this paper cites.
Vega-lite: A grammar of interactive graphics
Arvind Satyanarayan, Dominik Moritz, Kanit Wongsuphasawat, and Jeffrey Heer · 2016
Earlier work this paper cites.
"why should I trust you?": Explaining the predictions of any classifier
Marco Túlio Ribeiro, Sameer Singh, and Carlos Guestrin · 2016
Cited alongside, same era.
Visualizing and understanding neural models in NLP
Jiwei Li, Xinlei Chen, Eduard H. Hovy, and Dan Jurafsky · 2016
Cited alongside, same era.
Probing for semantic evidence of composition by means of simple classification tasks
Allyson Ettinger, Ahmed Elgohary, and Philip Resnik · 2016
Cited alongside, same era.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N. Gomez, Lukasz Kaiser, and Illia Polosukhin · 2017
Cited alongside, same era.
Conceptvector: Text visual analytics via interactive lexicon building using word embedding
Deokgun Park, Seungyeon Kim, Jurim Lee, Jaegul Choo, Nicholas Diakopoulos, and Niklas Elmqvist · 2017
Cited alongside, same era.
A unified approach to interpreting model predictions
BERT rediscovers the classical NLP pipeline
Ian Tenney, Dipanjan Das, and Ellie Pavlick · 2019
Later among the works it cites.
Visualizing and measuring the geometry of BERT
Emily Reif, Ann Yuan, Martin Wattenberg, Fernanda B. Viégas, Andy Coenen, Adam Pearce, and Been Kim · 2019
Later among the works it cites.
A structural probe for finding syntax in word representations
John Hewitt and Christopher D. Manning · 2019
Later among the works it cites.
Opennre: An open and extensible toolkit for neural relation extraction
Xu Han, Tianyu Gao, Yuan Yao, Deming Ye, Zhiyuan Liu, and Maosong Sun · 2019
Later among the works it cites.
Seeing things from a different angle: Discovering diverse perspectives about claims
Sihao Chen, Daniel Khashabi, Wenpeng Yin, Chris Callison-Burch, and Dan Roth · 2019
Later among the works it cites.
Entity, relation, and event extraction with contextualized span representations
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Scott M. Lundberg and Su-In Lee · 2017
Cited alongside, same era.
Scattertext: a browser-based tool for visualizing how corpora differ
Jason S. Kessler · 2017
Cited alongside, same era.
Lstmvis: A tool for visual analysis of hidden state dynamics in recurrent neural networks
Hendrik Strobelt, Sebastian Gehrmann, Hanspeter Pfister, and Alexander M. Rush · 2017
Cited alongside, same era.
Understanding hidden memories of recurrent neural networks
Yao Ming, Shaozu Cao, Ruixiang Zhang, Zhen Li, Yuanzhe Chen, Yangqiu Song, and Huamin Qu · 2017
Cited alongside, same era.
Activis: Visual exploration of industry-scale deep neural network models
Minsuk Kahng, Pierre Y. Andrews, Aditya Kalro, and Duen Horng (Polo) Chau · 2017
Cited alongside, same era.
Ray: A distributed framework for emerging AI applications
Philipp Moritz, Robert Nishihara, Stephanie Wang, Alexey Tumanov, Richard Liaw, Eric Liang, Melih Elibol, Zongheng Yang, William Paul, Michael I. Jordan, and Ion Stoica · 2018
Cited alongside, same era.
Tune: A research platform for distributed model selection and training
Richard Liaw, Eric Liang, Robert Nishihara, Philipp Moritz, Joseph E. Gonzalez, and Ion Stoica · 2018
Cited alongside, same era.
David Wadden, Ulme Wennberg, Yi Luan, and Hannaneh Hajishirzi · 2019
Later among the works it cites.
Fewrel 2.0: Towards more challenging few-shot relation classification
Tianyu Gao, Xu Han, Hao Zhu, Zhiyuan Liu, Peng Li, Maosong Sun, and Jie Zhou · 2019
Later among the works it cites.
Lipstick on a pig: Debiasing methods cover up systematic gender biases in word embeddings but do not remove them
Hila Gonen and Yoav Goldberg · 2019
Later among the works it cites.
Language models are few-shot learners
Tom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, Sandhini Agarwal, Ariel Herbert-Voss, Gretchen Krueger, Tom Henighan, Rewon Child, Aditya Ramesh, Daniel M. Ziegler, Jeffrey Wu, Clemens Winter, Christopher Hesse, Mark Chen, Eric Sigler, Mateusz Litwin, Scott Gray, Benjamin Chess, Jack Clark, Christopher Berner, Sam McCandlish, Alec Radford, Ilya Sutskever, and Dario Amodei · 2020
Later among the works it cites.
ALBERT: A lite BERT for self-supervised learning of language representations
Zhenzhong Lan, Mingda Chen, Sebastian Goodman, Kevin Gimpel, Piyush Sharma, and Radu Soricut · 2020
Later among the works it cites.
Reformer: The efficient transformer
Nikita Kitaev, Lukasz Kaiser, and Anselm Levskaya · 2020
Later among the works it cites.
Visualizing transformers for nlp: a brief survey
Adrian MP Braşoveanu and Răzvan Andonie · 2020
Later among the works it cites.
ex bert: A visual analysis tool to explore learned representations in transformer models
Benjamin Hoover, Hendrik Strobelt, and Sebastian Gehrmann · 2020
Later among the works it cites.
Do neural language models overcome reporting bias?
Vered Shwartz and Yejin Choi · 2020
Later among the works it cites.
The importance of interpretability and visualization in machine learning for applications in medicine and health care
Alfredo Vellido · 2020
Later among the works it cites.
Fooling LIME and SHAP: adversarial attacks on post hoc explanation methods
Dylan Slack, Sophie Hilgard, Emily Jia, Sameer Singh, and Himabindu Lakkaraju · 2020
Later among the works it cites.
Is BERT really robust? A strong baseline for natural language attack on text classification and entailment
Di Jin, Zhijing Jin, Joey Tianyi Zhou, and Peter Szolovits · 2020
Later among the works it cites.
Textattack: A framework for adversarial attacks, data augmentation, and adversarial training in NLP
John X. Morris, Eli Lifland, Jin Yong Yoo, Jake Grigsby, Di Jin, and Yanjun Qi · 2020
Later among the works it cites.
Gradient-based analysis of NLP models is manipulable
Junlin Wang, Jens Tuyls, Eric Wallace, and Sameer Singh · 2020
Later among the works it cites.
Every document owns its structure: Inductive text classification via graph neural networks
Yufeng Zhang, Xueli Yu, Zeyu Cui, Shu Wu, Zhongzhen Wen, and Liang Wang · 2020
Later among the works it cites.
Unsupervised domain adaptive graph convolutional networks
Man Wu, Shirui Pan, Chuan Zhou, Xiaojun Chang, and Xingquan Zhu · 2020
Later among the works it cites.
Quantifying attention flow in transformers
Samira Abnar and Willem H. Zuidema · 2020
Later among the works it cites.
Attention flows: Analyzing and comparing attention mechanisms in language models
Joseph F. DeRose, Jiayao Wang, and Matthew Berger · 2020
Later among the works it cites.
How to probe sentence embeddings in low-resource languages: On structural design choices for probing task evaluation
Steffen Eger, Johannes Daxenberger, and Iryna Gurevych · 2020
Later among the works it cites.
Syntaxgym: An online platform for targeted evaluation of language models
Jon Gauthier, Jennifer Hu, Ethan Wilcox, Peng Qian, and Roger Levy · 2020
Later among the works it cites.
VL-BERT: pre-training of generic visual-linguistic representations
Weijie Su, Xizhou Zhu, Yue Cao, Bin Li, Lewei Lu, Furu Wei, and Jifeng Dai · 2020
Later among the works it cites.
Information-theoretic probing with minimum description length
Elena Voita and Ivan Titov · 2020
Later among the works it cites.
Visbert: Hidden-state visualizations for transformers
Betty van Aken, Benjamin Winter, Alexander Löser, and Felix A. Gers · 2020
Later among the works it cites.
A comparison of architectures and pretraining methods for contextualized multilingual word embeddings
Niels van der Heijden, Samira Abnar, and Ekaterina Shutova · 2020
Later among the works it cites.
Large-scale adversarial training for vision-and-language representation learning
Zhe Gan, Yen-Chun Chen, Linjie Li, Chen Zhu, Yu Cheng, and Jingjing Liu · 2020
Later among the works it cites.
Behind the scene: Revealing the secrets of pre-trained vision-and-language models
Jize Cao, Zhe Gan, Yu Cheng, Licheng Yu, Yen-Chun Chen, and Jingjing Liu · 2020
Later among the works it cites.
A comparison of pre-trained vision-and-language models for multimodal representation learning across medical images and reports
Yikuan Li, Hanyin Wang, and Yuan Luo · 2020
Later among the works it cites.
Adversarial training for aspect-based sentiment analysis with BERT
Akbar Karimi, Leonardo Rossi, and Andrea Prati · 2021
Later among the works it cites.
Zeyu Yun, Yubei Chen, Bruno A. Olshausen, and Yann LeCun · 2021
Later among the works it cites.
Kai Han, An Xiao, Enhua Wu, Jianyuan Guo, Chunjing Xu, and Yunhe Wang · 2021
Later among the works it cites.
Vilt: Vision - and - language transformer without convolution or region supervision
Wonjae Kim, Bokyung Son, and Ildoo Kim · 2021
Later among the works it cites.