Fetching the paper…
Reading the bibliography…
The effectiveness of a language model is influenced by its token representations, which must encode contextual information and handle the same word form having a plurality of meanings (polysemy).
A semantic concordance
George A. Miller, Claudia Leacock, Ee Tengi, and Ross T. Bunker. 1993 · 1993
Earlier work this paper cites.
Wordnet: A lexical database for english
George A. Miller. 1995 · 1995
Earlier work this paper cites.
Recurrent neural network based language model
Tomas Mikolov, Martin Karafiát, Lukás Burget, Jan Cernocký, and Sanjeev Khudanpur. 2010 · 2010
Earlier work this paper cites.
FLOW: A first-language-oriented writing assistant system
Mei-Hua Chen, Shih-Ting Huang, Hung-Ting Hsieh, Ting-Hui Kao, and Jason S. Chang. 2012 · 2012
Earlier work this paper cites.
Improving word representations via global context and multiple word prototypes
Eric Huang, Richard Socher, Christopher Manning, and Andrew Ng. 2012 · 2012
Earlier work this paper cites.
Random walks for knowledge-based word sense disambiguation
Eneko Agirre, Oier López de Lacalle, and Aitor Soroa. 2014 · 2014
Earlier work this paper cites.
A unified model for word sense representation and disambiguation
Xinxiong Chen, Zhiyuan Liu, and Maosong Sun. 2014 · 2014
Earlier work this paper cites.
Entity linking meets word sense disambiguation: a unified approach
Andrea Moro, Alessandro Raganato, and Roberto Navigli. 2014 · 2014
Earlier work this paper cites.
Efficient non-parametric estimation of multiple embeddings per word in vector space
Arvind Neelakantan, Jeevan Shankar, Alexandre Passos, and Andrew McCallum. 2014 · 2014
Earlier work this paper cites.
Do multi-sense embeddings improve natural language understanding?
Jiwei Li and Dan Jurafsky. 2015 · 2015
Earlier work this paper cites.
AutoExtend: Extending word embeddings to embeddings for synsets and lexemes
Sascha Rothe and Hinrich Schütze. 2015 · 2015
Earlier work this paper cites.
context2vec: Learning generic context embedding with bidirectional LSTM
Oren Melamud, Jacob Goldberger, and Ido Dagan. 2016 · 2016
Cited alongside, same era.
Pointer Sentinel Mixture Models
Stephen Merity, Caiming Xiong, James Bradbury, and Richard Socher. 2016 · 2016
Cited alongside, same era.
Enriching word vectors with subword information
Piotr Bojanowski, Edouard Grave, Armand Joulin, and Tomas Mikolov. 2017 · 2017
Cited alongside, same era.
Inductive representation learning on large graphs
Will Hamilton, Zhitao Ying, and Jure Leskovec. 2017 · 2017
Cited alongside, same era.
Semi-Supervised Classification with Graph Convolutional Networks
Thomas N. Kipf and Max Welling. 2017 · 2017
Cited alongside, same era.
Neural sequence learning models for word sense disambiguation
Transformer-XL: Attentive language models beyond a fixed-length context
Zihang Dai, Zhilin Yang, Yiming Yang, Jaime Carbonell, Quoc Le, and Ruslan Salakhutdinov. 2019 · 2019
Later among the works it cites.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Later among the works it cites.
Fast graph representation learning with PyTorch Geometric
Matthias Fey and Jan E. Lenssen. 2019 · 2019
Later among the works it cites.
GlossBERT: BERT for word sense disambiguation with gloss knowledge
Luyao Huang, Chi Sun, Xipeng Qiu, and Xuanjing Huang. 2019 · 2019
Later among the works it cites.
LSTMEmbed: Learning word and sense representations from a large semantically annotated corpus with long short-term memories
Ignacio Iacobacci and Roberto Navigli. 2019 · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Alessandro Raganato, Claudio Delli Bovi, and Roberto Navigli. 2017 · 2017
Cited alongside, same era.
Attention is All you Need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Cited alongside, same era.
Universal language model fine-tuning for text classification
Jeremy Howard and Sebastian Ruder. 2018 · 2018
Cited alongside, same era.
Deep contextualized word representations
Matthew E. Peters, Mark Neumann, Mohit Iyyer, Matt Gardner, Christopher Clark, Kenton Lee, and Luke Zettlemoyer. 2018 · 2018
Cited alongside, same era.
Graph attention networks
Petar Veličković, Guillem Cucurull, Arantxa Casanova, Adriana Romero, Pietro Liò, and Yoshua Bengio. 2018 · 2018
Cited alongside, same era.
UFSAC: Unification of Sense Annotated Corpora and Tools
Loïc Vial, Benjamin Lecouteux, and Didier Schwab. 2018 · 2018
Cited alongside, same era.
Sensembert: Context-enhanced sense embeddings for multilingual word sense disambiguation
Bianca Scarlini, Tommaso Pasini, and Roberto Navigli. 2020a
Cited in the paper.
Sawan Kumar, Sharmistha Jat, Karan Saxena, and Partha Talukdar. 2019 · 2019
Later among the works it cites.
Barack’s wife hillary: Using knowledge graphs for fact-aware language modeling
Robert Logan, Nelson F. Liu, Matthew E. Peters, Matt Gardner, and Sameer Singh. 2019 · 2019
Later among the works it cites.
When is a bishop not like a rook? when it’s like a rabbi! multi-prototype BERT embeddings for estimating semantic relationships
Gabriella Chronis and Katrin Erk. 2020 · 2020
Closest in time.
SenseBERT: Driving some sense into BERT
Yoav Levine, Barak Lenz, Or Dagan, Ori Ram, Dan Padnos, Or Sharir, Shai Shalev-Shwartz, Amnon Shashua, and Yoav Shoham. 2020 · 2020
Closest in time.
With more contexts comes better performance: Contextualized sense embeddings for all-round word sense disambiguation
Bianca Scarlini, Tommaso Pasini, and Roberto Navigli. 2020b · 2020
Closest in time.