Fetching the paper…
Reading the bibliography…
Word vector specialisation (also known as retrofitting) is a portable, light-weight approach to fine-tuning arbitrary distributional word vector spaces by injecting external knowledge from rich lexical resources such as WordNet.
Distributional structure
Zellig S. Harris. 1954 · 1954
Earlier work this paper cites.
Long Short-Term Memory
Sepp Hochreiter and Jürgen Schmidhuber. 1997 · 1997
Earlier work this paper cites.
WordNet
Christiane Fellbaum. 1998 · 1998
Earlier work this paper cites.
Placing search in context: The concept revisited
Lev Finkelstein, Evgeniy Gabrilovich, Yossi Matias, Ehud Rivlin, Zach Solan, Gadi Wolfman, and Eytan Ruppin. 2002 · 2002
Earlier work this paper cites.
Roget’s 21st Century Thesaurus (3rd Edition)
Barbara Ann Kipfer. 2009 · 2009
Earlier work this paper cites.
Understanding the difficulty of training deep feedforward neural networks
Xavier Glorot and Yoshua Bengio. 2010 · 2010
Earlier work this paper cites.
Rectified linear units improve restricted Boltzmann machines
Vinod Nair and Geoffrey E. Hinton. 2010 · 2010
Earlier work this paper cites.
Cognitive User Interfaces
Steve Young. 2010 · 2010
Earlier work this paper cites.
Natural language processing (almost) from scratch
Ronan Collobert, Jason Weston, Léon Bottou, Michael Karlen, Koray Kavukcuoglu, and Pavel P. Kuksa. 2011 · 2011
Earlier work this paper cites.
Adaptive subgradient methods for online learning and stochastic optimization
John C. Duchi, Elad Hazan, and Yoram Singer. 2011 · 2011
Earlier work this paper cites.
WSABIE: Scaling up to large vocabulary image annotation
Jason Weston, Samy Bengio, and Nicolas Usunier. 2011 · 2011
Earlier work this paper cites.
BabelNet: The automatic construction, evaluation and application of a wide-coverage multilingual semantic network
Roberto Navigli and Simone Paolo Ponzetto. 2012 · 2012
Earlier work this paper cites.
Polyglot: Distributed word representations for multilingual NLP
Rami Al-Rfou, Bryan Perozzi, and Steven Skiena. 2013 · 2013
Earlier work this paper cites.
DeViSE: A deep visual-semantic embedding model
Andrea Frome, Gregory S. Corrado, Jonathon Shlens, Samy Bengio, Jeffrey Dean, Marc’Aurelio Ranzato, and Tomas Mikolov. 2013 · 2013
Earlier work this paper cites.
PPDB: The Paraphrase Database
Juri Ganitkevitch, Benjamin Van Durme, and Chris Callison-Burch. 2013 · 2013
Earlier work this paper cites.
Rectifier nonlinearities improve neural network acoustic models
Andrew L. Maas, Awni Y. Hannun, and Andrew Y. Ng. 2013 · 2013
Earlier work this paper cites.
Tailoring continuous word representations for dependency parsing
Mohit Bansal, Kevin Gimpel, and Karen Livescu. 2014 · 2014
Earlier work this paper cites.
Knowledge-powered deep learning for word embedding
Jiang Bian, Bin Gao, and Tie-Yan Liu. 2014 · 2014
Earlier work this paper cites.
Multimodal distributional semantics
Elia Bruni, Nam-Khanh Tran, and Marco Baroni. 2014 · 2014
Earlier work this paper cites.
A fast and accurate dependency parser using neural networks
Danqi Chen and Christopher D. Manning. 2014 · 2014
Earlier work this paper cites.
The Second Dialog State Tracking Challenge
Matthew Henderson, Blaise Thomson, and Jason D. Wiliams. 2014 · 2014
Earlier work this paper cites.
Learning a lexical simplifier using wikipedia
Colby Horn, Cathryn Manduca, and David Kauchak. 2014 · 2014
Earlier work this paper cites.
Dependency-based word embeddings
Omer Levy and Yoav Goldberg. 2014 · 2014
Earlier work this paper cites.
Glove: Global vectors for word representation
Jeffrey Pennington, Richard Socher, and Christopher Manning. 2014 · 2014
Earlier work this paper cites.
RC-NET: A general framework for incorporating knowledge into word representations
Chang Xu, Yalong Bai, Jiang Bian, Bin Gao, Gang Wang, Xiaoguang Liu, and Tie-Yan Liu. 2014 · 2014
Cited alongside, same era.
Improving lexical embeddings with semantic knowledge
Mo Yu and Mark Dredze. 2014 · 2014
Cited alongside, same era.
Word semantic representations using bayesian probabilistic tensor factorization
Jingwei Zhang, Jeremy Salwen, Michael Glass, and Alfio Gliozzo. 2014 · 2014
Cited alongside, same era.
Improving zero-shot learning by mitigating the hubness problem
Georgiana Dinu, Angeliki Lazaridou, and Marco Baroni. 2015 · 2015
Cited alongside, same era.
Retrofitting word vectors to semantic lexicons
Manaal Faruqui, Jesse Dodge, Sujay Kumar Jauhar, Chris Dyer, Eduard Hovy, and Noah A. Smith. 2015 · 2015
Cited alongside, same era.
Intent detection using semantically enriched word embeddings
Joo-Kyung Kim, Gokhan Tur, Asli Celikyilmaz, Bin Cao, and Ye-Yi Wang. 2016 · 2016
Later among the works it cites.
Sihan Li, Jiantao Jiao, Yanjun Han, and Tsachy Weissman. 2016 · 2016
Later among the works it cites.
Dmytro Mishkin and Jiri Matas. 2016 · 2016
Later among the works it cites.
Counter-fitting word vectors to linguistic constraints
Nikola Mrkšić, Diarmuid Ó Séaghdha, Blaise Thomson, Milica Gašić, Lina Maria Rojas-Barahona, Pei-Hao Su, David Vandyke, Tsung-Hsien Wen, and Steve Young. 2016 · 2016
Later among the works it cites.
Integrating distributional lexical contrast into word embeddings for antonym-synonym distinction
Kim Anh Nguyen, Sabine Schulte im Walde, and Ngoc Thang Vu. 2016 · 2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Goran Glavaš and Sanja Štajner. 2015 · 2015
Cited alongside, same era.
Delving deep into rectifiers: Surpassing human-level performance on ImageNet classification
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun. 2015 · 2015
Cited alongside, same era.
SimLex-999: Evaluating semantic models with (genuine) similarity estimation
Felix Hill, Roi Reichart, and Anna Korhonen. 2015 · 2015
Cited alongside, same era.
Adam: A method for stochastic optimization
Diederik P. Kingma and Jimmy Ba. 2015 · 2015
Cited alongside, same era.
An empirical analysis of optimization for max-margin NLP
Jonathan K. Kummerfeld, Taylor Berg-Kirkpatrick, and Dan Klein. 2015 · 2015
Cited alongside, same era.
Hubness and pollution: Delving into cross-space mapping for zero-shot learning
Angeliki Lazaridou, Georgiana Dinu, and Marco Baroni. 2015 · 2015
Cited alongside, same era.
Separated by an un-common language: Towards judgment language informed vector space modeling
Ira Leviant and Roi Reichart. 2015 · 2015
Cited alongside, same era.
Encoding prior knowledge with eigenword embeddings
Dominique Osborne, Shashi Narayan, and Shay Cohen. 2016 · 2016
Later among the works it cites.
Continuously learning neural dialogue management
Pei-Hao Su, Milica Gašić, Nikola Mrkšić, Lina Rojas-Barahona, Stefan Ultes, David Vandyke, Tsung-Hsien Wen, and Steve Young. 2016 · 2016
Later among the works it cites.
The Dialog State Tracking Challenge series: A review
Jason D. Williams, Antoine Raux, and Matthew Henderson. 2016 · 2016
Later among the works it cites.
Learning bilingual word embeddings with (almost) no bilingual data
Mikel Artetxe, Gorka Labaka, and Eneko Agirre. 2017 · 2017
Later among the works it cites.
Enriching word vectors with subword information
Piotr Bojanowski, Edouard Grave, Armand Joulin, and Tomas Mikolov. 2017 · 2017
Later among the works it cites.
Word translation without parallel data
Alexis Conneau, Guillaume Lample, Marc’Aurelio Ranzato, Ludovic Denoyer, and Hervé Jégou. 2017 · 2017
Later among the works it cites.
Sigmoid-weighted linear units for neural network function approximation in reinforcement learning
Stefan Elfwing, Eiji Uchibe, and Kenji Doya. 2017 · 2017
Later among the works it cites.
Dual tensor model for detecting asymmetric lexico-semantic relations
Goran Glavaš and Simone Paolo Ponzetto. 2017 · 2017
Later among the works it cites.
Self-normalizing neural networks
Günter Klambauer, Thomas Unterthiner, Andreas Mayr, and Sepp Hochreiter. 2017 · 2017
Later among the works it cites.
Neural belief tracker: Data-driven dialogue state tracking
Nikola Mrkšić, Diarmuid Ó Séaghdha, Tsung-Hsien Wen, Blaise Thomson, and Steve Young. 2017 · 2017
Later among the works it cites.
Semantic specialisation of distributional word vector spaces using monolingual and cross-lingual constraints
Nikola Mrkšić, Ivan Vulić, Diarmuid Ó Séaghdha, Ira Leviant, Roi Reichart, Milica Gašić, Anna Korhonen, and Steve Young. 2017 · 2017
Later among the works it cites.
Hierarchical embeddings for hypernymy detection and directionality
Kim Anh Nguyen, Maximilian Köper, Sabine Schulte im Walde, and Ngoc Thang Vu. 2017 · 2017
Later among the works it cites.
Poincaré embeddings for learning hierarchical representations
Maximilian Nickel and Douwe Kiela. 2017 · 2017
Later among the works it cites.
Searching for activation functions
Prajit Ramachandran, Barret Zoph, and Quoc V. Le. 2017 · 2017
Later among the works it cites.
A survey of cross-lingual embedding models
Sebastian Ruder, Ivan Vulić, and Anders Søgaard. 2017 · 2017
Later among the works it cites.
A network-based end-to-end trainable task-oriented dialogue system
Tsung-Hsien Wen, David Vandyke, Nikola Mrkšić, Milica Gašić, Lina M. Rojas-Barahona, Pei-Hao Su, Stefan Ultes, and Steve Young. 2017 · 2017
Later among the works it cites.
Specialising word vectors for lexical entailment
Ivan Vulić and Nikola Mrkšić. 2018 · 2018
Closest in time.
Specializing word embeddings for similarity or relatedness
Douwe Kiela, Felix Hill, and Stephen Clark. 2015 · 2048
Closest in time.