Fetching the paper…
Reading the bibliography…
Multi-view learning can provide self-supervision when different views are available of the same data.
Praktische verfahren der gleichungsauflösung
R. Mises and H. Pollaczek-Geiringer · 1929
Earlier work this paper cites.
Distributional structure
Z. S. Harris · 1954
Earlier work this paper cites.
A synopsis of linguistic theory
J. R. Firth · 1957
Earlier work this paper cites.
Learning classification with unlabeled data
V. R. de Sa · 1993
Earlier work this paper cites.
Learning question classifiers
X. Li and D. Roth · 2002
Earlier work this paper cites.
Unsupervised construction of large paraphrase corpora: Exploiting massively parallel news sources
W. B. Dolan, C. Quirk, and C. Brockett · 2004
Earlier work this paper cites.
Mining and summarizing customer reviews
M. Hu and B. Liu · 2004
Earlier work this paper cites.
A sentimental education: Sentiment analysis using subjectivity summarization based on minimum cuts
B. Pang and L. Lee · 2004
Earlier work this paper cites.
Right hemisphere sensitivity to word- and sentence-level context: evidence from event-related brain potentials
S. Coulson, K. D. Federmeier, C. van Petten, and M. Kutas · 2005
Earlier work this paper cites.
Seeing stars: Exploiting class relationships for sentiment categorization with respect to rating scales
B. Pang and L. Lee · 2005
Earlier work this paper cites.
Annotating expressions of opinions and emotions in language
J. Wiebe, T. Wilson, and C. Cardie · 2005
Earlier work this paper cites.
A special role for the right hemisphere in metaphor comprehension? erp evidence from hemifield presentation
S. Coulson and C. van Petten · 2007
Earlier work this paper cites.
From frequency to meaning: Vector space models of semantics
P. D. Turney and P. Pantel · 2010
Earlier work this paper cites.
Semeval-2012 task 6: A pilot on semantic textual similarity
E. Agirre, D. M. Cer, M. T. Diab, and A. Gonzalez-Agirre · 2012
Earlier work this paper cites.
Laterality functional asymmetry in the intact brain
M. Bryden · 2012
Earlier work this paper cites.
*sem 2013 shared task: Semantic textual similarity
E. Agirre, D. M. Cer, M. T. Diab, A. Gonzalez-Agirre, and W. Guo · 2013
Earlier work this paper cites.
Ppdb: The paraphrase database
J. Ganitkevitch, B. V. Durme, and C. Callison-Burch · 2013
Earlier work this paper cites.
Umbc_ebiquity-core: semantic textual similarity systems
L. Han, A. L. Kashyap, T. Finin, J. Mayfield, and J. Weese · 2013
Earlier work this paper cites.
Distributed representations of words and phrases and their compositionality
T. Mikolov, I. Sutskever, K. Chen, G. S. Corrado, and J. Dean · 2013
Earlier work this paper cites.
On the difficulty of training recurrent neural networks
R. Pascanu, T. Mikolov, and Y. Bengio · 2013
Cited alongside, same era.
Recursive deep models for semantic compositionality over a sentiment treebank
R. Socher, A. Perelygin, J. Wu, J. Chuang, C. D. Manning, A. Ng, and C. Potts · 2013
Cited alongside, same era.
Semeval-2014 task 10: Multilingual semantic textual similarity
E. Agirre, C. Banea, C. Cardie, D. M. Cer, M. T. Diab, A. Gonzalez-Agirre, W. Guo, R. Mihalcea, G. Rigau, and J. Wiebe · 2014
Cited alongside, same era.
Empirical evaluation of gated recurrent neural networks on sequence modeling
J. Chung, C. Gulcehre, K. Cho, and Y. Bengio · 2014
Cited alongside, same era.
Adam: A method for stochastic optimization
D. Kingma and J. Ba · 2014
Cited alongside, same era.
Semeval-2016 task 1: Semantic textual similarity, monolingual and cross-lingual evaluation
E. Agirre, C. Banea, D. M. Cer, M. T. Diab, A. Gonzalez-Agirre, R. Mihalcea, G. Rigau, and J. Wiebe · 2016
Later among the works it cites.
A latent variable model approach to pmi-based word embeddings
S. Arora, Y. Li, Y. Liang, T. Ma, and A. Risteski · 2016
Later among the works it cites.
J. Ba, R. Kiros, and G. E. Hinton · 2016
Later among the works it cites.
Learning distributed representations of sentences from unlabelled data
F. Hill, K. Cho, and A. Korhonen · 2016
Later among the works it cites.
Siamese cbow: Optimizing word embeddings for sentence representations
T. Kenter, A. Borisov, and M. de Rijke · 2016
Later among the works it cites.
Hierarchical attention networks for document classification
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Distributed representations of sentences and documents
Q. V. Le and T. Mikolov · 2014
Cited alongside, same era.
A model of coherence based on distributed sentence representation
J. Li and E. H. Hovy · 2014
Cited alongside, same era.
A sick cure for the evaluation of compositional distributional semantic models
M. Marelli, S. Menini, M. Baroni, L. Bentivogli, R. Bernardi, and R. Zamparelli · 2014
Cited alongside, same era.
Glove: Global vectors for word representation
J. Pennington, R. Socher, and C. D. Manning · 2014
Cited alongside, same era.
Semeval-2015 task 2: Semantic textual similarity, english, spanish and pilot on interpretability
E. Agirre, C. Banea, C. Cardie, D. M. Cer, M. T. Diab, A. Gonzalez-Agirre, W. Guo, I. Lopez-Gazpio, M. Maritxalar, R. Mihalcea, G. Rigau, L. Uria, and J. Wiebe · 2015
Cited alongside, same era.
A large annotated corpus for learning natural language inference
S. R. Bowman, G. Angeli, C. Potts, and C. D. Manning · 2015
Cited alongside, same era.
Delving deep into rectifiers: Surpassing human-level performance on imagenet classification
K. He, X. Zhang, S. Ren, and J. Sun · 2015
Cited alongside, same era.
Z. Yang, D. Yang, C. Dyer, X. He, A. J. Smola, and E. H. Hovy · 2016
Later among the works it cites.
A simple but tough-to-beat baseline for sentence embeddings
S. Arora, Y. Liang, and T. Ma · 2017
Later among the works it cites.
Enriching word vectors with subword information
P. Bojanowski, E. Grave, A. Joulin, and T. Mikolov · 2017
Later among the works it cites.
Supervised learning of universal sentence representations from natural language inference data
A. Conneau, D. Kiela, H. Schwenk, L. Barrault, and A. Bordes · 2017
Later among the works it cites.
Learning generic sentence representations using convolutional neural networks
Z. Gan, Y. Pu, R. Henao, C. Li, X. He, and L. Carin · 2017
Later among the works it cites.
Discourse-based objectives for fast unsupervised sentence representation learning
Y. Jernite, S. R. Bowman, and D. Sontag · 2017
Later among the works it cites.
Learned in translation: Contextualized word vectors
B. McCann, J. Bradbury, C. Xiong, and R. Socher · 2017
Later among the works it cites.
Advances in pre-training distributed word representations
T. Mikolov, E. Grave, P. Bojanowski, C. Puhrsch, and A. Joulin · 2017
Later among the works it cites.
Dissent: Sentence representation learning from explicit discourse relations
A. Nie, E. D. Bennett, and N. D. Goodman · 2017
Later among the works it cites.
Revisiting recurrent networks for paraphrastic sentence embeddings
J. Wieting and K. Gimpel · 2017
Later among the works it cites.
A broad-coverage challenge corpus for sentence understanding through inference
A. Williams, N. Nangia, and S. R. Bowman · 2017
Later among the works it cites.
An efficient framework for learning sentence representations
L. Logeswaran and H. Lee · 2018
Closest in time.
All-but-the-top: Simple and effective postprocessing for word representations
J. Mu, S. Bhat, and P. Viswanath · 2018
Closest in time.