Fetching the paper…
Reading the bibliography…
Deep learning methods employ multiple processing layers to learn hierarchical representations of data and have produced state-of-the-art results in many domains.
D. E. Rumelhart, G. E. Hinton, and R. J. Williams, “Learning internal representations by error propagation,” DTIC Document, Tech. Rep., 1985
1985
Earlier work this paper cites.
A. Waibel, T. Hanazawa, G. Hinton, K. Shikano, and K. J. Lang, “Phoneme recognition using time-delay neural networks,” IEEE transactions on acoustics, speech, and signal processing , vol. 37, no. 3, pp. 328–339, 1989
1989
Earlier work this paper cites.
J. L. Elman, “Finding structure in time,” Cognitive science , vol. 14, no. 2, pp. 179–211, 1990
1990
Earlier work this paper cites.
J. L. Elman, “Distributed representations, simple recurrent networks, and grammatical structure,” Machine learning , vol. 7, no. 2-3, pp. 195–225, 1991
1991
Earlier work this paper cites.
R. J. Williams, “Simple statistical gradient-following algorithms for connectionist reinforcement learning,” Machine learning , vol. 8, no. 3-4, pp. 229–256, 1992
1992
Earlier work this paper cites.
T. Robinson, M. Hochberg, and S. Renals, “The use of recurrent neural networks in continuous speech recognition,” in Automatic speech and speaker recognition . Springer, 1996, pp. 233–258
1996
Earlier work this paper cites.
S. Hochreiter and J. Schmidhuber, “Long short-term memory,” Neural computation , vol. 9, no. 8, pp. 1735–1780, 1997
1997
Earlier work this paper cites.
F. A. Gers, J. Schmidhuber, and F. Cummins, “Learning to forget: Continual prediction with lstm,” 9th International Conference on Artificial Neural Networks , pp. 850–855, 1999
1999
Earlier work this paper cites.
A. M. Glenberg and D. A. Robertson, “Symbol grounding and meaning: A comparison of high-dimensional and embodied theories of meaning,” Journal of memory and language , vol. 43, no. 3, pp. 379–401, 2000
2000
Earlier work this paper cites.
K. Papineni, S. Roukos, T. Ward, and W.-J. Zhu, “Bleu: a method for automatic evaluation of machine translation,” in Proceedings of the 40th annual meeting on association for computational linguistics . Association for Computational Linguistics, 2002, pp. 311–318
2002
Earlier work this paper cites.
Y. Bengio, R. Ducharme, P. Vincent, and C. Jauvin, “A neural probabilistic language model,” Journal of machine learning research , vol. 3, no. Feb, pp. 1137–1155, 2003
2003
Earlier work this paper cites.
D. M. Blei, A. Y. Ng, and M. I. Jordan, “Latent dirichlet allocation,” Journal of machine Learning research , vol. 3, no. Jan, pp. 993–1022, 2003
2003
Earlier work this paper cites.
P. Koehn, F. J. Och, and D. Marcu, “Statistical phrase-based translation,” in Proceedings of the 2003 Conference of the North American Chapter of the Association for Computational Linguistics on Human Language Technology-Volume 1 . Association for Computational Linguistics, 2003, pp. 48–54
2003
Earlier work this paper cites.
E. F. Tjong Kim Sang and F. De Meulder, “Introduction to the conll-2003 shared task: Language-independent named entity recognition,” in Proceedings of the seventh conference on Natural language learning at HLT-NAACL 2003-Volume 4 . Association for Computational Linguistics, 2003, pp. 142–147
2003
Earlier work this paper cites.
S. T. Dumais, “Latent semantic analysis,” Annual review of information science and technology , vol. 38, no. 1, pp. 188–230, 2004
2004
Earlier work this paper cites.
B. Taskar, C. Guestrin, and D. Koller, “Max-margin markov networks,” in Advances in neural information processing systems , 2004, pp. 25–32
2004
Earlier work this paper cites.
J. Giménez and L. Marquez, “Fast and accurate part-of-speech tagging: The svm approach revisited,” Recent Advances in Natural Language Processing III , pp. 153–162, 2004
2004
Earlier work this paper cites.
B. Pang and L. Lee, “Seeing stars: Exploiting class relationships for sentiment categorization with respect to rating scales,” in Proceedings of the 43rd annual meeting on association for computational linguistics . Association for Computational Linguistics, 2005, pp. 115–124
2005
Earlier work this paper cites.
S. Petrov, L. Barrett, R. Thibaux, and D. Klein, “Learning accurate, compact, and interpretable tree annotation,” in Proceedings of the 21st International Conference on Computational Linguistics and the 44th annual meeting of the Association for Computational Linguistics . Association for Computational Linguistics, 2006, pp. 433–440
2006
Earlier work this paper cites.
R. Collobert and J. Weston, “A unified architecture for natural language processing: Deep neural networks with multitask learning,” in Proceedings of the 25th international conference on Machine learning . ACM, 2008, pp. 160–167
2008
Earlier work this paper cites.
L. Bentivogli, P. Clark, I. Dagan, and D. Giampiccolo, “The fifth pascal recognizing textual entailment challenge.” in TAC , 2009
2009
Earlier work this paper cites.
T. Mikolov, M. Karafiát, L. Burget, J. Cernockỳ, and S. Khudanpur, “Recurrent neural network based language model.” in Interspeech , vol. 2, 2010, p. 3
2010
Earlier work this paper cites.
P. D. Turney and P. Pantel, “From frequency to meaning: Vector space models of semantics,” Journal of artificial intelligence research , vol. 37, pp. 141–188, 2010
2010
Earlier work this paper cites.
S. Young, M. Gašić, S. Keizer, F. Mairesse, J. Schatzmann, B. Thomson, and K. Yu, “The hidden information state model: A practical framework for pomdp-based spoken dialogue management,” Computer Speech & Language , vol. 24, no. 2, pp. 150–174, 2010
2010
Earlier work this paper cites.
R. Collobert, J. Weston, L. Bottou, M. Karlen, K. Kavukcuoglu, and P. Kuksa, “Natural language processing (almost) from scratch,” Journal of Machine Learning Research , vol. 12, no. Aug, pp. 2493–2537, 2011
2011
Earlier work this paper cites.
J. Weston, S. Bengio, and N. Usunier, “Wsabie: Scaling up to large vocabulary image annotation,” in IJCAI , vol. 11, 2011, pp. 2764–2770
2011
Earlier work this paper cites.
R. Socher, C. C. Lin, C. Manning, and A. Y. Ng, “Parsing natural scenes and natural language with recursive neural networks,” in Proceedings of the 28th international conference on machine learning (ICML-11) , 2011, pp. 129–136
2011
Earlier work this paper cites.
X. Glorot, A. Bordes, and Y. Bengio, “Domain adaptation for large-scale sentiment classification: A deep learning approach,” in Proceedings of the 28th international conference on machine learning (ICML-11) , 2011, pp. 513–520
2011
Earlier work this paper cites.
R. Socher, J. Pennington, E. H. Huang, A. Y. Ng, and C. D. Manning, “Semi-supervised recursive autoencoders for predicting sentiment distributions,” in Proceedings of the conference on empirical methods in natural language processing . Association for Computational Linguistics, 2011, pp. 151–161
2011
Earlier work this paper cites.
T. Mikolov, S. Kombrink, L. Burget, J. Černockỳ, and S. Khudanpur, “Extensions of recurrent neural network language model,” in Acoustics, Speech and Signal Processing (ICASSP), 2011 IEEE International Conference on . IEEE, 2011, pp. 5528–5531
2011
Earlier work this paper cites.
I. Sutskever, J. Martens, and G. E. Hinton, “Generating text with recurrent neural networks,” in Proceedings of the 28th International Conference on Machine Learning (ICML-11) , 2011, pp. 1017–1024
2011
Earlier work this paper cites.
P.-h. Su, V. David, D. Kim, T.-h. Wen, and S. Young, “Learning from real users: Rating dialogue success with neural networks for reinforcement learning in spoken dialogue systems,” in in Proceedings of Interspeech . Citeseer, 2015, pp. 2007–2011
2011
Earlier work this paper cites.
A. Ritter, C. Cherry, and W. B. Dolan, “Data-driven response generation in social media,” in Proceedings of the conference on empirical methods in natural language processing . Association for Computational Linguistics, 2011, pp. 583–593
2011
Earlier work this paper cites.
A. Krizhevsky, I. Sutskever, and G. E. Hinton, “Imagenet classification with deep convolutional neural networks,” in Advances in neural information processing systems , 2012, pp. 1097–1105
2012
Earlier work this paper cites.
A. Mukherjee and B. Liu, “Aspect extraction through semi-supervised modeling,” in Proceedings of the 50th Annual Meeting of the Association for Computational Linguistics: Long Papers-Volume 1 . Association for Computational Linguistics, 2012, pp. 339–348
2012
Earlier work this paper cites.
R. Socher, B. Huval, C. D. Manning, and A. Y. Ng, “Semantic compositionality through recursive matrix-vector spaces,” in Proceedings of the 2012 joint conference on empirical methods in natural language processing and computational natural language learning . Association for Computational Linguistics, 2012, pp. 1201–1211
2012
Earlier work this paper cites.
S. Pradhan, A. Moschitti, N. Xue, O. Uryupina, and Y. Zhang, “Conll-2012 shared task: Modeling multilingual unrestricted coreference in ontonotes,” in Joint Conference on EMNLP and CoNLL-Shared Task . Association for Computational Linguistics, 2012, pp. 1–40
2012
Earlier work this paper cites.
T. Mikolov, I. Sutskever, K. Chen, G. S. Corrado, and J. Dean, “Distributed representations of words and phrases and their compositionality,” in Advances in neural information processing systems , 2013, pp. 3111–3119
2013
Earlier work this paper cites.
R. Socher, A. Perelygin, J. Y. Wu, J. Chuang, C. D. Manning, A. Y. Ng, C. Potts et al. , “Recursive deep models for semantic compositionality over a sentiment treebank,” in Proceedings of the conference on empirical methods in natural language processing (EMNLP) , vol. 1631, 2013, p. 1642
2013
Earlier work this paper cites.
2013
Earlier work this paper cites.
K. M. Hermann and P. Blunsom, “The role of syntax in vector space models of compositional semantics,” in Proceedings of the 51st Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) . Association for Computational Linguistics, 2013
2013
Earlier work this paper cites.
I. Labutov and H. Lipson, “Re-embedding words.” in ACL (2) , 2013, pp. 489–493
2013
Earlier work this paper cites.
X. Zheng, H. Chen, and T. Xu, “Deep learning for chinese word segmentation and pos tagging.” in EMNLP , 2013, pp. 647–657
2013
Earlier work this paper cites.
M. Auli, M. Galley, C. Quirk, and G. Zweig, “Joint language and translation modeling with recurrent neural networks.” in EMNLP , 2013, pp. 1044–1054
2013
Earlier work this paper cites.
A. Graves, A.-r. Mohamed, and G. Hinton, “Speech recognition with deep recurrent neural networks,” in Acoustics, speech and signal processing (icassp), 2013 ieee international conference on . IEEE, 2013, pp. 6645–6649
2013
Earlier work this paper cites.
2013
Earlier work this paper cites.
S. Young, M. Gašić, B. Thomson, and J. D. Williams, “Pomdp-based statistical spoken dialog systems: A review,” Proceedings of the IEEE , vol. 101, no. 5, pp. 1160–1179, 2013
2013
Earlier work this paper cites.
2013
Earlier work this paper cites.
M. Zhu, Y. Zhang, W. Chen, M. Zhang, and J. Zhu, “Fast and accurate shift-reduce constituent parsing.” in ACL (1) , 2013, pp. 434–443
2013
Earlier work this paper cites.
A. Fader, L. S. Zettlemoyer, and O. Etzioni, “Paraphrase-driven learning for open question answering.” in ACL (1) , 2013, pp. 1608–1618
2013
Earlier work this paper cites.
E. Cambria and B. White, “Jumping NLP curves: A review of natural language processing research,” IEEE Computational Intelligence Magazine , vol. 9, no. 2, pp. 48–57, 2014
2014
Earlier work this paper cites.
J. Pennington, R. Socher, and C. D. Manning, “Glove: Global vectors for word representation.” in EMNLP , vol. 14, 2014, pp. 1532–1543
2014
Earlier work this paper cites.
X. Rong, “word2vec parameter learning explained,” arXiv preprint arXiv:1411.2738 , 2014
2014
Earlier work this paper cites.
D. Tang, F. Wei, N. Yang, M. Zhou, T. Liu, and B. Qin, “Learning sentiment-specific word embedding for twitter sentiment classification.” in ACL (1) , 2014, pp. 1555–1565
2014
Earlier work this paper cites.
C. N. Dos Santos and M. Gatti, “Deep convolutional neural networks for sentiment analysis of short texts.” in COLING , 2014, pp. 69–78
2014
Earlier work this paper cites.
C. D. Santos and B. Zadrozny, “Learning character-level representations for part-of-speech tagging,” in Proceedings of the 31st International Conference on Machine Learning (ICML-14) , 2014, pp. 1818–1826
2014
Earlier work this paper cites.
A. Sharif Razavian, H. Azizpour, J. Sullivan, and S. Carlsson, “Cnn features off-the-shelf: an astounding baseline for recognition,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition Workshops , 2014, pp. 806–813
2014
Earlier work this paper cites.
Y. Jia, E. Shelhamer, J. Donahue, S. Karayev, J. Long, R. Girshick, S. Guadarrama, and T. Darrell, “Caffe: Convolutional architecture for fast feature embedding,” in Proceedings of the 22nd ACM international conference on Multimedia . ACM, 2014, pp. 675–678
2014
Earlier work this paper cites.
N. Kalchbrenner, E. Grefenstette, and P. Blunsom, “A convolutional neural network for modelling sentences,” Proceedings of the 52nd Annual Meeting of the Association for Computational Linguistics , June 2014. [Online]. Available: http://goo.gl/EsQCuC
2014
Earlier work this paper cites.
2014
Earlier work this paper cites.
M. Denil, A. Demiraj, N. Kalchbrenner, P. Blunsom, and N. de Freitas, “Modelling, visualising and summarising documents with a single convolutional neural network,” 26th International Conference on Computational Linguistics , pp. 1601–1612, 2014
2014
Earlier work this paper cites.
Y. Shen, X. He, J. Gao, L. Deng, and G. Mesnil, “A latent semantic model with convolutional-pooling structure for information retrieval,” in Proceedings of the 23rd ACM International Conference on Conference on Information and Knowledge Management . ACM, 2014, pp. 101–110
2014
Earlier work this paper cites.
W.-t. Yih, X. He, and C. Meek, “Semantic parsing for single-relation question answering.” in ACL (2) . Citeseer, 2014, pp. 643–648
2014
Earlier work this paper cites.
O. Abdel-Hamid, A.-r. Mohamed, H. Jiang, L. Deng, G. Penn, and D. Yu, “Convolutional neural networks for speech recognition,” IEEE/ACM Transactions on audio, speech, and language processing , vol. 22, no. 10, pp. 1533–1545, 2014
2014
Earlier work this paper cites.
S. Liu, N. Yang, M. Li, and M. Zhou, “A recursive recurrent neural network for statistical machine translation,” Proceedings of the 52nd Annual Meeting of the Association for Computational Linguistics , pp. 1491–1500, 2014
2014
Cited alongside, same era.
I. Sutskever, O. Vinyals, and Q. V. Le, “Sequence to sequence learning with neural networks,” in Advances in neural information processing systems , 2014, pp. 3104–3112
2014
Cited alongside, same era.
A. Graves and N. Jaitly, “Towards end-to-end speech recognition with recurrent neural networks,” in Proceedings of the 31st International Conference on Machine Learning (ICML-14) , 2014, pp. 1764–1772
2014
Cited alongside, same era.
2014
Cited alongside, same era.
2016
Later among the works it cites.
S. Poria, E. Cambria, D. Hazarika, and P. Vij, “A deeper look into sarcastic tweets using deep convolutional neural networks,” in COLING , 2016, pp. 1601–1612
2016
Later among the works it cites.
2016
Later among the works it cites.
2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2014
Cited alongside, same era.
2014
Cited alongside, same era.
M. Sundermeyer, T. Alkhouli, J. Wuebker, and H. Ney, “Translation modeling with bidirectional recurrent neural networks.” in EMNLP , 2014, pp. 14–25
2014
Cited alongside, same era.
2014
Cited alongside, same era.
J. Weston, S. Chopra, and A. Bordes, “Memory networks,” arXiv preprint arXiv:1410.3916 , 2014
2014
Cited alongside, same era.
2014
Cited alongside, same era.
I. Goodfellow, J. Pouget-Abadie, M. Mirza, B. Xu, D. Warde-Farley, S. Ozair, A. Courville, and Y. Bengio, “Generative adversarial nets,” in Advances in neural information processing systems , 2014, pp. 2672–2680
2014
Cited alongside, same era.
D. Chen and C. D. Manning, “A fast and accurate dependency parser using neural networks.” in EMNLP , 2014, pp. 740–750
2014
Cited alongside, same era.
2016
Later among the works it cites.
2016
Later among the works it cites.
Y. Wang, M. Huang, X. Zhu, and L. Zhao, “Attention-based lstm for aspect-level sentiment classification.” in EMNLP , 2016, pp. 606–615
2016
Later among the works it cites.
2016
Later among the works it cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2016, pp. 770–778
2016
Later among the works it cites.
2016
Later among the works it cites.
2016
Later among the works it cites.
2016
Later among the works it cites.
Y. Zhang, Z. Gan, and L. Carin, “Generating text via adversarial training,” in NIPS workshop on Adversarial Training , 2016
2016
Later among the works it cites.
C. Xiong, S. Merity, and R. Socher, “Dynamic memory networks for visual and textual question answering,” arXiv , vol. 1603, 2016
2016
Later among the works it cites.
2016
Later among the works it cites.
A. Zadeh, R. Zellers, E. Pincus, and L.-P. Morency, “Multimodal sentiment intensity analysis in videos: Facial gestures and verbal messages,” IEEE Intelligent Systems , vol. 31, no. 6, pp. 82–88, 2016
2016
Later among the works it cites.
2016
Later among the works it cites.
2016
Later among the works it cites.
X. Zhou, D. Dong, H. Wu, S. Zhao, D. Yu, H. Tian, X. Liu, and R. Yan, “Multi-view response selection for human-computer conversation.” in EMNLP , 2016, pp. 372–381
2016
Later among the works it cites.
I. V. Serban, A. Sordoni, Y. Bengio, A. C. Courville, and J. Pineau, “Building end-to-end dialogue systems using generative hierarchical neural network models.” in AAAI , 2016, pp. 3776–3784
2016
Later among the works it cites.
E. Cambria, S. Poria, A. Gelbukh, and M. Thelwall, “Sentiment analysis is a big suitcase,” IEEE Intelligent Systems , vol. 32, no. 6, pp. 74–80, 2017
2017
Closest in time.
A. Gittens, D. Achlioptas, and M. W. Mahoney, “Skip-gram-zipf+ uniform= vector additivity,” in Proceedings of the 55th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , vol. 1, 2017, pp. 69–76
2017
Closest in time.
2017
Closest in time.
H. Peng, E. Cambria, and X. Zou, “Radical-based hierarchical embeddings for chinese sentiment analysis at sentence level,” in FLAIRS , 2017, pp. 347–352
2017
Closest in time.
2017
Closest in time.
2017
Closest in time.
2017
Closest in time.
A. Mousa and B. Schuller, “Contextual bidirectional long short-term memory recurrent neural network language models: A generative approach to sentiment analysis,” in Proceedings of the 15th Conference of the European Chapter of the Association for Computational Linguistics: Volume 1, Long Papers , vol. 1, 2017, pp. 1023–1032
2017
Closest in time.
G. Chen, D. Ye, E. Cambria, J. Chen, and Z. Xing, “Ensemble application of convolutional and recurrent neural networks for multi-label text categorization,” in IJCNN , 2017, pp. 2377–2383
2017
Closest in time.
S. Poria, E. Cambria, D. Hazarika, N. Mazumder, A. Zadeh, and L.-P. Morency, “Context-dependent sentiment analysis in user-generated videos,” in ACL , 2017, pp. 873–883
2017
Closest in time.
A. Zadeh, M. Chen, S. Poria, E. Cambria, and L.-P. Morency, “Tensor fusion network for multimodal sentiment analysis,” in Empirical Methods in NLP , 2017
2017
Closest in time.
E. Tong, A. Zadeh, and L.-P. Morency, “Combating human trafficking with deep multimodal models,” in Association for Computational Linguistics , 2017
2017
Closest in time.
I. Chaturvedi, E. Ragusa, P. Gastaldo, R. Zunino, and E. Cambria, “Bayesian network based extreme learning machine for subjectivity detection,” Journal of The Franklin Institute , 2017
2017
Closest in time.
2017
Closest in time.
2017
Closest in time.
2017
Closest in time.
2017
Closest in time.
2017
Closest in time.
L. Yu, W. Zhang, J. Wang, and Y. Yu, “Seqgan: sequence generative adversarial nets with policy gradient,” in Thirty-First AAAI Conference on Artificial Intelligence , 2017
2017
Closest in time.
2017
Closest in time.
H. Zhou, Y. Zhang, C. Cheng, S. Huang, X. Dai, and J. Chen, “A neural probabilistic structured-prediction method for transition-based natural language processing,” Journal of Artificial Intelligence Research , vol. 58, pp. 703–729, 2017
2017
Closest in time.
2017
Closest in time.
L. He, K. Lee, M. Lewis, and L. Zettlemoyer, “Deep semantic role labeling: What works and what’s next,” in Proceedings of the Annual Meeting of the Association for Computational Linguistics , 2017
2017
Closest in time.
L.-C. Yu, J. Wang, K. R. Lai, and X. Zhang, “Refining word embeddings for sentiment analysis,” in Proceedings of the 2017 Conference on Empirical Methods in Natural Language Processing , 2017, pp. 545–550
2017
Closest in time.
2017
Closest in time.
2017
Closest in time.
Y. Shen, P.-S. Huang, J. Gao, and W. Chen, “Reasonet: Learning to stop reading in machine comprehension,” in Proceedings of the 23rd ACM SIGKDD International Conference on Knowledge Discovery and Data Mining . ACM, 2017, pp. 1047–1055
2017
Closest in time.
2017
Closest in time.
C. Qian, Z. Xiao-Dan, L. Zhen-Hua, W. Si, J. Hui, and I. Diana, “Enhanced lstm for natural language inference,” In ACL , 2017
2017
Closest in time.
H. Luheng, L. Kenton, L. Mike, and S. Z. Luke, “Deep semantic role labeling: What works and what’s next,” In ACL , 2017
2017
Closest in time.
L. Kenton, H. Luheng, L. Mike, and S. Z. Luke, “End-to-end neural coreference resolution,” In EMNLP , 2017
2017
Closest in time.
E. P. Matthew, A. Waleed, B. Chandra, and P. Russell, “Semi-supervised sequence tagging with bidirectional language models,” In ACL , 2017
2017
Closest in time.
M. Bryan, B. James, X. Caiming, and S. Richard, “Learned in translation: Contextualized word vectors,” In NIPS 2017 , 2017
2017
Closest in time.
2017
Closest in time.
2017
Closest in time.
2018
Closest in time.
A. Radford, K. Narasimhan, T. Salimans, and I. Sutskever, “Improving language understanding by generative pre-training,” URL https://s3-us-west-2. amazonaws. com/openai-assets/research-covers/language-unsupervised/language_ understanding_paper. pdf , 2018
2018
Closest in time.
2018
Closest in time.
Y. Ma, H. Peng, and E. Cambria, “Targeted aspect-based sentiment analysis via embedding commonsense knowledge into an attentive lstm,” in AAAI , 2018
2018
Closest in time.
2018
Closest in time.
2018
Closest in time.
B. Hu, Z. Lu, H. Li, and Q. Chen, “Convolutional neural network architectures for matching natural language sentences,” in Advances in neural information processing systems , 2014, pp. 2042–2050
2050
Closest in time.
K. Xu, J. Ba, R. Kiros, K. Cho, A. Courville, R. Salakhudinov, R. Zemel, and Y. Bengio, “Show, attend and tell: Neural image caption generation with visual attention,” in International Conference on Machine Learning , 2015, pp. 2048–2057
2057
Closest in time.