Learning Word Representations from Scarce and Noisy Data with Embedding Sub-spaces
Astudillo, Ramon F, Amir, Silvio, Lin, Wang, Silva, Mário, and Trancoso, Isabel · 2015
Later among the works it cites.
Neural Machine Translation by Jointly Learning to Align and Translate
Bahdanau, Dzmitry, Cho, Kyunghyun, and Bengio, Yoshua · 2015
Later among the works it cites.
A Large Annotated Corpus for Learning Natural Language Inference
Bowman, Samuel R, Angeli, Gabor, Potts, Christopher, and Manning, Christopher D · 2015
Later among the works it cites.
Attention-based Models for Speech Recognition
Chorowski, Jan K, Bahdanau, Dzmitry, Serdyuk, Dmitriy, Cho, Kyunghyun, and Bengio, Yoshua · 2015
Later among the works it cites.
An Exploration of Softmax Alternatives Belonging to the Spherical Loss Family
Original
de Brébisson, Alexandre and Vincent, Pascal · 2015
Later among the works it cites.
Learning to Transduce with Unbounded Memory
Grefenstette, Edward, Hermann, Karl Moritz, Suleyman, Mustafa, and Blunsom, Phil · 2015
Later among the works it cites.
Statistical Learning with Sparsity: the Lasso and Generalizations
Hastie, Trevor, Tibshirani, Robert, and Wainwright, Martin · 2015
Later among the works it cites.
Teaching Machines to Read and Comprehend
Hermann, Karl Moritz, Kocisky, Tomas, Grefenstette, Edward, Espeholt, Lasse, Kay, Will, Suleyman, Mustafa, and Blunsom, Phil · 2015
Later among the works it cites.
Consistent Multilabel Classification
Koyejo, Sanmi, Natarajan, Nagarajan, Ravikumar, Pradeep K, and Dhillon, Inderjit S · 2015
Later among the works it cites.
Reasoning about Entailment with Neural Attention
Original
Rocktäschel, Tim, Grefenstette, Edward, Hermann, Karl Moritz, Kočiskỳ, Tomáš, and Blunsom, Phil · 2015
Later among the works it cites.
A Neural Attention Model for Abstractive Sentence Summarization
Rush, Alexander M, Chopra, Sumit, and Weston, Jason · 2015
Later among the works it cites.
End-to-End Memory Networks
Sukhbaatar, Sainbayar, Szlam, Arthur, Weston, Jason, and Fergus, Rob · 2015
Later among the works it cites.
Efficient Exact Gradient Update for Training Deep Networks with Very Large Sparse Targets
Vincent, Pascal · 2015
Later among the works it cites.
Show, Attend and Tell: Neural Image Caption Generation with Visual Attention
Xu, Kelvin, Ba, Jimmy, Kiros, Ryan, Courville, Aaron, Salakhutdinov, Ruslan, Zemel, Richard, and Bengio, Yoshua · 2015
Later among the works it cites.
Deep Learning
Goodfellow, Ian, Bengio, Yoshua, and Courville, Aaron · 2016
Closest in time.