Fetching the paper…
Reading the bibliography…
Existing deep multitask learning (MTL) approaches align layers shared between tasks in a parallel ordering.
Long short-term memory
S. Hochreiter and J. Schmidhuber · 1997
Earlier work this paper cites.
Multitask learning
R. Caruana · 1998
Earlier work this paper cites.
A unified architecture for natural language processing: Deep neural networks with multitask learning
R. Collobert and J. Weston · 2008
Earlier work this paper cites.
Sparse deep belief net model for visual area v2
H. Lee, C. Ekanadham, and A. Y. Ng · 2008
Earlier work this paper cites.
Transfer learning using Kolmogorov complexity: Basic theory and empirical evaluations
M. M. Mahmud and S. Ray · 2008
Earlier work this paper cites.
On universal transfer learning
M. H. Mahmud · 2009
Earlier work this paper cites.
Parsing natural scenes and natural language with recursive neural networks
R. Socher, C. C.-Y. Lin, A. Y. Ng, and C. D. Manning · 2011
Earlier work this paper cites.
Adaptive dropout for training deep neural networks
J. Ba and B. Frey · 2013
Earlier work this paper cites.
Cross-language knowledge transfer using multilingual deep neural network with shared hidden layers
J. T. Huang, J. Li, D. Yu, L. Deng, and Y. Gong · 2013
Earlier work this paper cites.
UCI machine learning repository, 2013
M. Lichman · 2013
Earlier work this paper cites.
Exact solutions to the nonlinear dynamics of learning in deep linear neural networks
A. M. Saxe, J. L. McClelland, and S. Ganguli · 2013
Earlier work this paper cites.
Multi-task learning in deep neural networks for improved phoneme recognition
M. L. Seltzer and J. Droppo · 2013
Earlier work this paper cites.
Adam: A method for stochastic optimization
D. P. Kingma and J. Ba · 2014
Earlier work this paper cites.
Recurrent convolutional neural networks for scene labeling
P. Pinheiro and R. Collobert · 2014
Earlier work this paper cites.
Dropout: A Simple Way to Prevent Neural Networks from Overfitting
N. Srivastava, G. Hinton, A. Krizhevsky, I. Sutskever, and R. Salakhutdinov · 2014
Earlier work this paper cites.
Facial landmark detection by deep multi-task learning
Z. Zhang, P. Luo, C. C. Loy, and X. Tang · 2014
Earlier work this paper cites.
TensorFlow: Large-scale machine learning on heterogeneous systems, 2015
M. Abadi, A. Agarwal, P. Barham, E. Brevdo, Z. Chen, C. Citro, G. S. Corrado, A. Davis, J. Dean, M. Devin, S. Ghemawat, I. Goodfellow, A. Harp, G. Irving, M. Isard, Y. Jia, R. Jozefowicz, L. Kaiser, M. Kudlur, J. Levenberg, D. Mané, R. Monga, S. Moore, D. Murray, C. Olah, M. Schuster, J. Shlens, B. Steiner, I. Sutskever, K. Talwar, P. Tucker, V. Vanhoucke, V. Vasudevan, F. Viégas, O. Vinyals, P. Warden, M. Wattenberg, M. Wicke, Y. Yu, and X. Zheng · 2015
Earlier work this paper cites.
Keras, 2015
F. Chollet et al · 2015
Cited alongside, same era.
Multi-task learning for multiple language translation
D. Dong, H. Wu, W. He, D. Yu, and H. Wang · 2015
Cited alongside, same era.
Rapid adaptation for deep neural networks through multi-task learning
Z. Huang, J. Li, S. M. Siniscalchi, I.-F. Chen, J. Wu, and C.-H. Lee · 2015
Cited alongside, same era.
Human-level concept learning through probabilistic program induction
B. M. Lake, R. Salakhutdinov, and J. B. Tenenbaum · 2015
Cited alongside, same era.
Deep learning
Y. Lecun, Y. Bengio, and G. Hinton · 2015
Cited alongside, same era.
Recurrent convolutional neural network for object recognition
M. Liang and X. Hu · 2015
Cited alongside, same era.
R. Ranjan, V. M. Patel, and R. Chellappa · 2016
Later among the works it cites.
One-shot generalization in deep generative models
D. Rezende, Shakir, I. Danihelka, K. Gregor, and D. Wierstra · 2016
Later among the works it cites.
MOON: A mixed objective optimization network for the recognition of facial attributes
E. M. Rudd, M. Günther, and T. E. Boult · 2016
Later among the works it cites.
A. A. Rusu, N. C. Rabinowitz, G. Desjardins, H. Soyer, et al · 2016
Later among the works it cites.
Residual networks behave like ensembles of relatively shallow networks
A. Veit, M. J. Wilber, and S. Belongie · 2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Deep neural networks employing multi-task learning and stacked bottleneck features for speech synthesis
Z. Wu, C. Valentini-Botinhao, O. Watts, and S. King · 2015
Cited alongside, same era.
Integrated perception with recurrent multi-task neural networks
H. Bilen and A. Vedaldi · 2016
Cited alongside, same era.
Learning modular neural network policies for multi-task and multi-robot transfer
C. Devin, A. Gupta, T. Darrell, P. Abbeel, and S. Levine · 2016
Cited alongside, same era.
D. Ha, A. M. Dai, and Q. V. Le · 2016
Cited alongside, same era.
A joint many-task model: Growing a neural network for multiple NLP tasks
K. Hashimoto, C. Xiong, Y. Tsuruoka, and R. Socher · 2016
Cited alongside, same era.
Deep residual learning for image recognition
K. He, X. Zhang, S. Ren, and J. Sun · 2016
Cited alongside, same era.
A. R. Zamir, T. Wu, L. Sun, W. Shen, J. Malik, and S. Saverese · 2016
Later among the works it cites.
Stack-propagation: Improved representation learning for syntax
Y. Zhang and D. Weiss · 2016
Later among the works it cites.
Universal representations: The missing link between faces, text, planktons, and cat breeds
H. Bilen and A. Vedaldi · 2017
Closest in time.
Pathnet: Evolution channels gradient descent in super neural networks
C. Fernando, D. Banarse, C. Blundell, Y. Zwols, D. Ha, A. A. Rusu, A. Pritzel, and D. Wierstra · 2017
Closest in time.
AFFACT - alignment free facial attribute classification technique
M. Günther, A. Rozsa, and T. E. Boult · 2017
Closest in time.
Adaptively weighted multi-task deep network for person attribute classification
K. He, Z. Wang, Y. Fu, R. Feng, Y.-G. Jiang, and X. Xue · 2017
Closest in time.
Reinforcement learning with unsupervised auxiliary tasks
M. Jaderberg, V. Mnih, W. M. Czarnecki, T. Schaul, J. Z. Leibo, D. Silver, and K. Kavukcuoglu · 2017
Closest in time.
Overcoming catastrophic forgetting in neural networks
J. Kirkpatrick, R. Pascanu, N. Rabinowitz, J. Veness, et al · 2017
Closest in time.
Fully-adaptive feature sharing in multi-task networks with applications in person attribute classification
Y. Lu, A. Kumar, S. Zhai, Y. Cheng, T. Javidi, and R. S. Feris · 2017
Closest in time.
Outrageously large neural networks: The sparsely-gated mixture-of-experts layer
N. Shazeer, A. Mirhoseini, K. Maziarz, A. Davis, Q. V. Le, G. E. Hinton, and J. Dean · 2017
Closest in time.
Multitask Learning with Low-Level Auxiliary Tasks for Encoder-Decoder Based Speech Recognition
S. Toshniwal, H. Tang, L. Lu, and K. Livescu · 2017
Closest in time.
Deep multi-task representation learning: A tensor factorisation approach
Y. Yang and T. Hospedales · 2017
Closest in time.