Fetching the paper…
Reading the bibliography…
Transfer learning is a vital technique that generalizes models trained for one setting or task to other settings or tasks.
D. Gentner, “Structure-mapping: A theoretical framework for analogy,”
1983
Earlier work this paper cites.
J. G. Carbonell,
1983
Earlier work this paper cites.
P. E. Utgoff, “Incremental induction of decision trees,”
1989
Earlier work this paper cites.
J.-L. Gauvain and C.-H. Lee, “Maximum a posteriori estimation for multivariate Gaussian mixture observations of Markov chains,”
1994
Earlier work this paper cites.
“NIPS 95 workshop on learning to learn: Knowledge consolidation and transfer in inductive systems,” 1995. [Online]. Available:
1995
Earlier work this paper cites.
C. J. Leggetter and P. Woodland, “Maximum likelihood linear regression for speaker adaptation of continuous density hidden Markov models,”
1995
Earlier work this paper cites.
J. Neto, L. Almeida, M. Hochberg, C. Martins, L. Nunes, S. Renals, and T. Robinson, “Speaker-adaptation for hybrid HMM-ANN continuous speech recognition system,” in
1995
Earlier work this paper cites.
R. Caruana, “Multitask learning,”
1997
Earlier work this paper cites.
D. Gentner and K. J. Holyoak, “Reasoning and learning by analogy: Introduction.”
1997
Earlier work this paper cites.
A. Blum and T. Mitchell, “Combining labeled and unlabeled data with co-training,” in
1998
Earlier work this paper cites.
J. H. Martin and D. Jurafsky, “Speech and language processing,”
2000
Earlier work this paper cites.
J. Baxter, “A model of inductive bias learning,”
2000
Earlier work this paper cites.
T. Schultz and A. Waibel, “Language-independent and language-adaptive acoustic modeling for speech recognition,”
2001
Earlier work this paper cites.
M. Tamura, T. Masuko, K. Tokuda, and T. Kobayashi, “Adaptation of pitch and spectrum for HMM-based speech synthesis using MLLR,” in
2001
Earlier work this paper cites.
X. Zhu, “Semi-supervised learning literature survey,” Computer Sciences TRP 1530, University of Wisconsin ¨C Madison, 2005
2005
Earlier work this paper cites.
O. Arandjelovic and R. Cipolla, “Incremental learning of temporally-coherent Gaussian mixture models,”
2006
Earlier work this paper cites.
J. Blitzer, R. McDonald, and F. Pereira, “Domain adaptation with structural correspondence learning,” in
2006
Earlier work this paper cites.
G. E. Hinton, S. Osindero, and Y.-W. Teh, “A fast learning algorithm for deep belief nets,”
2006
Earlier work this paper cites.
G. E. Hinton and R. R. Salakhutdinov, “Reducing the dimensionality of data with neural networks,”
2006
Earlier work this paper cites.
L. Fei-Fei, R. Fergus, and P. Perona, “One-shot learning of object categories,”
2006
Earlier work this paper cites.
Y. Bengio, H. Schwenk, J.-S. Senécal, F. Morin, and J.-L. Gauvain, “Neural probabilistic language models,” in
2006
Earlier work this paper cites.
R. Raina, A. Battle, H. Lee, B. Packer, and A. Y. Ng, “Self-taught learning: transfer learning from unlabeled data,” in
2007
Earlier work this paper cites.
Y. Bengio, P. Lamblin, D. Popovici, H. Larochelle
2007
Earlier work this paper cites.
R. Gemello, F. Mana, S. Scanzio, P. Laface, and R. De Mori, “Linear hidden transformations for adaptation of hybrid ANN/HMM models,”
2007
Earlier work this paper cites.
J. Yamagishi and T. Kobayashi, “Average-voice-based speech synthesis using HSMM-based speaker adaptation and adaptive training,”
2007
Earlier work this paper cites.
J. Benesty,
2008
Earlier work this paper cites.
A. Declercq and J. H. Piater, “Online learning of Gaussian mixture models-a two-level approach.” in
2008
Earlier work this paper cites.
W. Dai, Y. Chen, G.-R. Xue, Q. Yang, and Y. Yu, “Translated learning: Transfer learning across different feature spaces,” in
2008
Earlier work this paper cites.
R. Collobert and J. Weston, “A unified architecture for natural language processing: Deep neural networks with multitask learning,” in
2008
Earlier work this paper cites.
H. Larochelle, D. Erhan, and Y. Bengio, “Zero-data learning of new tasks.” in
2008
Earlier work this paper cites.
A. Mnih and G. E. Hinton, “A scalable hierarchical distributed language model,” in
2008
Earlier work this paper cites.
M. E. Taylor and P. Stone, “Transfer learning for reinforcement learning domains: A survey,”
2009
Earlier work this paper cites.
R. Salakhutdinov and G. E. Hinton, “Deep boltzmann machines,” in
2009
Earlier work this paper cites.
Y.-J. Wu, Y. Nankaku, and K. Tokuda, “State mapping based method for cross-lingual speaker adaptation in HMM-based speech synthesis.” in
2009
Earlier work this paper cites.
J. Yamagishi, T. Kobayashi, Y. Nakano, K. Ogata, and J. Isogai, “Analysis of speaker adaptation algorithms for HMM-based speech synthesis and a constrained SMAPLR adaptation algorithm,”
2009
Earlier work this paper cites.
P. Koehn,
2009
Earlier work this paper cites.
S. J. Pan and Q. Yang, “A survey on transfer learning,”
2010
Earlier work this paper cites.
X. Shi, Q. Liu, W. Fan, P. S. Yu, and R. Zhu, “Transfer learning on heterogenous feature spaces via spectral transformation,” in
2010
Earlier work this paper cites.
P. Vincent, H. Larochelle, I. Lajoie, Y. Bengio, and P.-A. Manzagol, “Stacked denoising autoencoders: Learning useful representations in a deep network with a local denoising criterion,”
2010
Earlier work this paper cites.
S. M. Gutstein,
2010
Earlier work this paper cites.
B. Li and K. C. Sim, “Comparison of discriminative input and output transformations for speaker adaptation in the hybrid NN/HMM systems,” in
2010
Earlier work this paper cites.
H. Liang, J. Dines, and L. Saheer, “A comparison of supervised and unsupervised cross-lingual speaker adaptation approaches for HMM-based speech synthesis,” in
2010
Earlier work this paper cites.
L. Shi, R. Mihalcea, and M. Tian, “Cross language text classification by model translation and semi-supervised learning,” in
2010
Earlier work this paper cites.
J. Turian, L. Ratinov, and Y. Bengio, “Word representations: a simple and general method for semi-supervised learning,” in
2010
Earlier work this paper cites.
R. Socher and L. Fei-Fei, “Connecting modalities: Semi-supervised segmentation and annotation of images using unaligned text corpora,” in
2010
Earlier work this paper cites.
C. Wang and S. Mahadevan, “Heterogeneous domain adaptation using manifold alignment,” in
2011
Earlier work this paper cites.
Y. Zhu, Y. Chen, Z. Lu, S. J. Pan, G.-R. Xue, Y. Yu, and Q. Yang, “Heterogeneous transfer learning for image classification.” in
2011
Earlier work this paper cites.
H.-Y. Wang and Q. Yang, “Transfer learning by structural analogy,” in
2011
Earlier work this paper cites.
S. J. Pan, I. W. Tsang, J. T. Kwok, and Q. Yang, “Domain adaptation via transfer component analysis,”
2011
Earlier work this paper cites.
P. Prettenhofer and B. Stein, “Cross-lingual adaptation using structural correspondence learning,”
2011
Earlier work this paper cites.
B. Kulis, K. Saenko, and T. Darrell, “What you saw is not what you get: Domain adaptation using asymmetric kernel transforms,” in
2011
Earlier work this paper cites.
B. Wei and C. J. Pal, “Heterogeneous transfer learning with RBMs.” in
2011
Cited alongside, same era.
J. Guinney, Q. Wu, and S. Mukherjee, “Estimating variable structure and dependence in multitask learning via gradients,”
2011
Cited alongside, same era.
Y. Bengio and O. Delalleau, “On the expressive power of deep architectures,” in
2011
Cited alongside, same era.
J. Ngiam, A. Khosla, M. Kim, J. Nam, H. Lee, and A. Y. Ng, “Multimodal deep learning,” in
2011
Cited alongside, same era.
X. Glorot, A. Bordes, and Y. Bengio, “Domain adaptation for large-scale sentiment classification: A deep learning approach,” in
2011
Cited alongside, same era.
N. T. Vu, F. Kraus, and T. Schultz, “Cross-language bootstrapping based on completely unsupervised training using multilingual A-stabil,” in
A. Frome, G. S. Corrado, J. Shlens, S. Bengio, J. Dean, T. Mikolov
2013
Later among the works it cites.
G. E. Hinton, O. Vinyals, and J. Dean, “Distilling the knowledge in a neural network,” in
2014
Later among the works it cites.
J. T. Zhou, S. J. Pan, I. W. Tsang, and Y. Yan, “Hybrid heterogeneous transfer learning through deep learning,” in
2014
Later among the works it cites.
X. He, J. Gao, and L. Deng, “Deep learning for natural language processing and related applications (Tutorial at ICASSP),” in
2014
Later among the works it cites.
M. Oquab, L. Bottou, I. Laptev, and J. Sivic, “Learning and transferring mid-level image representations using convolutional neural networks,” in
2014
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2011
Cited alongside, same era.
N. Dehak, P. Kenny, R. Dehak, P. Dumouchel, and P. Ouellet, “Front-end factor analysis for speaker verification,”
2011
Cited alongside, same era.
M. L. Seltzer and A. Acero, “Separating speaker and environmental variability using factored transforms.” in
2011
Cited alongside, same era.
K. Shinoda, “Speaker adaptation techniques for automatic speech recognition,”
2011
Cited alongside, same era.
M. Gibson and W. Byrne, “Unsupervised intralingual and cross-lingual speaker adaptation for HMM-based speech synthesis using two-pass decision tree construction,”
2011
Cited alongside, same era.
W. De Smet, J. Tang, and M.-F. Moens, “Knowledge transfer across multilingual corpora via latent topics,” in
2011
Cited alongside, same era.
R. Collobert, J. Weston, L. Bottou, M. Karlen, K. Kavukcuoglu, and P. Kuksa, “Natural language processing (almost) from scratch,”
2011
Cited alongside, same era.
K. M. Knill, M. J. Gales, A. Ragni, and S. P. Rath, “Language independent and unsupervised acoustic models for speech recognition and keyword spotting,” in
2014
Later among the works it cites.
D. Chen, B. Mak, C.-C. Leung, and S. Sivadas, “Joint acoustic modeling of triphones and trigraphemes by multi-task learning deep neural networks for low-resource speech recognition,” in
2014
Later among the works it cites.
V. Ehsan, L. Xin, M. Erik, L. M. Ignacio, and G.-D. Javier, “Deep neural networks for small footprint text-dependent speaker verification,”
2014
Later among the works it cites.
Y. Xu, J. Du, L.-R. Dai, and C.-H. Lee, “Cross-language transfer learning for deep neural network based speech enhancement,” in
2014
Later among the works it cites.
S. Xue, O. Abdel-Hamid, H. Jiang, and L. Dai, “Direct adaptation of hybrid DNN/HMM model for fast speaker adaptation in LVCSR based on speaker code,” in
2014
Later among the works it cites.
P. Karanasou, Y. Wang, M. J. Gales, and P. C. Woodland, “Adaptation of deep neural network acoustic models using factorised i-vectors,” in
2014
Later among the works it cites.
A. Senior and I. Lopez-Moreno, “Improving DNN speaker independence with i-vector inputs,” in
2014
Later among the works it cites.
V. Gupta, P. Kenny, P. Ouellet, and T. Stafylakis, “I-vector-based speaker adaptation of deep neural networks for french broadcast audio transcription,” in
2014
Later among the works it cites.
M. Rouvier and B. Favre, “Speaker adaptation of DNN-based ASR with i-vectors: Does it actually adapt models to speakers?” in
2014
Later among the works it cites.
P. Swietojanski and S. Renals, “Learning hidden unit contributions for unsupervised speaker adaptation of neural network acoustic models,” in
2014
Later among the works it cites.
J. Xue, J. Li, D. Yu, M. Seltzer, and Y. Gong, “Singular value decomposition based low-footprint speaker adaptation and personalization for deep neural network,” in
2014
Later among the works it cites.
S. Xue, H. Jiang, and L. Dai, “Speaker adaptation of hybrid NN/HMM model for speech recognition based on singular value decomposition,” in
2014
Later among the works it cites.
Y. Tang, A. Mohan, R. C. Rose, and C. Ma, “Deep neural network trained with speaker representation for speaker normalization,” in
2014
Later among the works it cites.
Y. Miao, H. Zhang, and F. Metze, “Towards speaker adaptive training of deep neural network acoustic models,” in
2014
Later among the works it cites.
H. Zen and A. Senior, “Deep mixture density networks for acoustic modeling in statistical parametric speech synthesis,” in
2014
Later among the works it cites.
J. Ba and R. Caruana, “Do deep nets really need to be deep?” in
2014
Later among the works it cites.
J. Li, R. Zhao, J.-T. Huang, and Y. Gong, “Learning small-size DNN with output-distribution-based criteria,” in
2014
Later among the works it cites.
2014
Later among the works it cites.
M. Faruqui and C. Dyer, “Improving vector space word representations using multilingual correlation,” in
2014
Later among the works it cites.
K. M. Hermann and P. Blunsom, “Multilingual models for compositional distributed semantics,”
2014
Later among the works it cites.
2014
Later among the works it cites.
R. Kiros, R. Salakhutdinov, and R. Zemel, “Multimodal neural language models,” in
2014
Later among the works it cites.
R. Socher, A. Karpathy, Q. V. Le, C. D. Manning, and A. Y. Ng, “Grounded compositional semantics for finding and describing images with sentences,”
2014
Later among the works it cites.
M. Long, J. Wang, G. Ding, D. Shen, and Q. Yang, “Transfer learning with graph co-regularization,”
2014
Later among the works it cites.
J. Lu, V. Behbood, P. Hao, H. Zuo, S. Xue, and G. Zhang, “Transfer learning using computational intelligence: A survey,”
2015
Closest in time.
2015
Closest in time.
Z. Tang, D. Wang, Y. Pan, and Z. Zhang, “Knowledge transfer pre-training,”
2015
Closest in time.
J. Hirschberg and C. D. Manning, “Advances in natural language processing,”
2015
Closest in time.
Y. Bengio, I. J. Goodfellow, and A. Courville,
2015
Closest in time.
W. Zhang, R. Li, T. Zeng, Q. Sun, S. Kumar, J. Ye, and S. Ji, “Deep model based transfer and multi-task learning for biological image analysis,” in
2015
Closest in time.
X. Z. Zhiyuan Tang, “Speech recognition with pronunciation vecotrs,” CSLT, Tsinghua University, 2015. [Online]. Available:
2015
Closest in time.
M. Zhao, D. Wang, Z. Zhang, and X. Zhang, “Music removal by convolutional denoising autoencoder in speech recognition,” in
2015
Closest in time.
Y. Miao and F. Metze, “On speaker adaptation of long short-term memory recurrent neural networks,” in
2015
Closest in time.
K. Hashimoto, K. Oura, Y. Nankaku, and K. Tokuda, “The effect of neural networks in statistical parametric speech synthesis,” in
2015
Closest in time.
B. Potard, P. Motlicek, and D. Imseng, “Preliminary work on speaker adaptation for DNN-based speech synthesis,” Idiap, Tech. Rep., 2015
2015
Closest in time.
Z. Wu, P. Swietojanski, C. Veaux, S. Renals, and S. King, “A study of speaker adaptation for DNN-based speech synthesis,” in
2015
Closest in time.
W. Chan, N. R. Ke, and I. Lane, “Transferring knowledge from a RNN to a DNN,”
2015
Closest in time.
M. Long and J. Wang, “Learning transferable features with deep adaptation networks,”
2015
Closest in time.
Y. Lu, “Unsupervised learning of neural network outputs,”
2015
Closest in time.
N. Chen, Y. Qian, and K. Yu, “Multi-task learning for text-dependent speaker verification,” in
2015
Closest in time.
R. Fér, P. Matějka, F. Grézl, O. Plchot, and J. Černockỳ, “Multilingual bottleneck features for language recognition,” in
2015
Closest in time.
X. Ma, X. Wang, and D. Wang, “Recognize foreign low-frequency words with similar pairs,” in
2015
Closest in time.
C. Xing, D. Wang, C. Liu, and Y. Lin, “Normalized word embedding and orthogonal transform for bilingual word translation,” in
2015
Closest in time.