J. Pearl, “Probabilistic reasoning in intelligent systems: Networks of plausible inference. morgan kauffman pub,” 1988
1988
Earlier work this paper cites.
T. Vámos, “Judea pearl: Probabilistic reasoning in intelligent systems,” Decision Support Systems , vol. 8, no. 1, pp. 73–75, 1992
1992
Earlier work this paper cites.
H. Schmid, “Part-of-speech tagging with neural networks,” in Proceedings of the 15th Conference on Computational Linguistics - Volume 1 , ser. COLING ’94. Stroudsburg, PA, USA: Association for Computational Linguistics, 1994, pp. 172–176. [Online]. Available: http://dx.doi.org/10.3115/991886.991915
1994
Earlier work this paper cites.
J. H. Searcy and J. C. Bartlett, “Inversion and processing of component and spatial-relational information in faces.” Journal of experimental psychology. Human perception and performance , vol. 22, no. 4, pp. 904–915, Aug. 1996
1996
Earlier work this paper cites.
Z. Ghahramani, G. E. Hinton et al. , “The em algorithm for mixtures of factor analyzers,” Technical Report CRG-TR-96-1, University of Toronto, Tech. Rep., 1996
1996
Earlier work this paper cites.
S. Hochreiter and J. Schmidhuber, “Long short-term memory,” Neural computation , vol. 9, no. 8, pp. 1735–1780, 1997
1997
Earlier work this paper cites.
Y. LeCun, L. Bottou, Y. Bengio, and P. Haffner, “Gradient-based learning applied to document recognition,” Proceedings of the IEEE , vol. 86, no. 11, pp. 2278–2324, 1998
1998
Earlier work this paper cites.
M. Jordan, Learning in Graphical Models , ser. Adaptive computation and machine learning. London, 1998. [Online]. Available: https://books.google.com/books?id=zac7L4LbNtUC
1998
Earlier work this paper cites.
M. I. Jordan and T. J. Sejnowski, Graphical models: Foundations of neural computation . MIT press, 2001
2001
Earlier work this paper cites.
F. R. Kschischang, B. J. Frey, and H.-A. Loeliger, “Factor graphs and the sum-product algorithm,” Information Theory, IEEE Transactions on , vol. 47, no. 2, pp. 498–519, 2001
2001
Earlier work this paper cites.
S. Roweis and Z. Ghahramani, “Learning nonlinear dynamical systems using the expectation–maximization algorithm,” Kalman filtering and neural networks , p. 175, 2001
2001
Earlier work this paper cites.
B. M. Wilamowski, S. Iplikci, O. Kaynak, and M. Ö. Efe, “An algorithm for fast convergence in training neural networks,” in Proceedings of the international joint conference on neural networks , vol. 2, 2001, pp. 1778–1782
2001
Earlier work this paper cites.
L. Breiman, “Random forests,” Machine learning , vol. 45, no. 1, pp. 5–32, 2001
2001
Earlier work this paper cites.
A. Jordan, “On discriminative vs. generative classifiers: A comparison of logistic regression and naive bayes,” Advances in neural information processing systems , vol. 14, p. 841, 2002
2002
Earlier work this paper cites.
R. Hartley and A. Zisserman, Multiple view geometry in computer vision . Cambridge university press, 2003
2003
Earlier work this paper cites.
D. Griffiths and M. Tenenbaum, “Hierarchical topic models and the nested chinese restaurant process,” Advances in neural information processing systems , vol. 16, p. 17, 2004
2004
Earlier work this paper cites.
A. Hyvärinen, J. Karhunen, and E. Oja, Independent component analysis . John Wiley & Sons, 2004, vol. 46
2004
Earlier work this paper cites.
P. F. Felzenszwalb and D. P. Huttenlocher, “Efficient belief propagation for early vision,” International journal of computer vision , vol. 70, no. 1, pp. 41–54, 2006
2006
Earlier work this paper cites.
C. M. Bishop et al. , Pattern recognition and machine learning . springer New York, 2006, vol. 4, no. 4
2006
Earlier work this paper cites.
L. Wiskott, “How does our visual system achieve shift and size invariance,” JL van Hemmen and TJ Sejnowski, editors , vol. 23, pp. 322–340, 2006
2006
Earlier work this paper cites.
C. M. Bishop, J. Lasserre et al. , “Generative or discriminative? getting the best of both worlds,” Bayesian Statistics , vol. 8, pp. 3–24, 2007
2007
Earlier work this paper cites.