Fetching the paper…
Reading the bibliography…
We present a method for gesture detection and localisation based on multi-scale and multi-modal deep learning.
Y. LeCun, L. Bottou, Y. Bengio, and P. Haffner, “Gradient-based learning applied to document recognition,” Proceedings of the IEEE , vol. 86, no. 11, pp. 2278–2324, 1998
1998
Earlier work this paper cites.
E. L. Lehmann, “Elements of Large-Sample Theory,” in ICML , 1998
1998
Earlier work this paper cites.
Y. LeCun, L. Bottou, Y. Bengio, and P. Haffner, “Gradient-based learning applied to document recognition,” in Proceedings of the IEEE , vol. 86(11), 1998, pp. 2278–2324
1998
Earlier work this paper cites.
L. A. Alexandre, A. C. Campilho, and M. Kamel, “On combining classifiers using sum and product rules,” in Pattern Recognition Letters , no. 22, 2001, pp. 1283–1289
2001
Earlier work this paper cites.
A. Lee, T. Kawahara, and K. Shikano, “Julius - an open source real-time large vocabulary recognition engine,” in Interspeech , 2001
2001
Earlier work this paper cites.
F. Bach, G. Lanckriet, and M. Jordan, “Multiple Kernel Learning, Conic Duality, and the SMO Algorithm,” in ICML , 2004
2004
Earlier work this paper cites.
P. Geurts, D. Ernst, and L. Wehenkel, “Extremely randomized trees,” in Machine learning, 63(1), 3-42 , 2006
2006
Earlier work this paper cites.
M. Ranzato, F. J. Huang, Y.-L. Boureau, and Y. LeCun, “Unsupervised Learning of Invariant Feature Hierarchies with Applications to Object Recognition,” in CVPR , 2007
2007
Earlier work this paper cites.
K. Fang, W. Gao, and D. Zhao, “Large-Vocabulary Continuous Sign Language Recognition Based on Transition-Movement Models,” in IEEE Transactions on Systems, Man, and Cybernetics , 2007
2007
Earlier work this paper cites.
P. Gehler and S. Nowozin, “On Feature Combination for Multiclass Object Classification,” in ICCV , 2009
2009
Earlier work this paper cites.
B. Chen, J.-A. Ting, B. Marlin, and N. de Freitas, “Deep learning of invariant Spatio-Temporal Features from Video,” in NIPSW , 2010
2010
Earlier work this paper cites.
J. Bergstra et al. , “Theano: A CPU and GPU Math Expression Compiler,” in Proceedings of the Scipy Conference , 2010
2010
Earlier work this paper cites.
J. Shotton, A. Fitzgibbon, M. Cook, T. Sharp, M. Finocchio, R. Moore, A. Kipman, and A. Blake, “Real-time human pose recognition in parts from single depth images,” in CVPR , 2011
2011
Earlier work this paper cites.
C. Keskin, F. Kiraç, Y. Kara, and L. Akarun, “Real time hand pose estimation using depth sensors,” in ICCV Workshop , 2011
2011
Earlier work this paper cites.
I. Oikonomidis, N. Kyriazis, and A. Argyros, “Efficient model-based 3D tracking of hand articulations using Kinect,” in BMVC , 2011
2011
Earlier work this paper cites.
Q. V. Le, W. Y. Zou, S. Y. Yeung, and A. Y. Ng, “Learning hierarchical invariant spatio-temporal features for action recognition with independent subspace analysis,” in CVPR , 2011
2011
Earlier work this paper cites.
J. Ngiam, A. Khosla, M. Kin, J. Nam, H. Lee, and A. Y. Ng, “Multimodal deep learning,” in ICML , 2011
2011
Earlier work this paper cites.
A. Krizhevsky, I. Sutskever, and G. Hinton, “ImageNet Classification with Deep Convolutional Neural Networks,” in NIPS , 2012
2012
Earlier work this paper cites.
M. Baccouche, F. Mamalet, C. Wolf, C. Garcia, and A. Baskurt, “Spatio-Temporal Convolutional Sparse Auto-Encoder for Sequence Classification,” in BMVC , 2012
2012
Earlier work this paper cites.
J. Wang, Z. Liu, Y. Wu, and J. Yuan, “Mining actionlet ensemble for action recognition with depth cameras,” in CVPR , 2012
2012
Earlier work this paper cites.
J. Sung, C. Ponce, B. Selman, and A. Saxena, “Unstructured Human Activity Detection from RGBD Images,” in ICRA , 2012
2012
Earlier work this paper cites.
G. Ye, D. Liu, I.-H. Jhuo, and S.-F. Chang, “Robust Late Fusion With Rank Minimization,” in CVPR , 2012
2012
Earlier work this paper cites.
P. Natarajan, S. Wu, S. Vitaladevuni, X. Zhuang, S. Tsakalidis, U. Park, R. Prasad, and P. Natarajan, “Multimodal Feature Fusion for Robust Event Detection in Web Videos,” in CVPR , 2012
2012
Cited alongside, same era.
2012
Cited alongside, same era.
C. Farabet, C. Couprie, L. Najman, and Y. LeCun, “Learning Hierarchical Features for Scene Labeling,” in PAMI , 2013
2013
Cited alongside, same era.
S. E. Kahou et al. , “Combining modality specific deep neural networks for emotion recognition in video,” in ICMI , 2013
2013
Cited alongside, same era.
F. Zhou, F. De la Torre, and J.-K. Hodgins, “Hierarchical Aligned Cluster Analysis for Temporal Clustering of Human Motion,” in PAMI , 2013
A. Karpathy, G. Toderici, S. Shetty, T. Leung, R. Sukthankar, and F.-F. Li, “Large-scale Video Classification with Convolutional Neural Networks,” in CVPR , 2014
2014
Closest in time.
2014
Closest in time.
A. Jain, J. Tompson, Y. LeCun, and C. Bregler, “MoDeep: A Deep Learning Framework Using Motion Features for Human Pose Estimation,” in ACCV , 2014
2014
Closest in time.
S. Escalera, X. Baró, J. Gonzàlez, M. Bautista, M. Madadi, M. Reyes, V. Ponce, H. Escalante, J. Shotton, and I. Guyon, “ChaLearn Looking at People Challenge 2014: Dataset and Results,” in ECCVW , 2014
2014
Closest in time.
D. Tang, H. J. Chang, A. Tejani, and T.-K. Kim, “Latent Regression Forest: Structured Estimation of 3D Articulated Hand Posture,” in CVPR , 2014
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2013
Cited alongside, same era.
H. Wang, A. Kläser, C. Schmid, and C.-L. Liu, “Dense trajectories and motion boundary descriptors for action recognition,” IJCV , 2013
2013
Cited alongside, same era.
D. Tang, T.-H. Yu, and T.-K. Kim, “Real-time Articulated Hand Pose Estimation using Semi-supervised Transductive Regression Forests,” in ICCV , 2013
2013
Cited alongside, same era.
F. Wang and Y. Li, “Beyond Physical Connections: Tree Models in Human Pose Estimation,” in CVPR , 2013
2013
Cited alongside, same era.
X. Chen and M. Koskela, “Online RGB-D gesture recognition with extreme learning machines,” in ICMI , 2013
2013
Cited alongside, same era.
K. Nandakumar et al. , “A Multi-modal Gesture Recognition System Using Audio, Video, and Skeletal Joint Data Categories and Subject Descriptors,” in ICMI Workshop , 2013
2013
Cited alongside, same era.
S. Ji, W. Xu, M. Yang, and K. Yu, “3D Convolutional Neural Networks for Human Action Recognition,” PAMI , 2013
2013
Cited alongside, same era.
N. Srivastava and R. Salakhutdinov, “Multimodal learning with Deep Boltzmann Machines,” in NIPS , 2013
2013
Cited alongside, same era.
2014
Closest in time.
J. Tompson, M. Stein, Y. LeCun, and K. Perlin, “Real-Time Continuous Pose Recovery of Human Hands Using Convolutional Networks,” in ACM Transaction on Graphics , 2014
2014
Closest in time.
N. Neverova, C. Wolf, G. Taylor, and F. Nebout, “Hand segmentation with structured convolutional learning,” in ACCV , 2014
2014
Closest in time.
C. Qian, X. Sun, Y. Wei, X. Tang, and J. Sun, “Realtime and Robust Hand Tracking from Depth,” in CVPR , 2014
2014
Closest in time.
X. Chen, R. Mottaghi, X. Liu, S. Fidler, R. Urtasun, and A. Yuille, “Detect What You Can: Detecting and Representing Objects using Holistic Models and Body Parts,” in CVPR , 2014
2014
Closest in time.
C. Monnier, S. German, and A. Ost, “A Multi-scale Boosted Detector for Efficient and Robust Gesture Recognition,” in ECCVW , 2014
2014
Closest in time.
J. Y. Chang, “Nonparametric Gesture Labeling from Multi-modal Data,” in ECCV Workshop , 2014
2014
Closest in time.
L. Pigou, S. Dieleman, and P.-J. Kindermans, “Sign Language Recognition Using Convolutional Neural Networks,” in ECCVW , 2014
2014
Closest in time.
D. Wu, “Deep Dynamic Neural Networks for Gesture Segmentation and Recognition,” in ECCV Workshop , 2014
2014
Closest in time.
Z. Wu, Y.-G. Jiang, J. Wang, J. Pu, and X. Xue, “Exploring Inter-feature and Inter-class Relationships with Deep Neural Networks for Video Classification,” in ACM Multimedia , 2014
2014
Closest in time.
A. Hernandez-Vela et al. , “Probability-based Dynamic Time Warping and Bag-of-Visual-and-Depth-Words for Human Gesture Recognition in RGB-D,” in Pattern Recognition Letters , 2014
2014
Closest in time.
N. Neverova, C. Wolf, G. Taylor, and F. Nebout, “Multi-scale deep learning for gesture detection and localization,” in ECCVW , 2014
2014
Closest in time.
P. Baldi and P. Sadowski, “The dropout learning algorithm,” Journal of Artificial Intelligence , vol. 210, pp. 78–122, 2014
2014
Closest in time.
N. Camgoz, A. Kindiroglu, and L. Akarun, “Gesture Recognition using Template Based Random Forest Classifiers,” in ECCVW , 2014
2014
Closest in time.
G. Evangelidis, G. Singh, and R. Horaud, “Continuous gesture recognition from articulated poses,” in ECCV Workshop , 2014
2014
Closest in time.
X. Peng, L. Wang, and Z. Cai, “Action and Gesture Temporal Spotting with Super Vector Representation,” in ECCVW , 2014
2014
Closest in time.
G. Chen, D. Clarke, M. Giuliani, D. Weikersdorfer, and A. Knoll, “Multi-modality Gesture Detection and Recognition With Un-supervision, Randomization and Discrimination,” in ECCVW , 2014
2014
Closest in time.