Fetching the paper…
Reading the bibliography…
We propose a technique that tackles action detection in multimodal videos under a realistic and challenging condition in which only limited training data and partially observed modalities are available.
Caruana, R.: Multitask learning. In: Learning to learn, pp. 95–133. Springer (1998)
1998
Earlier work this paper cites.
Noury, N., Fleury, A., Rumeau, P., Bourke, A., Laighin, G., Rialle, V., Lundy, J.: Fall detection-principles and methods. In: Engineering in Medicine and Biology Society (2007)
2007
Earlier work this paper cites.
Zach, C., Pock, T., Bischof, H.: A duality based approach for realtime tv-l 1 optical flow. Pattern Recognition pp. 214–223 (2007)
2007
Earlier work this paper cites.
Bengio, Y., Louradour, J., Collobert, R., Weston, J.: Curriculum learning. In: International Conference on Machine Learning (ICML) (2009)
2009
Earlier work this paper cites.
Deng, J., Dong, W., Socher, R., Li, L.J., Li, K., Fei-Fei, L.: Imagenet: A large-scale hierarchical image database. In: Computer Vision and Pattern Recognition (CVPR) (2009)
2009
Earlier work this paper cites.
Pan, S.J., Yang, Q.: A survey on transfer learning. IEEE Transactions on Knowledge and Data Engineering 22
2009
Earlier work this paper cites.
Vapnik, V., Vashist, A.: A new learning paradigm: Learning using privileged information. Neural networks 22
2009
Earlier work this paper cites.
Pan, S.J., Yang, Q.: A survey on transfer learning. IEEE Transactions on knowledge and data engineering 22
2010
Earlier work this paper cites.
Sung, J., Ponce, C., Selman, B., Saxena, A.: Human activity detection from rgbd images. In: AAAI workshop on Pattern, Activity and Intent Recognition (2011)
2011
Earlier work this paper cites.
Krizhevsky, A., Sutskever, I., Hinton, G.E.: Imagenet classification with deep convolutional neural networks. In: Advances in neural information processing systems (NIPS) (2012)
2012
Earlier work this paper cites.
Soomro, K., Zamir, A.R., Shah, M.: Ucf101: A dataset of 101 human actions classes from videos in the wild. CRCV-TR-12-01 (2012)
2012
Earlier work this paper cites.
Wang, J., Liu, Z., Wu, Y., Yuan, J.: Mining actionlet ensemble for action recognition with depth cameras. In: Computer Vision and Pattern Recognition (CVPR) (2012)
2012
Earlier work this paper cites.
Fernando, B., Habrard, A., Sebban, M., Tuytelaars, T.: Unsupervised visual domain adaptation using subspace alignment. In: International Conference on Computer Vision (ICCV). pp. 2960–2967 (2013)
2013
Earlier work this paper cites.
Koppula, H.S., Gupta, R., Saxena, A.: Learning human activities and object affordances from rgb-d videos. The International Journal of Robotics Research 32
2013
Earlier work this paper cites.
Ni, B., Wang, G., Moulin, P.: Rgbd-hudaact: A color-depth video database for human daily activity recognition. In: Consumer Depth Cameras for Computer Vision (2013)
2013
Earlier work this paper cites.
Chung, J., Gulcehre, C., Cho, K., Bengio, Y.: Empirical evaluation of gated recurrent neural networks on sequence modeling (2014)
2014
Earlier work this paper cites.
Jiang, L., Meng, D., Mitamura, T., Hauptmann, A.G.: Easy samples first: Self-paced reranking for zero-example multimedia search. In: MM (2014)
2014
Earlier work this paper cites.
Simonyan, K., Zisserman, A.: Two-stream convolutional networks for action recognition in videos. In: Advances in neural information processing systems (NIPS) (2014)
2014
Earlier work this paper cites.
Yosinski, J., Clune, J., Bengio, Y., Lipson, H.: How transferable are features in deep neural networks? In: Advances in neural information processing systems (NIPS) (2014)
2014
Earlier work this paper cites.
Chen, X., Gupta, A.: Webly supervised learning of convolutional networks. In: International Conference on Computer Vision (ICCV) (2015)
2015
Earlier work this paper cites.
Ding, Z., Shao, M., Fu, Y.: Missing modality transfer learning via latent low-rank constraint. IEEE Transactions on Image Processing 24
2015
Earlier work this paper cites.
Du, Y., Wang, W., Wang, L.: Hierarchical recurrent neural network for skeleton based action recognition. In: Computer Vision and Pattern Recognition (CVPR) (2015)
2015
Earlier work this paper cites.
Girshick, R.: Fast r-cnn. In: International Conference on Computer Vision (ICCV) (2015)
2015
Cited alongside, same era.
Gorban, A., Idrees, H., Jiang, Y., Zamir, A.R., Laptev, I., Shah, M., Sukthankar, R.: Thumos challenge: Action recognition with a large number of classes. In: Computer Vision and Pattern Recognition (CVPR) Workshop (2015)
2015
Cited alongside, same era.
Hinton, G., Vinyals, O., Dean, J.: Distilling the knowledge in a neural network. In: NIPS workshop (2015)
2015
Cited alongside, same era.
Kingma, P.K., Ba, J.: Adam: A method for stochastic optimization (2015)
2015
Cited alongside, same era.
Ren, S., He, K., Girshick, R., Sun, J.: Faster r-cnn: Towards real-time object detection with region proposal networks. In: Neural Information Processing Systems (NIPS) (2015)
2015
Cited alongside, same era.
2017
Closest in time.
Haque, A., Guo, M., Alahi, A., Yeung, S., Luo, Z., Rege, A., Jopling, J., Downing, N.L., Beninati, W., Singh, A., Platchek, T., Milstein, A., Fei-Fei, L.: Towards vision-based smart hospitals: A system for tracking and monitoring hand hygiene compliance. Proceedings of Machine Learning for Healthcare 2017 (2017)
2017
Closest in time.
2017
Closest in time.
Li, W., Chen, L., Xu, D., Gool, L.V.: Visual recognition in rgb images and videos by learning from rgb-d data. IEEE Transactions on Pattern Analysis and Machine Intelligence 40
2017
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Tran, D., Bourdev, L., Fergus, R., Torresani, L., Paluri, M.: Learning spatiotemporal features with 3d convolutional networks. In: International Conference on Computer Vision (ICCV) (2015)
2015
Cited alongside, same era.
Wang, Z., Ji, Q.: Classifier learning with hidden information. In: Computer Vision and Pattern Recognition (CVPR) (2015)
2015
Cited alongside, same era.
Zhang, Z., Conly, C., Athitsos, V.: A survey on vision-based fall detection. In: Conference on PErvasive Technologies Related to Assistive Environments (PETRA) (2015)
2015
Cited alongside, same era.
Escorcia, V., Heilbron, F.C., Niebles, J.C., Ghanem, B.: Daps: Deep action proposals for action understanding. In: European Conference on Computer Vision (ECCV) (2016)
2016
Cited alongside, same era.
Gupta, S., Hoffman, J., Malik, J.: Cross modal distillation for supervision transfer. In: Computer Vision and Pattern Recognition (CVPR) (2016)
2016
Cited alongside, same era.
He, K., Zhang, X., Ren, S., Sun, J.: Deep residual learning for image recognition. In: Computer Vision and Pattern Recognition (CVPR) (2016)
2016
Cited alongside, same era.
Hoffman, J., Gupta, S., Darrell, T.: Learning with side information through modality hallucination. In: Computer Vision and Pattern Recognition (CVPR) (2016)
2016
Cited alongside, same era.
Li, Y., Yang, J., Song, Y., Cao, L., Luo, J., Li, J.: Learning from noisy labels with distillation. In: International Conference on Computer Vision (ICCV) (2017)
2017
Closest in time.
2017
Closest in time.
2017
Closest in time.
Liu, J., Wang, G., Hu, P., Duan, L.Y., Kot, A.C.: Global context-aware attention lstm networks for 3d action recognition. In: Computer Vision and Pattern Recognition (CVPR) (2017)
2017
Closest in time.
Liu, M., Liu, H., Chen, C.: Enhanced skeleton visualization for view invariant human action recognition. Pattern Recognition 68
2017
Closest in time.
Luo, Z., Peng, B., Huang, D.A., Alahi, A., Fei-Fei, L.: Unsupervised learning of long-term motion dynamics for videos. In: Computer Vision and Pattern Recognition (CVPR) (2017)
2017
Closest in time.
Luo, Z., Zou, Y., Hoffman, J., Fei-Fei, L.: Label efficient learning of transferable representations across domains and tasks. In: Advances in neural information processing systems (NIPS) (2017)
2017
Closest in time.
Qin, Z., Shelton, C.R.: Event detection in continuous video: An inference in point process approach. IEEE Transactions on Image Processing 26
2017
Closest in time.
Shahroudy, A., Ng, T.T., Gong, Y., Wang, G.: Deep multimodal feature analysis for action recognition in rgb+ d videos. IEEE Transactions on Pattern Analysis and Machine Intelligence (TPAMI) (2017)
2017
Closest in time.
Shao, L., Cai, Z., Liu, L., Lu, K.: Performance evaluation of deep feature learning for rgb-d image/video classification. Information Sciences 385
2017
Closest in time.
Shi, Z., Kim, T.K.: Learning and refining of privileged information-based rnns for action recognition from depth sequences. In: Computer Vision and Pattern Recognition (CVPR) (2017)
2017
Closest in time.
Wang, H., Wang, L.: Learning robust representations using recurrent neural networks for skeleton based action classification and detection. In: International Conference on Multimedia & Expo Workshops (ICMEW) (2017)
2017
Closest in time.
Xu, D., Ouyang, W., Ricci, E., Wang, X., Sebe, N.: Learning cross-modal deep representations for robust pedestrian detection. In: Computer Vision and Pattern Recognition (CVPR) (2017)
2017
Closest in time.
Yang, H., Zhou, J.T., Cai, J., Ong, Y.S.: Miml-fcn+: Multi-instance multi-label learning via fully convolutional networks with privileged information. In: Computer Vision and Pattern Recognition (CVPR) (2017)
2017
Closest in time.
Yeung, S., Ramanathan, V., Russakovsky, O., Shen, L., Mori, G., Fei-Fei, L.: Learning to learn from noisy web videos (2017)
2017
Closest in time.
Zhang, S., Liu, X., Xiao, J.: On geometric features for skeleton-based action recognition using multilayer lstm networks. In: IEEE Winter Conference on Applications of Computer Vision (WACV) (2017)
2017
Closest in time.
2017
Closest in time.
Luo*, Z., Hsieh*, J.T., Balachandar, N., Yeung, S., Pusiol, G., Luxenberg, J., Li, G., Li, L.J., Downing, N.L., Milstein, A., Fei-Fei, L.: Computer vision-based descriptive analytics of seniors’ daily activities for long-term health monitoring. Machine Learning for Healthcare (MLHC) (2018)
2018
Closest in time.