Fetching the paper…
Reading the bibliography…
This paper introduces video domain generalization where most video classification networks degenerate due to the lack of exposure to the target domains of divergent distributions.
B. Schölkopf, R. C. Williamson, A. J. Smola, J. Shawe-Taylor, J. C. Platt et al. , “Support vector method for novelty detection.” in NIPS , vol. 12, 1999, pp. 582–588
1999
Earlier work this paper cites.
B. Jiang, M. Wang, W. Gan, W. Wu, and J. Yan, “STM: Spatiotemporal and motion encoding for action recognition,” in ICCV , 2019, pp. 2000–2009
2009
Earlier work this paper cites.
S. Ben-David, J. Blitzer, K. Crammer, A. Kulesza, F. Pereira, and J. W. Vaughan, “A theory of learning from different domains,” Machine Learning , pp. 151–175, 2010
2010
Earlier work this paper cites.
H. Kuehne, H. Jhuang, E. Garrote, T. Poggio, and T. Serre, “HMDB: A large video database for human motion recognition,” in ICCV , 2011, pp. 2556–2563
2011
Earlier work this paper cites.
K. Soomro, A. R. Zamir, and M. Shah, “UCF101: A dataset of 101 human actions classes from videos in the wild,” CoRR , 2012
2012
Earlier work this paper cites.
J. Yosinski, J. Clune, Y. Bengio, and H. Lipson, “How transferable are features in deep neural networks?” in NeurIPS , 2014, pp. 3320–3328
2014
Earlier work this paper cites.
K. Simonyan and A. Zisserman, “Two-stream convolutional networks for action recognition in videos,” in NeurIPS , 2014, pp. 568–576
2014
Earlier work this paper cites.
A. Karpathy, G. Toderici, S. Shetty, T. Leung, R. Sukthankar, and L. Fei-Fei, “Large-scale video classification with convolutional neural networks,” in CVPR , 2014, pp. 1725–1732
2014
Earlier work this paper cites.
M. Ghifary, W. B. Kleijn, M. Zhang, and D. Balduzzi, “Domain generalization for object recognition with multi-task autoencoders,” in ICCV , 2015, pp. 2551–2559
2015
Earlier work this paper cites.
D. Tran, L. Bourdev, R. Fergus, L. Torresani, and M. Paluri, “Learning spatiotemporal features with 3D convolutional networks,” in ICCV , 2015, pp. 4489–4497
2015
Earlier work this paper cites.
L. Wang, Y. Xiong, and et al., “Temporal segment networks: Towards good practices for deep action recognition,” in ECCV , 2016, pp. 20–36
2016
Earlier work this paper cites.
C. Feichtenhofer, A. Pinz, and R. Wildes, “Spatiotemporal residual networks for video action recognition,” in NeurIPS , 2016, pp. 3468–3476
2016
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in CVPR , 2016, pp. 770–778
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
A. Shahroudy, J. Liu, T.-T. Ng, and G. Wang, “NTU RGB+D: A large scale dataset for 3d human activity analysis,” in CVPR , 2016, pp. 1010–1019
2016
Earlier work this paper cites.
J. Carreira and A. Zisserman, “Quo vadis, action recognition? A new model and the Kinetics dataset,” in CVPR , 2017, pp. 4724–4733
2017
Earlier work this paper cites.
E. Tzeng, J. Hoffman, K. Saenko, and T. Darrell, “Adversarial discriminative domain adaptation,” in CVPR , 2017, pp. 2962–2971
2017
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin, “Attention is all you need,” in NeurIPS , 2017, pp. 5998–6008
2017
Earlier work this paper cites.
Z. Qiu, T. Yao, and T. Mei, “Learning spatio-temporal representation with pseudo-3D residual networks,” in ICCV , 2017, pp. 5534–5542
2017
Earlier work this paper cites.
R. Girdhar and D. Ramanan., “Attentional pooling for action recognition,” in NeurIPS , 2017, pp. 34–45
2017
Earlier work this paper cites.
2017
Cited alongside, same era.
2017
Cited alongside, same era.
X. Wang, R. Girshick, A. Gupta, and K. He, “Non-local neural networks,” in CVPR , 2018, pp. 7794–7803
2018
Cited alongside, same era.
B. Zhou, A. Andonian, A. Oliva, and A. Torralba, “Temporal relational reasoning in videos,” in ECCV , 2018, pp. 831–846
2018
Cited alongside, same era.
M. Long, Z. Cao, J. Wang, and M. I. Jordan, “Conditional adversarial domain adaptation,” in NeurIPS , 2018, pp. 1647–1657
C. Wu, C. Feichtenhofer, H. Fan, K. He, P. Krähenbühl, and R. B. Girshick, “Long-term feature banks for detailed video understanding,” in CVPR , 2019, pp. 284–293
2019
Closest in time.
J. Lin, C. Gan, and S. Han, “TSM: Temporal shift module for efficient video understanding,” in ICCV , 2019, pp. 7082–7092
2019
Closest in time.
D. Tran, H. Wang, L. Torresani, and M. Feiszli, “Video classification with channel-separated convolutional networks,” in ICCV , 2019, pp. 5551–5560
2019
Closest in time.
Y. Wang, L. Jiang, M.-H. Yang, L.-J. Li, M. Long, and L. Fei-Fei, “Eidetic 3D LSTM: A model for video prediction and beyond,” in ICLR , 2019
2019
Closest in time.
Y. Li, Y. Yang, W. Zhou, and T. M. Hospedales, “Feature-critic networks for heterogeneous domain generalization,” in ICML , 2019, pp. 3915–3924
2019
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2018
Cited alongside, same era.
A. Jamal, V. P. Namboodiri, D. Deodhare, and K. S. Venkatesh, “Deep domain adaptation in action space,” in BMVC , 2018, p. 264
2018
Cited alongside, same era.
R. Volpi, H. Namkoong, O. Sener, J. C. Duchi, V. Murino, and S. Savarese, “Generalizing to unseen domains via adversarial data augmentation,” in NeurIPS , 2018, pp. 5339–5349
2018
Cited alongside, same era.
S. Sun, Z. Kuang, W. Ouyang, L. Sheng, and W. Zhang, “Optical flow guided feature: A fast and robust motion representation for video action recognition,” in CVPR , 2018, pp. 1390–1399
2018
Cited alongside, same era.
S. Xie, C. Sun, J. Huang, Z. Tu, and K. Murphy, “Rethinking spatiotemporal feature learning for video understanding,” in ECCV , 2018
2018
Cited alongside, same era.
L. Zhang, G. Zhu, L. Mei, P. Shen, S. A. A. Shah, and M. Bennamoun, “Attention in convolutional LSTM for gesture recognition,” in NeurIPS , 2018, pp. 1323–1335
2018
Cited alongside, same era.
H. Li, S. Jialin Pan, S. Wang, and A. C. Kot, “Domain generalization with adversarial feature learning,” in CVPR , 2018, pp. 5400–5409
2018
Cited alongside, same era.
S. Shankar, V. Piratla, S. Chakrabarti, S. Chaudhuri, P. Jyothi, and S. Sarawagi, “Generalizing across domains via cross-gradient training,” in ICLR , 2018
2018
Cited alongside, same era.
D. Li, J. Zhang, Y. Yang, C. Liu, Y.-Z. Song, and T. M. Hospedales, “Episodic training for domain generalization,” in ICCV , 2019, pp. 1446–1455
2019
Closest in time.
R. Volpi and V. Murino, “Addressing model vulnerability to distributional shifts over image transformation sets,” in ICCV , 2019, pp. 7980–7989
2019
Closest in time.
P. T. Jackson, A. A. Abarghouei, S. Bonner, T. P. Breckon, and B. Obara, “Style augmentation: data augmentation via style randomization.” in CVPR Workshops , 2019, pp. 83–92
2019
Closest in time.
H. Zhang, Y. Yu, J. Jiao, E. P. Xing, L. E. Ghaoui, and M. I. Jordan, “Theoretically principled trade-off between robustness and accuracy,” in ICML , 2019, pp. 7472–7482
2019
Closest in time.
U. Ahsan, R. Madhok, and I. Essa, “Video jigsaw: Unsupervised learning of spatiotemporal context for video action recognition,” in WACV , 2019, pp. 179–189
2019
Closest in time.
R. Goyal, S. E. Kahou, V. Michalski, J. Materzynska, S. Westphal, H. Kim, V. Haenel, I. Fruend, P. Yianilos, M. Mueller-Freitag et al. , “The “Something Something” video database for learning and evaluating visual common sense,” in ICCV , 2017, pp. 5843–5851
2019
Closest in time.
C. Yang, Y. Xu, J. Shi, B. Dai, and B. Zhou, “Temporal pyramid network for action recognition,” in CVPR , 2020, pp. 591–600
2020
Closest in time.
F. Qiao, L. Zhao, and X. Peng, “Learning to learn single domain generalization,” in CVPR , 2020, pp. 12 556–12 565
2020
Closest in time.
B. Pan, Z. Cao, E. Adeli, and J. C. Niebles, “Adversarial cross-domain action recognition with co-attention.” in AAAI , 2020, pp. 11 815–11 822
2020
Closest in time.
J. Munro and D. Damen, “Multi-modal domain adaptation for fine-grained action recognition,” in CVPR , 2020, pp. 122–132
2020
Closest in time.
2020
Closest in time.
2021
Closest in time.
Y. Shu, Z. Cao, C. Wang, J. Wang, and M. Long, “Open domain generalization with domain-augmented meta-learning,” in CVPR , 2021
2021
Closest in time.
Y. Wang, M. Long, J. Wang, and P. S. Yu, “Spatiotemporal pyramid network for video action recognition.” in CVPR , 2017, pp. 2097–2106
2097
Closest in time.