Fetching the paper…
Reading the bibliography…
First person action recognition is an increasingly researched topic because of the growing popularity of wearable cameras.
Imagenet: A large-scale hierarchical image database
J. Deng, W. Dong, R. Socher, L. Li, Kai Li, and Li Fei-Fei · 2009
Earlier work this paper cites.
Unbiased look at dataset bias
Antonio Torralba and Alexei A Efros · 2011
Earlier work this paper cites.
Domain generalization via invariant feature representation
Krikamol Muandet, David Balduzzi, and Bernhard Schölkopf · 2013
Earlier work this paper cites.
Two-stream convolutional networks for action recognition in videos
Karen Simonyan and Andrew Zisserman · 2014
Earlier work this paper cites.
Unsupervised domain adaptation by backpropagation
Yaroslav Ganin and Victor Lempitsky · 2015
Earlier work this paper cites.
Batch normalization: Accelerating deep network training by reducing internal covariate shift
Sergey Ioffe and Christian Szegedy · 2015
Earlier work this paper cites.
Learning transferable features with deep adaptation networks
Mingsheng Long, Yue Cao, Jianmin Wang, and Michael Jordan · 2015
Earlier work this paper cites.
Learning spatiotemporal features with 3d convolutional networks
Du Tran, Lubomir Bourdev, Rob Fergus, Lorenzo Torresani, and Manohar Paluri · 2015
Earlier work this paper cites.
Soundnet: Learning sound representations from unlabeled video
Yusuf Aytar, Carl Vondrick, and Antonio Torralba · 2016
Earlier work this paper cites.
Out of time: automated lip sync in the wild
Joon Son Chung and Andrew Zisserman · 2016
Earlier work this paper cites.
Going deeper into first-person activity recognition
Minghuang Ma, Haoqi Fan, and Kris M Kitani · 2016
Earlier work this paper cites.
First person action recognition using deep learned descriptors
Suriya Singh, Chetan Arora, and CV Jawahar · 2016
Earlier work this paper cites.
Temporal segment networks: Towards good practices for deep action recognition
Limin Wang, Yuanjun Xiong, Zhe Wang, Yu Qiao, Dahua Lin, Xiaoou Tang, and Luc Van Gool · 2016
Earlier work this paper cites.
Look, listen and learn
Relja Arandjelovic and Andrew Zisserman · 2017
Earlier work this paper cites.
Quo vadis, action recognition? a new model and the kinetics dataset
Joao Carreira and Andrew Zisserman · 2017
Earlier work this paper cites.
Revisiting batch normalization for practical domain adaptation
Yanghao Li, Naiyan Wang, Jianping Shi, Jiaying Liu, and Xiaodi Hou · 2017
Earlier work this paper cites.
Convolutional long short-term memory networks for recognizing first person interactions
Swathikiran Sudhakaran and Oswald Lanz · 2017
Earlier work this paper cites.
Objects that sound
Relja Arandjelovic and Andrew Zisserman · 2018
Earlier work this paper cites.
Scaling egocentric vision: The epic-kitchens dataset
Dima Damen, Hazel Doughty, Giovanni Maria Farinella, Sanja Fidler, Antonino Furnari, Evangelos Kazakos, Davide Moltisanti, Jonathan Munro, Toby Perrett, Will Price, et al · 2018
Earlier work this paper cites.
Cycada: Cycle-consistent adversarial domain adaptation
Judy Hoffman, Eric Tzeng, Taesung Park, Jun-Yan Zhu, Phillip Isola, Kate Saenko, Alexei Efros, and Trevor Darrell · 2018
Earlier work this paper cites.
Deep domain adaptation in action space
Arshad Jamal, Vinay P Namboodiri, Dipti Deodhare, and KS Venkatesh · 2018
Earlier work this paper cites.
Cooperative learning of audio and video models from self-supervised synchronization
Bruno Korbar, Du Tran, and Lorenzo Torresani · 2018
Earlier work this paper cites.
Motion feature network: Fixed motion filter for action recognition
Myunggi Lee, Seungeui Lee, Sungjoon Son, Gyutae Park, and Nojun Kwak · 2018
Cited alongside, same era.
Domain generalization with adversarial feature learning
Haoliang Li, Sinno Jialin Pan, Shiqi Wang, and Alex C Kot · 2018
Cited alongside, same era.
Deep domain generalization via conditional invariant adversarial networks
Ya Li, Xinmei Tian, Mingming Gong, Yajing Liu, Tongliang Liu, Kun Zhang, and Dacheng Tao · 2018
Cited alongside, same era.
Adaptive batch normalization for practical domain adaptation
Yanghao Li, Naiyan Wang, Jianping Shi, Xiaodi Hou, and Jiaying Liu · 2018
Cited alongside, same era.
Audio-visual scene analysis with self-supervised multisensory features
Andrew Owens and Alexei A Efros · 2018
Cited alongside, same era.
Maximum classifier discrepancy for unsupervised domain adaptation
Kuniaki Saito, Kohei Watanabe, Yoshitaka Ushiku, and Tatsuya Harada · 2018
Epic-fusion: Audio-visual temporal binding for egocentric action recognition
Evangelos Kazakos, Arsha Nagrani, Andrew Zisserman, and Dima Damen · 2019
Later among the works it cites.
Tsm: Temporal shift module for efficient video understanding
Ji Lin, Chuang Gan, and Song Han · 2019
Later among the works it cites.
Transferable adversarial training: A general approach to adapting deep classifiers
Hong Liu, Mingsheng Long, Jianmin Wang, and Michael Jordan · 2019
Later among the works it cites.
Deep attention network for egocentric action recognition
M. Lu, Z. Li, Y. Wang, and G. Pan · 2019
Later among the works it cites.
Learning spatiotemporal attention for egocentric action recognition
Minlong Lu, Danping Liao, and Ze-Nian Li · 2019
Later among the works it cites.
Hierarchical feature aggregation networks for video action recognition
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Attention is all we need: Nailing down object-centric attention for egocentric activity recognition
Swathikiran Sudhakaran and Oswald Lanz · 2018
Cited alongside, same era.
Optical flow guided feature: A fast and robust motion representation for video action recognition
Shuyang Sun, Zhanghui Kuang, Lu Sheng, Wanli Ouyang, and Wei Zhang · 2018
Cited alongside, same era.
Generalizing to unseen domains via adversarial data augmentation
Riccardo Volpi, Hongseok Namkoong, Ozan Sener, John C Duchi, Vittorio Murino, and Silvio Savarese · 2018
Cited alongside, same era.
The sound of pixels
Hang Zhao, Chuang Gan, Andrew Rouditchenko, Carl Vondrick, Josh McDermott, and Antonio Torralba · 2018
Cited alongside, same era.
Temporal relational reasoning in videos
Bolei Zhou, Alex Andonian, Aude Oliva, and Antonio Torralba · 2018
Cited alongside, same era.
Domain generalization by solving jigsaw puzzles
Fabio M Carlucci, Antonio D’Innocente, Silvia Bucci, Barbara Caputo, and Tatiana Tommasi · 2019
Cited alongside, same era.
Swathikiran Sudhakaran, Sergio Escalera, and Oswald Lanz · 2019
Later among the works it cites.
Lsta: Long short-term attention for egocentric action recognition
Swathikiran Sudhakaran, Sergio Escalera, and Oswald Lanz · 2019
Later among the works it cites.
Long-term feature banks for detailed video understanding
Chao-Yuan Wu, Christoph Feichtenhofer, Haoqi Fan, Kaiming He, Philipp Krahenbuhl, and Ross Girshick · 2019
Later among the works it cites.
Larger norm more transferable: An adaptive feature norm approach for unsupervised domain adaptation
Ruijia Xu, Guanbin Li, Jihan Yang, and Liang Lin · 2019
Later among the works it cites.
Adversarial pyramid network for video domain generalization
Zhiyu Yao, Yunbo Wang, Xingqiang Du, Mingsheng Long, and Jianmin Wang · 2019
Later among the works it cites.
Dance with flow: Two-in-one stream action detection
Jiaojiao Zhao and Cees GM Snoek · 2019
Later among the works it cites.
Self-supervised learning of audio-visual objects from video
Triantafyllos Afouras, Andrew Owens, Joon Son Chung, and Andrew Zisserman · 2020
Later among the works it cites.
Unsupervised and semi-supervised domain adaptation for action recognition from drones
Jinwoo Choi, Gaurav Sharma, Manmohan Chandraker, and Jia-Bin Huang · 2020
Later among the works it cites.
Rolling-unrolling lstms for action anticipation from first-person video
Antonino Furnari and Giovanni Farinella · 2020
Later among the works it cites.
Listen to look: Action recognition by previewing audio
Ruohan Gao, Tae-Hyun Oh, Kristen Grauman, and Lorenzo Torresani · 2020
Later among the works it cites.
Multi-modal domain adaptation for fine-grained action recognition
Jonathan Munro and Dima Damen · 2020
Later among the works it cites.
Knowing what, where and when to look: Efficient video action modeling with attention
Juan-Manuel Perez-Rua, Brais Martinez, Xiatian Zhu, Antoine Toisoul, Victor Escorcia, and Tao Xiang · 2020
Later among the works it cites.
Mirco Planamente, Andrea Bottino, and Barbara Caputo · 2020
Later among the works it cites.
Discriminative adversarial domain adaptation
Hui Tang and Kui Jia · 2020
Later among the works it cites.
Symbiotic attention: Uts-baidu submission to the epic-kitchens 2020 action recognition challenge
Xiaohan Wang, Yu Wu, Linchao Zhu, Yi Yang, and Yueting Zhuang · 2020
Later among the works it cites.
Self-supervised learning across domains
Silvia Bucci, Antonio D’Innocente, Yujun Liao, Fabio Maria Carlucci, Barbara Caputo, and Tatiana Tommasi · 2021
Closest in time.