Fetching the paper…
Reading the bibliography…
Human action recognition is a challenging problem, particularly when there is high variability in factors such as subject appearance, backgrounds and viewpoint.
J. Hoffman, E. Tzeng, T. Park, J.-Y. Zhu, P. Isola, K. Saenko, A. Efros, and T. Darrell, “Cycada: Cycle-consistent adversarial domain adaptation,” in International conference on machine learning . Pmlr, 2018, pp. 1989–1998
1998
Earlier work this paper cites.
2007
Earlier work this paper cites.
K. Saenko, B. Kulis, M. Fritz, and T. Darrell, “Adapting visual category models to new domains,” in European conference on computer vision . Springer, 2010, pp. 213–226
2010
Earlier work this paper cites.
2012
Earlier work this paper cites.
K. Simonyan and A. Zisserman, “Two-stream convolutional networks for action recognition in videos,” Advances in neural information processing systems , vol. 27, 2014
2014
Earlier work this paper cites.
A. Karpathy, G. Toderici, S. Shetty, T. Leung, R. Sukthankar, and L. Fei-Fei, “Large-scale video classification with convolutional neural networks,” in Proceedings of the IEEE conference on Computer Vision and Pattern Recognition , 2014, pp. 1725–1732
2014
Earlier work this paper cites.
J. Yue-Hei Ng, M. Hausknecht, S. Vijayanarasimhan, O. Vinyals, R. Monga, and G. Toderici, “Beyond short snippets: Deep networks for video classification,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2015, pp. 4694–4702
2015
Earlier work this paper cites.
M. Long, Y. Cao, J. Wang, and M. I. Jordan, “Learning transferable features with deep adaptation networks,” in Proceedings of the 32nd International Conference on International Conference on Machine Learning - Volume 37 , ser. ICML’15. JMLR.org, 2015, p. 97–105
2015
Earlier work this paper cites.
F. Negin and F. Bremond, “Human action recognition in videos: A survey,” INRIA Technical Report , 2016
2016
Earlier work this paper cites.
Z. Zhang, H. Rebecq, C. Forster, and D. Scaramuzza, “Benefit of large field-of-view cameras for visual odometry,” in 2016 IEEE International Conference on Robotics and Automation (ICRA) , 2016, pp. 801–808
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
Y. Ganin, E. Ustinova, H. Ajakan, P. Germain, H. Larochelle, F. Laviolette, M. Marchand, and V. Lempitsky, “Domain-adversarial training of neural networks,” The journal of machine learning research , vol. 17, no. 1, pp. 2096–2030, 2016
2016
Earlier work this paper cites.
A. Gaidon, Q. Wang, Y. Cabon, and E. Vig, “Virtual worlds as proxy for multi-object tracking analysis,” in Proc. of CVPR’16 , 2016
2016
Earlier work this paper cites.
A. Shafaei and J. Little, “Real-time human motion capture with multiple depth cameras,” in Proc. of CRV’16 , 2016
2016
Earlier work this paper cites.
X. Peng, B. Usman, N. Kaushik, J. Hoffman, D. Wang, and K. Saenko, “Visda: The visual domain adaptation challenge,” 2017
2017
Earlier work this paper cites.
J. Carreira and A. Zisserman, “Quo vadis, action recognition? a new model and the kinetics dataset,” in proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2017, pp. 6299–6308
2017
Earlier work this paper cites.
R. Hou, C. Chen, and M. Shah, “Tube convolutional neural network (t-cnn) for action detection in videos,” in Proceedings of the IEEE international conference on computer vision , 2017, pp. 5822–5831
2017
Earlier work this paper cites.
S. Saha, G. Singh, and F. Cuzzolin, “Amtnet: Action-micro-tube regression by end-to-end trainable deep architecture,” in Proceedings of the IEEE International Conference on Computer Vision , 2017, pp. 4414–4423
2017
Earlier work this paper cites.
M. Barekatain, M. Martí, H.-F. Shih, S. Murray, K. Nakayama, Y. Matsuo, and H. Prendinger, “Okutama-action: An aerial view video dataset for concurrent human action detection,” in Proceedings of the IEEE conference on computer vision and pattern recognition workshops , 2017, pp. 28–35
2017
Earlier work this paper cites.
A. Dosovitskiy, G. Ros, F. Codevilla, A. Lopez, and V. Koltun, “CARLA: An open urban driving simulator,” in Proceedings of the 1st Annual Conference on Robot Learning , 2017, pp. 1–16
2017
Earlier work this paper cites.
C. de Souza, A. Gaidon, Y. Cabon, and A. López, “Procedural generation of videos to train deep action recognition networks,” in Proc. of CVPR’17 , 2017
2017
Cited alongside, same era.
V. Choutas, P. Weinzaepfel, J. Revaud, and C. Schmid, “Potion: Pose motion representation for action recognition,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2018, pp. 7024–7033
2018
Cited alongside, same era.
S. Sankaranarayanan, Y. Balaji, A. Jain, S. N. Lim, and R. Chellappa, “Learning from synthetic data: Addressing domain shift for semantic segmentation,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2018, pp. 3752–3761
2018
Cited alongside, same era.
A. G. Perera, Y. Wei Law, and J. Chahl, “Uav-gesture: A dataset for uav control and gesture recognition,” in Proceedings of the European Conference on Computer Vision (ECCV) Workshops , 2018, pp. 0–0
2018
Cited alongside, same era.
C. de Melo, B. Rothrock, O. U. P. Gurram, and B. Manjunath, “Vision-based gesture recognition in human-robot teams using synthetic data,” in Proc. of the IROS’20 , 2020
2020
Later among the works it cites.
M. Contributors, “Openmmlab’s next generation video understanding toolbox and benchmark,” https://github.com/open-mmlab/mmaction2 , 2020
2020
Later among the works it cites.
C. Feichtenhofer, “X3d: Expanding architectures for efficient video recognition,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2020, pp. 203–213
2020
Later among the works it cites.
E. D. Cubuk, B. Zoph, J. Shlens, and Q. Le, “Randaugment: Practical automated data augmentation with a reduced search space,” in Advances in Neural Information Processing Systems , H. Larochelle, M. Ranzato, R. Hadsell, M. Balcan, and H. Lin, Eds., vol. 33. Curran Associates, Inc., 2020, pp. 18 613–18 624. [Online]. Available: https://proceedings.neurips.cc/paper/2020/file/d85b63ef0ccb114d0a3bb7b7d808028f-Paper.pdf
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
S. Sankaranarayanan, Y. Balaji, C. D. Castillo, and R. Chellappa, “Generate to adapt: Aligning domains using generative adversarial networks,” in 2018 IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2018, pp. 8503–8512
2018
Cited alongside, same era.
K. Bousmalis, A. Irpan, P. Wohlhart, Y. Bai, M. Kelcey, M. Kalakrishnan, L. Downs, J. Ibarz, P. Pastor, K. Konolige, S. Levine, and V. Vanhoucke, “Using simulation and domain adaptation to improve efficiency of deep robotic grasping,” 2018 IEEE International Conference on Robotics and Automation (ICRA) , pp. 4243–4250, 2018
2018
Cited alongside, same era.
A. Yan, Y. Wang, Z. Li, and Y. Qiao, “Pa3d: Pose-action 3d machine for video recognition,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2019, pp. 7922–7931
2019
Cited alongside, same era.
X. Peng, Q. Bai, X. Xia, Z. Huang, K. Saenko, and B. Wang, “Moment matching for multi-source domain adaptation,” in Proceedings of the IEEE/CVF international conference on computer vision , 2019, pp. 1406–1415
2019
Cited alongside, same era.
S. Nikolenko, “Synthetic data for deep learning,” 2019
2019
Cited alongside, same era.
S. Wang, J. Yue, Y. Dong, S. He, H. Wang, and S. Ning, “A synthetic dataset for visual slam evaluation,” Robot. Auton. Syst. , vol. 124, no. C, feb 2020. [Online]. Available: https://doi.org/10.1016/j.robot.2019.103336
2019
Cited alongside, same era.
M. Savva, A. Kadian, O. Maksymets, Y. Zhao, E. Wijmans, B. Jain, J. Straub, J. Liu, V. Koltun, J. Malik, D. Parikh, and D. Batra, “Habitat: A Platform for Embodied AI Research,” in Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV) , 2019
2019
Cited alongside, same era.
M. Cao, X. Zhou, Y. Xu, Y. Pang, and B. Yao, “Adversarial domain adaptation with semantic consistency for cross-domain image classification,” in Proceedings of the 28th ACM International Conference on Information and Knowledge Management , ser. CIKM ’19. New York, NY, USA: Association for Computing Machinery, 2019, p. 259–268. [Online]. Available: https://doi.org/10.1145/3357384.3357918
2019
Cited alongside, same era.
2020
Later among the works it cites.
P. Weinzaepfel and G. Rogez, “Mimetics: Towards understanding human actions out of context,” International Journal of Computer Vision , vol. 129, no. 5, pp. 1675–1690, 2021
2021
Later among the works it cites.
C. M. de Melo, A. Torralba, L. Guibas, J. DiCarlo, R. Chellappa, and J. Hodgins, “Next-generation deep learning based on simulators and synthetic data,” Trends in Cognitive Sciences , vol. 26, pp. 174–187, 2021
2021
Later among the works it cites.
L. Zherdeva, E. Minaev, D. Zherdev, and V. Fursov, “Synthetic dataset for navigation tasks of autonomous systems and ground robots,” in 2021 International Conference on Information Technology and Nanotechnology (ITNT) , 2021, pp. 1–4
2021
Later among the works it cites.
2021
Later among the works it cites.
T. Li, J. Liu, W. Zhang, Y. Ni, W. Wang, and Z. Li, “Uav-human: A large benchmark for human behavior understanding with unmanned aerial vehicles,” in Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2021, pp. 16 266–16 275
2021
Later among the works it cites.
V. Makoviychuk, L. Wawrzyniak, Y. Guo, M. Lu, K. Storey, M. Macklin, D. Hoeller, N. Rudin, A. Allshire, A. Handa, and G. State, “Isaac gym: High performance gpu based physics simulation for robot learning,” in Proceedings of the Neural Information Processing Systems Track on Datasets and Benchmarks , J. Vanschoren and S. Yeung, Eds., vol. 1, 2021. [Online]. Available: https://datasets-benchmarks-proceedings.neurips.cc/paper/2021/file/28dd2c7955ce926456240b2ff0100bde-Paper-round2.pdf
2021
Later among the works it cites.
A. Szot, A. Clegg, E. Undersander, E. Wijmans, Y. Zhao, J. Turner, N. Maestre, M. Mukadam, D. Chaplot, O. Maksymets, A. Gokaslan, V. Vondrus, S. Dharur, F. Meier, W. Galuba, A. Chang, Z. Kira, V. Koltun, J. Malik, M. Savva, and D. Batra, “Habitat 2.0: Training home assistants to rearrange their habitat,” in Advances in Neural Information Processing Systems (NeurIPS) , 2021
2021
Later among the works it cites.
L. Weihs, M. Deitke, A. Kembhavi, and R. Mottaghi, “Visual room rearrangement,” in IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , June 2021
2021
Later among the works it cites.
C. Li, F. Xia, R. Martín-Martín, M. Lingelbach, S. Srivastava, B. Shen, K. E. Vainio, C. Gokmen, G. Dharan, T. Jain, A. Kurenkov, K. Liu, H. Gweon, J. Wu, L. Fei-Fei, and S. Savarese, “igibson 2.0: Object-centric simulation for robot learning of everyday household tasks,” in 5th Annual Conference on Robot Learning , 2021. [Online]. Available: https://openreview.net/forum?id=2uGN5jNJROR
2021
Later among the works it cites.
D. Ho, K. Rao, Z. Xu, E. Jang, M. Khansari, and Y. Bai, “Retinagan: An object-aware approach to sim-to-real transfer,” in 2021 IEEE International Conference on Robotics and Automation (ICRA) , 2021, pp. 10 920–10 926
2021
Later among the works it cites.
“Visual signals: Field manual 21-60,” 1987. [Online]. Available: https://www.radford.edu/content/dam/colleges/chbs/rotc/Forms/fm/Visual%20Signals%20FM%2021-60.pdf
2021
Later among the works it cites.
2022
Later among the works it cites.
K. Dimitropoulos, I. Hatzilygeroudis, and K. Chatzilygeroudis, “A brief survey of sim2real methods for robot learning,” in Advances in Service and Industrial Robotics , A. Müller and M. Brandstötter, Eds. Cham: Springer International Publishing, 2022, pp. 133–140
2022
Later among the works it cites.
K. Weerakoon, A. Sathyamoorthy, and D. Manocha, “Sim-to-real strategy for spatially aware robot navigation in uneven outdoor environments,” 05 2022
2022
Later among the works it cites.
V. G. T. da Costa, G. Zara, P. Rota, T. Oliveira-Santos, N. Sebe, V. Murino, and E. Ricci, “Dual-head contrastive domain adaptation for video action recognition,” in Proceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision , 2022, pp. 1181–1190
2022
Later among the works it cites.