Fetching the paper…
Reading the bibliography…
Recent achievements in language models have showcased their extraordinary capabilities in bridging visual information with semantic language understanding.
C. H. Lampert, H. Nickisch, and S. Harmeling, “Learning to detect unseen object classes by between-class attribute transfer,” in 2009 IEEE conference on computer vision and pattern recognition . IEEE, 2009, pp. 951–958
2009
Earlier work this paper cites.
W. Li, Z. Zhang, and Z. Liu, “Action recognition based on a bag of 3d points,” in 2010 IEEE computer society conference on computer vision and pattern recognition-workshops . IEEE, 2010, pp. 9–14
2010
Earlier work this paper cites.
H. Jhuang, J. Gall, S. Zuffi, C. Schmid, and M. J. Black, “Towards understanding action recognition,” in Proceedings of the IEEE international conference on computer vision , 2013, pp. 3192–3199
2013
Earlier work this paper cites.
J. Zhu, B. Wang, X. Yang, W. Zhang, and Z. Tu, “Action recognition with actons,” in Proceedings of the IEEE International Conference on Computer Vision , 2013, pp. 3559–3566
2013
Earlier work this paper cites.
X. Yang and Y. Tian, “Effective 3d action recognition using eigenjoints,” Journal of Visual Communication and Image Representation , vol. 25, no. 1, pp. 2–11, 2014
2014
Earlier work this paper cites.
J. Qin, L. Liu, L. Shao, F. Shen, B. Ni, J. Chen, and Y. Wang, “Zero-shot action recognition with error-correcting output codes,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2017, pp. 2833–2842
2017
Earlier work this paper cites.
J. Yang, H. Zou, H. Jiang, and L. Xie, “Device-free occupant activity sensing using wifi-enabled iot devices for smart homes,” IEEE Internet of Things Journal , vol. 5, no. 5, pp. 3991–4002, 2018
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
M. Wang, Y. D. Zhang, and G. Cui, “Human motion recognition exploiting radar with stacked recurrent neural network,” Digital Signal Processing , vol. 87, pp. 125–131, 2019
2019
Earlier work this paper cites.
D. Mandal, S. Narayan, S. K. Dwivedi, V. Gupta, S. Ahmed, F. S. Khan, and L. Shao, “Out-of-distribution detection for generalized zero-shot action recognition,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2019, pp. 9985–9993
2019
Earlier work this paper cites.
T. Brown, B. Mann, N. Ryder, M. Subbiah, J. D. Kaplan, P. Dhariwal, A. Neelakantan, P. Shyam, G. Sastry, A. Askell et al. , “Language models are few-shot learners,” Advances in neural information processing systems , vol. 33, pp. 1877–1901, 2020
2020
Earlier work this paper cites.
2020
Earlier work this paper cites.
R. Mishra, A. Gupta, H. P. Gupta, and T. Dutta, “A sensors based deep learning model for unseen locomotion mode identification using multiple semantic matrices,” IEEE Transactions on Mobile Computing , vol. 21, no. 3, pp. 799–810, 2020
2020
Earlier work this paper cites.
L. M. Dang, K. Min, H. Wang, M. J. Piran, C. H. Lee, and H. Moon, “Sensor-based and vision-based human activity recognition: A comprehensive survey,” Pattern Recognition , vol. 108, p. 107561, 2020
2020
Earlier work this paper cites.
F. Luo, S. Poslad, and E. Bodanese, “Temporal convolutional networks for multiperson activity recognition using a 2-d lidar,” IEEE Internet of Things Journal , vol. 7, no. 8, pp. 7432–7442, 2020
2020
Earlier work this paper cites.
D. Banerjee, S. Rani, A. M. George, A. Chowdhury, S. Dey, A. Mukherjee, T. Chakravarty, and A. Pal, “Application of spiking neural networks for action recognition from radar data,” in 2020 International Joint Conference on Neural Networks (IJCNN) . IEEE, 2020, pp. 1–10
2020
Earlier work this paper cites.
Q. Zhang, Z. Lei, Z. Zhang, and S. Z. Li, “Context-aware attention network for image-text retrieval,” in Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2020, pp. 3536–3545
2020
Earlier work this paper cites.
H. Chen, G. Ding, X. Liu, Z. Lin, J. Liu, and J. Han, “Imram: Iterative matching with recurrent attention memory for cross-modal image-text retrieval,” in Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2020, pp. 12 655–12 663
2020
Cited alongside, same era.
W. Xu, Y. Xian, J. Wang, B. Schiele, and Z. Akata, “Attribute prototype network for zero-shot learning,” Advances in Neural Information Processing Systems , vol. 33, pp. 21 969–21 980, 2020
2020
Cited alongside, same era.
A. Radford, J. W. Kim, C. Hallacy, A. Ramesh, G. Goh, S. Agarwal, G. Sastry, A. Askell, P. Mishkin, J. Clark et al. , “Learning transferable visual models from natural language supervision,” in International conference on machine learning . PMLR, 2021, pp. 8748–8763
2021
Cited alongside, same era.
F. Luo, S. Khan, Y. Huang, and K. Wu, “Binarized neural network for edge intelligence of sensor-based human activity recognition,” IEEE Transactions on Mobile Computing , 2021
2021
H. Luo, L. Ji, M. Zhong, Y. Chen, W. Lei, N. Duan, and T. Li, “Clip4clip: An empirical study of clip for end to end video clip retrieval and captioning,” Neurocomputing , vol. 508, pp. 293–304, 2022
2022
Later among the works it cites.
R. Zhang, Z. Guo, W. Zhang, K. Li, X. Miao, B. Cui, Y. Qiao, P. Gao, and H. Li, “Pointclip: Point cloud understanding by clip,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2022, pp. 8552–8562
2022
Later among the works it cites.
L. Chen, Y. Zhang, S. Miao, S. Zhu, R. Hu, L. Peng, and M. Lv, “Salience: An unsupervised user adaptation model for multiple wearable sensors based human activity recognition,” IEEE Transactions on Mobile Computing , 2022
2022
Later among the works it cites.
J. Yang, X. Chen, H. Zou, D. Wang, and L. Xie, “Autofi: Towards automatic wifi human sensing via geometric self-supervised learning,” IEEE Internet of Things Journal , 2022
2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
J. Roche, V. De-Silva, J. Hook, M. Moencks, and A. Kondoz, “A multimodal data processing system for lidar-based human activity recognition,” IEEE Transactions on Cybernetics , vol. 52, no. 10, pp. 10 027–10 040, 2021
2021
Cited alongside, same era.
C. Jia, Y. Yang, Y. Xia, Y.-T. Chen, Z. Parekh, H. Pham, Q. Le, Y.-H. Sung, Z. Li, and T. Duerig, “Scaling up visual and vision-language representation learning with noisy text supervision,” in International conference on machine learning . PMLR, 2021, pp. 4904–4916
2021
Cited alongside, same era.
2021
Cited alongside, same era.
S. Chen and D. Huang, “Elaborative rehearsal for zero-shot action recognition,” in Proceedings of the IEEE/CVF International Conference on Computer Vision , 2021, pp. 13 638–13 647
2021
Cited alongside, same era.
G. Bertasius, H. Wang, and L. Torresani, “Is space-time attention all you need for video understanding?” in ICML , vol. 2, no. 3, 2021, p. 4
2021
Cited alongside, same era.
H. Zhao, L. Jiang, J. Jia, P. H. Torr, and V. Koltun, “Point transformer,” in Proceedings of the IEEE/CVF international conference on computer vision , 2021, pp. 16 259–16 268
2021
Cited alongside, same era.
Y. Du, F. Wei, Z. Zhang, M. Shi, Y. Gao, and G. Li, “Learning to prompt for open-vocabulary object detection with vision-language model,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2022, pp. 14 084–14 093
2022
Cited alongside, same era.
M. Xu, Z. Zhang, F. Wei, Y. Lin, Y. Cao, H. Hu, and X. Bai, “A simple baseline for open-vocabulary semantic segmentation with pre-trained vision-language model,” in European Conference on Computer Vision . Springer, 2022, pp. 736–753
2022
Cited alongside, same era.
J. Yang, X. Chen, H. Zou, D. Wang, Q. Xu, and L. Xie, “Efficientfi: Toward large-scale lightweight wifi sensing via csi compression,” IEEE Internet of Things Journal , vol. 9, no. 15, pp. 13 086–13 095, 2022
2022
Later among the works it cites.
C.-C. Lin, K. Lin, L. Wang, Z. Liu, and L. Li, “Cross-modal representation learning for zero-shot action recognition,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2022, pp. 19 978–19 988
2022
Later among the works it cites.
K. Zhou, J. Yang, C. C. Loy, and Z. Liu, “Learning to prompt for vision-language models,” International Journal of Computer Vision , vol. 130, no. 9, pp. 2337–2348, 2022
2022
Later among the works it cites.
R. Taori, I. Gulrajani, T. Zhang, Y. Dubois, X. Li, C. Guestrin, P. Liang, and T. B. Hashimoto, “Alpaca: A strong, replicable instruction-following model,” Stanford Center for Research on Foundation Models. https://crfm. stanford. edu/2023/03/13/alpaca. html , vol. 3, no. 6, p. 7, 2023
2023
Closest in time.
S. Yun, S. H. Park, P. H. Seo, and J. Shin, “Ifseg: Image-free semantic segmentation via vision-language model,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2023, pp. 2967–2977
2023
Closest in time.
2023
Closest in time.
Y. Zhou, H. Huang, S. Yuan, H. Zou, L. Xie, and J. Yang, “Metafi++: Wifi-enabled transformer-based human pose estimation for metaverse avatar simulation,” IEEE Internet of Things Journal , 2023
2023
Closest in time.
F. Luo, S. Khan, A. Li, Y. Huang, and K. Wu, “Edgeactnet: Edge intelligence-enabled human activity recognition using radar point cloud,” IEEE Transactions on Mobile Computing , 2023
2023
Closest in time.
R. Girdhar, A. El-Nouby, Z. Liu, M. Singh, K. V. Alwala, A. Joulin, and I. Misra, “Imagebind: One embedding space to bind them all,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2023, pp. 15 180–15 190
2023
Closest in time.
J. Xu, S. Liu, A. Vahdat, W. Byeon, X. Wang, and S. De Mello, “Open-vocabulary panoptic segmentation with text-to-image diffusion models,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2023, pp. 2955–2966
2023
Closest in time.
2023
Closest in time.
J. Yang, H. Huang, Y. Zhou, X. Chen, Y. Xu, S. Yuan, H. Zou, C. X. Lu, and L. Xie, “Mm-fi: Multi-modal non-intrusive 4d human dataset for versatile wireless sensing,” in NeurIPS-23 Datasets and Benchmarks Track , 2023
2023
Closest in time.