Fetching the paper…
Reading the bibliography…
Recognizing pedestrian attributes is an important task in the computer vision community due to it plays an important role in video surveillance.
D. R. Reddy et al. , “Speech understanding systems: A summary of results of the five-year research effort,” Department of Computer Science. Camegie-Mell University, Pittsburgh, PA , 1977
1977
Earlier work this paper cites.
S. Z. Li, “Markov random field models in computer vision,” in European conference on computer vision . Springer, 1994, pp. 361–370
1994
Earlier work this paper cites.
Y. LeCun, L. Bottou, Y. Bengio, and P. Haffner, “Gradient-based learning applied to document recognition,” Proceedings of the IEEE , vol. 86, no. 11, pp. 2278–2324, 1998
1998
Earlier work this paper cites.
R. A. Rensink, “The dynamic representation of scenes,” Visual cognition , vol. 7, no. 1-3, pp. 17–42, 2000
2000
Earlier work this paper cites.
J. Lafferty, A. McCallum, and F. C. Pereira, “Conditional random fields: Probabilistic models for segmenting and labeling sequence data,” 2001
2001
Earlier work this paper cites.
A. Clare and R. D. King, “Knowledge discovery in multi-label phenotype data,” in European Conference on Principles of Data Mining and Knowledge Discovery . Springer, 2001, pp. 42–53
2001
Earlier work this paper cites.
J. Lafferty, A. McCallum, and F. C. Pereira, “Conditional random fields: Probabilistic models for segmenting and labeling sequence data,” 2001
2001
Earlier work this paper cites.
A. Elisseeff and J. Weston, “A kernel method for multi-labelled classification,” in Advances in neural information processing systems , 2002, pp. 681–687
2002
Earlier work this paper cites.
M. F. Tappen and W. T. Freeman, “Comparison of graph cuts with belief propagation for stereo, using identical mrf parameters,” in null . IEEE, 2003, p. 900
2003
Earlier work this paper cites.
D. G. Lowe, “Distinctive image features from scale-invariant keypoints,” International journal of computer vision , vol. 60, no. 2, pp. 91–110, 2004
2004
Earlier work this paper cites.
M. R. Boutell, J. Luo, X. Shen, and C. M. Brown, “Learning multi-label scene classification,” Pattern recognition , vol. 37, no. 9, pp. 1757–1771, 2004
2004
Earlier work this paper cites.
N. Dalal and B. Triggs, “Histograms of oriented gradients for human detection,” in Computer Vision and Pattern Recognition, 2005. CVPR 2005. IEEE Computer Society Conference on , vol. 1. IEEE, 2005, pp. 886–893
2005
Earlier work this paper cites.
N. Ghamrawi and A. McCallum, “Collective multi-label classification,” in Proceedings of the 14th ACM international conference on Information and knowledge management . ACM, 2005, pp. 195–200
2005
Earlier work this paper cites.
M. Varma and A. Zisserman, “A statistical approach to texture classification from single images,” International journal of computer vision , vol. 62, no. 1-2, pp. 61–81, 2005
2005
Earlier work this paper cites.
S. M. Bileschi, “Streetscenes: Towards scene understanding in still images,” MASSACHUSETTS INST OF TECH CAMBRIDGE, Tech. Rep., 2006
2006
Earlier work this paper cites.
S. Lazebnik, C. Schmid, and J. Ponce, “Beyond bags of features: Spatial pyramid matching for recognizing natural scene categories,” in null . IEEE, 2006, pp. 2169–2178
2006
Earlier work this paper cites.
A. Graves, S. Fernández, F. Gomez, and J. Schmidhuber, “Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks,” in Proceedings of the 23rd international conference on Machine learning . ACM, 2006, pp. 369–376
2006
Earlier work this paper cites.
J. A. Tropp, A. C. Gilbert, and M. J. Strauss, “Algorithms for simultaneous sparse approximation. part i: Greedy pursuit,” Signal processing , vol. 86, no. 3, pp. 572–588, 2006
2006
Earlier work this paper cites.
M.-L. Zhang and Z.-H. Zhou, “Ml-knn: A lazy learning approach to multi-label learning,” Pattern recognition , vol. 40, no. 7, pp. 2038–2048, 2007
2007
Earlier work this paper cites.
S.-C. Zhu, D. Mumford et al. , “A stochastic grammar of images,” Foundations and Trends® in Computer Graphics and Vision , vol. 2, no. 4, pp. 259–362, 2007
2007
Earlier work this paper cites.
J. Fürnkranz, E. Hüllermeier, E. L. Mencía, and K. Brinker, “Multilabel classification via calibrated label ranking,” Machine learning , vol. 73, no. 2, pp. 133–153, 2008
2008
Earlier work this paper cites.
L. Bourdev and J. Malik, “Poselets: Body part detectors trained using 3d human pose annotations,” in Computer Vision, 2009 IEEE 12th International Conference on . IEEE, 2009, pp. 1365–1372
2009
Earlier work this paper cites.
A. Graves and J. Schmidhuber, “Offline handwriting recognition with multidimensional recurrent neural networks,” in Advances in neural information processing systems , 2009, pp. 545–552
2009
Earlier work this paper cites.
M. Everingham, L. Van Gool, C. K. I. Williams, J. Winn, and A. Zisserman, “The PASCAL Visual Object Classes Challenge 2010 (VOC2010) Results,” http://www.pascal-network.org/challenges/VOC/voc2010/workshop/index.html
2010
Earlier work this paper cites.
P. F. Felzenszwalb, R. B. Girshick, D. McAllester, and D. Ramanan, “Object detection with discriminatively trained part-based models,” IEEE transactions on pattern analysis and machine intelligence , vol. 32, no. 9, pp. 1627–1645, 2010
2010
Earlier work this paper cites.
L. Bourdev, S. Maji, T. Brox, and J. Malik, “Detecting people using mutually consistent poselet activations,” in European conference on computer vision . Springer, 2010, pp. 168–181
2010
Earlier work this paper cites.
M. P. Kumar, B. Packer, and D. Koller, “Self-paced learning for latent variable models,” in Advances in Neural Information Processing Systems , 2010, pp. 1189–1197
2010
Earlier work this paper cites.
M. Eichner, M. Marin-Jimenez, A. Zisserman, and V. Ferrari, “Articulated human pose estimation and search in (almost) unconstrained still images,” ETH Zurich, D-ITET, BIWI, Technical Report No , vol. 272, 2010
2010
Earlier work this paper cites.
C.-C. Chang and C.-J. Lin, “Libsvm: a library for support vector machines,” ACM transactions on intelligent systems and technology (TIST) , vol. 2, no. 3, p. 27, 2011
2011
Earlier work this paper cites.
L. Bourdev, S. Maji, and J. Malik, “Describing people: A poselet-based approach to attribute classification,” in Computer Vision (ICCV), 2011 IEEE International Conference on . IEEE, 2011, pp. 1543–1550
2011
Earlier work this paper cites.
G. Sharma and F. Jurie, “Learning discriminative spatial representation for image classification,” in BMVC 2011-British Machine Vision Conference . BMVA Press, 2011, pp. 1–11
2011
Earlier work this paper cites.
J. Read, B. Pfahringer, G. Holmes, and E. Frank, “Classifier chains for multi-label classification,” Machine learning , vol. 85, no. 3, p. 333, 2011
2011
Earlier work this paper cites.
G. Tsoumakas, I. Katakis, and I. Vlahavas, “Random k-labelsets for multilabel classification,” IEEE Transactions on Knowledge and Data Engineering , vol. 23, no. 7, pp. 1079–1089, 2011
2011
Earlier work this paper cites.
J. Liu, B. Kuipers, and S. Savarese, “Recognizing human actions by attributes,” in Computer Vision and Pattern Recognition (CVPR), 2011 IEEE Conference on . IEEE, 2011, pp. 3337–3344
2011
Earlier work this paper cites.
H. Chen, A. Gallagher, and B. Girod, “Describing clothing by semantic attributes,” in Computer Vision – ECCV 2012 , A. Fitzgibbon, S. Lazebnik, P. Perona, Y. Sato, and C. Schmid, Eds. Berlin, Heidelberg: Springer Berlin Heidelberg, 2012, pp. 609–623
2012
Earlier work this paper cites.
H. Chen, A. Gallagher, and B. Girod, “Describing clothing by semantic attributes,” in European conference on computer vision . Springer, 2012, pp. 609–623
2012
Earlier work this paper cites.
A. Geiger, P. Lenz, and R. Urtasun, “Are we ready for autonomous driving? the kitti vision benchmark suite,” in Computer Vision and Pattern Recognition (CVPR), 2012 IEEE Conference on . IEEE, 2012, pp. 3354–3361
2012
Earlier work this paper cites.
J. Yan, Z. Lei, D. Yi, and S. Z. Li, “Multi-pedestrian detection in crowded scenes: A global view,” in Computer Vision and Pattern Recognition (CVPR), 2012 IEEE Conference on . IEEE, 2012, pp. 3124–3129
2012
Earlier work this paper cites.
A. Krizhevsky, I. Sutskever, and G. E. Hinton, “Imagenet classification with deep convolutional neural networks,” in Advances in neural information processing systems , 2012, pp. 1097–1105
2012
Earlier work this paper cites.
R. Layne, T. M. Hospedales, and S. Gong, “Towards person identification and re-identification with attributes,” in European Conference on Computer Vision . Springer, 2012, pp. 402–412
2012
Earlier work this paper cites.
N. Silberman, D. Hoiem, P. Kohli, and R. Fergus, “Indoor segmentation and support inference from rgbd images,” in European Conference on Computer Vision . Springer, 2012, pp. 746–760
2012
Earlier work this paper cites.
R. Layne, T. M. Hospedales, and S. Gong, “Towards person identification and re-identification with attributes,” in European Conference on Computer Vision . Springer, 2012, pp. 402–412
2012
Earlier work this paper cites.
J. Joo, S. Wang, and S.-C. Zhu, “Human attribute recognition by rich appearance dictionary,” in Proceedings of the IEEE International Conference on Computer Vision , 2013, pp. 721–728
2013
Earlier work this paper cites.
J. Zhu, S. Liao, Z. Lei, D. Yi, and S. Li, “Pedestrian attribute classification in surveillance: Database and evaluation,” in Proceedings of the IEEE International Conference on Computer Vision Workshops , 2013, pp. 331–338
2013
Earlier work this paper cites.
M. Lin, Q. Chen, and S. Yan, “Network in network,” arXiv preprint arXiv:1312.4400 , 2013
2013
Earlier work this paper cites.
M. Lin, Q. Chen, and S. Yan, “Network in network,” arXiv preprint arXiv:1312.4400 , 2013
2013
Earlier work this paper cites.
X. Wang, T. Zhang, D. R. Tretter, and Q. Lin, “Personal clothing retrieval on photo collections by color and attributes,” IEEE Transactions on Multimedia , vol. 15, no. 8, pp. 2035–2045, 2013
2013
Earlier work this paper cites.
C. Thornton, F. Hutter, H. H. Hoos, and K. Leyton-Brown, “Auto-WEKA: Combined selection and hyperparameter optimization of classification algorithms,” in Proc. of KDD-2013 , 2013, pp. 847–855
2013
Earlier work this paper cites.
N. Zhang, M. Paluri, M. Ranzato, T. Darrell, and L. Bourdev, “Panda: Pose aligned networks for deep attribute modeling,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2014, pp. 1637–1644
2014
Earlier work this paper cites.
Y. Deng, P. Luo, C. C. Loy, and X. Tang, “Pedestrian attribute recognition at far distance,” in Proceedings of the 22nd ACM international conference on Multimedia . ACM, 2014, pp. 789–792
2014
Earlier work this paper cites.
M.-L. Zhang and Z.-H. Zhou, “A review on multi-label learning algorithms,” IEEE transactions on knowledge and data engineering , vol. 26, no. 8, pp. 1819–1837, 2014
2014
Earlier work this paper cites.
2014
Earlier work this paper cites.
2014
Earlier work this paper cites.
I. Goodfellow, J. Pouget-Abadie, M. Mirza, B. Xu, D. Warde-Farley, S. Ozair, A. Courville, and Y. Bengio, “Generative adversarial nets,” in Advances in neural information processing systems , 2014, pp. 2672–2680
2014
Earlier work this paper cites.
2014
Earlier work this paper cites.
2014
Earlier work this paper cites.
C. L. Zitnick and P. Dollár, “Edge boxes: Locating object proposals from edges,” in European conference on computer vision . Springer, 2014, pp. 391–405
2014
Earlier work this paper cites.
B. Zhou, A. Lapedriza, J. Xiao, A. Torralba, and A. Oliva, “Learning deep features for scene recognition using places database,” in Advances in neural information processing systems , 2014, pp. 487–495
2014
Earlier work this paper cites.
S. Khamis, C.-H. Kuo, V. K. Singh, V. D. Shet, and L. S. Davis, “Joint learning for attribute-consistent person re-identification,” in European Conference on Computer Vision . Springer, 2014, pp. 134–146
2014
Earlier work this paper cites.
R. Feris, R. Bobbitt, L. Brown, and S. Pankanti, “Attribute-based people search: Lessons learnt from a practical surveillance system,” in Proceedings of International Conference on Multimedia Retrieval . ACM, 2014, p. 153
2014
Earlier work this paper cites.
V. Mnih, N. Heess, A. Graves et al. , “Recurrent models of visual attention,” in Advances in neural information processing systems , 2014, pp. 2204–2212
2014
Earlier work this paper cites.
S. Gupta, R. Girshick, P. Arbeláez, and J. Malik, “Learning rich features from rgb-d images for object detection and segmentation,” in European Conference on Computer Vision . Springer, 2014, pp. 345–360
2014
Cited alongside, same era.
M. Danelljan, F. Shahbaz Khan, M. Felsberg, and J. Van de Weijer, “Adaptive color attributes for real-time visual tracking,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2014, pp. 1090–1097
2014
Cited alongside, same era.
P. Sudowe, H. Spitzer, and B. Leibe, “Person attribute recognition with a jointly-trained holistic cnn model,” in Proceedings of the IEEE International Conference on Computer Vision Workshops , 2015, pp. 87–95
2015
Cited alongside, same era.
D. Li, X. Chen, and K. Huang, “Multi-attribute learning for pedestrian attribute recognition in surveillance scenarios,” in Pattern Recognition (ACPR), 2015 3rd IAPR Asian Conference on . IEEE, 2015, pp. 111–115
2015
Cited alongside, same era.
K. He, Z. Wang, Y. Fu, R. Feng, Y.-G. Jiang, and X. Xue, “Adaptively weighted multi-task deep network for person attribute classification,” in Proceedings of the 2017 ACM on Multimedia Conference . ACM, 2017, pp. 1636–1644
2017
Later among the works it cites.
Y. Lu, A. Kumar, S. Zhai, Y. Cheng, T. Javidi, and R. Feris, “Fully-adaptive feature sharing in multi-task networks with applications in person attribute classification,” in CVPR , vol. 1, no. 2, 2017, p. 6
2017
Later among the works it cites.
M. Fabbri, S. Calderara, and R. Cucchiara, “Generative adversarial models for people attribute recognition in surveillance,” in Advanced Video and Signal Based Surveillance (AVSS), 2017 14th IEEE International Conference on . IEEE, 2017, pp. 1–6
2017
Later among the works it cites.
2017
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
A. H. Abdulnabi, G. Wang, J. Lu, and K. Jia, “Multi-task cnn model for attribute prediction,” IEEE Transactions on Multimedia , vol. 17, no. 11, pp. 1949–1959, 2015
2015
Cited alongside, same era.
J. Zhu, S. Liao, D. Yi, Z. Lei, and S. Z. Li, “Multi-label cnn based pedestrian attribute learning for soft biometrics,” in Biometrics (ICB), 2015 International Conference on . IEEE, 2015, pp. 535–540
2015
Cited alongside, same era.
G. Gkioxari, R. Girshick, and J. Malik, “Actions and attributes from wholes and parts,” in The IEEE International Conference on Computer Vision (ICCV) , December 2015
2015
Cited alongside, same era.
D. Hall and P. Perona, “Fine-grained classification of pedestrians in video: Benchmark and state of the art,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2015, pp. 5482–5491
2015
Cited alongside, same era.
Y. Xiong, K. Zhu, D. Lin, and X. Tang, “Recognize complex events from static images by fusing deep channels,” in 2015 IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , June 2015, pp. 1600–1609
2015
Cited alongside, same era.
L. Duong, T. Cohn, S. Bird, and P. Cook, “Low resource dependency parsing: Cross-lingual parameter sharing in a neural network parser,” in Proceedings of the 53rd Annual Meeting of the Association for Computational Linguistics and the 7th International Joint Conference on Natural Language Processing (Volume 2: Short Papers) , vol. 2, 2015, pp. 845–850
2015
Cited alongside, same era.
C. Szegedy, W. Liu, Y. Jia, P. Sermanet, S. Reed, D. Anguelov, D. Erhan, V. Vanhoucke, A. Rabinovich et al. , “Going deeper with convolutions.” Cvpr, 2015
2015
Cited alongside, same era.
C.-Y. Lee, S. Xie, P. Gallagher, Z. Zhang, and Z. Tu, “Deeply-supervised nets,” in Artificial Intelligence and Statistics , 2015, pp. 562–570
2015
Cited alongside, same era.
Later among the works it cites.
2017
Later among the works it cites.
2017
Later among the works it cites.
G. Huang, Z. Liu, K. Q. Weinberger, and L. van der Maaten, “Densely connected convolutional networks,” in Proceedings of the IEEE conference on computer vision and pattern recognition , vol. 1, no. 2, 2017, p. 3
2017
Later among the works it cites.
S. Sabour, N. Frosst, and G. E. Hinton, “Dynamic routing between capsules,” in Advances in Neural Information Processing Systems , 2017, pp. 3859–3869
2017
Later among the works it cites.
J. Gehring, M. Auli, D. Grangier, D. Yarats, and Y. N. Dauphin, “Convolutional sequence to sequence learning,” in International Conference on Machine Learning , 2017, pp. 1243–1252
2017
Later among the works it cites.
H.-S. Chang, E. Learned-Miller, and A. McCallum, “Active bias: Training more accurate neural networks by emphasizing high variance samples,” in Advances in Neural Information Processing Systems , 2017, pp. 1002–1012
2017
Later among the works it cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin, “Attention is all you need,” in Advances in Neural Information Processing Systems , 2017, pp. 5998–6008
2017
Later among the works it cites.
Q. Dong, S. Gong, and X. Zhu, “Multi-task curriculum transfer deep learning of clothing attributes,” in Applications of Computer Vision (WACV), 2017 IEEE Winter Conference on . IEEE, 2017, pp. 520–529
2017
Later among the works it cites.
N. Sarafianos, T. Giannakopoulos, C. Nikou, and I. A. Kakadiaris, “Curriculum learning for multi-task classification of visual attributes,” in Proceedings of the IEEE International Conference on Computer Vision , 2017, pp. 2608–2615
2017
Later among the works it cites.
J. Zhu, S. Liao, Z. Lei, and S. Z. Li, “Multi-label convolutional neural network based pedestrian attribute classification,” Image and Vision Computing , vol. 58, pp. 224–229, 2017
2017
Later among the works it cites.
D. Zhang, D. Meng, and J. Han, “Co-saliency detection via a self-paced multiple-instance learning framework,” IEEE transactions on pattern analysis and machine intelligence , vol. 39, no. 5, pp. 865–878, 2017
2017
Later among the works it cites.
Z. Wang, T. Chen, G. Li, R. Xu, and L. Lin, “Multi-label image recognition by recurrently discovering attentional regions,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2017, pp. 464–472
2017
Later among the works it cites.
2017
Later among the works it cites.
2017
Later among the works it cites.
L. Ma, X. Jia, Q. Sun, B. Schiele, T. Tuytelaars, and L. Van Gool, “Pose guided person image generation,” in Advances in Neural Information Processing Systems , 2017, pp. 406–416
2017
Later among the works it cites.
H. Zhang, T. Xu, H. Li, S. Zhang, X. Wang, X. Huang, and D. N. Metaxas, “Stackgan: Text to photo-realistic image synthesis with stacked generative adversarial networks,” in Proceedings of the IEEE International Conference on Computer Vision , 2017, pp. 5907–5915
2017
Later among the works it cites.
X. Wang, A. Shrivastava, and A. Gupta, “A-fast-rcnn: Hard positive generation via adversary for object detection,” in IEEE Conference on Computer Vision and Pattern Recognition , 2017
2017
Later among the works it cites.
Z. Zheng, L. Zheng, and Y. Yang, “Unlabeled samples generated by gan improve the person re-identification baseline in vitro,” in Proceedings of the IEEE International Conference on Computer Vision , 2017, pp. 3754–3762
2017
Later among the works it cites.
Q. Wu, C. Shen, P. Wang, A. Dick, and d. H. A. Van, “Image captioning and visual question answering based on attributes and external knowledge.” IEEE Transactions on Pattern Analysis and Machine Intelligence , vol. PP, no. 99, pp. 1–1, 2017
2017
Later among the works it cites.
C. Li, X. Sun, X. Wang, L. Zhang, and J. Tang, “Grayscale-thermal object tracking via multitask laplacian sparse representation,” IEEE Transactions on Systems, Man, and Cybernetics: Systems , vol. 47, no. 4, pp. 673–681, 2017
2017
Later among the works it cites.
C. Li, X. Wang, L. Zhang, J. Tang, H. Wu, and L. Lin, “Weighted low-rank decomposition for robust grayscale-thermal foreground detection,” IEEE Transactions on Circuits and Systems for Video Technology , vol. 27, no. 4, pp. 725–738, 2017
2017
Later among the works it cites.
A. Wu, W.-S. Zheng, H.-X. Yu, S. Gong, and J. Lai, “Rgb-infrared cross-modality person re-identification,” in 2017 IEEE International Conference on Computer Vision (ICCV) . IEEE, 2017, pp. 5390–5399
2017
Later among the works it cites.
B. Kresnaraman, Y. Kawanishi, D. Deguchi, T. Takahashi, Y. Mekada, I. Ide, and H. Murase, “Headgear recognition by decomposing human images in the thermal infrared spectrum,” in Quality in Research (QiR): International Symposium on Electrical and Computer Engineering, 2017 15th International Conference on . IEEE, 2017, pp. 164–168
2017
Later among the works it cites.
D. Li, X. Chen, Z. Zhang, and K. Huang, “Pose guided deep model for pedestrian attribute recognition in surveillance scenarios,” in 2018 IEEE International Conference on Multimedia and Expo (ICME) . IEEE, 2018, pp. 1–6
2018
Later among the works it cites.
P. Liu, X. Liu, J. Yan, and J. Shao, “Localization guided learning for pedestrian attribute recognition,” 2018
2018
Later among the works it cites.
N. Sarafianos, X. Xu, and I. A. Kakadiaris, “Deep imbalanced attribute classification using visual attention aggregation,” in European Conference on Computer Vision . Springer, 2018, pp. 708–725
2018
Later among the works it cites.
X. Zhao, L. Sang, G. Ding, Y. Guo, and X. Jin, “Grouping attribute recognition for pedestrian with joint recurrent learning.” in IJCAI , 2018, pp. 3177–3183
2018
Later among the works it cites.
2018
Later among the works it cites.
S. Park, B. X. Nie, and S.-C. Zhu, “Attribute and-or grammar for joint parsing of human pose, parts and attributes,” IEEE transactions on pattern analysis and machine intelligence , vol. 40, no. 7, pp. 1555–1569, 2018
2018
Later among the works it cites.
G. Hinton, N. Frosst, and S. Sabour, “Matrix capsules with em routing,” 2018
2018
Later among the works it cites.
2018
Later among the works it cites.
T. Yang and A. B. Chan, “Learning dynamic memory networks for object tracking,” in The European Conference on Computer Vision (ECCV) , September 2018
2018
Later among the works it cites.
C. Ma, C. Shen, A. Dick, Q. Wu, P. Wang, A. van den Hengel, and I. Reid, “Visual question answering with memory-augmented networks,” in The IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , June 2018
2018
Later among the works it cites.
T.-Y. Lin, P. Goyal, R. Girshick, K. He, and P. Dollár, “Focal loss for dense object detection,” IEEE transactions on pattern analysis and machine intelligence , 2018
2018
Later among the works it cites.
A. Mathews, L. Xie, and X. He, “Semstyle: Learning to generate stylised image captions using unaligned text,” in The IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , June 2018
2018
Later among the works it cites.
N. Sarafianos, T. Giannakopoulos, C. Nikou, and I. A. Kakadiaris, “Curriculum learning of visual attribute clusters for multi-task classification,” Pattern Recognition , vol. 80, pp. 94–108, 2018
2018
Later among the works it cites.
C. Li, F. Wei, J. Yan, X. Zhang, Q. Liu, and H. Zha, “A self-paced regularization framework for multilabel learning,” IEEE transactions on neural networks and learning systems , vol. 29, no. 6, pp. 2660–2666, 2018
2018
Later among the works it cites.
X. Liu, Y. Xu, L. Zhu, and Y. Mu, “A stochastic attribute grammar for robust cross-view human tracking,” IEEE transactions on circuits and systems for video technology , 2018
2018
Later among the works it cites.
2018
Later among the works it cites.
X. Qian, Y. Fu, T. Xiang, W. Wang, J. Qiu, Y. Wu, Y.-G. Jiang, and X. Xue, “Pose-normalized image generation for person re-identification,” in Proceedings of the European Conference on Computer Vision (ECCV) , 2018, pp. 650–667
2018
Later among the works it cites.
X. Wang, C. Li, B. Luo, and J. Tang, “Sint++: Robust visual tracking via adversarial positive instance generation,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2018, pp. 4864–4873
2018
Later among the works it cites.
2018
Later among the works it cites.
Y. He, J. Lin, Z. Liu, H. Wang, L.-J. Li, and S. Han, “Amc: Automl for model compression and acceleration on mobile devices,” in European Conference on Computer Vision . Springer, 2018, pp. 815–832
2018
Later among the works it cites.
X. L. L. L. ChenHan Jiang, Hang Xu, “Hybrid knowledge routed modules for large-scale object detection,” in NIPS , 2018
2018
Later among the works it cites.
2018
Later among the works it cites.
Q. L. X. Z. R. H. K. HUANG, “Visual-semantic graph reasoning for pedestrian attribute recognition,” in Association for the Advancement of Artificial Intelligence, AAAI , 2019
2019
Closest in time.
D. Li, Z. Zhang, X. Chen, and K. Huang, “A richly annotated pedestrian dataset for person retrieval in real surveillance scenarios,” IEEE transactions on image processing , vol. 28, no. 4, pp. 1575–1590, 2019
2019
Closest in time.
Z. Chen, A. Li, and Y. Wang, “A temporal attentive approach for video-based pedestrian attribute recognition,” in Pattern Recognition and Computer Vision: Second Chinese Conference, PRCV 2019, Xi’an, China, November 8–11, 2019, Proceedings, Part II 2 . Springer, 2019, pp. 209–220
2019
Closest in time.
G. D. J. H. N. D. Xin Zhao, Liufang Sang and C. Yan, “Recurrent attention model for pedestrian attribute recognition ∗ * ,” in Association for the Advancement of Artificial Intelligence, AAAI , 2019
2019
Closest in time.
2019
Closest in time.
2019
Closest in time.
2019
Closest in time.
T. Li, J. Liu, W. Zhang, Y. Ni, W. Wang, and Z. Li, “Uav-human: A large benchmark for human behavior understanding with unmanned aerial vehicles,” in Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2021, pp. 16 266–16 275
2021
Closest in time.
A. Specker, M. Cormier, and J. Beyerer, “Upar: Unified pedestrian attribute recognition and person retrieval,” in Proceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision , 2023, pp. 981–990
2023
Closest in time.
M. Jaderberg, K. Simonyan, A. Zisserman et al. , “Spatial transformer networks,” in Advances in neural information processing systems , 2015, pp. 2017–2025
2025
Closest in time.
O. Irsoy and C. Cardie, “Deep recursive neural networks for compositionality in language,” in Advances in neural information processing systems , 2014, pp. 2096–2104
2096
Closest in time.