Fetching the paper…
Reading the bibliography…
Visual semantic information comprises two important parts: the meaning of each visual semantic unit and the coherent visual semantic relation conveyed by these visual semantic units.
C. Liu, J. Yuen, and A. Torralba, “Nonparametric scene parsing: Label transfer via dense scene alignment,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2009, pp. 1972–1979
1979
Earlier work this paper cites.
R. Battiti, “First-and second-order methods for learning: between steepest descent and newton’s method,” Neural Computation , vol. 4, no. 2, pp. 141–166, 1992
1992
Earlier work this paper cites.
A. Georges, G. Kotliar, W. Krauth, and M. J. Rozenberg, “Dynamical mean-field theory of strongly correlated fermion systems and the limit of infinite dimensions,” Reviews of Modern Physics , vol. 68, no. 1, p. 13, 1996
1996
Earlier work this paper cites.
A.-L. Barabási, R. Albert, and H. Jeong, “Mean-field theory for scale-free random networks,” Physica A: Statistical Mechanics and its Applications , vol. 272, no. 1-2, pp. 173–187, 1999
1999
Earlier work this paper cites.
K. P. Murphy, Y. Weiss, and M. I. Jordan, “Loopy belief propagation for approximate inference: An empirical study,” in Proceedings of the Fifteenth Conference on Uncertainty in Artificial Intelligence (UAI) , 1999, pp. 467–475
1999
Earlier work this paper cites.
J. D. Lafferty, A. McCallum, and F. C. N. Pereira, “Conditional random fields: Probabilistic models for segmenting and labeling sequence data,” in Proceedings of the Eighteenth International Conference on Machine Learning (ICML) , 2001, pp. 282–289
2001
Earlier work this paper cites.
J. S. Yedidia, W. T. Freeman, and Y. Weiss, “Generalized belief propagation,” in Advances in Neural Information Processing Systems (NIPS) , 2001, pp. 689–695
2001
Earlier work this paper cites.
M. J. Wainwright, T. S. Jaakkola, and A. S. Willsky, “A new class of upper bounds on the log partition function,” IEEE Transactions on Information Theory , vol. 51, no. 7, pp. 2313–2335, 2005
2005
Earlier work this paper cites.
T. Werner, “A linear programming approach to max-sum problem: A review,” IEEE Transactions on Pattern Analysis and Machine Intelligence , vol. 29, no. 7, pp. 1165–1179, 2007
2007
Earlier work this paper cites.
M. Everingham, L. Van Gool, C. K. Williams, J. Winn, and A. Zisserman, “The pascal visual object classes challenge 2007 (voc 2007) results (2007),” 2008
2008
Earlier work this paper cites.
S. Ross, D. Munoz, M. Hebert, and J. A. Bagnell, “Learning message-passing inference machines for structured prediction,” in Proceedings of IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2011, pp. 2737–2744
2011
Earlier work this paper cites.
P. Krähenbühl and V. Koltun, “Efficient inference in fully connected crfs with gaussian edge potentials,” in Advances in Neural Information Processing Systems (NIPS) , 2011, pp. 109–117
2011
Earlier work this paper cites.
X. Glorot, A. Bordes, and Y. Bengio, “Domain adaptation for large-scale sentiment classification: A deep learning approach,” in Proceedings of the 28th International Conference on Machine Learning (ICML) , 2011, pp. 513–520
2011
Earlier work this paper cites.
S. J. Pan, I. W. Tsang, J. T. Kwok, and Q. Yang, “Domain adaptation via transfer component analysis,” IEEE Transactions on Neural Networks , vol. 22, no. 2, pp. 199–210, 2011
2011
Earlier work this paper cites.
A. Krizhevsky, I. Sutskever, and G. E. Hinton, “Imagenet classification with deep convolutional neural networks,” in Advances in Neural Information Processing Systems (NIPS) , 2012, pp. 1097–1105
2012
Earlier work this paper cites.
B. Alexe, T. Deselaers, and V. Ferrari, “Measuring the objectness of image windows,” IEEE Transactions on Pattern Analysis and Machine Intelligence , vol. 34, no. 11, pp. 2189–2202, 2012
2012
Earlier work this paper cites.
Q. Liu and A. Ihler, “Variational algorithms for marginal map,” The Journal of Machine Learning Research , vol. 14, no. 1, pp. 3165–3200, 2013
2013
Earlier work this paper cites.
2013
Earlier work this paper cites.
T. Mikolov, I. Sutskever, K. Chen, G. S. Corrado, and J. Dean, “Distributed representations of words and phrases and their compositionality,” in Advances in Neural Information Processing Systems (NIPS) , 2013, pp. 3111–3119
2013
Earlier work this paper cites.
J. Kappes, B. Andres, F. Hamprecht, C. Schnorr, S. Nowozin, D. Batra, S. Kim, B. Kausler, J. Lellmann, N. Komodakis et al. , “A comparative study of modern inference techniques for discrete energy minimization problems,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2013, pp. 1328–1335
2013
Earlier work this paper cites.
R. Socher, M. Ganjoo, C. D. Manning, and A. Ng, “Zero-shot learning through cross-modal transfer,” in Advances in Neural Information Processing Systems (NIPS) , 2013, pp. 935–943
2013
Earlier work this paper cites.
C. Farabet, C. Couprie, L. Najman, and Y. LeCun, “Learning hierarchical features for scene labeling,” IEEE Transactions on Pattern Analysis and Machine Intelligence , vol. 35, no. 8, pp. 1915–1929, 2013
2013
Earlier work this paper cites.
J. Tighe and S. Lazebnik, “Finding things: Image parsing with regions and per-exemplar detectors,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2013, pp. 3001–3008
2013
Earlier work this paper cites.
2014
Earlier work this paper cites.
2014
Earlier work this paper cites.
T.-Y. Lin, M. Maire, S. Belongie, J. Hays, P. Perona, D. Ramanan, P. Dollár, and C. L. Zitnick, “Microsoft coco: Common objects in context,” in Proceedings of European Conference on Computer Vision (ECCV) , 2014, pp. 740–755
2014
Earlier work this paper cites.
R. Mottaghi, X. Chen, X. Liu, N.-G. Cho, S.-W. Lee, S. Fidler, R. Urtasun, and A. Yuille, “The role of context for object detection and semantic segmentation in the wild,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2014, pp. 891–898
2014
Earlier work this paper cites.
P. H. Pinheiro and R. Collobert, “Recurrent convolutional neural networks for scene labeling,” in Proceedings of International Conference on Machine Learning (ICML) , 2014, pp. 82–90
2014
Earlier work this paper cites.
A. Sharma, O. Tuzel, and M.-Y. Liu, “Recursive context propagation network for semantic scene labeling,” in Advances in Neural Information Processing Systems (NIPS) , 2014, pp. 2447–2455
2014
Earlier work this paper cites.
J. Yang, B. Price, S. Cohen, and M.-H. Yang, “Context driven scene parsing with attention to rare classes,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2014, pp. 3294–3301
2014
Earlier work this paper cites.
S. Zheng, S. Jayasumana, B. Romera-Paredes, V. Vineet, Z. Su, D. Du, C. Huang, and P. H. Torr, “Conditional random fields as recurrent neural networks,” in Proceedings of the IEEE International Conference on Computer Vision (ICCV) , 2015, pp. 1529–1537
2015
Earlier work this paper cites.
A. Karpathy and L. Fei-Fei, “Deep visual-semantic alignments for generating image descriptions,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2015, pp. 3128–3137
2015
Earlier work this paper cites.
J. Long, E. Shelhamer, and T. Darrell, “Fully convolutional networks for semantic segmentation,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2015, pp. 3431–3440
2015
Earlier work this paper cites.
S. Ren, K. He, R. Girshick, and J. Sun, “Faster r-cnn: Towards real-time object detection with region proposal networks,” in Advances in Neural Information Processing Systems (NIPS) , 2015, pp. 91–99
2015
Earlier work this paper cites.
R. Girshick, “Fast r-cnn,” in Proceedings of the IEEE International Conference on Computer Vision (ICCV) , 2015, pp. 1440–1448
2015
Earlier work this paper cites.
O. Russakovsky, J. Deng, H. Su, J. Krause, S. Satheesh, S. Ma, Z. Huang, A. Karpathy, A. Khosla, M. Bernstein et al. , “Imagenet large scale visual recognition challenge,” International Journal of Computer Vision , vol. 115, no. 3, pp. 211–252, 2015
2015
Cited alongside, same era.
I. K. K. M. L.-C. Chen, G. Papandreou and A. L. Yuille, “Semantic image segmentation with deep convolutional nets and fully connected crfs,” in Proceedings of the International Conference on Learning Representations (ICLR) , 2015
2015
Cited alongside, same era.
B. Romera-Paredes and P. Torr, “An embarrassingly simple approach to zero-shot learning,” in Proceedings of International Conference on Machine Learning (ICML) , 2015, pp. 2152–2161
2015
Cited alongside, same era.
2015
Cited alongside, same era.
X. He and L. Deng, “Deep learning for image-to-text generation: a technical overview,” IEEE Signal Processing Magazine , vol. 34, no. 6, pp. 109–116, 2017
2017
Later among the works it cites.
K. Fu, J. Jin, R. Cui, F. Sha, and C. Zhang, “Aligning where to see and what to tell: Image captioning with region-based attention and scene-specific contexts,” IEEE Transactions on Pattern Analysis and Machine Intelligence , vol. 39, no. 12, pp. 2321–2334, 2017
2017
Later among the works it cites.
M. Malinowski, M. Rohrbach, and M. Fritz, “Ask your neurons: A deep learning approach to visual question answering,” International Journal of Computer Vision , vol. 125, no. 1-3, pp. 110–135, 2017
2017
Later among the works it cites.
A. Agrawal, J. Lu, S. Antol, M. Mitchell, C. L. Zitnick, D. Parikh, and D. Batra, “Vqa: Visual question answering,” International Journal of Computer Vision , vol. 123, no. 1, pp. 4–31, 2017
2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
H. Bilen, M. Pedersoli, and T. Tuytelaars, “Weakly supervised object detection with convex clustering,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2015, pp. 1081–1089
2015
Cited alongside, same era.
Y. Ganin and V. Lempitsky, “Unsupervised domain adaptation by backpropagation,” in Proceedings of International Conference on Machine Learning (ICML) , 2015, pp. 1180–1189
2015
Cited alongside, same era.
V. M. Patel, R. Gopalan, R. Li, and R. Chellappa, “Visual domain adaptation: A survey of recent advances,” IEEE Signal Processing Magazine , vol. 32, no. 3, pp. 53–69, 2015
2015
Cited alongside, same era.
J. Dai, K. He, and J. Sun, “Convolutional feature masking for joint object and stuff segmentation,” in Proceedings of IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2015, pp. 3992–4000
2015
Cited alongside, same era.
W. Byeon, T. M. Breuel, F. Raue, and M. Liwicki, “Scene labeling with lstm recurrent neural networks,” in Proceedings of IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2015, pp. 3547–3555
2015
Cited alongside, same era.
C. Lu, R. Krishna, M. Bernstein, and L. Fei-Fei, “Visual relationship detection with language priors,” in Proceedings of European Conference on Computer Vision (ECCV) . Springer, 2016, pp. 852–869
2016
Cited alongside, same era.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2016, pp. 770–778
2016
Cited alongside, same era.
B. Shuai, Z. Zuo, B. Wang, and G. Wang, “Dag-recurrent neural networks for scene labeling,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2016, pp. 3620–3629
2016
Cited alongside, same era.
D. Teney, Q. Wu, and A. van den Hengel, “Visual question answering: A tutorial,” IEEE Signal Processing Magazine , vol. 34, no. 6, pp. 63–75, 2017
2017
Later among the works it cites.
2017
Later among the works it cites.
Y. Wei, X. Liang, Y. Chen, X. Shen, M.-M. Cheng, J. Feng, Y. Zhao, and S. Yan, “Stc: A simple to complex framework for weakly-supervised semantic segmentation,” IEEE Transactions on Pattern Analysis and Machine Intelligence , vol. 39, no. 11, pp. 2314–2320, 2017
2017
Later among the works it cites.
R. G. Cinbis, J. Verbeek, and C. Schmid, “Weakly supervised object localization with multi-fold multiple instance learning,” IEEE Transactions on Pattern Analysis and Machine Intelligence , vol. 39, no. 1, pp. 189–203, 2017
2017
Later among the works it cites.
K. T. Schütt, F. Arbabzadah, S. Chmiela, K. R. Müller, and A. Tkatchenko, “Quantum-chemical insights from deep tensor neural networks,” Nature Communications , vol. 8, p. 13890, 2017
2017
Later among the works it cites.
J. Gilmer, S. S. Schoenholz, P. F. Riley, O. Vinyals, and G. E. Dahl, “Neural message passing for quantum chemistry,” in Proceedings of International Conference on Machine Learning (ICML) , 2017, pp. 1263–1272
2017
Later among the works it cites.
J. Redmon and A. Farhadi, “Yolo9000: Better, faster, stronger,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2017, pp. 6517–6525
2017
Later among the works it cites.
R. Krishna, Y. Zhu, O. Groth, J. Johnson, K. Hata, J. Kravitz, S. Chen, Y. Kalantidis, L.-J. Li, D. A. Shamma et al. , “Visual genome: Connecting language and vision using crowdsourced dense image annotations,” International Journal of Computer Vision , vol. 123, no. 1, pp. 32–73, 2017
2017
Later among the works it cites.
E. Shelhamer, J. Long, and T. Darrell, “Fully convolutional networks for semantic segmentation.” IEEE Transactions on Pattern Analysis and Machine Intelligence , vol. 39, no. 4, pp. 640–651, 2017
2017
Later among the works it cites.
H. Zhang, Z. Kyaw, J. Yu, and S.-F. Chang, “Ppr-fcn: Weakly supervised visual relation detection via parallel pairwise r-fcn,” in Proceedings of the IEEE International Conference on Computer Vision (ICCV) , 2017, pp. 4233–4241
2017
Later among the works it cites.
J. Peyre, I. Laptev, C. Schmid, and J. Sivic, “Weakly-supervised learning of visual relations,” in Proceedings of IEEE International Conference on Computer Vision (ICCV) , 2017, pp. 5179–5188
2017
Later among the works it cites.
X. Chen, L.-J. Li, L. Fei-Fei, and A. Gupta, “Iterative visual reasoning beyond convolutions,” in Proceedings of IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2018, pp. 7239–7248
2018
Later among the works it cites.
X. Zeng, W. Ouyang, J. Yan, H. Li, T. Xiao, K. Wang, Y. Liu, Y. Zhou, B. Yang, Z. Wang et al. , “Crafting gbd-net for object detection,” IEEE Transactions on Pattern Analysis and Machine Intelligence , vol. 40, no. 9, pp. 2109–2123, 2018
2018
Later among the works it cites.
Y. Liu, R. Wang, S. Shan, and X. Chen, “Structure inference net: Object detection using scene-level context and instance-level relationships,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2018, pp. 6985–6994
2018
Later among the works it cites.
J. Han, D. Zhang, G. Cheng, N. Liu, and D. Xu, “Advanced deep-learning techniques for salient and category-specific object detection: a survey,” IEEE Signal Processing Magazine , vol. 35, no. 1, pp. 84–100, 2018
2018
Later among the works it cites.
Z. Liu, X. Li, P. Luo, C. C. Loy, and X. Tang, “Deep learning markov random field for semantic segmentation,” IEEE Transactions on Pattern Analysis and Machine Intelligence , vol. 40, no. 8, pp. 1814–1828, 2018
2018
Later among the works it cites.
B. Shuai, Z. Zuo, B. Wang, and G. Wang, “Scene segmentation with dag-recurrent neural networks,” IEEE Transactions on Pattern Analysis and Machine Intelligence , vol. 40, no. 6, pp. 1480–1493, 2018
2018
Later among the works it cites.
2018
Later among the works it cites.
2018
Later among the works it cites.
R. Zellers, M. Yatskar, S. Thomson, and Y. Choi, “Neural motifs: Scene graph parsing with global context,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2018, pp. 5831–5840
2018
Later among the works it cites.
J. Yang, J. Lu, S. Lee, D. Batra, and D. Parikh, “Graph r-cnn for scene graph generation,” in Proceedings of European Conference on Computer Vision (ECCV) , September 2018, pp. 690–706
2018
Later among the works it cites.
Y. Li, W. Ouyang, B. Zhou, J. Shi, C. Zhang, and X. Wang, “Factorizable net: An efficient subgraph-based framework for scene graph generation,” in Proceedings of European Conference on Computer Vision (ECCV) , September 2018, pp. 346–363
2018
Later among the works it cites.
R. Herzig, M. Raboh, G. Chechik, J. Berant, and A. Globerson, “Mapping images to scene graphs with permutation-invariant structured prediction,” in Advances in Neural Information Processing Systems (NIPS) , 2018, pp. 7211–7221
2018
Later among the works it cites.
S. Woo, D. Kim, D. Cho, and I. S. Kweon, “Linknet: Relational embedding for scene graph,” in Advances in Neural Information Processing Systems (NIPS) , 2018, pp. 558–568
2018
Later among the works it cites.
A. Arnab, S. Zheng, S. Jayasumana, B. Romera-Paredes, M. Larsson, A. Kirillov, B. Savchynskyy, C. Rother, F. Kahl, and P. H. Torr, “Conditional random fields meet deep neural networks for semantic segmentation: Combining probabilistic graphical models with deep learning for structured prediction,” IEEE Signal Processing Magazine , vol. 35, no. 1, pp. 37–52, 2018
2018
Later among the works it cites.
L.-C. Chen, G. Papandreou, I. Kokkinos, K. Murphy, and A. L. Yuille, “Deeplab: Semantic image segmentation with deep convolutional nets, atrous convolution, and fully connected crfs,” IEEE Transactions on Pattern Analysis and Machine Intelligence , vol. 40, no. 4, pp. 834–848, 2018
2018
Later among the works it cites.
C. Zhang, J. Butepage, H. Kjellstrom, and S. Mandt, “Advances in variational inference,” IEEE Transactions on Pattern Analysis and Machine Intelligence , 2018
2018
Later among the works it cites.
D. Zhang, J. Han, L. Zhao, and D. Meng, “Leveraging prior-knowledge for weakly supervised object detection under a collaborative self-paced curriculum learning framework,” International Journal of Computer Vision , pp. 1–18, 2018
2018
Later among the works it cites.
G. Fazelnia and J. Paisley, “Crvi: Convex relaxation for variational inference,” in Proceedings of International Conference on Machine Learning (ICML) , 2018, pp. 1476–1484
2018
Later among the works it cites.
Z. Yang, J. Zhao, B. Dhingra, K. He, W. W. Cohen, R. R. Salakhutdinov, and Y. LeCun, “Glomo: Unsupervised learning of transferable relational graphs,” in Advances in Neural Information Processing Systems (NIPS) , 2018, pp. 8964–8975
2018
Later among the works it cites.
G. Yin, L. Sheng, B. Liu, N. Yu, X. Wang, J. Shao, and C. Change Loy, “Zoom-net: Mining deep feature interactions for visual relationship recognition,” in Proceedings of European Conference on Computer Vision (ECCV) , 2018, pp. 330–347
2018
Later among the works it cites.