Fetching the paper…
Reading the bibliography…
The performance of current Scene Graph Generation (SGG) models is severely hampered by hard-to-distinguish predicates, e.g., woman-on/standing on/walking on-beach.
K. Papineni, S. Roukos, T. Ward, and W. Zhu, “Bleu: a method for automatic evaluation of machine translation,” in ACL , 2002
2002
Earlier work this paper cites.
L. J. Old, “An analysis of semantic overlap among english prepositions in roget’s thesaurus,” in ACL-SIGSEM , 2003
2003
Earlier work this paper cites.
T. Lin, M. Maire, S. J. Belongie, J. Hays, P. Perona, D. Ramanan, P. Dollár, and C. L. Zitnick, “Microsoft COCO: common objects in context,” in ECCV , 2014
2014
Earlier work this paper cites.
M. J. Denkowski and A. Lavie, “Meteor universal: Language specific translation evaluation for any target language,” in ACL , 2014
2014
Earlier work this paper cites.
R. Vedantam, C. L. Zitnick, and D. Parikh, “Cider: Consensus-based image description evaluation,” in CVPR , 2015
2015
Earlier work this paper cites.
S. Schuster, R. Krishna, A. X. Chang, L. Fei-Fei, and C. D. Manning, “Generating semantically precise scene graphs from textual descriptions for improved image retrieval,” in EMNLP , 2015
2015
Earlier work this paper cites.
R. B. Girshick, J. Donahue, T. Darrell, and J. Malik, “Region-based convolutional networks for accurate object detection and segmentation,” in PAMI , 2016
2016
Earlier work this paper cites.
P. Anderson, B. Fernando, M. Johnson, and S. Gould, “SPICE: semantic propositional image caption evaluation,” in ECCV , 2016
2016
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, L. Kaiser, and I. Polosukhin, “Attention is all you need,” in NeurIPS , 2017
2017
Earlier work this paper cites.
D. Xu, Y. Zhu, C. B. Choy, and L. Fei-Fei, “Scene graph generation by iterative message passing,” in CVPR , 2017
2017
Earlier work this paper cites.
C. Guo, G. Pleiss, Y. Sun, and K. Q. Weinberger, “On calibration of modern neural networks,” in ICML , 2017
2017
Earlier work this paper cites.
T. Lin, P. Goyal, R. B. Girshick, K. He, and P. Dollár, “Focal loss for dense object detection,” in ICCV , 2017
2017
Earlier work this paper cites.
S. Ren, K. He, R. B. Girshick, and J. Sun, “Faster R-CNN: towards real-time object detection with region proposal networks,” in PAMI , 2017
2017
Earlier work this paper cites.
R. Vedantam, S. Bengio, K. Murphy, D. Parikh, and G. Chechik, “Context-aware captions from context-agnostic supervision,” in CVPR , 2017
2017
Earlier work this paper cites.
R. Zellers, M. Yatskar, S. Thomson, and Y. Choi, “Neural motifs: Scene graph parsing with global context,” in CVPR , 2018
2018
Earlier work this paper cites.
C. Yu, X. Zhao, Q. Zheng, P. Zhang, and X. You, “Hierarchical bilinear pooling for fine-grained visual recognition,” in ECCV , 2018
2018
Earlier work this paper cites.
R. Luo, B. L. Price, S. Cohen, and G. Shakhnarovich, “Discriminability objective for training descriptive captions,” in CVPR , 2018
2018
Earlier work this paper cites.
L. Ruotian, “A scene graph generation codebase in pytorch,” 2018, https://github.com/ruotianluo/self-critical.pytorch
2018
Earlier work this paper cites.
J. Gu, S. R. Joty, J. Cai, H. Zhao, X. Yang, and G. Wang, “Unpaired image captioning via scene graph alignments,” in ICCV , 2019
2019
Earlier work this paper cites.
X. Yang, K. Tang, H. Zhang, and J. Cai, “Auto-encoding scene graphs for image captioning,” in CVPR , 2019
2019
Earlier work this paper cites.
K. Tang, H. Zhang, B. Wu, W. Luo, and W. Liu, “Learning to compose dynamic tree structures for visual contexts,” in CVPR , 2019
2019
Earlier work this paper cites.
D. A. Hudson and C. D. Manning, “GQA: A new dataset for real-world visual reasoning and compositional question answering,” in CVPR , 2019
2019
Earlier work this paper cites.
J. Shi, H. Zhang, and J. Li, “Explainable and explicit visual reasoning over scene graphs,” in CVPR , 2019
2019
Earlier work this paper cites.
Y. Liang, Y. Bai, W. Zhang, X. Qian, L. Zhu, and T. Mei, “Vrr-vg: Refocusing visually-relevant relationships,” in ICCV , 2019
2019
Cited alongside, same era.
T. Chen, W. Yu, R. Chen, and L. Lin, “Knowledge-embedded routing network for scene graph generation,” in CVPR , 2019
2019
Cited alongside, same era.
Y. Cui, M. Jia, T. Lin, Y. Song, and S. J. Belongie, “Class-balanced loss based on effective number of samples,” in CVPR , 2019
2019
Cited alongside, same era.
W. Ge, X. Lin, and Y. Yu, “Weakly supervised complementary parts models for fine-grained image classification from the bottom up,” in CVPR , 2019
2019
Cited alongside, same era.
W. Luo, X. Yang, X. Mo, Y. Lu, L. Davis, J. Li, J. Yang, and S. Lim, “Cross-x learning for fine-grained visual categorization,” in ICCV , 2019
2019
Cited alongside, same era.
R. Li, S. Zhang, B. Wan, and X. He, “Bipartite graph network with adaptive message passing for unbiased scene graph generation,” in CVPR , 2021
2021
Later among the works it cites.
A. Desai, T.-Y. Wu, S. Tripathi, and N. Vasconcelos, “Learning of visual relations: The devil is in the tails,” in ICCV , 2021
2021
Later among the works it cites.
Z. Zhong, J. Cui, S. Liu, and J. Jia, “Improving calibration for long-tailed recognition,” in CVPR , 2021
2021
Later among the works it cites.
T. Wang, Y. Zhu, C. Zhao, W. Zeng, J. Wang, and M. Tang, “Adaptive class suppression loss for long-tail object detection,” in CVPR , 2021
2021
Later among the works it cites.
C. Feng, Y. Zhong, and W. Huang, “Exploring classification equilibrium in long-tailed object detection,” in ICCV , 2021
2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Y. Ding, Y. Zhou, Y. Zhu, Q. Ye, and J. Jiao, “Selective sparse sampling for fine-grained image recognition,” in ICCV , 2019
2019
Cited alongside, same era.
H. Zheng, J. Fu, Z. Zha, and J. Luo, “Looking for the devil in the details: Learning trilinear attention sampling network for fine-grained image recognition,” in CVPR , 2019
2019
Cited alongside, same era.
A. Gupta, P. Dollár, and R. B. Girshick, “LVIS: A dataset for large vocabulary instance segmentation,” in CVPR , 2019
2019
Cited alongside, same era.
B. Schroeder and S. Tripathi, “Structured query-based image retrieval using scene graphs,” in CVPR , 2020
2020
Cited alongside, same era.
M. Hildebrandt, H. Li, R. Koner, V. Tresp, and S. Günnemann, “Scene graph reasoning for visual question answering,” in CoRR , 2020
2020
Cited alongside, same era.
K. Tang, “A scene graph generation codebase in pytorch,” 2020, https://github.com/KaihuaTang/Scene-Graph-Benchmark.pytorch
2020
Cited alongside, same era.
X. Lin, C. Ding, J. Zeng, and D. Tao, “Gps-net: Graph property sensing network for scene graph generation,” in CVPR , 2020
2020
Cited alongside, same era.
A. K. Menon, S. Jayasumana, A. S. Rawat, H. Jain, A. Veit, and S. Kumar, “Long-tail learning via logit adjustment,” in ICLR , 2021
2021
Later among the works it cites.
J. Wang, W. Zhang, Y. Zang, Y. Cao, J. Pang, T. Gong, K. Chen, Z. Liu, C. C. Loy, and D. Lin, “Seesaw loss for long-tailed instance segmentation,” in CVPR , 2021
2021
Later among the works it cites.
S. Yang, S. Liu, C. Yang, and C. Wang, “Re-rank coarse classification with local region enhanced features for fine-grained image recognition,” in CoRR , 2021
2021
Later among the works it cites.
S. Khandelwal, M. Suhail, and L. Sigal, “Segmentation-grounded scene graph generation,” in ICCV , 2021
2021
Later among the works it cites.
G. Yang, J. Zhang, Y. Zhang, B. Wu, and Y. Yang, “Probabilistic modeling of semantic ambiguity for scene graph generation,” in CVPR , 2021
2021
Later among the works it cites.
M. Chiou, H. Ding, H. Yan, C. Wang, R. Zimmermann, and J. Feng, “Recovering the unbiased scene graphs from the biased ones,” in ACMMM , 2021
2021
Later among the works it cites.
L. Li, L. Chen, Y. Huang, Z. Zhang, S. Zhang, and J. Xiao, “The devil is in the labels: Noisy label correction for robust scene graph generation,” in CVPR , 2021
2021
Later among the works it cites.
M. Chen, X. Lyu, Y. Guo, J. Liu, L. Gao, and J. Song, “Multi-scale graph attention network for scene graph generation,” in ICME , 2022
2022
Closest in time.
C. Zheng, X. Lyu, Y. Guo, P. Zeng, J. Song, and L. Gao, “Rouge: A package for automatic evaluation of summaries,” in ICME , 2022
2022
Closest in time.
X. Dong, T. Gan, X. Song, J. Wu, Y. Cheng, and L. Nie, “Stacked hybrid-attention and group collaborative learning for unbiased scene graph generation,” in CVPR , 2022
2022
Closest in time.
X. Lyu, L. Gao, Y. Guo, Z. Zhao, H. Huang, H. T. Shen, and J. Song, “Fine-grained predicates learning for scene graph generation,” in CVPR , 2022
2022
Closest in time.
L. Tao, L. Mi, N. Li, X. Cheng, Y. Hu, and Z. Chen, “Predicate correlation learning for scene graph generation,” in TIP , 2022
2022
Closest in time.
Y. Deng, Y. Li, Y. Zhang, X. Xiang, J. Wang, J. Chen, and J. Ma, “Hierarchical memory learning for fine-grained scene graph generation,” in CoRR , 2022
2022
Closest in time.
A. Zhang, Y. Yao, Q. Chen, W. Ji, Z. Liu, M. Sun, and T. Chua, “Fine-grained scene graph generation with data transfer,” in CoRR , 2022
2022
Closest in time.
A. Goel, B. Fernando, F. Keller, and H. Bilen, “Not all relations are equal: Mining informative labels for scene graph generation,” in CVPR , 2022
2022
Closest in time.
W. Li, H. Zhang, Q. Bai, G. Zhao, N. Jiang, and X. Yuan, “Ppdl: Predicate probability distribution based loss for unbiased scene graph generation,” in CVPR , 2022
2022
Closest in time.
C. Chen, Y. Zhan, B. Yu, L. Liu, Y. Luo, and B. Du, “Resistance training using prior bias: toward unbiased scene graph generation,” in AAAI , 2022
2022
Closest in time.
X. Chang, T. Wang, C. Sun, and W. Cai, “Biasing like human: A cognitive bias framework for scene graph generation,” in CoRR , 2022
2022
Closest in time.