Fetching the paper…
Reading the bibliography…
In this paper, we propose a novel model called SGFormer, Semantic Graph TransFormer for point cloud-based 3D scene graph generation.
Language models are few-shot learners
Brown, T.; Mann, B.; Ryder, N.; Subbiah, M.; Kaplan, J. D.; Dhariwal, P.; Neelakantan, A.; Shyam, P.; Sastry, G.; Askell, A.; et al. 2020 · 1901
Earlier work this paper cites.
Visual relationship detection with internal and external linguistic knowledge distillation
Yu, R.; Li, A.; Morariu, V. I.; and Davis, L. S. 2017 · 1982
Earlier work this paper cites.
Large-scale object classification using label relation graphs
Deng, J.; Ding, N.; Jia, Y.; Frome, A.; Murphy, K.; Bengio, S.; Li, Y.; Neven, H.; and Adam, H. 2014 · 2014
Earlier work this paper cites.
Glove: Global vectors for word representation
Pennington, J.; Socher, R.; and Manning, C. D. 2014 · 2014
Earlier work this paper cites.
Distilling the knowledge in a neural network
Hinton, G.; Vinyals, O.; Dean, J.; et al. 2015 · 2015
Earlier work this paper cites.
Image retrieval using scene graphs
Johnson, J.; Krishna, R.; Stark, M.; Li, L.-J.; Shamma, D.; Bernstein, M.; and Fei-Fei, L. 2015 · 2015
Earlier work this paper cites.
Semi-supervised classification with graph convolutional networks
Kipf, T. N.; and Welling, M. 2016 · 2016
Earlier work this paper cites.
Visual relationship detection with language priors
Lu, C.; Krishna, R.; Bernstein, M.; and Fei-Fei, L. 2016 · 2016
Earlier work this paper cites.
Exploring spatial context for 3D semantic segmentation of point clouds
Engelmann, F.; Kontogianni, T.; Hermans, A.; and Leibe, B. 2017 · 2017
Earlier work this paper cites.
Visual genome: Connecting language and vision using crowdsourced dense image annotations
Krishna, R.; Zhu, Y.; Groth, O.; Johnson, J.; Hata, K.; Kravitz, J.; Chen, S.; Kalantidis, Y.; Li, L.-J.; Shamma, D. A.; et al. 2017 · 2017
Earlier work this paper cites.
Incorporating external knowledge to answer open-domain visual questions with dynamic memory networks
Li, G.; Su, H.; and Zhu, W. 2017 · 2017
Earlier work this paper cites.
Scene graph generation from objects, phrases and region captions
Li, Y.; Ouyang, W.; Zhou, B.; Wang, K.; and Wang, X. 2017 · 2017
Earlier work this paper cites.
Focal loss for dense object detection
Lin, T.-Y.; Goyal, P.; Girshick, R.; He, K.; and Dollár, P. 2017 · 2017
Earlier work this paper cites.
Pointnet: Deep learning on point sets for 3d classification and segmentation
Qi, C. R.; Su, H.; Mo, K.; and Guibas, L. J. 2017 · 2017
Earlier work this paper cites.
Conceptnet 5.5: An open multilingual graph of general knowledge
Speer, R.; Chin, J.; and Havasi, C. 2017 · 2017
Earlier work this paper cites.
Attention is all you need
Vaswani, A.; Shazeer, N.; Parmar, N.; Uszkoreit, J.; Jones, L.; Gomez, A. N.; Kaiser, Ł.; and Polosukhin, I. 2017 · 2017
Cited alongside, same era.
Scene graph generation by iterative message passing
Xu, D.; Zhu, Y.; Choy, C. B.; and Fei-Fei, L. 2017 · 2017
Cited alongside, same era.
Mapping images to scene graphs with permutation-invariant structured prediction
Herzig, R.; Raboh, M.; Chechik, G.; Berant, J.; and Globerson, A. 2018 · 2018
Cited alongside, same era.
Deeper insights into graph convolutional networks for semi-supervised learning
Li, Q.; Han, Z.; and Wu, X.-M. 2018 · 2018
Cited alongside, same era.
Visual relationship detection with deep structural ranking
Liang, K.; Guo, Y.; Chang, H.; and Chen, X. 2018 · 2018
Cited alongside, same era.
stagnet: An attentive semantic rnn for group activity recognition
Qi, M.; Qin, J.; Li, A.; Wang, Y.; Luo, J.; and Van Gool, L. 2018 · 2018
Hybrid topological and 3d dense mapping through autonomous exploration for large indoor environments
Gomez, C.; Fehr, M.; Millane, A.; Hernandez, A. C.; Nieto, J.; Barber, R.; and Siegwart, R. 2020 · 2020
Later among the works it cites.
Global context reasoning for semantic segmentation of 3D point clouds
Ma, Y.; Guo, Y.; Liu, H.; Lei, Y.; and Wen, G. 2020 · 2020
Later among the works it cites.
STC-GAN: Spatio-temporally coupled generative adversarial networks for predictive scene parsing
Qi, M.; Wang, Y.; Li, A.; and Luo, J. 2020 · 2020
Later among the works it cites.
Retargetable AR: Context-aware augmented reality in indoor scenes based on 3D scene graph
Tahara, T.; Seno, T.; Narita, G.; and Ishikawa, T. 2020 · 2020
Later among the works it cites.
Unbiased scene graph generation from biased training
Tang, K.; Niu, Y.; Huang, J.; Shi, J.; and Zhang, H. 2020 · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Fully-convolutional point networks for large-scale point clouds
Rethage, D.; Wald, J.; Sturm, J.; Navab, N.; and Tombari, F. 2018 · 2018
Cited alongside, same era.
Graph r-cnn for scene graph generation
Yang, J.; Lu, J.; Lee, S.; Batra, D.; and Parikh, D. 2018 · 2018
Cited alongside, same era.
Neural motifs: Scene graph parsing with global context
Zellers, R.; Yatskar, M.; Thomson, S.; and Choi, Y. 2018 · 2018
Cited alongside, same era.
Graph-based global reasoning networks
Chen, Y.; Rohrbach, M.; Yan, Z.; Shuicheng, Y.; Feng, J.; and Kalantidis, Y. 2019 · 2019
Cited alongside, same era.
3d-sis: 3d semantic instance segmentation of rgb-d scans
Hou, J.; Dai, A.; and Nießner, M. 2019 · 2019
Cited alongside, same era.
Pytorch: An imperative style, high-performance deep learning library
Paszke, A.; Gross, S.; Massa, F.; Lerer, A.; Bradbury, J.; Chanan, G.; Killeen, T.; Lin, Z.; Gimelshein, N.; Antiga, L.; et al. 2019 · 2019
Cited alongside, same era.
Wald, J.; Dhamo, H.; Navab, N.; and Tombari, F. 2020 · 2020
Later among the works it cites.
Graph-to-3d: End-to-end generation and manipulation of 3d scenes using scene graphs
Dhamo, H.; Manhardt, F.; Navab, N.; and Tombari, F. 2021 · 2021
Later among the works it cites.
A Generalization of Transformer Networks to Graphs
Dwivedi, V. P.; and Bresson, X. 2021 · 2021
Later among the works it cites.
Learning transferable visual models from natural language supervision
Radford, A.; Kim, J. W.; Hallacy, C.; Ramesh, A.; Goh, G.; Agarwal, S.; Sastry, G.; Askell, A.; Mishkin, P.; Clark, J.; et al. 2021 · 2021
Later among the works it cites.
Do transformers really perform badly for graph representation?
Ying, C.; Cai, T.; Luo, S.; Zheng, S.; Ke, G.; He, D.; Shen, Y.; and Liu, T.-Y. 2021 · 2021
Later among the works it cites.
Exploring Contextual Relationships in 3D Cloud Points by Semantic Knowledge Mining
Chen, L.; Lu, J.; Cai, Y.; Wang, C.; and He, G. 2022 · 2022
Later among the works it cites.
Softgroup for 3d instance segmentation on point clouds
Vu, T.; Kim, K.; Luu, T. M.; Nguyen, T.; and Yoo, C. D. 2022 · 2022
Later among the works it cites.
HyperDet3D: Learning a Scene-conditioned 3D Object Detector
Zheng, Y.; Duan, Y.; Lu, J.; Zhou, J.; and Tian, Q. 2022 · 2022
Later among the works it cites.
Disentangled Counterfactual Learning for Physical Audiovisual Commonsense Reasoning
Lv, C.; Zhang, S.; Tian, Y.; Qi, M.; and Ma, H. 2023 · 2023
Closest in time.
Llama: Open and efficient foundation language models
Touvron, H.; Lavril, T.; Izacard, G.; Martinet, X.; Lachaux, M.-A.; Lacroix, T.; Rozière, B.; Goyal, N.; Hambro, E.; Azhar, F.; et al. 2023 · 2023
Closest in time.