Fetching the paper…
Reading the bibliography…
Transformer, which originates from machine translation, is particularly powerful at modeling long-range dependencies.
W. K. Hastings, “Monte Carlo sampling methods using Markov chains and their applications,”
1970
Earlier work this paper cites.
M. H. DeGroot and S. E. Fienberg, “The comparison and evaluation of forecasters,”
1983
Earlier work this paper cites.
L. Itti, C. Koch, and E. Niebur, “A model of saliency-based visual attention for rapid scene analysis,”
1998
Earlier work this paper cites.
W. Wright, “Bayesian approach to neural-network modeling with input uncertainty,”
1999
Earlier work this paper cites.
M. I. Jordan, Z. Ghahramani, and et al., “An introduction to variational methods for graphical models,” in
1999
Earlier work this paper cites.
C. Rother, V. Kolmogorov, and A. Blake, “”grabcut” interactive foreground extraction using iterated graph cuts,”
2004
Earlier work this paper cites.
PhD thesis, Stanford University, 2005
R. Ng, M. Levoy, M. Brédif, G. Duval, M. Horowitz, and P. Hanrahan, · 2005
Earlier work this paper cites.
Y. LeCun, S. Chopra, R. Hadsell, M. Ranzato, and F. Huang, “A tutorial on energy-based learning,”
2006
Earlier work this paper cites.
M. J. Wainwright and M. I. Jordan, “Graphical models, exponential families, and variational inference,”
2008
Earlier work this paper cites.
J. Deng, W. Dong, R. Socher, L.-J. Li, K. Li, and L. Fei-Fei, “Imagenet: A large-scale hierarchical image database,” in
2009
Earlier work this paper cites.
C. Liu, J. Yuen, and A. Torralba, “Sift flow: Dense correspondence across scenes and its applications,”
2010
Earlier work this paper cites.
D. Sun, S. Roth, and M. J. Black, “Secrets of optical flow estimation and their principles,” in
2010
Earlier work this paper cites.
V. Movahedi and J. H. Elder, “Design and perceptual validation of performance measures for salient object segmentation,” in
2010
Earlier work this paper cites.
R. Neal, “Mcmc using hamiltonian dynamics,”
2011
Earlier work this paper cites.
Y. Niu, Y. Geng, X. Li, and F. Liu, “Leveraging stereopsis for saliency analysis,” in
2012
Earlier work this paper cites.
Z. Zhang, “Microsoft kinect sensor and its effect,”
2012
Earlier work this paper cites.
Q. Yan, L. Xu, J. Shi, and J. Jia, “Hierarchical saliency detection,” in
2013
Earlier work this paper cites.
C. Yang, L. Zhang, H. Lu, X. Ruan, and M.-H. Yang, “Saliency detection via graph-based manifold ranking,” in
2013
Earlier work this paper cites.
A. Damianou and N. D. Lawrence, “Deep Gaussian processes,” in
2013
Earlier work this paper cites.
H. Peng, B. Li, W. Xiong, W. Hu, and R. Ji, “RGBD salient object detection: A benchmark and algorithms,” in
2014
Earlier work this paper cites.
Y. Cheng, H. Fu, X. Wei, J. Xiao, and X. Cao, “Depth enhanced saliency detection method,” in
2014
Earlier work this paper cites.
N. Li, J. Ye, Y. Ji, H. Ling, and J. Yu, “Saliency detection on light field,” in
2014
Earlier work this paper cites.
I. Goodfellow, J. Pouget-Abadie, M. Mirza, B. Xu, D. Warde-Farley, S. Ozair, A. Courville, and Y. Bengio, “Generative adversarial nets,” in
2014
Earlier work this paper cites.
M.-M. Cheng, N. J. Mitra, X. Huang, P. H. Torr, and S.-M. Hu, “Global contrast based salient region detection,”
2014
Earlier work this paper cites.
D. Kingma and M. Welling, “Auto-encoding variational bayes,” in
2014
Earlier work this paper cites.
P. Arbeláez, J. Pont-Tuset, J. T. Barron, F. Marques, and J. Malik, “Multiscale combinatorial grouping,” in
2014
Earlier work this paper cites.
Y. Li, X. Hou, C. Koch, J. M. Rehg, and A. L. Yuille, “The secrets of salient object segmentation,” in
2014
Earlier work this paper cites.
I. J. Goodfellow, J. Shlens, and C. Szegedy, “Explaining and harnessing adversarial examples,” in
2014
Earlier work this paper cites.
C. Szegedy, W. Zaremba, I. Sutskever, J. Bruna, D. Erhan, I. Goodfellow, and R. Fergus, “Intriguing properties of neural networks,” in
2014
Earlier work this paper cites.
R. Ju, Y. Liu, T. Ren, L. Ge, and G. Wu, “Depth-aware salient object detection using anisotropic center-surround difference,”
2015
Earlier work this paper cites.
K. Simonyan and A. Zisserman, “Very deep convolutional networks for large-scale image recognition,” in
2015
Earlier work this paper cites.
J. Kim, D. Han, Y.-W. Tai, and J. Kim, “Salient region detection via high-dimensional color transform and local spatial support,”
2015
Earlier work this paper cites.
O. Ronneberger, P. Fischer, and T. Brox, “U-net: Convolutional networks for biomedical image segmentation,” in
2015
Earlier work this paper cites.
S. Xie and Z. Tu, “Holistically-nested edge detection,” in
2015
Earlier work this paper cites.
K. Sohn, H. Lee, and X. Yan, “Learning structured output representation using deep conditional generative models,” in
2015
Earlier work this paper cites.
J. Dai, K. He, and J. Sun, “Boxsup: Exploiting bounding boxes to supervise convolutional networks for semantic segmentation,” in
2015
Earlier work this paper cites.
G. Li and Y. Yu, “Visual saliency based on multiscale deep features,” in
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in
2016
Earlier work this paper cites.
D. Lin, J. Dai, J. Jia, K. He, and J. Sun, “Scribblesup: Scribble-supervised convolutional networks for semantic segmentation,” in
2016
Earlier work this paper cites.
A. Bearman, O. Russakovsky, V. Ferrari, and L. Fei-Fei, “What’s the point: Semantic segmentation with point supervision,” in
2016
Earlier work this paper cites.
Y. Gal and Z. Ghahramani, “Dropout as a bayesian approximation: Representing model uncertainty in deep learning,” in
2016
Earlier work this paper cites.
L. Wang, H. Lu, Y. Wang, M. Feng, D. Wang, B. Yin, and X. Ruan, “Learning to detect salient objects with image-level supervision,” in
2017
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, L. Kaiser, and I. Polosukhin, “Attention is all you need,” in
2017
Earlier work this paper cites.
C. Guo, G. Pleiss, Y. Sun, and K. Q. Weinberger, “On calibration of modern neural networks,” in
2017
Earlier work this paper cites.
T. Han, Y. Lu, S.-C. Zhu, and Y. N. Wu, “Alternating back-propagation for generator network,” in
2017
Earlier work this paper cites.
A. Kendall and Y. Gal, “What uncertainties do we need in bayesian deep learning for computer vision?,” in
2017
Earlier work this paper cites.
Z. Luo, A. Mishra, A. Achkar, J. Eichel, S. Li, and P.-M. Jodoin, “Non-local deep features for salient object detection,” in
2017
Earlier work this paper cites.
Q. Hou, M.-M. Cheng, X. Hu, A. Borji, Z. Tu, and P. H. Torr, “Deeply supervised salient object detection with short connections,” in
2017
Earlier work this paper cites.
T. Wang, A. Borji, L. Zhang, P. Zhang, and H. Lu, “A stagewise refinement model for detecting salient objects in images,” in
2017
Earlier work this paper cites.
Y. Liu, M.-M. Cheng, X. Hu, K. Wang, and X. Bai, “Richer convolutional features for edge detection,” in
2017
Earlier work this paper cites.
J. Han, H. Chen, N. Liu, C. Yan, and X. Li, “Cnns-based rgb-d saliency detection via cross-view transfer and multiview fusion,”
2017
Earlier work this paper cites.
L. Qu, S. He, J. Zhang, J. Tian, Y. Tang, and Q. Yang, “RGBD salient object detection via deep fusion,”
2017
Earlier work this paper cites.
J. Han, H. Chen, N. Liu, C. Yan, and X. Li, “CNNs-based RGB-D saliency detection via cross-view transfer and multiview fusion,”
2017
Earlier work this paper cites.
N. Souly, C. Spampinato, and M. Shah, “Semi supervised semantic segmentation using generative adversarial network,” in
2017
Cited alongside, same era.
K.-J. Hsu12, Y.-Y. Lin, and Y.-Y. Chuang, “Weakly supervised saliency detection with a category-driven map generator,”
2017
Cited alongside, same era.
P. Vernaza and M. Chandraker, “Learning random-walk label propagation for weakly-supervised semantic segmentation,” in
2017
Cited alongside, same era.
C. Godard, O. Mac Aodha, and G. J. Brostow, “Unsupervised monocular depth estimation with left-right consistency,” in
2017
Cited alongside, same era.
I. Higgins, L. Matthey, A. Pal, C. Burgess, X. Glorot, M. Botvinick, S. Mohamed, and A. Lerchner, “beta-VAE: Learning basic visual concepts with a constrained variational framework,” in
2017
Cited alongside, same era.
C. Li, R. Cong, Y. Piao, Q. Xu, and C. C. Loy, “RGB-D salient object detection with cross-modality modulation and selection,” in
2020
Later among the works it cites.
G. Li, Z. Liu, L. Ye, Y. Wang, and H. Ling, “Cross-modal weighting network for RGB-D salient object detection,” in
2020
Later among the works it cites.
A. Luo, X. Li, F. Yang, Z. Jiao, H. Cheng, and S. Lyu, “Cascade graph neural networks for RGB-D salient object detection,” in
2020
Later among the works it cites.
S. Chen and Y. Fu, “Progressively guided alternate refinement network for RGB-D salient object detection,” in
2020
Later among the works it cites.
M. Zhang, S. X. Fei, J. Liu, S. Xu, Y. Piao, and H. Lu, “Asymmetric two-stream architecture for accurate RGB-D saliency detection,” in
2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
D.-P. Fan, M.-M. Cheng, Y. Liu, T. Li, and A. Borji, “Structure-measure: A new way to evaluate foreground maps,” in
2017
Cited alongside, same era.
J. Zhang, T. Zhang, Y. Dai, M. Harandi, and R. Hartley, “Deep unsupervised saliency detection: A multiple noisy labeling perspective,” in
2018
Cited alongside, same era.
N. Liu, J. Han, and M.-H. Yang, “Picanet: Learning pixel-wise contextual attention for saliency detection,” in
2018
Cited alongside, same era.
X. Zhang, T. Wang, J. Qi, H. Lu, and G. Wang, “Progressive attention guided recurrent network for salient object detection,” in
2018
Cited alongside, same era.
H. Chen and Y. Li, “Progressively complementarity-aware fusion network for RGB-D salient object detection,” in
2018
Cited alongside, same era.
S. Kohl, B. Romera-Paredes, C. Meyer, J. De Fauw, J. R. Ledsam, K. Maier-Hein, S. M. A. Eslami, D. Jimenez Rezende, and O. Ronneberger, “A probabilistic u-net for segmentation of ambiguous images,” in
2018
Cited alongside, same era.
W.-C. Hung, Y.-H. Tsai, Y.-T. Liou, Y.-Y. Lin, and M.-H. Yang, “Adversarial learning for semi-supervised semantic segmentation,” in
2018
Cited alongside, same era.
N. Carion, F. Massa, G. Synnaeve, N. Usunier, A. Kirillov, and S. Zagoruyko, “End-to-end object detection with transformers,” in
2020
Later among the works it cites.
R. Groenendijk, S. Karaoglu, T. Gevers, and T. Mensink, “On the benefit of adversarial training for monocular depth estimation,”
2020
Later among the works it cites.
Q. H. Le, K. Youcef-Toumi, D. Tsetserukou, and A. Jahanian, “Gan mask r-cnn:instance semantic segmentation benefits from generative adversarial networks,” in
2020
Later among the works it cites.
V. Kulharia, S. Chandra, A. Agrawal, P. Torr, and A. Tyagi, “Box2seg: Attention weighted loss and discriminative feature learning for weakly supervised segmentation,” in
2020
Later among the works it cites.
J. Zhang, J. Xie, and N. Barnes, “Learning noise-aware encoder-decoder from noisy labels by alternating back-propagation for saliency detection,” in
2020
Later among the works it cites.
H. Zhou, X. Xie, J.-H. Lai, Z. Chen, and L. Yang, “Interactive two-stream decoder for accurate and fast saliency detection,” in
2020
Later among the works it cites.
M. Vadera, B. Jalaian, and B. Marlin, “Generalized bayesian posterior expectation distillation for deep neural networks,” in
2020
Later among the works it cites.
M. Zhang, T. Liu, Y. Piao, S. Yao, and H. Lu, “Auto-msfnet: Search multi-scale fusion network for salient object detection,” in
2021
Closest in time.
B. Xu, H. Liang, R. Liang, and P. Chen, “Locate globally, segment locally: A progressive architecture with knowledge review network for salient object detection,” in
2021
Closest in time.
J. Zhang, D.-P. Fan, Y. Dai, X. Yu, Y. Zhong, N. Barnes, and L. Shao, “Rgb-d saliency detection via cascaded mutual information minimization,” in
2021
Closest in time.
R. Ranftl, A. Bochkovskiy, and V. Koltun, “Vision transformers for dense prediction,” in
2021
Closest in time.
Z. Liu, Y. Lin, Y. Cao, H. Hu, Y. Wei, Z. Zhang, S. Lin, and B. Guo, “Swin transformer: Hierarchical vision transformer using shifted windows,” in
2021
Closest in time.
A. Dosovitskiy, L. Beyer, A. Kolesnikov, D. Weissenborn, X. Zhai, T. Unterthiner, M. Dehghani, M. Minderer, G. Heigold, S. Gelly, J. Uszkoreit, and N. Houlsby, “An image is worth 16x16 words: Transformers for image recognition at scale,” in
2021
Closest in time.
L. Tang, B. Li, Y. Zhong, S. Ding, and M. Song, “Disentangled high quality salient object detection,” in
2021
Closest in time.
M. Zhang, T. Liu, Y. Piao, S. Yao, and H. Lu, “Auto-msfnet: Search multi-scale fusion network for salient object detection,” in
2021
Closest in time.
Z. Zhang, Z. Lin, J. Xu, W.-D. Jin, S.-P. Lu, and D.-P. Fan, “Bilateral attention network for RGB-D salient object detection,”
2021
Closest in time.
G. Li, Z. Liu, M. Chen, Z. Bai, W. Lin, and H. Ling, “Hierarchical alternate interaction network for rgb-d salient object detection,”
2021
Closest in time.
X. Zhu, W. Su, L. Lu, B. Li, X. Wang, and J. Dai, “Deformable DETR: Deformable transformers for end-to-end object detection,” in
2021
Closest in time.
Z. Dai, B. Cai, Y. Lin, and J. Chen, “UP-DETR: Unsupervised pre-training for object detection with transformers,” in
2021
Closest in time.
W. Wang, E. Xie, X. Li, D.-P. Fan, K. Song, D. Liang, T. Lu, P. Luo, and L. Shao, “Pyramid vision transformer: A versatile backbone for dense prediction without convolutions,” in
2021
Closest in time.
B. Yan, H. Peng, J. Fu, D. Wang, and H. Lu, “Learning spatio-temporal transformer for visual tracking,” in
2021
Closest in time.
S. Jiang, D. Campbell, Y. Lu, H. Li, and R. Hartley, “Learning to estimate hidden motions with global motion aggregation,” in
2021
Closest in time.
S. Zheng, J. Lu, H. Zhao, X. Zhu, Z. Luo, Y. Wang, Y. Fu, J. Feng, T. Xiang, P. H. Torr, and L. Zhang, “Rethinking semantic segmentation from a sequence-to-sequence perspective with transformers,” in
2021
Closest in time.
N. Liu, N. Zhang, K. Wan, L. Shao, and J. Han, “Visual saliency transformer,” in
2021
Closest in time.
J. Zhang, J. Xie, N. Barnes, and P. Li, “Learning generative vision transformer with energy-based latent space for saliency prediction,” in
2021
Closest in time.
J. Lee, J. Yi, C. Shin, and S. Yoon, “Bbam: Bounding box attribution map for weakly supervised semantic and instance segmentation,” in
2021
Closest in time.
Z. Tian, C. Shen, X. Wang, and H. Chen, “Boxinst: High-performance instance segmentation with box annotations,” in
2021
Closest in time.
S. Yu, B. Zhang, J. Xiao, and E. G. Lim, “Structure-consistent weakly supervised salient object detection with local saliency coherence,” in
2021
Closest in time.
H. Chen, J. Wang, H. C. Chen, X. Zhen, F. Zheng, R. Ji, and L. Shao, “Seminar learning for click-level weakly supervised semantic segmentation,” in
2021
Closest in time.
Z. Zhao, C. Xia, C. Xie, and J. Li, “Complementary trilateral decoder for fast and accurate salient object detection,” in
2021
Closest in time.
M. Naseer, K. Ranasinghe, S. Khan, M. Hayat, F. S. Khan, and M.-H. Yang, “Intriguing properties of vision transformers,” in
2021
Closest in time.
J. Zhang, D.-P. Fan, Y. Dai, S. Anwar, F. Saleh, S. Aliakbarian, and N. Barnes, “Uncertainty inspired rgb-d saliency detection,”
2022
Closest in time.
W. Wang, Q. Lai, H. Fu, J. Shen, H. Ling, and R. Yang, “Salient object detection in the deep learning era: An in-depth survey,”
2022
Closest in time.
Y.-H. Wu, Y. Liu, L. Zhang, M.-M. Cheng, and B. Ren, “Edn: Salient object detection via extremely-downsampled network,”
2022
Closest in time.
Z. Yang, S. Soltanian-Zadeh, and S. Farsiu, “Biconnet: an edge-preserved connectivity-based approach for salient object detection,”
2022
Closest in time.
S. Cao and Z. Zhang, “Deep hybrid models for out-of-distribution detection,” in
2022
Closest in time.
W. Zhang, L. Zheng, H. Wang, X. Wu, and X. Li, “Saliency hierarchy modeling via generative kernels for salient object detection,” in
2022
Closest in time.
M. Lee, C. Park, S. Cho, and S. Lee, “Spsn: Superpixel prototype sampling network for rgb-d salient object detection,” in
2022
Closest in time.
N. Liu, N. Zhang, L. Shao, and J. Han, “Learning selective mutual attention and contrast for rgb-d saliency detection,”
2022
Closest in time.
K. Fu, D.-P. Fan, G.-P. Ji, Q. Zhao, J. Shen, and C. Zhu, “Siamese network for rgb-d salient object detection and beyond,”
2022
Closest in time.
G. Zhang, Z. Luo, K. Cui, S. Lu, and E. P. Xing, “Meta-detr: Image-level few-shot detection with inter-class correlation exploitation,”
2022
Closest in time.
Y. Xu, Y. Ban, G. Delorme, C. Gan, D. Rus, and X. Alameda-Pineda, “Transcenter: Transformers with dense representations for multiple-object tracking,”
2022
Closest in time.
W. Mao, Y. Ge, C. Shen, Z. Tian, X. Wang, Z. Wang, and A. v. den Hengel, “Poseur: Direct human pose estimation with transformers,” in
2022
Closest in time.
Y. Chen, Q. Gao, X. Wang,
2022
Closest in time.
H. Zhang, Y. Zeng, H. Lu, L. Zhang, J. Li, and J. Qi, “Learning to detect salient object with multi-source weak supervision,”
2022
Closest in time.
G. Franchi, X. Yu, A. Bursuc, E. Aldea, S. Dubuisson, and D. Filliat, “Latent discriminant deterministic uncertainty,” in
2022
Closest in time.
S. Y. Lee, “Gibbs sampler and coordinate ascent variational inference: A set-theoretical review,”
2022
Closest in time.
M. Zhuge, D.-P. Fan, N. Liu, D. Zhang, D. Xu, and L. Shao, “Salient object detection via integrity learning,”
2023
Closest in time.