Fetching the paper…
Reading the bibliography…
Learning from a large corpus of data, pre-trained models have achieved impressive progress nowadays.
Maximal flow through a network
Lester Randolph Ford and Delbert R Fulkerson · 1956
Earlier work this paper cites.
Algorithms for the reduction of the number of points required to represent a digitized line or its caricature
David H Douglas and Thomas K Peucker · 1973
Earlier work this paper cites.
Least squares quantization in pcm
Stuart Lloyd · 1982
Earlier work this paper cites.
Normalized cuts and image segmentation
Jianbo Shi and Jitendra Malik · 2000
Earlier work this paper cites.
The PASCAL visual object classes challenge 2007 (VOC2007) results, 2007
Mark Everingham, Luc Van Gool, Christopher KI Williams, John Winn, and Andrew Zisserman · 2007
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database
Jia Deng, Wei Dong, Richard Socher, Li-Jia Li, Kai Li, and Li Fei-Fei · 2009
Earlier work this paper cites.
Learning to predict where humans look
Tilke Judd, Krista Ehinger, Frédo Durand, and Antonio Torralba · 2009
Earlier work this paper cites.
Unsupervised detection of regions of interest using iterative link analysis
Gunhee Kim and Antonio Torralba · 2009
Earlier work this paper cites.
Sun database: Large-scale scene recognition from abbey to zoo
Jianxiong Xiao, James Hays, Krista A Ehinger, Aude Oliva, and Antonio Torralba · 2010
Earlier work this paper cites.
Efficient inference in fully connected crfs with gaussian edge potentials
Philipp Krähenbühl and Vladlen Koltun · 2011
Earlier work this paper cites.
The caltech-ucsd birds-200-2011 dataset
Catherine Wah, Steve Branson, Peter Welinder, Pietro Perona, and Serge Belongie · 2011
Earlier work this paper cites.
The PASCAL Visual Object Classes Challenge 2012 (VOC2012) Results
Mark Everingham, Luc Van Gool, Christopher K. I. Williams, John Winn, and Andrew Zisserman · 2012
Earlier work this paper cites.
Geodesic saliency using background priors
Yichen Wei, Fang Wen, Wangjiang Zhu, and Jian Sun · 2012
Earlier work this paper cites.
Selective search for object recognition
Jasper Uijlings, Karin van de Sande, Theo Gevers, and Arnold Smeulders · 2013
Earlier work this paper cites.
Hierarchical saliency detection
Qiong Yan, Li Xu, Jianping Shi, and Jiaya Jia · 2013
Earlier work this paper cites.
Saliency detection via graph-based manifold ranking
Chuan Yang, Lihe Zhang, Huchuan Lu, Xiang Ruan, and Ming-Hsuan Yang · 2013
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba · 2014
Earlier work this paper cites.
Microsoft COCO: common objects in context
Tsung-Yi Lin, Michael Maire, Serge Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Dollár, and Lawrence Zitnick · 2014
Earlier work this paper cites.
Conditional generative adversarial nets
Mehdi Mirza and Simon Osindero · 2014
Earlier work this paper cites.
Saliency optimization from robust background detection
Wangjiang Zhu, Shuang Liang, Yichen Wei, and Jian Sun · 2014
Earlier work this paper cites.
Edge boxes: Locating object proposals from edges
Lawrence Zitnick and Piotr Dollár · 2014
Earlier work this paper cites.
A weighted sparse coding framework for saliency detection
Nianyi Li, Bilin Sun, and Jingyi Yu · 2015
Earlier work this paper cites.
U-net: Convolutional networks for biomedical image segmentation
Olaf Ronneberger, Philipp Fischer, and Thomas Brox · 2015
Earlier work this paper cites.
Hierarchical image saliency detection on extended cssd
Jianping Shi, Qiong Yan, Li Xu, and Jiaya Jia · 2015
Earlier work this paper cites.
Deep unsupervised learning using nonequilibrium thermodynamics
Jascha Sohl-Dickstein, Eric Weiss, Niru Maheswaranathan, and Surya Ganguli · 2015
Earlier work this paper cites.
The fast bilateral solver
Jonathan T Barron and Ben Poole · 2016
Earlier work this paper cites.
Deeply supervised salient object detection with short connections
Qibin Hou, Ming-Ming Cheng, Xiaowei Hu, A. Borji, Z. Tu, and Philip H. S. Torr · 2016
Earlier work this paper cites.
Rethinking atrous convolution for semantic image segmentation
Liang-Chieh Chen, George Papandreou, Florian Schroff, and Hartwig Adam · 2017
Earlier work this paper cites.
A point set generation network for 3d object reconstruction from a single image
Haoqiang Fan, Hao Su, and Leonidas J Guibas · 2017
Earlier work this paper cites.
Image-to-image translation with conditional adversarial networks
Phillip Isola, Jun-Yan Zhu, Tinghui Zhou, and Alexei A Efros · 2017
Earlier work this paper cites.
Learning to detect salient objects with image-level supervision
Lijun Wang, Huchuan Lu, Yifan Wang, Mengyang Feng, Dong Wang, Baocai Yin, and Xiang Ruan · 2017
Earlier work this paper cites.
Unpaired image-to-image translation using cycle-consistent adversarial networks
Jun-Yan Zhu, Taesung Park, Phillip Isola, and Alexei A Efros · 2017
Cited alongside, same era.
Large scale gan training for high fidelity natural image synthesis
Andrew Brock, Jeff Donahue, and Karen Simonyan · 2018
Cited alongside, same era.
Multimodal unsupervised image-to-image translation
Xun Huang, Ming-Yu Liu, Serge Belongie, and Jan Kautz · 2018
Cited alongside, same era.
Emergence of object segmentation in perturbed generative models
Adam Bielski and Paolo Favaro · 2019
Cited alongside, same era.
Unsupervised object segmentation by redrawing
Mickaël Chen, Thierry Artières, and Ludovic Denoyer · 2019
Cited alongside, same era.
Invariant information clustering for unsupervised image classification and segmentation
Learning transferable visual models from natural language supervision
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, et al · 2021
Later among the works it cites.
Localizing objects with self-supervised transformers and no labels
Oriane Siméoni, Gilles Puy, Huy V Vo, Simon Roburin, Spyros Gidaris, Andrei Bursuc, Patrick Pérez, Renaud Marlet, and Jean Ponce · 2021
Later among the works it cites.
Large-scale unsupervised object discovery
Huy V. Vo, Elena Sizikova, Cordelia Schmid, Patrick Pérez, and Jean Ponce · 2021
Later among the works it cites.
Object segmentation without labels with large-scale generative models
Andrey Voynov, Stanislav Morozov, and Artem Babenko · 2021
Later among the works it cites.
Open-vocabulary object detection using captions
Alireza Zareian, Kevin Dela Rosa, Derek Hao Hu, and Shih-Fu Chang · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Xu Ji, Joao F Henriques, and Andrea Vedaldi · 2019
Cited alongside, same era.
A style-based generator architecture for generative adversarial networks
Tero Karras, Samuli Laine, and Timo Aila · 2019
Cited alongside, same era.
Deepusps: Deep robust unsupervised saliency prediction with self-supervision
Duc Tam Nguyen, Maximilian Dax, Chaithanya Kumar Mummadi, Thi-Phuong-Nhung Ngo, Thi Hoai Phuong Nguyen, Zhongyu Lou, and Thomas Brox · 2019
Cited alongside, same era.
Generative modeling by estimating gradients of the data distribution
Yang Song and Stefano Ermon · 2019
Cited alongside, same era.
Unsupervised object discovery and co-localization by deep descriptor transforming
Xiu-Shen Wei, Chen-Lin Zhang, Jianxin Wu, Chunhua Shen, and Zhi-Hua Zhou · 2019
Cited alongside, same era.
Pyramid feature attention network for saliency detection
Ting Zhao and Xiangqian Wu · 2019
Cited alongside, same era.
Onegan: Simultaneous unsupervised learning of conditional image generation, foreground segmentation, and fine-grained clustering
Yaniv Benny and Lior Wolf · 2020
Cited alongside, same era.
Datasetgan: Efficient labeled data factory with minimal human effort
Yuxuan Zhang, Huan Ling, Jun Gao, Kangxue Yin, Jean-Francois Lafleche, Adela Barriuso, Antonio Torralba, and Sanja Fidler · 2021
Later among the works it cites.
Masked autoencoders are scalable vision learners
Kaiming He, Xinlei Chen, Saining Xie, Yanghao Li, Piotr Dollár, and Ross Girshick · 2022
Later among the works it cites.
Is synthetic data from generative models ready for image recognition?
Ruifei He, Shuyang Sun, Xin Yu, Chuhui Xue, Wenqing Zhang, Philip Torr, Song Bai, and Xiaojuan Qi · 2022
Later among the works it cites.
Prompting visual-language models for efficient video understanding
Chen Ju, Tengda Han, Kunhao Zheng, Ya Zhang, and Weidi Xie · 2022
Later among the works it cites.
Adaptive mutual supervision for weakly-supervised temporal action localization
Chen Ju, Peisen Zhao, Siheng Chen, Ya Zhang, Xiaoyun Zhang, Yanfeng Wang, and Qi Tian · 2022
Later among the works it cites.
Chen Ju, Kunhao Zheng, Jinxiang Liu, Peisen Zhao, Ya Zhang, Jianlong Chang, Yanfeng Wang, and Qi Tian · 2022
Later among the works it cites.
Bigdatasetgan: Synthesizing imagenet with pixel-wise annotations
Daiqing Li, Huan Ling, Seung Wook Kim, Karsten Kreis, Sanja Fidler, and Antonio Torralba · 2022
Later among the works it cites.
Repaint: Inpainting using denoising diffusion probabilistic models
Andreas Lugmayr, Martin Danelljan, Andres Romero, Fisher Yu, Radu Timofte, and Luc Van Gool · 2022
Later among the works it cites.
Open-vocabulary semantic segmentation with frozen vision-language models
Chaofan Ma, Yuhuan Yang, Yanfeng Wang, Ya Zhang, and Weidi Xie · 2022
Later among the works it cites.
Deep spectral methods: A surprisingly strong baseline for unsupervised semantic segmentation and localization
Luke Melas-Kyriazi, C. Rupprecht, Iro Laina, and Andrea Vedaldi · 2022
Later among the works it cites.
Chatgpt: Optimizing language models for dialogue, 2022
OpenAI · 2022
Later among the works it cites.
Hierarchical text-conditional image generation with clip latents
Aditya Ramesh, Prafulla Dhariwal, Alex Nichol, Casey Chu, and Mark Chen · 2022
Later among the works it cites.
High-resolution image synthesis with latent diffusion models
Robin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser, and Björn Ommer · 2022
Later among the works it cites.
Palette: Image-to-image diffusion models
Chitwan Saharia, William Chan, Huiwen Chang, Chris Lee, Jonathan Ho, Tim Salimans, David Fleet, and Mohammad Norouzi · 2022
Later among the works it cites.
Photorealistic text-to-image diffusion models with deep language understanding
Chitwan Saharia, William Chan, Saurabh Saxena, Lala Li, Jay Whang, Emily Denton, Seyed Kamyar Seyed Ghasemipour, Burcu Karagol Ayan, S Sara Mahdavi, Rapha Gontijo Lopes, et al · 2022
Later among the works it cites.
Image super-resolution via iterative refinement
Chitwan Saharia, Jonathan Ho, William Chan, Tim Salimans, David J Fleet, and Mohammad Norouzi · 2022
Later among the works it cites.
Laion-5b: An open large-scale dataset for training next generation image-text models
Christoph Schuhmann, Romain Beaumont, Richard Vencu, Cade Gordon, Ross Wightman, Mehdi Cherti, Theo Coombes, Aarush Katta, Clayton Mullis, Mitchell Wortsman, et al · 2022
Later among the works it cites.
Bridging the gap to real-world object-centric learning
Maximilian Seitzer, Max Horn, Andrii Zadaianchuk, Dominik Zietlow, Tianjun Xiao, Carl-Johann Simon-Gabriel, Tong He, Zheng Zhang, Bernhard Schölkopf, Thomas Brox, and Francesco Locatello · 2022
Later among the works it cites.
Learning co-segmentation by segment swapping for retrieval and discovery
Xi Shen, Alexei A Efros, Armand Joulin, and Mathieu Aubry · 2022
Later among the works it cites.
Unsupervised salient object detection with spectral cluster voting
Gyungin Shin, Samuel Albanie, and Weidi Xie · 2022
Later among the works it cites.
Unsupervised object localization: Observing the background to discover objects
Oriane Siméoni, Chloé Sekkat, Gilles Puy, Antonin Vobecky, Éloi Zablocki, and Patrick Pérez · 2022
Later among the works it cites.
Clip-nerf: Text-and-image driven manipulation of neural radiance fields
Can Wang, Menglei Chai, Mingming He, Dongdong Chen, and Jing Liao · 2022
Later among the works it cites.
Freesolo: Learning to segment objects without annotations
Xinlong Wang, Zhiding Yu, Shalini De Mello, Jan Kautz, Anima Anandkumar, Chunhua Shen, and José Manuel Álvarez · 2022
Later among the works it cites.
Self-supervised transformers for unsupervised object discovery using normalized cut
Yangtao Wang, Xi Shen, Shell Xu Hu, Yuan Yuan, James L Crowley, and Dominique Vaufreydaz · 2022
Later among the works it cites.
Selfreformer: Self-refined network with transformer for salient object detection
Yi Ke Yun and Weisi Lin · 2022
Later among the works it cites.
Constraint and union for partially-supervised temporal sentence grounding
Chen Ju, Haicheng Wang, Jinxiang Liu, Chaofan Ma, Ya Zhang, Peisen Zhao, Jianlong Chang, and Qi Tian · 2023
Closest in time.