Fetching the paper…
Reading the bibliography…
We introduce the Segment Anything (SA) project: a new task, model, and dataset for image segmentation.
A computational approach to edge detection
John Canny · 1986
Earlier work this paper cites.
The validity and practicality of sun-reactive skin types i through vi
Thomas B. Fitzpatrick · 1988
Earlier work this paper cites.
Snakes: Active contour models
Michael Kass, Andrew Witkin, and Demetri Terzopoulos · 1988
Earlier work this paper cites.
Is learning the n-th thing any easier than learning the first?
Sebastian Thrun · 1995
Earlier work this paper cites.
Adaptive background mixture models for real-time tracking
Chris Stauffer and W Eric L Grimson · 1999
Earlier work this paper cites.
On seeing stuff: the perception of materials by humans and machines
Edward H Adelson · 2001
Earlier work this paper cites.
A database of human segmented natural images and its application to evaluating segmentation algorithms and measuring ecological statistics
David Martin, Charless Fowlkes, Doron Tal, and Jitendra Malik · 2001
Earlier work this paper cites.
Learning a classification model for segmentation
Xiaofeng Ren and Jitendra Malik · 2003
Earlier work this paper cites.
Efficient graph-based image segmentation
Pedro F Felzenszwalb and Daniel P Huttenlocher · 2004
Earlier work this paper cites.
TextonBoost: Joint appearance, shape and context modeling for mulit-class object recognition and segmentation
Jamie Shotton, John Winn, Carsten Rother, and Antonio Criminisi · 2006
Earlier work this paper cites.
Automatic image colorization via multimodal predictions
Guillaume Charpiat, Matthias Hofmann, and Bernhard Schölkopf · 2008
Earlier work this paper cites.
What is an object?
Bogdan Alexe, Thomas Deselaers, and Vittorio Ferrari · 2010
Earlier work this paper cites.
Contour detection and hierarchical image segmentation
Pablo Arbeláez, Michael Maire, Charless Fowlkes, and Jitendra Malik · 2010
Earlier work this paper cites.
SUN database: Large-scale scene recognition from abbey to zoo
Jianxiong Xiao, James Hays, Krista Ehinger, Aude Oliva, and Antonio Torralba · 2010
Earlier work this paper cites.
Learning to recognize objects in egocentric activities
Alireza Fathi, Xiaofeng Ren, and James M. Rehg · 2011
Earlier work this paper cites.
Segmentation as selective search for object recognition
Koen EA van de Sande, Jasper RR Uijlings, Theo Gevers, and Arnold WM Smeulders · 2011
Earlier work this paper cites.
Learning parameterized skills
Bruno da Silva, George Konidaris, and Andrew Barto · 2012
Earlier work this paper cites.
Multiple choice learning: Learning to produce multiple structured outputs
Abner Guzman-Rivera, Dhruv Batra, and Pushmeet Kohli · 2012
Earlier work this paper cites.
Fast edge detection using structured forests
Piotr Dollár and C Lawrence Zitnick · 2014
Earlier work this paper cites.
Rich feature hierarchies for accurate object detection and semantic segmentation
Ross Girshick, Jeff Donahue, Trevor Darrell, and Jitendra Malik · 2014
Earlier work this paper cites.
Microsoft COCO: Common objects in context
Tsung-Yi Lin, Michael Maire, Serge Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Dollár, and C Lawrence Zitnick · 2014
Earlier work this paper cites.
Gelatinous zooplankton biomass in the global oceans: geographic variation and environmental drivers
Cathy H Lucas, Daniel OB Jones, Catherine J Hollyhead, Robert H Condon, Carlos M Duarte, William M Graham, Kelly L Robinson, Kylie A Pitt, Mark Schildhauer, and Jim Regetz · 2014
Earlier work this paper cites.
Dropout: A simple way to prevent neural networks from overfitting
Nitish Srivastava, Geoffrey Hinton, Alex Krizhevsky, Ilya Sutskever, and Ruslan Salakhutdinov · 2014
Earlier work this paper cites.
Delving into egocentric actions
Yin Li, Zhefan Ye, and James M. Rehg · 2015
Earlier work this paper cites.
Faster R-CNN: Towards real-time object detection with region proposal networks
Shaoqing Ren, Kaiming He, Ross Girshick, and Jian Sun · 2015
Earlier work this paper cites.
Holistically-nested edge detection
Saining Xie and Zhuowen Tu · 2015
Earlier work this paper cites.
Jimmy Lei Ba, Jamie Ryan Kiros, and Geoffrey E Hinton · 2016
Earlier work this paper cites.
Object-proposal evaluation protocol is’ gameable’
Neelima Chavali, Harsh Agrawal, Aroma Mahendru, and Dhruv Batra · 2016
Earlier work this paper cites.
The Cityscapes dataset for semantic urban scene understanding
Marius Cordts, Mohamed Omran, Sebastian Ramos, Timo Rehfeld, Markus Enzweiler, Rodrigo Benenson, Uwe Franke, Stefan Roth, and Bernt Schiele · 2016
Earlier work this paper cites.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Earlier work this paper cites.
Gaussian error linear units (gelus)
Dan Hendrycks and Kevin Gimpel · 2016
Earlier work this paper cites.
Deep networks with stochastic depth
Gao Huang, Yu Sun, Zhuang Liu, Daniel Sedra, and Kilian Q Weinberger · 2016
Earlier work this paper cites.
V-Net: Fully convolutional neural networks for volumetric medical image segmentation
Fausto Milletari, Nassir Navab, and Seyed-Ahmad Ahmadi · 2016
Earlier work this paper cites.
Finely-grained annotated datasets for image-based plant phenotyping
Massimo Minervini, Andreas Fischbach, Hanno Scharr, and Sotirios A. Tsaftaris · 2016
Earlier work this paper cites.
Deep interactive object selection
Ning Xu, Brian Price, Scott Cohen, Jimei Yang, and Thomas S Huang · 2016
Earlier work this paper cites.
Accurate, large minibatch SGD: Training ImageNet in 1 hour
Priya Goyal, Piotr Dollár, Ross Girshick, Pieter Noordhuis, Lukasz Wesolowski, Aapo Kyrola, Andrew Tulloch, Yangqing Jia, and Kaiming He · 2017
Earlier work this paper cites.
Mask R-CNN
Kaiming He, Georgia Gkioxari, Piotr Dollár, and Ross Girshick · 2017
Earlier work this paper cites.
Focal loss for dense object detection
Tsung-Yi Lin, Priya Goyal, Ross Girshick, Kaiming He, and Piotr Dollár · 2017
Earlier work this paper cites.
Extreme clicking for efficient object annotation
Dim P Papadopoulos, Jasper RR Uijlings, Frank Keller, and Vittorio Ferrari · 2017
Earlier work this paper cites.
Semi-supervised sequence tagging with bidirectional language models
Matthew E Peters, Waleed Ammar, Chandra Bhagavatula, and Russell Power · 2017
Cited alongside, same era.
Action recognition in RGB-D egocentric videos
Yansong Tang, Yi Tian, Jiwen Lu, Jianjiang Feng, and Jie Zhou · 2017
Cited alongside, same era.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Lukasz Kaiser, and Illia Polosukhin · 2017
Cited alongside, same era.
Men also like shopping: Reducing gender bias amplification using corpus-level constraints
Jieyu Zhao, Tianlu Wang, Mark Yatskar, Vicente Ordonez, and Kai-Wei Chang · 2017
Cited alongside, same era.
Places: A 10 million image database for scene recognition
Bolei Zhou, Agata Lapedriza, Aditya Khosla, Aude Oliva, and Antonio Torralba · 2017
Cited alongside, same era.
Simple copy-paste is a strong data augmentation method for instance segmentation
Golnaz Ghiasi, Yin Cui, Aravind Srinivas, Rui Qian, Tsung-Yi Lin, Ekin D Cubuk, Quoc V Le, and Barret Zoph · 2021
Later among the works it cites.
Scaling up visual and vision-language representation learning with noisy text supervision
Chao Jia, Yinfei Yang, Ye Xia, Yi-Ting Chen, Zarana Parekh, Hieu Pham, Quoc Le, Yun-Hsuan Sung, Zhen Li, and Tom Duerig · 2021
Later among the works it cites.
Carbon emissions and large neural network training
David Patterson, Joseph Gonzalez, Quoc Le, Chen Liang, Lluis-Miquel Munguia, Daniel Rothchild, David So, Maud Texier, and Jeff Dean · 2021
Later among the works it cites.
Learning transferable visual models from natural language supervision
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, et al · 2021
Later among the works it cites.
Zero-shot text-to-image generation
Aditya Ramesh, Mikhail Pavlov, Gabriel Goh, Scott Gray, Chelsea Voss, Alec Radford, Mark Chen, and Ilya Sutskever · 2021
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Gustav Bredell, Christine Tanner, and Ender Konukoglu · 2018
Cited alongside, same era.
Cascade R-CNN: Delving into high quality object detection
Zhaowei Cai and Nuno Vasconcelos · 2018
Cited alongside, same era.
Interactive image segmentation with latent diversity
Zhuwen Li, Qifeng Chen, and Vladlen Koltun · 2018
Cited alongside, same era.
Iteratively trained interactive segmentation
Sabarinath Mahadevan, Paul Voigtlaender, and Bastian Leibe · 2018
Cited alongside, same era.
Deep extreme cut: From extreme points to object segmentation
Kevis-Kokitsi Maninis, Sergi Caelles, Jordi Pont-Tuset, and Luc Van Gool · 2018
Cited alongside, same era.
ilastik: interactive machine learning for (bio)image analysis
Stuart Berg, Dominik Kutra, Thorben Kroeger, Christoph N. Straehle, Bernhard X. Kausler, Carsten Haubold, Martin Schiegg, Janez Ales, Thorsten Beier, Markus Rudy, Kemal Eren, Jaime I. Cervantes, Buote Xu, Fynn Beuttenmueller, Adrian Wolny, Chong Zhang, Ullrich Koethe, Fred A. Hamprecht, and Anna Kreshuk · 2019
Cited alongside, same era.
Nucleus segmentation across imaging experiments: the 2018 data science bowl
Juan C. Caicedo, Allen Goodman, Kyle W. Karhohs, Beth A. Cimini, Jeanelle Ackerman, Marzieh Haghighi, CherKeng Heng, Tim Becker, Minh Doan, Claire McQuin, Mohammad Rohban, Shantanu Singh, and Anne E. Carpenter · 2019
Cited alongside, same era.
Later among the works it cites.
Hypersim: A photorealistic synthetic dataset for holistic indoor scene understanding
Mike Roberts, Jason Ramapuram, Anurag Ranjan, Atulit Kumar, Miguel Angel Bautista, Nathan Paczan, Russ Webb, and Joshua M. Susskind · 2021
Later among the works it cites.
A step toward more inclusive people annotations for fairness
Candice Schumann, Susanna Ricco, Utsav Prabhu, Vittorio Ferrari, and Caroline Pantofaru · 2021
Later among the works it cites.
HyperExtended LightFace: A facial attribute analysis framework
Sefik Ilkin Serengil and Alper Ozpinar · 2021
Later among the works it cites.
Towards real-world prohibited item detection: A large-scale x-ray benchmark
Boying Wang, Libo Zhang, Longyin Wen, Xianglong Liu, and Yanjun Wu · 2021
Later among the works it cites.
iShape: A first step towards irregular shape instance segmentation
Lei Yang, Yan Zi Wei, Yisheng HE, Wei Sun, Zhenhang Huang, Haibin Huang, and Haoqiang Fan · 2021
Later among the works it cites.
K-Net: Towards unified image segmentation
Wenwei Zhang, Jiangmiao Pang, Kai Chen, and Chen Change Loy · 2021
Later among the works it cites.
ZeroWaste dataset: Towards deformable object segmentation in cluttered scenes
Dina Bashkirova, Mohamed Abdelfattah, Ziliang Zhu, James Akl, Fadi Alladkani, Ping Hu, Vitaly Ablavsky, Berk Calli, Sarah Adel Bargal, and Kate Saenko · 2022
Later among the works it cites.
3D instance segmentation of MVS buildings
Jiazhou Chen, Yanghui Xu, Shufang Lu, Ronghua Liang, and Liangliang Nan · 2022
Later among the works it cites.
FocalClick: towards practical interactive image segmentation
Xi Chen, Zhiyan Zhao, Yilei Zhang, Manni Duan, Donglian Qi, and Hengshuang Zhao · 2022
Later among the works it cites.
Masked-attention mask transformer for universal image segmentation
Bowen Cheng, Ishan Misra, Alexander G Schwing, Alexander Kirillov, and Rohit Girdhar · 2022
Later among the works it cites.
PaLM: Scaling language modeling with pathways
Aakanksha Chowdhery, Sharan Narang, Jacob Devlin, Maarten Bosma, Gaurav Mishra, Adam Roberts, Paul Barham, Hyung Won Chung, Charles Sutton, Sebastian Gehrmann, et al · 2022
Later among the works it cites.
Night and day instance segmented park (NDISPark) dataset: a collection of images taken by day and by night for vehicle detection, segmentation and counting in parking areas
Luca Ciampi, Carlos Santiago, Joao Costeira, Claudio Gennaro, and Giuseppe Amato · 2022
Later among the works it cites.
Semantic segmentation in art paintings
Nadav Cohen, Yael Newman, and Ariel Shamir · 2022
Later among the works it cites.
Rescaling egocentric vision: Collection, pipeline and challenges for EPIC-KITCHENS-100
Dima Damen, Hazel Doughty, Giovanni Maria Farinella, Antonino Furnari, Jian Ma, Evangelos Kazakos, Davide Moltisanti, Jonathan Munro, Toby Perrett, Will Price, and Michael Wray · 2022
Later among the works it cites.
EPIC-KITCHENS VISOR benchmark: Video segmentations and object relations
Ahmad Darkhalil, Dandan Shan, Bin Zhu, Jian Ma, Amlan Kar, Richard Higgins, Sanja Fidler, David Fouhey, and Dima Damen · 2022
Later among the works it cites.
CrowdWorkSheets: Accounting for individual and collective identities underlying crowdsourced dataset annotation
Mark Díaz, Ian Kivlichan, Rachel Rosen, Dylan Baker, Razvan Amironesei, Vinodkumar Prabhakaran, and Emily Denton · 2022
Later among the works it cites.
Instance segmentation for autonomous log grasping in forestry operations
Jean-Michel Fortin, Olivier Gamache, Vincent Grondin, François Pomerleau, and Philippe Giguère · 2022
Later among the works it cites.
SOCRATES: Introducing depth in visual wildlife monitoring using stereo vision
Timm Haucke, Hjalmar S. Kühl, and Volker Steinhage · 2022
Later among the works it cites.
Masked autoencoders are scalable vision learners
Kaiming He, Xinlei Chen, Saining Xie, Yanghao Li, Piotr Dollár, and Ross Girshick · 2022
Later among the works it cites.
Training compute-optimal large language models
Jordan Hoffmann, Sebastian Borgeaud, Arthur Mensch, Elena Buchatskaya, Trevor Cai, Eliza Rutherford, Diego de Las Casas, Lisa Anne Hendricks, Johannes Welbl, Aidan Clark, et al · 2022
Later among the works it cites.
Oneformer: One transformer to rule universal image segmentation
Jitesh Jain, Jiachen Li, MangTik Chiu, Ali Hassani, Nikita Orlov, and Humphrey Shi · 2022
Later among the works it cites.
Learning open-world object proposals without learning to classify
Dahun Kim, Tsung-Yi Lin, Anelia Angelova, In So Kweon, and Weicheng Kuo · 2022
Later among the works it cites.
Exploring plain vision transformer backbones for object detection
Yanghao Li, Hanzi Mao, Ross Girshick, and Kaiming He · 2022
Later among the works it cites.
SimpleClick: Interactive image segmentation with simple vision transformers
Qin Liu, Zhenlin Xu, Gedas Bertasius, and Marc Niethammer · 2022
Later among the works it cites.
EDTER: Edge detection with transformer
Mengyang Pu, Yaping Huang, Yuming Liu, Qingji Guan, and Haibin Ling · 2022
Later among the works it cites.
DOORS: Dataset fOr bOuldeRs Segmentation
Mattia Pugliatti and Francesco Topputo · 2022
Later among the works it cites.
Occluded video instance segmentation: A benchmark
Jiyang Qi, Yan Gao, Yao Hu, Xinggang Wang, Xiaoyu Liu, Xiang Bai, Serge Belongie, Alan Yuille, Philip Torr, and Song Bai · 2022
Later among the works it cites.
Reviving iterative training with mask guidance for interactive segmentation
Konstantin Sofiiuk, Ilya A Petrov, and Anton Konushin · 2022
Later among the works it cites.
The world by income and regions, 2022
The World Bank · 2022
Later among the works it cites.
Greenhouse Gas Equivalencies Calculator
United States Environmental Protection Agency · 2022
Later among the works it cites.
Open-world instance segmentation: Exploiting pseudo ground truth from learned pairwise affinity
Weiyao Wang, Matt Feiszli, Heng Wang, Jitendra Malik, and Du Tran · 2022
Later among the works it cites.
Fine-grained egocentric hand-object segmentation: Dataset, model, and applications
Lingzhi Zhang, Shenghao Zhou, Simon Stent, and Jianbo Shi · 2022
Later among the works it cites.
Multiview compressive coding for 3D reconstruction
Chao-Yuan Wu, Justin Johnson, Jitendra Malik, Christoph Feichtenhofer, and Georgia Gkioxari · 2023
Closest in time.