Fetching the paper…
Reading the bibliography…
Interaction and collaboration between humans and intelligent machines has become increasingly important as machine learning methods move into real-world applications that involve end users.
Snakes, shapes, and gradient vector flow
C. Xu and J. L. Prince · 1998
Earlier work this paper cites.
Interactive graph cuts for optimal boundary & region segmentation of objects in nd images
Y. Y. Boykov and M.-P. Jolly · 2001
Earlier work this paper cites.
Grabcut: Interactive foreground extraction using iterated graph cuts
C. Rother, V. Kolmogorov, and A. Blake · 2004
Earlier work this paper cites.
Random walks for image segmentation
L. Grady · 2006
Earlier work this paper cites.
Image segmentation with a bounding box prior
V. Lempitsky, P. Kohli, C. Rother, and T. Sharp · 2009
Earlier work this paper cites.
icoseg: Interactive co-segmentation with intelligent scribble guidance
D. Batra, A. Kowdle, D. Parikh, J. Luo, and T. Chen · 2010
Earlier work this paper cites.
Geodesic graph cut for interactive image segmentation
B. L. Price, B. Morse, and S. Cohen · 2010
Earlier work this paper cites.
Segmentation from a box
L. Grady, M.-P. Jolly, and A. Seitz · 2011
Earlier work this paper cites.
The PASCAL Visual Object Classes Challenge 2012 (VOC2012) Results
M. Everingham, L. Van Gool, C. K. I. Williams, J. Winn, and A. Zisserman · 2012
Earlier work this paper cites.
How interaction methods affect image segmentation: user experience in the task
R. Hebbalaguppe, K. McGuinness, J. Kuklyte, G. Healy, N. O’Connor, and A. Smeaton · 2013
Earlier work this paper cites.
The ignorant led by the blind: A hybrid human–machine vision system for fine-grained categorization
S. Branson, G. Van Horn, C. Wah, P. Perona, and S. Belongie · 2014
Earlier work this paper cites.
On the properties of neural machine translation: Encoder-decoder approaches
K. Cho, B. Van Merriënboer, D. Bahdanau, and Y. Bengio · 2014
Earlier work this paper cites.
Empirical evaluation of gated recurrent neural networks on sequence modeling
J. Chung, C. Gulcehre, K. Cho, and Y. Bengio · 2014
Earlier work this paper cites.
Referit game: Referring to objects in photographs of natural scenes
S. Kazemzadeh, V. Ordonez, M. Matten, and T. L. Berg · 2014
Earlier work this paper cites.
Microsoft coco: Common objects in context
T.-Y. Lin, M. Maire, S. Belongie, J. Hays, P. Perona, D. Ramanan, P. Dollár, and C. L. Zitnick · 2014
Earlier work this paper cites.
Deep captioning with multimodal recurrent neural networks (m-rnn)
J. Mao, W. Xu, Y. Yang, J. Wang, Z. Huang, and A. Yuille · 2014
Earlier work this paper cites.
Glove: Global vectors for word representation
J. Pennington, R. Socher, and C. Manning · 2014
Earlier work this paper cites.
Segnet: A deep convolutional encoder-decoder architecture for image segmentation
V. Badrinarayanan, A. Kendall, and R. Cipolla · 2015
Earlier work this paper cites.
Boxsup: Exploiting bounding boxes to supervise convolutional networks for semantic segmentation
J. Dai, K. He, and J. Sun · 2015
Earlier work this paper cites.
Long-term recurrent convolutional networks for visual recognition and description
J. Donahue, L. Anne Hendricks, S. Guadarrama, M. Rohrbach, S. Venugopalan, K. Saenko, and T. Darrell · 2015
Earlier work this paper cites.
Batch normalization: Accelerating deep network training by reducing internal covariate shift
S. Ioffe and C. Szegedy · 2015
Earlier work this paper cites.
Deep visual-semantic alignments for generating image descriptions
A. Karpathy and L. Fei-Fei · 2015
Cited alongside, same era.
Fully convolutional networks for semantic segmentation
J. Long, E. Shelhamer, and T. Darrell · 2015
Cited alongside, same era.
Ask your neurons: A neural-based approach to answering questions about images
M. Malinowski, M. Rohrbach, and M. Fritz · 2015
Cited alongside, same era.
Image segmentation in twenty questions
C. Rupprecht, L. Peter, and N. Navab · 2015
Cited alongside, same era.
Show and tell: A neural image caption generator
O. Vinyals, A. Toshev, S. Bengio, and D. Erhan · 2015
Cited alongside, same era.
Show, attend and tell: Neural image caption generation with visual attention
K. Xu, J. Ba, R. Kiros, K. Cho, A. Courville, R. Salakhudinov, R. Zemel, and Y. Bengio · 2015
Cited alongside, same era.
Generation and comprehension of unambiguous object descriptions
J. Mao, J. Huang, A. Toshev, O. Camburu, A. Yuille, and K. Murphy · 2016
Later among the works it cites.
Full-resolution residual networks for semantic segmentation in street scenes
T. Pohlen, A. Hermans, M. Mathias, and B. Leibe · 2016
Later among the works it cites.
Ask, attend and answer: Exploring question-guided spatial attention for visual question answering
H. Xu and K. Saenko · 2016
Later among the works it cites.
Deep interactive object selection
N. Xu, B. Price, S. Cohen, J. Yang, and T. S. Huang · 2016
Later among the works it cites.
Modeling context in referring expressions
L. Yu, P. Poirson, S. Yang, A. C. Berg, and T. L. Berg · 2016
Later among the works it cites.
Vqa: Visual question answering
A. Agrawal, J. Lu, S. Antol, M. Mitchell, C. L. Zitnick, D. Parikh, and D. Batra · 2017
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Simple baseline for visual question answering
B. Zhou, Y. Tian, S. Sukhbaatar, A. Szlam, and R. Fergus · 2015
Cited alongside, same era.
Learning to compose neural networks for question answering
J. Andreas, M. Rohrbach, T. Darrell, and D. Klein · 2016
Cited alongside, same era.
Neural module networks
J. Andreas, M. Rohrbach, T. Darrell, and D. Klein · 2016
Cited alongside, same era.
Coco-stuff: Thing and stuff classes in context
H. Caesar, J. Uijlings, and V. Ferrari · 2016
Cited alongside, same era.
Interactive segmentation from 1-bit feedback
D.-J. Chen, H.-T. Chen, and L.-W. Chang · 2016
Cited alongside, same era.
L.-C. Chen, G. Papandreou, I. Kokkinos, K. Murphy, and A. L. Yuille · 2016
Cited alongside, same era.
Later among the works it cites.
M. Amrehn, S. Gaube, M. Unberath, F. Schebesch, T. Horz, M. Strumia, S. Steidl, M. Kowarschik, and A. Maier · 2017
Later among the works it cites.
Bottom-up and top-down attention for image captioning and vqa
P. Anderson, X. He, C. Buehler, D. Teney, M. Johnson, S. Gould, and L. Zhang · 2017
Later among the works it cites.
Rethinking atrous convolution for semantic image segmentation
L.-C. Chen, G. Papandreou, F. Schroff, and H. Adam · 2017
Later among the works it cites.
Towards diverse and natural image descriptions via a conditional gan
B. Dai, D. Lin, R. Urtasun, and S. Fidler · 2017
Later among the works it cites.
Guesswhat?! visual object discovery through multi-modal dialogue
H. De Vries, F. Strub, S. Chandar, O. Pietquin, H. Larochelle, and A. Courville · 2017
Later among the works it cites.
Modulating early visual processing by language
H. De Vries, F. Strub, J. Mary, H. Larochelle, O. Pietquin, and A. C. Courville · 2017
Later among the works it cites.
Learning to reason: End-to-end module networks for visual question answering
R. Hu, J. Andreas, M. Rohrbach, T. Darrell, and K. Saenko · 2017
Later among the works it cites.
The one hundred layers tiramisu: Fully convolutional densenets for semantic segmentation
S. Jégou, M. Drozdzal, D. Vazquez, A. Romero, and Y. Bengio · 2017
Later among the works it cites.
Inferring and executing programs for visual reasoning
J. Johnson, B. Hariharan, L. van der Maaten, J. Hoffman, L. Fei-Fei, C. L. Zitnick, and R. Girshick · 2017
Later among the works it cites.
Ask your neurons: A deep learning approach to visual question answering
M. Malinowski, M. Rohrbach, and M. Fritz · 2017
Later among the works it cites.
Deepcut: Object segmentation from bounding box annotations using convolutional neural networks
M. Rajchl, M. C. Lee, O. Oktay, K. Kamnitsas, J. Passerat-Palmbach, W. Bai, M. Damodaram, M. A. Rutherford, J. V. Hajnal, B. Kainz, et al · 2017
Later among the works it cites.
Deepigeos: A deep interactive geodesic framework for medical image segmentation
G. Wang, M. A. Zuluaga, W. Li, R. Pratt, P. A. Patel, M. Aertsen, T. Doel, A. L. David, J. Deprest, S. Ourselin, et al · 2017
Later among the works it cites.
Real-time user-guided image colorization with learned deep priors
R. Zhang, J.-Y. Zhu, P. Isola, X. Geng, A. S. Lin, T. Yu, and A. A. Efros · 2017
Later among the works it cites.
Film: Visual reasoning with a general conditioning layer
E. Perez, F. Strub, H. de Vries, V. Dumoulin, and A. C. Courville · 2018
Closest in time.