Fetching the paper…
Reading the bibliography…
It is common to implicitly assume access to intelligently captured inputs (e.g., photos from a human photographer), yet autonomously capturing good observations is itself a major challenge.
Canonical perspective and the perception of objects
S. Palmer, E. Rosch, and P. Chase · 1981
Earlier work this paper cites.
Perception of partly occluded objects in infancy
P. J. Kellman and E. S. Spelke · 1983
Earlier work this paper cites.
Active vision
J. Aloimonos, I. Weiss, and A. Bandyopadhyay · 1988
Earlier work this paper cites.
Active perception
R. Bajcsy · 1988
Earlier work this paper cites.
Learning stochastic feedforward networks
R. M. Neal · 1990
Earlier work this paper cites.
Animate vision
D. Ballard · 1991
Earlier work this paper cites.
Orientation dependence in the recognition of familiar and novel views of three-dimensional objects
S. Edelman and H. H. Bülthoff · 1992
Earlier work this paper cites.
Active object recognition
D. Wilkes and J. Tsotsos · 1992
Earlier work this paper cites.
Simple statistical gradient-following algorithms for connectionist reinforcement learning
R. J. Williams · 1992
Earlier work this paper cites.
Long short-term memory
S. Hochreiter and J. Schmidhuber · 1997
Earlier work this paper cites.
Simultaneous localization and map-building using active vision
A. J. Davison and D. W. Murray · 2002
Earlier work this paper cites.
Sensor placement for effective coverage and surveillance in distributed sensor networks
S. S. Dhillon and K. Chakrabarty · 2003
Earlier work this paper cites.
Multiple view geometry in computer vision
R. Hartley and A. Zisserman · 2003
Earlier work this paper cites.
Neurobiology of attention, chapter contextual influences on saliency
A. Torralba · 2005
Earlier work this paper cites.
Graph-based visual saliency
J. Harel, C. Koch, and P. Perona · 2006
Earlier work this paper cites.
Contextual guidance of eye movements and attention in real-world scenes: the role of global features in object search
A. Torralba, A. Oliva, M. S. Castelhano, and J. M. Henderson · 2006
Earlier work this paper cites.
Peripheral-foveal vision for real-time object recognition and tracking in video
S. Gould, J. Arfvidsson, A. Kaehler, B. Sapp, M. Messner, G. Bradski, P. Baumstarck, S. Chung, and A. Ng · 2007
Earlier work this paper cites.
Scene completion using millions of photographs
J. Hays and A. A. Efros · 2007
Earlier work this paper cites.
Near-optimal observation selection using submodular functions
A. Krause and C. Guestrin · 2007
Earlier work this paper cites.
Active policy learning for robot planning and exploration under uncertainty
R. Martinez-Cantin, N. de Freitas, A. Doucet, and J. A. Castellanos · 2007
Earlier work this paper cites.
Trajectory optimization using reinforcement learning for map exploration
T. Kollar and N. Roy · 2008
Earlier work this paper cites.
Development of three-dimensional object completion in infancy
K. C. Soska and S. P. Johnson · 2008
Earlier work this paper cites.
Frequency-tuned salient region detection
R. Achanta, S. Hemami, F. Estrada, and S. Susstrunk · 2009
Earlier work this paper cites.
Optimal scanning for faster object detection
N. Butko and J. Movellan · 2009
Earlier work this paper cites.
Make3d: Learning 3d scene structure from a single still image
A. Saxena, M. Sun, and A. Y. Ng · 2009
Earlier work this paper cites.
Actionable information in vision
S. Soatto · 2009
Cited alongside, same era.
Systems in development: motor skill acquisition facilitates three-dimensional object completion
K. C. Soska, K. E. Adolph, and S. P. Johnson · 2010
Cited alongside, same era.
Learning attentional policies for tracking and recognition in video with deep networks
L. Bazzani, H. Larochelle, V. Murino, J.-A. Ting, and N. d. Freitas · 2011
Cited alongside, same era.
Learning to detect a salient object
T. Liu, Z. Yuan, J. Sun, J. Wang, N. Zheng, X. Tang, and H.-Y. Shum · 2011
Cited alongside, same era.
Coverage problems in sensor networks: A survey
B. Wang · 2011
Cited alongside, same era.
Timely object recognition
S. Karayev, T. Baumgartner, M. Fritz, and T. Darrell · 2012
Cited alongside, same era.
Deep Q-learning for active recognition of GERMS
M. Malmir, K. Sikka, D. Forster, J. Movellan, and G. W. Cottrell · 2015
Later among the works it cites.
3d shapenets: A deep representation for volumetric shapes
Z. Wu, S. Song, A. Khosla, F. Yu, L. Zhang, X. Tang, and J. Xiao · 2015
Later among the works it cites.
Show, attend and tell: Neural image caption generation with visual attention
K. Xu, J. Ba, R. Kiros, K. Cho, A. Courville, R. Salakhutdinov, R. Zemel, and Y. Bengio · 2015
Later among the works it cites.
3D-R2N2: A unified approach for single and multi-view 3d object reconstrution
C. Choy, D. Xu, J. Gwak, K. Chen, and S. Savarese · 2016
Later among the works it cites.
Look-ahead before you leap: end-to-end active recognition by forecasting the effect of motion
D. Jayaraman and K. Grauman · 2016
Later among the works it cites.
Pairwise decomposition of image sequences for active multi-view recognition
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Imagenet classification with deep convolutional neural networks
A. Krizhevsky, I. Sutskever, and G. E. Hinton · 2012
Cited alongside, same era.
Saliency filters: Contrast based filtering for salient region detection
F. Perazzi, P. Krähenbühl, Y. Pritch, and A. Hornung · 2012
Cited alongside, same era.
Recognizing scene viewpoint using panoramic place representation
J. Xiao, K. A. Ehinger, A. Oliva, and A. Torralba · 2012
Cited alongside, same era.
50 years of object recognition: Directions forward
A. Andreopoulos and J. Tsotsos · 2013
Cited alongside, same era.
Perception-driven navigation: Active visual slam for robotic area coverage
A. Kim and R. M. Eustice · 2013
Cited alongside, same era.
Learning to predict gaze in egocentric video
Y. Li, A. Fathi, and J. M. Rehg · 2013
Cited alongside, same era.
E. Johns, S. Leutenegger, and A. Davison · 2016
Later among the works it cites.
Learning purposeful behaviour in the absence of rewards
M. Machado and M. Bowling · 2016
Later among the works it cites.
Reinforcement learning for visual object detection
S. Mathe, A. Pirinen, and C. Sminchisescu · 2016
Later among the works it cites.
Context encoders: Feature learning by inpainting
D. Pathak, P. Krahenbuhl, J. Donahue, T. Darrell, and A. A. Efros · 2016
Later among the works it cites.
Pano2vid: Automatic cinematography for watching 360 videos
Y.-C. Su, D. Jayaraman, and K. Grauman · 2016
Later among the works it cites.
Single image 3d interpreter network
J. Wu, T. Xue, J. Lim, Y. Tian, J. Tenenbaum, A. Torralba, and W. Freeman · 2016
Later among the works it cites.
Perspective transformer nets: Learning single-view 3d object reconstruction without 3d supervision
X. Yan, J. Yang, E. Yumer, Y. Guo, and H. Lee · 2016
Later among the works it cites.
End-to-end learning of action detection from frame glimpses in videos
S. Yeung, O. Russakovsky, G. Mori, and L. Fei-Fei · 2016
Later among the works it cites.
Split-brain autoencoders: Unsupervised learning by cross-channel prediction
R. Zhang, P. Isola, and A. A. Efros · 2016
Later among the works it cites.
View synthesis by appearance flow
T. Zhou, S. Tulsiani, W. Sun, J. Malik, and A. Efros · 2016
Later among the works it cites.
A dataset for developing and benchmarking active vision
P. Ammirato, P. Poirson, E. Park, J. Kosecka, and A. C. Berg · 2017
Closest in time.
Cognitive mapping and planning for visual navigation
S. Gupta, J. Davidson, S. Levine, R. Sukthankar, and J. Malik · 2017
Closest in time.
Hierarchical surface prediction for 3d object reconstruction
C. Häne, S. Tulsiani, and J. Malik · 2017
Closest in time.
Deep 360 pilot: Learning a deep agent for piloting through 360 sports videos
H. Hu, Y. Lin, M. Liu, H. Cheng, Y. Chang, and M. Sun · 2017
Closest in time.
Unsupervised learning through one-shot image-based shape reconstruction
D. Jayaraman, R. Gao, and K. Grauman · 2017
Closest in time.
Colorization as a proxy task for visual understanding
G. Larsson, M. Maire, and G. Shakhnarovich · 2017
Closest in time.
Learning to navigate in complex environments
P. Mirowski, R. Pascanu, F. Viola, H. Soyer, A. Ballard, A. Banino, M. Denil, R. Goroshin, L. Sifre, K. Kavukcuoglu, et al · 2017
Closest in time.
Curiosity-driven exploration by self-supervised prediction
D. Pathak, P. Agrawal, A. Efros, and T. Darrell · 2017
Closest in time.
Target-driven visual navigation in indoor scenes using deep reinforcement learning
Y. Zhu, R. Mottaghi, E. Kolve, J. J. Lim, A. Gupta, L. Fei-Fei, and A. Farhadi · 2017
Closest in time.