Ambient sound provides supervision for visual learning
A. Owens, J. Wu, J. H. McDermott, W. T. Freeman, and A. Torralba · 2016
Later among the works it cites.
Context encoders: Feature learning by inpainting
D. Pathak, P. Krahenbuhl, J. Donahue, T. Darrell, and A. A. Efros · 2016
Later among the works it cites.
The curious robot: Learning visual representations via physical interactions
L. Pinto, D. Gandhi, Y. Han, Y.-L. Park, and A. Gupta · 2016
Later among the works it cites.
Generating videos with scene dynamics
C. Vondrick, H. Pirsiavash, and A. Torralba · 2016
Later among the works it cites.
Colorful image colorization
R. Zhang, P. Isola, and A. A. Efros · 2016
Later among the works it cites.
Deep unsupervised similarity learning using partially ordered sets
M. A. Bautista, A. Sanakoyeu, and B. Ommer · 2017
Later among the works it cites.
Lstm self-supervision for detailed behavior analysis
B. Brattoli, U. Buchler, A.-S. Wahl, M. E. Schwab, and B. Ommer · 2017
Later among the works it cites.
Multi-task self-supervised visual learning
C. Doersch and A. Zisserman · 2017
Later among the works it cites.
Self-supervised video representation learning with odd-one-out networks
B. Fernando, H. Bilen, E. Gavves, and S. Gould · 2017
Later among the works it cites.
The kinetics human action video dataset
Original
W. Kay, J. Carreira, K. Simonyan, B. Zhang, C. Hillier, S. Vijayanarasimhan, F. Viola, T. Green, T. Back, P. Natsev, et al · 2017
Later among the works it cites.
Unsupervised representation learning by sorting sequences
H.-Y. Lee, J.-B. Huang, M. Singh, and M.-H. Yang · 2017
Later among the works it cites.
Unsupervised video understanding by reconciliation of posture similarities
T. Milbich, M. Bautista, E. Sutter, and B. Ommer · 2017
Later among the works it cites.
Representation learning by learning to count
M. Noroozi, H. Pirsiavash, and P. Favaro · 2017
Later among the works it cites.
Learning features by watching objects move
D. Pathak, R. Girshick, P. Dollar, T. Darrell, and B. Hariharan · 2017
Later among the works it cites.
Supervision via competition: Robot adversaries for learning tasks
L. Pinto, J. Davidson, and A. Gupta · 2017
Later among the works it cites.
Learning to push by grasping: Using multiple tasks for effective learning
L. Pinto and A. Gupta · 2017
Later among the works it cites.
Cross-domain self-supervised multi-task feature learning using synthetic imagery
Original
Z. Ren and Y. J. Lee · 2017
Later among the works it cites.
What actions are needed for understanding human actions in videos?
Original
G. A. Sigurdsson, O. Russakovsky, and A. Gupta · 2017
Later among the works it cites.
Self-supervised learning of pose embeddings from spatiotemporal relations in videos
O. Sumer, T. Dencker, and B. Ommer · 2017
Later among the works it cites.
Transitive invariance for self-supervised visual representation learning
Original
X. Wang, K. He, and A. Gupta · 2017
Later among the works it cites.
Split-brain autoencoders: Unsupervised learning by cross-channel prediction
R. Zhang, P. Isola, and A. A. Efros · 2017
Later among the works it cites.