Fetching the paper…
Reading the bibliography…
A steady momentum of innovations and breakthroughs has convincingly pushed the limits of unsupervised image representation learning.
Unsupervised Learning from Video with Deep Neural Embeddings
Zhuang, C.; She, T.; Andonian, A.; and Yamins, D. 2019 · 1905
Earlier work this paper cites.
Learning Video Representations using Contrastive Bidirectional Transformer
Sun, C.; Baradel, F.; Murphy, K.; and Schmid, C. 2019 · 1906
Earlier work this paper cites.
Tian, Y.; Krishnan, D.; and Isola, P. 2019 · 1906
Earlier work this paper cites.
Momentum Contrast for Unsupervised Visual Representation Learning
He, K.; Fan, H.; Wu, Y.; Xie, S.; and Girshick, R. B. 2019 · 1911
Earlier work this paper cites.
Watching the World Go By: Representation Learning from Unlabeled Videos
Gordon, D.; Ehsani, K.; Fox, D.; and Farhadi, A. 2020 · 2003
Earlier work this paper cites.
Deep learning from temporal coherence in video
Mobahi, H.; Collobert, R.; and Weston, J. 2009 · 2009
Earlier work this paper cites.
HMDB: A large video database for human motion recognition
Kuehne, H.; Jhuang, H.; Garrote, E.; Poggio, T.; and Serre, T. 2011 · 2011
Earlier work this paper cites.
UCF101: A Dataset of 101 Human Actions Classes From Videos in The Wild
Soomro, K.; Zamir, A. R.; and Shah, M. 2012 · 2012
Earlier work this paper cites.
Deep Learning of Invariant Features via Simulated Fixations in Video
Zou, W.; Zhu, S.; Yu, K.; and Ng, A. Y. 2012 · 2012
Earlier work this paper cites.
Learning to See by Moving
Agrawal, P.; Carreira, J.; and Malik, J. 2015 · 2015
Earlier work this paper cites.
Learning Large-Scale Automatic Image Colorization
Deshpande, A.; Rock, J.; and Forsyth, D. 2015 · 2015
Earlier work this paper cites.
Unsupervised Visual Representation Learning by Context Prediction
Doersch, C.; Gupta, A.; and Efros, A. A. 2015 · 2015
Earlier work this paper cites.
ActivityNet: A large-scale video benchmark for human activity understanding
Heilbron, F. C.; Escorcia, V.; Ghanem, B.; and Niebles, J. C. 2015 · 2015
Earlier work this paper cites.
Learning image representations tied to ego-motion
Jayaraman, D.; and Grauman, K. 2015 · 2015
Earlier work this paper cites.
Unsupervised Learning of Video Representations using LSTMs
Srivastava, N.; Mansimov, E.; and Salakhutdinov, R. 2015 · 2015
Earlier work this paper cites.
Unsupervised Learning of Visual Representations Using Videos
Wang, X.; and Gupta, A. 2015 · 2015
Cited alongside, same era.
Object Tracking Benchmark
Wu, Y.; Lim, J.; and Yang, M.-H. 2015 · 2015
Cited alongside, same era.
Unsupervised Learning for Physical Interaction through Video Prediction
Finn, C.; Goodfellow, I. J.; and Levine, S. 2016 · 2016
Cited alongside, same era.
Shuffle and Learn: Unsupervised Learning Using Temporal Order Verification
Misra, I.; et al. 2016 · 2016
Cited alongside, same era.
Unsupervised Learning of Visual Representations by Solving Jigsaw Puzzles
Noroozi, M.; and Favaro, P. 2016 · 2016
Cited alongside, same era.
Learning deep intrinsic video representation by exploring temporal coherence and graph structure
Pan, Y.; Li, Y.; Yao, T.; Mei, T.; Li, H.; and Rui, Y. 2016 · 2016
Representation learning with contrastive predictive coding
Oord, A. v. d.; Li, Y.; and Vinyals, O. 2018 · 2018
Later among the works it cites.
Temporal segment networks for action recognition in videos
Wang, L.; Xiong, Y.; Wang, Z.; Qiao, Y.; Lin, D.; Tang, X.; and Van Gool, L. 2018 · 2018
Later among the works it cites.
Unsupervised Feature Learning via Non-parametric Instance Discrimination
Wu, Z.; Xiong, Y.; Yu, S. X.; and Lin, D. 2018 · 2018
Later among the works it cites.
Video Jigsaw: Unsupervised Learning of Spatiotemporal Context for Video Action Recognition
Ahsan, U.; Madhok, R.; and Essa, I. 2019 · 2019
Later among the works it cites.
Learning Representations by Maximizing Mutual Information Across Views
Bachman, P.; Hjelm, R. D.; and Buchwalter, W. 2019 · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Generating Videos with Scene Dynamics
Vondrick, C.; Pirsiavash, H.; and Torralba, A. 2016 · 2016
Cited alongside, same era.
Colorful Image Colorization
Zhang, R.; Isola, P.; and Efros, A. A. 2016 · 2016
Cited alongside, same era.
The Kinetics Human Action Video Dataset
Kay, W.; Carreira, J.; Simonyan, K.; Zhang, B.; Hillier, C.; Vijayanarasimhan, S.; Viola, F.; Green, T.; Back, T.; Natsev, P.; Suleyman, M.; and Zisserman, A. 2017 · 2017
Cited alongside, same era.
Unsupervised Representation Learning by Sorting Sequences
Lee, H.-Y.; Huang, J.-B.; Singh, M.; and Yang, M.-H. 2017 · 2017
Cited alongside, same era.
Video Frame Synthesis Using Deep Voxel Flow
Liu, Z.; Yeh, R. A.; Tang, X.; Liu, Y.; and Agarwala, A. 2017 · 2017
Cited alongside, same era.
Unsupervised Learning of Long-Term Motion Dynamics for Videos
Luo, Z.; Peng, B.; Huang, D.-A.; Alahi, A.; and Fei-Fei, L. 2017 · 2017
Cited alongside, same era.
Scaling and Benchmarking Self-Supervised Visual Representation Learning
Goyal, P.; Mahajan, D.; Gupta, A.; and Misra, I. 2019 · 2019
Later among the works it cites.
Video Representation Learning by Dense Predictive Coding
Han, T.; Xie, W.; et al. 2019 · 2019
Later among the works it cites.
Learning deep representations by mutual information estimation and maximization
Hjelm, R. D.; Fedorov, A.; Lavoie-Marchildon, S.; Grewal, K.; Bachman, P.; Trischler, A.; and Bengio, Y. 2019 · 2019
Later among the works it cites.
GOT-10k: A Large High-Diversity Benchmark for Generic Object Tracking in the Wild
Huang, L.; Zhao, X.; and Huang, K. 2019 · 2019
Later among the works it cites.
Learning spatio-temporal representation with local and global diffusion
Qiu, Z.; Yao, T.; Ngo, C.-W.; Tian, X.; and Mei, T. 2019 · 2019
Later among the works it cites.
Self-Supervised Spatio-Temporal Representation Learning for Videos by Predicting Motion and Appearance Statistics
Wang, J.; Jiao, J.; Bao, L.; He, S.; Liu, Y.; and Liu, W. 2019 · 2019
Later among the works it cites.
Self-Supervised Spatiotemporal Learning via Video Clip Order Prediction
Xu, D.; Xiao, J.; Zhao, Z.; Shao, J.; Xie, D.; and Zhuang, Y. 2019 · 2019
Later among the works it cites.
Joint Contrastive Learning with Infinite Possibilities
Cai, Q.; Wang, Y.; Pan, Y.; Yao, T.; and Mei, T. 2020 · 2020
Closest in time.
Temporal Contrastive Pretraining for Video Action Recognition
Lorre, G.; Rabarisoa, J.; Orcesi, A.; Ainouz, S.; and Canu, S. 2020 · 2020
Closest in time.