Fetching the paper…
Reading the bibliography…
This paper proposes a simple self-supervised approach for learning a representation for visual correspondence from raw video.
Laws of organization in perceptual forms
Max Wertheimer · 1938
Earlier work this paper cites.
Concerning nonnegative matrices and doubly stochastic matrices
Richard Sinkhorn and Paul Knopp · 1967
Earlier work this paper cites.
An iterative image registration technique with an application to stereo vision
Bruce D Lucas and Takeo Kanade · 1981
Earlier work this paper cites.
Self-organizing neural network that discovers surfaces in random-dot stereograms
Suzanna Becker and Geoffrey E Hinton · 1992
Earlier work this paper cites.
Learning classification with unlabeled data
Virginia R de Sa · 1994
Earlier work this paper cites.
Analyzing gait with spatiotemporal surfaces
Sourabh A Niyogi and Edward H Adelson · 1994
Earlier work this paper cites.
Motion segmentation and tracking using normalized cuts
Jianbo Shi and Jitendra Malik · 1998
Earlier work this paper cites.
Normalized cuts and image segmentation
Jianbo Shi and Jitendra Malik · 2000
Earlier work this paper cites.
Learning segmentation by random walks
Meila, Marina and Shi, Jianbo · 2001
Earlier work this paper cites.
Event-based analysis of video
Lihi Zelnik-Manor and Michal Irani · 2001
Earlier work this paper cites.
What went where
Josh Wills, Sameer Agarwal, and Serge Belongie · 2003
Earlier work this paper cites.
Spectral grouping using the nystrom method
Charless Fowlkes, Serge Belongie, Fan Chung, and Jitendra Malik · 2004
Earlier work this paper cites.
Learning a similarity metric discriminatively, with application to face verification
Sumit Chopra, Raia Hadsell, and Yann LeCun · 2005
Earlier work this paper cites.
Strike a pose: Tracking people by finding stylized poses
Deva Ramanan, David A Forsyth, and Andrew Zisserman · 2005
Earlier work this paper cites.
Graph cuts and efficient nd image segmentation
Yuri Boykov and Gareth Funka-Lea · 2006
Earlier work this paper cites.
Global data association for multi-object tracking using network flows
Li Zhang, Yuan Li, and Ramakant Nevatia · 2008
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database
Jia Deng, Wei Dong, Richard Socher, Li-Jia Li, Kai Li, and Li Fei-Fei · 2009
Earlier work this paper cites.
Filter flow
Steven M Seitz and Simon Baker · 2009
Earlier work this paper cites.
Noise-contrastive estimation: A new estimation principle for unnormalized statistical models
Michael Gutmann and Aapo Hyvärinen · 2010
Earlier work this paper cites.
Supervised random walks: predicting and recommending links in social networks
Lars Backstrom and Jure Leskovec · 2011
Earlier work this paper cites.
Multiple object tracking using k-shortest paths optimization
Jerome Berclaz, Francois Fleuret, Engin Turetken, and Pascal Fua · 2011
Earlier work this paper cites.
Temporally consistent multi-class video-object segmentation with the video graph-shifts algorithm
Albert YC Chen and Jason J Corso · 2011
Earlier work this paper cites.
Globally-optimal greedy algorithms for tracking a variable number of objects
Hamed Pirsiavash, Deva Ramanan, and Charless C. Fowlkes · 2011
Earlier work this paper cites.
Computer Vision - A Modern Approach, Second Edition
David A. Forsyth and Jean Ponce · 2012
Earlier work this paper cites.
Gmcp-tracker: Global multi-object tracking using generalized minimum clique graphs
Amir Roshan Zamir, Afshin Dehghan, and Mubarak Shah · 2012
Earlier work this paper cites.
Sinkhorn distances: Lightspeed computation of optimal transport
Marco Cuturi · 2013
Earlier work this paper cites.
Efficient estimation of word representations in vector space
Tomas Mikolov, Kai Chen, Greg Corrado, and Jeffrey Dean · 2013
Earlier work this paper cites.
Efficient image and video co-localization with frank-wolfe algorithm
Armand Joulin, Kevin Tang, and Li Fei-Fei · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba · 2014
Earlier work this paper cites.
Deepwalk
Bryan Perozzi, Rami Al-Rfou, and Steven Skiena · 2014
Earlier work this paper cites.
Dropout: A simple way to prevent neural networks from overfitting
Nitish Srivastava, Geoffrey Hinton, Alex Krizhevsky, Ilya Sutskever, and Ruslan Salakhutdinov · 2014
Earlier work this paper cites.
Learning to see by moving
Pulkit Agrawal, Joao Carreira, and Jitendra Malik · 2015
Earlier work this paper cites.
Unsupervised visual representation learning by context prediction
Carl Doersch, Abhinav Gupta, and Alexei A. Efros · 2015
Earlier work this paper cites.
Discriminative unsupervised feature learning with exemplar convolutional neural networks
Alexey Dosovitskiy, Philipp Fischer, Jost Tobias Springenberg, Martin Riedmiller, and Thomas Brox · 2015
Earlier work this paper cites.
A neural algorithm of artistic style
Leon A Gatys, Alexander S Ecker, and Matthias Bethge · 2015
Earlier work this paper cites.
Unsupervised learning of spatiotemporally coherent metrics
Ross Goroshin, Joan Bruna, Jonathan Tompson, David Eigen, and Yann LeCun · 2015
Earlier work this paper cites.
Learning visual groups from co-occurrences in space and time
Phillip Isola, Daniel Zoran, Dilip Krishnan, and Edward H Adelson · 2015
Earlier work this paper cites.
Learning image representations tied to egomotion
Dinesh Jayaraman and Kristen Grauman · 2015
Earlier work this paper cites.
Real-time hyperlapse creation via optimal frame selection
Neel Joshi, Wolf Kienzle, Mike Toelle, Matt Uyttendaele, and Michael F Cohen · 2015
Earlier work this paper cites.
Fully convolutional networks for semantic segmentation
Jonathan Long, Evan Shelhamer, and Trevor Darrell · 2015
Earlier work this paper cites.
Deep multi-scale video prediction beyond mean square error
Michaël Mathieu, Camille Couprie, and Yann LeCun · 2015
Earlier work this paper cites.
Unsupervised learning of video representations using LSTMs
Nitish Srivastava, Elman Mansimov, and Ruslan Salakhutdinov · 2015
Cited alongside, same era.
Line: Large-scale information network embedding
Jian Tang, Meng Qu, Mingzhe Wang, Ming Zhang, Jun Yan, and Qiaozhu Mei · 2015
Cited alongside, same era.
Unsupervised learning of visual representations using videos
Xiaolong Wang and Abhinav Gupta · 2015
Cited alongside, same era.
Object-centric representation learning from unlabeled videos
Ruohan Gao, Dinesh Jayaraman, and Kristen Grauman · 2016
Cited alongside, same era.
node2vec: Scalable feature learning for networks
Aditya Grover and Jure Leskovec · 2016
Cited alongside, same era.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Cited alongside, same era.
Long-term tracking in the wild: A benchmark
Jack Valmadre, Luca Bertinetto, Joao F Henriques, Ran Tao, Andrea Vedaldi, Arnold WM Smeulders, Philip HS Torr, and Efstratios Gavves · 2018
Later among the works it cites.
Representation learning with contrastive predictive coding
Aäron van den Oord, Yazhe Li, and Oriol Vinyals · 2018
Later among the works it cites.
Graph attention networks
Petar Velickovic, Guillem Cucurull, Arantxa Casanova, Adriana Romero, Pietro Lio, and Yoshua Bengio · 2018
Later among the works it cites.
Non-local neural networks
Xiaolong Wang, Ross Girshick, Abhinav Gupta, and Kaiming He · 2018
Later among the works it cites.
Learning and using the arrow of time
Donglai Wei, Joseph Lim, Andrew Zisserman, and William T. Freeman · 2018
Later among the works it cites.
Unsupervised feature learning via non-parametric instance discrimination
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Unsupervised learning of edges
Yin Li, Manohar Paluri, James M. Rehg, and Piotr Dollár · 2016
Cited alongside, same era.
Deep predictive coding networks for video prediction and unsupervised learning
William Lotter, Gabriel Kreiman, and David Cox · 2016
Cited alongside, same era.
Shuffle and Learn: Unsupervised Learning using Temporal Order Verification
Ishan Misra, C. Lawrence Zitnick, and Martial Hebert · 2016
Cited alongside, same era.
Unsupervised learning of visual representations by solving jigsaw puzzles
Mehdi Noroozi and Paolo Favaro · 2016
Cited alongside, same era.
Visually indicated sounds
Andrew Owens, Phillip Isola, Josh McDermott, Antonio Torralba, Edward H Adelson, and William T Freeman · 2016
Cited alongside, same era.
A benchmark dataset and evaluation methodology for video object segmentation
F. Perazzi, J. Pont-Tuset, B. McWilliams, L. Van Gool, M. Gross, and A. Sorkine-Hornung · 2016
Cited alongside, same era.
Zhirong Wu, Yuanjun Xiong, Stella X Yu, and Dahua Lin · 2018
Later among the works it cites.
Efficient video object segmentation via network modulation
Linjie Yang, Yanran Wang, Xuehan Xiong, Jianchao Yang, and Aggelos K Katsaggelos · 2018
Later among the works it cites.
Adaptive temporal encoding network for video instance-level human parsing
Qixian Zhou, Xiaodan Liang, Ke Gong, and Liang Lin · 2018
Later among the works it cites.
Learning representations by maximizing mutual information across views
Philip Bachman, R Devon Hjelm, and William Buchwalter · 2019
Later among the works it cites.
Temporal cycle-consistency learning
Debidatta Dwibedi, Yusuf Aytar, Jonathan Tompson, Pierre Sermanet, and Andrew Zisserman · 2019
Later among the works it cites.
Democracy index 2019 a year of democratic setbacks and popular protest
EIU.com · 2019
Later among the works it cites.
Genesis: Generative scene inference and sampling with object-centric latent representations
Martin Engelcke, Adam R Kosiorek, Oiwi Parker Jones, and Ingmar Posner · 2019
Later among the works it cites.
Slowfast networks for video recognition
Christoph Feichtenhofer, Haoqi Fan, Jitendra Malik, and Kaiming He · 2019
Later among the works it cites.
Multi-object representation learning with iterative variational inference
Klaus Greff, Raphaël Lopez Kaufman, Rishabh Kabra, Nick Watters, Chris Burgess, Daniel Zoran, Loic Matthey, Matthew Botvinick, and Alexander Lerchner · 2019
Later among the works it cites.
Video representation learning by dense predictive coding
Tengda Han, Weidi Xie, and Andrew Zisserman · 2019
Later among the works it cites.
Data-efficient image recognition with contrastive predictive coding
Olivier J Hénaff, Aravind Srinivas, Jeffrey De Fauw, Ali Razavi, Carl Doersch, SM Eslami, and Aaron van den Oord · 2019
Later among the works it cites.
Multigrid predictive filter flow for unsupervised learning on videos
Shu Kong and Charless Fowlkes · 2019
Later among the works it cites.
Self-supervised learning for video correspondence flow
Zihang Lai and Weidi Xie · 2019
Later among the works it cites.
Joint-task self-supervised learning for temporal correspondence
Xueting Li, Sifei Liu, Shalini De Mello, Xiaolong Wang, Jan Kautz, and Ming-Hsuan Yang · 2019
Later among the works it cites.
Video object segmentation without temporal information
K. K. Maninis, S. Caelles, Y. Chen, J. Pont-Tuset, L. Leal-Taixé, D. Cremers, and L. Van Gool · 2019
Later among the works it cites.
Online model distillation for efficient video inference
Ravi Teja Mullapudi, Steven Chen, Keyi Zhang, Deva Ramanan, and Kayvon Fatahalian · 2019
Later among the works it cites.
Video object segmentation using space-time memory networks
Seoung Wug Oh, Joon-Young Lee, Ning Xu, and Seon Joo Kim · 2019
Later among the works it cites.
Online object representations with contrastive learning
Soren Pirk, Mohi Khansari, Yunfei Bai, Corey Lynch, and Pierre Sermanet · 2019
Later among the works it cites.
Yonglong Tian, Dilip Krishnan, and Phillip Isola · 2019
Later among the works it cites.
MOTS: multi-object tracking and segmentation
Paul Voigtlaender, Michael Krause, Aljosa Osep, Jonathon Luiten, Berin Balachandar Gnana Sekar, Andreas Geiger, and Bastian Leibe · 2019
Later among the works it cites.
Unsupervised deep tracking
Ning Wang, Yibing Song, Chao Ma, Wengang Zhou, Wei Liu, and Houqiang Li · 2019
Later among the works it cites.
Fast online object tracking and segmentation: A unifying approach
Qiang Wang, Li Zhang, Luca Bertinetto, Weiming Hu, and Philip HS Torr · 2019
Later among the works it cites.
Learning correspondence from the cycle-consistency of time
Xiaolong Wang, Allan Jabri, and Alexei A Efros · 2019
Later among the works it cites.
Gromov-wasserstein learning for graph matching and node embedding
Hongteng Xu, Dixin Luo, Hongyuan Zha, and Lawrence Carin · 2019
Later among the works it cites.
Learning a neural solver for multiple object tracking
Guillem Brasó and Laura Leal-Taixé · 2020
Closest in time.
End-to-end object detection with transformers
Nicolas Carion, Francisco Massa, Gabriel Synnaeve, Nicolas Usunier, Alexander Kirillov, and Sergey Zagoruyko · 2020
Closest in time.
A simple framework for contrastive learning of visual representations
Ting Chen, Simon Kornblith, Mohammad Norouzi, and Geoffrey Hinton · 2020
Closest in time.
Watching the world go by: Representation learning from unlabeled videos, 2020
Daniel Gordon, Kiana Ehsani, Dieter Fox, and Ali Farhadi · 2020
Closest in time.
Momentum contrast for unsupervised visual representation learning
Kaiming He, Haoqi Fan, Yuxin Wu, Saining Xie, and Ross Girshick · 2020
Closest in time.
Contrastive learning of structured world models
Thomas Kipf, Elise van der Pol, and Max Welling · 2020
Closest in time.
Mast: A memory-augmented self-supervised tracker
Zihang Lai, Erika Lu, and Weidi Xie · 2020
Closest in time.
Large-scale optimal transport via adversarial training with cycle-consistency
Guansong Lu, Zhiming Zhou, Jian Shen, Cheng Chen, Weinan Zhang, and Yong Yu · 2020
Closest in time.
Superglue: Learning feature matching with graph neural networks
Paul-Edouard Sarlin, Daniel DeTone, Tomasz Malisiewicz, and Andrew Rabinovich · 2020
Closest in time.
Test-time training with self-supervision for generalization under distribution shifts
Yu Sun, Xiaolong Wang, Zhuang Liu, John Miller, Alexei A. Efros, and Moritz Hardt · 2020
Closest in time.
Self-supervised learning of video-induced visual invariances
Michael Tschannen, Josip Djolonga, Marvin Ritter, Aravindh Mahendran, Neil Houlsby, Sylvain Gelly, and Mario Lucic · 2020
Closest in time.