Fetching the paper…
Reading the bibliography…
In this work, we consider the problem of sequence-to-sequence alignment for signals containing outliers.
S. B. Needleman and C. D. Wunsch, “A general method applicable to the search for similarities in the amino acid sequence of two proteins,”
1970
Earlier work this paper cites.
H. Sakoe and S. Chiba, “Dynamic programming algorithm optimization for spoken processing recognition,”
1978
Earlier work this paper cites.
T. F. Smith and M. S. Waterman, “Identification of common molecular subsequences,”
1981
Earlier work this paper cites.
N. Nakatsu, Y. Kambayashi, and S. Yajima, “A longest common subsequence algorithm suitable for similar text strings,”
1982
Earlier work this paper cites.
Y. LeCun, L. Bottou, Y. Bengio, and P. Haffner, “Gradient-based learning applied to document recognition,”
1998
Earlier work this paper cites.
Berlin, Heidelberg: Springer-Verlag, 2007
M. Müller, · 2007
Earlier work this paper cites.
Y. Sakurai, C. Faloutsos, and M. Yamamuro, “Stream monitoring under the time warping distance,” in
2007
Earlier work this paper cites.
F. Zhou and F. Torre, “Canonical time warping for alignment of human behavior,”
2009
Earlier work this paper cites.
W. Zhang, M. Zhu, and K. G. Derpanis, “From actemes to action: A strongly-supervised representation for detailed action understanding,” in
2013
Earlier work this paper cites.
X. Wang and A. Gupta, “Unsupervised learning of visual representations using videos,” in
2015
Earlier work this paper cites.
B. G. Fabian Caba Heilbron, Victor Escorcia and J. C. Niebles, “ActivityNet: A large-scale video benchmark for human activity understanding,” in
2015
Earlier work this paper cites.
N. Srivastava, E. Mansimov, and R. Salakhudinov, “Unsupervised learning of video representations using LSTMs,” in
2015
Earlier work this paper cites.
G. Trigeorgis, M. A. Nicolaou, S. Zafeiriou, and B. W. Schuller, “Deep canonical time warping,” in
2016
Earlier work this paper cites.
I. Misra, C. L. Zitnick, and M. Hebert, “Shuffle and learn: Unsupervised learning using temporal order verification,” in
2016
Earlier work this paper cites.
J. J. Yu, A. W. Harley, and K. G. Derpanis, “Back to basics: Unsupervised learning of optical flow via brightness constancy and motion smoothness,” in
2016
Earlier work this paper cites.
D. Jayaraman and K. Grauman, “Slow and steand feature analysis: Higher order temporal coherence in video,” in
2016
Earlier work this paper cites.
C. Vondrick, H. Pirsiavash, and A. Torralba, “Anticipating visual representations from unlabeled video,” in
2016
Earlier work this paper cites.
M. Ma, H. Fan, and K. M. Kitani, “Going deeper into first-person activity recognition,” in
2016
Earlier work this paper cites.
D. Huang, L. Fei-Fei, and J. C. Niebles, “Connectionist temporal modeling for weakly supervised action labeling,” in
2016
Earlier work this paper cites.
M. Cuturi and M. Blondel, “Soft-DTW: A differentiable loss function for time-series,” in
2017
Cited alongside, same era.
B. Fernando, H. Bilen, E. Gavves, and S. Gould, “Self-supervised video representation learning with odd-one-out networks,” in
2017
Cited alongside, same era.
H. Lee, J. Huang, M. Singh, and M. Yang, “Unsupervised representation learning by sorting sequences,” in
2017
Cited alongside, same era.
L. Zhou, C. Xu, and J. J. Corso, “Towards automatic learning of procedures from web instructional videos,” in
2018
Cited alongside, same era.
Y. Tian, J. Shi, B. Li, Z. Duan, and C. Xu, “Audio-visual event localization in unconstrained videos,” in
2018
Cited alongside, same era.
D. Zhukov, J.-B. Alayrac, R. G. Cinbis, D. Fouhey, I. Laptev, and J. Sivic, “Cross-task weakly supervised learning from instructional videos,” in
2019
Later among the works it cites.
Y. Tang, D. Ding, Y. Rao, Y. Zheng, D. Zhang, L. Zhao, J. Lu, and J. Zhou, “COIN: A large-scale dataset for comprehensive instructional video analysis,” in
2019
Later among the works it cites.
D. Dwibedi, Y. Aytar, J. Tompson, P. Sermanet, and A. Zisserman, “Temporal cycle-consistency learning,” in
2019
Later among the works it cites.
X. Cai, T. Xu, J. Yi, J. Huang, and S. Rajasekaran, “DTWNet: A dynamic time warping network,” in
2019
Later among the works it cites.
D. Xu, J. Xiao, Z. Zhao, J. Shao, D. Xie, and Y. Zhuang, “Self-supervised spatiotemporal learning via video clip order prediction,” in
2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2018
Cited alongside, same era.
G. A. Sigurdsson, A. Gupta, C. Schmid, A. Farhadi, and K. Alahari, “Actor and observer: Joint modeling of first and third-person videos,” in
2018
Cited alongside, same era.
U. Büchler, B. Brattoli, and B. Ommer, “Improving spatiotemporal self-supervision by deep reinforcement learning,” in
2018
Cited alongside, same era.
D. Wei, J. J. Lim, A. Zisserman, and W. T. Freeman, “Learning and using the arrow of time,” in
2018
Cited alongside, same era.
S. Meister, J. Hur, and S. Roth, “UnFlow: Unsupervised learning of optical flow with a bidirectional census loss,” in
2018
Cited alongside, same era.
C. Vondrick, A. Shrivastava, A. Fathi, S. Guadarrama, and K. Murphy, “Tracking emerges by colorizing videos,” in
2018
Cited alongside, same era.
J. Janai, F. Güney, A. Ranjan, M. J. Black, and A. Geiger, “Unsupervised learning of multi-frame optical flow with occlusions,” in
2018
Cited alongside, same era.
X. Wang, A. Jabri, and A. A. Efros, “Learning correspondence from the cycle-consistency of time,” in
2019
Later among the works it cites.
J. Wang, J. Jiao, L. Bao, S. He, Y. Liu, and W. Liu, “Self-supervised spatio-temporal representation learning for videos by predicting motion and appearance statistics,” in
2019
Later among the works it cites.
T. Han, W. Xie, and A. Zisserman, “Video representation learning by dense predictive coding,” in
2019
Later among the works it cites.
A. Miech, D. Zhukov, J.-B. Alayrac, M. Tapaswi, I. Laptev, and J. Sivic, “HowTo100M: Learning a Text-Video Embedding by Watching Hundred Million Narrated Video Clips,” in
2019
Later among the works it cites.
C. Sun, A. Myers, C. Vondrick, K. Murphy, and C. Schmid, “VideoBERT: A joint model for video and language representation learning,” in
2019
Later among the works it cites.
P. X. Nguyen, D. Ramanan, and C. C. Fowlkes, “Weakly-supervised action localization with background modeling,” in
2019
Later among the works it cites.
K. Cao, J. Ji, Z. Cao, C. Chang, and J. C. Niebles, “Few-shot video classification via temporal alignment,” in
2020
Later among the works it cites.
A. Jabri, A. Owens, and A. A. Efros, “Space-time correspondence as a contrastive random walk,” in
2020
Later among the works it cites.
S. Benaim, A. Ephrat, O. Lang, I. Mosseri, W. T. Freeman, M. Rubinstein, M. Irani, and T. Dekel, “SpeedNet: Learning the speediness in videos,” in
2020
Later among the works it cites.
A. Miech, J.-B. Alayrac, L. Smaira, I. Laptev, J. Sivic, and A. Zisserman, “End-to-End Learning of Visual Representations from Uncurated Instructional Videos,” in
2020
Later among the works it cites.
2020
Later among the works it cites.
I. Hadji, K. G. Derpanis, and A. D. Jepson, “Representation learning via global temporal alignment and cycle-consistency,” in
2021
Closest in time.
X. Chang, F. Tung, and G. Mori, “Learning discriminative prototypes with dynamic time warping,” in
2021
Closest in time.