Fetching the paper…
Reading the bibliography…
Recently, Self-Supervised Representation Learning (SSRL) has attracted much attention in the field of computer vision, speech, natural language processing (NLP), and recently, with other types of modalities, including time series from sensors.
Learning classification with unlabeled data
Virginia R de Sa. 1994 · 1994
Earlier work this paper cites.
Measures of degeneracy and redundancy in biological networks
Giulio Tononi, Olaf Sporns, and Gerald M Edelman. 1999 · 1999
Earlier work this paper cites.
Learning a similarity metric discriminatively, with application to face verification. In 2005 IEEE Computer Society Conference on Computer Vision and Pattern Recognition (CVPR’05) , Vol. 1. IEEE, 539–546
Sumit Chopra, Raia Hadsell, and Yann LeCun. 2005a · 2005
Earlier work this paper cites.
Learning a similarity metric discriminatively, with application to face verification. In 2005 IEEE Computer Society Conference on Computer Vision and Pattern Recognition (CVPR’05) , Vol. 1. 539–546 vol. 1
S. Chopra, R. Hadsell, and Y. LeCun. 2005b · 2005
Earlier work this paper cites.
Multisensory contributions to low-level,‘unisensory’processing
Charles E Schroeder and John Foxe. 2005 · 2005
Earlier work this paper cites.
The development of embodied cognition: Six lessons from babies
Linda Smith and Michael Gasser. 2005 · 2005
Earlier work this paper cites.
Dimensionality reduction by learning an invariant mapping. In 2006 IEEE Computer Society Conference on Computer Vision and Pattern Recognition (CVPR’06) , Vol. 2. IEEE, 1735–1742
Raia Hadsell, Sumit Chopra, and Yann LeCun. 2006 · 2006
Earlier work this paper cites.
A fast learning algorithm for deep belief nets
Geoffrey E Hinton, Simon Osindero, and Yee-Whye Teh. 2006 · 2006
Earlier work this paper cites.
Greedy layer-wise training of deep networks. In Advances in neural information processing systems . 153–160
Yoshua Bengio, Pascal Lamblin, Dan Popovici, and Hugo Larochelle. 2007 · 2007
Earlier work this paper cites.
Large Scale Online Learning of Image Similarity Through Ranking
Gal Chechik, Varun Sharma, Uri Shalit, and Samy Bengio. 2010 · 2010
Earlier work this paper cites.
Noise-contrastive estimation: A new estimation principle for unnormalized statistical models. In Proceedings of 13th International Conference on Artificial Intelligence and Statistics . 297–304
Michael Gutman and Aapo Hyvarinen. 2010 · 2010
Earlier work this paper cites.
Representation learning: A review and new perspectives
Yoshua Bengio, Aaron Courville, and Pascal Vincent. 2013 · 2013
Earlier work this paper cites.
Sinkhorn distances: Lightspeed computation of optimal transport
Marco Cuturi. 2013 · 2013
Earlier work this paper cites.
Discriminative unsupervised feature learning with convolutional neural networks
Alexey Dosovitskiy, Jost Tobias Springenberg, Martin Riedmiller, and Thomas Brox. 2014 · 2014
Earlier work this paper cites.
Siamese neural networks for one-shot image recognition. In ICML deep learning workshop , Vol. 2. Lille, 0
Gregory Koch, Richard Zemel, Ruslan Salakhutdinov, et al · 2015
Earlier work this paper cites.
Facenet: A unified embedding for face recognition and clustering. In Proceedings of the IEEE conference on computer vision and pattern recognition . 815–823
Florian Schroff, Dmitry Kalenichenko, and James Philbin. 2015 · 2015
Earlier work this paper cites.
Discriminative learning of deep convolutional feature point descriptors. In Proceedings of the IEEE International Conference on Computer Vision . 118–126
Edgar Simo-Serra, Eduard Trulls, Luis Ferraz, Iasonas Kokkinos, Pascal Fua, and Francesc Moreno-Noguer. 2015 · 2015
Earlier work this paper cites.
Soundnet: Learning sound representations from unlabeled video
Yusuf Aytar, Carl Vondrick, and Antonio Torralba. 2016 · 2016
Earlier work this paper cites.
Jeff Donahue, Philipp Krähenbühl, and Trevor Darrell. 2016 · 2016
Earlier work this paper cites.
Shuffle and learn: unsupervised learning using temporal order verification. In European Conference on Computer Vision . Springer, 527–544
Ishan Misra, C Lawrence Zitnick, and Martial Hebert. 2016 · 2016
Earlier work this paper cites.
Deep metric learning via lifted structured feature embedding. In Proceedings of the IEEE conference on computer vision and pattern recognition . 4004–4012
Hyun Oh Song, Yu Xiang, Stefanie Jegelka, and Silvio Savarese. 2016 · 2016
Earlier work this paper cites.
Ambient sound provides supervision for visual learning. In European conference on computer vision . Springer, 801–816
Andrew Owens, Jiajun Wu, Josh H McDermott, William T Freeman, and Antonio Torralba. 2016 · 2016
Earlier work this paper cites.
Improved deep metric learning with multi-class n-pair loss objective. In Advances in neural information processing systems
Kihyuk Sohn. 2016 · 2016
Earlier work this paper cites.
Look, listen and learn. In Proceedings of the IEEE International Conference on Computer Vision
Relja Arandjelovic and Andrew Zisserman. 2017 · 2017
Earlier work this paper cites.
Self-supervised video representation learning with odd-one-out networks. In Proceedings of the IEEE conference on computer vision and pattern recognition . 3636–3645
Basura Fernando, Hakan Bilen, Efstratios Gavves, and Stephen Gould. 2017 · 2017
Earlier work this paper cites.
Multiresolution wavelet transform based feature extraction and ECG classification to detect cardiac abnormalities
Santanu Sahoo, Bhupen Kanungo, Suresh Behera, and Sukanta Sabut. 2017 · 2017
Earlier work this paper cites.
Time-Contrastive Networks: Self-Supervised Learning from Multi-view Observation. In 2017 IEEE Conference on Computer Vision and Pattern Recognition Workshops (CVPRW)
Pierre Sermanet, Corey Lynch, Jasmine Hsu, and Sergey Levine. 2017 · 2017
Earlier work this paper cites.
Sampling matters in deep embedding learning. In Proceedings of the IEEE International Conference on Computer Vision
Chao-Yuan Wu, R Manmatha, Alexander J Smola, and Philipp Krahenbuhl. 2017 · 2017
Earlier work this paper cites.
Rethinking spatiotemporal feature learning for video understanding
Saining Xie, Chen Sun, Jonathan Huang, Zhuowen Tu, and Kevin Murphy. 2017 · 2017
Earlier work this paper cites.
Deep multimodal representation learning from temporal data. In Proceedings of the IEEE conference on computer vision and pattern recognition . 5447–5455
Xitong Yang, Palghat Ramesh, Radha Chitta, Sriganesh Madhvanath, Edgar A Bernal, and Jiebo Luo. 2017 · 2017
Earlier work this paper cites.
Objects that sound. In Proceedings of the European conference on computer vision (ECCV) . 435–451
Relja Arandjelovic and Andrew Zisserman. 2018 · 2018
Earlier work this paper cites.
Deep clustering for unsupervised learning of visual features. In Proceedings of the European Conference on Computer Vision (ECCV) . 132–149
Mathilde Caron, Piotr Bojanowski, Armand Joulin, and Matthijs Douze. 2018 · 2018
Earlier work this paper cites.
Learning to separate object sounds by watching unlabeled video. In Proceedings of the European Conference on Computer Vision (ECCV) . 35–53
Ruohan Gao, Rogerio Feris, and Kristen Grauman. 2018 · 2018
Earlier work this paper cites.
Cooperative learning of audio and video models from self-supervised synchronization
Bruno Korbar, Du Tran, and Lorenzo Torresani. 2018 · 2018
Earlier work this paper cites.
Learnable pins: Cross-modal embeddings for person identity. In Proceedings of the European Conference on Computer Vision (ECCV) . 71–88
Arsha Nagrani, Samuel Albanie, and Andrew Zisserman. 2018 · 2018
Earlier work this paper cites.
Representation learning with contrastive predictive coding
Aaron van den Oord, Yazhe Li, and Oriol Vinyals. 2018 · 2018
Earlier work this paper cites.
Audio-visual scene analysis with self-supervised multisensory features. In Proceedings of the European Conference on Computer Vision (ECCV)
Andrew Owens and Alexei A Efros. 2018 · 2018
Earlier work this paper cites.
Time-Contrastive Networks: Self-Supervised Learning from Video
Pierre Sermanet, Corey Lynch, Yevgen Chebotar, Jasmine Hsu, Eric Jang, Stefan Schaal, and Sergey Levine. 2018 · 2018
Earlier work this paper cites.
Learning and using the arrow of time. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition . 8052–8060
Donglai Wei, Joseph J Lim, Andrew Zisserman, and William T Freeman. 2018 · 2018
Earlier work this paper cites.
Learning to detect and retrieve objects from unlabeled videos. In 2019 IEEE/CVF International Conference on Computer Vision Workshop (ICCVW) . IEEE, 3713–3717
Elad Amrani, Rami Ben-Ari, Tal Hakim, and Alex Bronstein. 2019 · 2019
Earlier work this paper cites.
Effectiveness of self-supervised pre-training for speech recognition
Alexei Baevski, Michael Auli, and Abdelrahman Mohamed. 2019 · 2019
Earlier work this paper cites.
Selective sensor fusion for neural visual-inertial odometry. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition . 10542–10551
Changhao Chen, Stefano Rosa, Yishu Miao, Chris Xiaoxuan Lu, Wei Wu, Andrew Markham, and Niki Trigoni. 2019 · 2019
Earlier work this paper cites.
Gru-ode-bayes: Continuous modeling of sporadically-observed time series
Edward De Brouwer, Jaak Simm, Adam Arany, and Yves Moreau. 2019 · 2019
Earlier work this paper cites.
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding. In Proc. of the Conf. of the North American Chapter of the Assoc. for Computational Linguistics: Human Lang. Tech., NAACL-HLT 2019
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Earlier work this paper cites.
Deep embedding learning with discriminative sampling policy. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition . 4964–4973
Yueqi Duan, Lei Chen, Jiwen Lu, and Jie Zhou. 2019 · 2019
Earlier work this paper cites.
Unsupervised scalable representation learning for multivariate time series. In Advances in Neural Information Processing Systems . 4650–4661
Jean-Yves Franceschi, Aymeric Dieuleveut, and Martin Jaggi. 2019 · 2019
Earlier work this paper cites.
Deep Multimodal Representation Learning: A Survey
Wenzhong Guo, Jianwen Wang, and Shiping Wang. 2019 · 2019
Earlier work this paper cites.
Learning deep representations by mutual information estimation and maximization. In International Conference on Learning Representations
R Devon Hjelm, Alex Fedorov, Samuel Lavoie-Marchildon, Karan Grewal, Phil Bachman, Adam Trischler, and Yoshua Bengio. 2019 · 2019
Earlier work this paper cites.
Vilbert: Pretraining task-agnostic visiolinguistic representations for vision-and-language tasks
Jiasen Lu, Dhruv Batra, Devi Parikh, and Stefan Lee. 2019 · 2019
Earlier work this paper cites.
Howto100m: Learning a text-video embedding by watching hundred million narrated video clips. In Proceedings of the IEEE/CVF International Conference on Computer Vision . 2630–2640
Antoine Miech, Dimitri Zhukov, Jean-Baptiste Alayrac, Makarand Tapaswi, Ivan Laptev, and Josef Sivic. 2019 · 2019
Earlier work this paper cites.
Learning Problem-Agnostic Speech Representations from Multiple Self-Supervised Tasks. In Interspeech 2019, 20th Annual Conference of the International Speech Communication Association 2019
Santiago Pascual, Mirco Ravanelli, Joan Serrà, Antonio Bonafonte, and Yoshua Bengio. [n.d.] · 2019
Earlier work this paper cites.
Continual Unsupervised Representation Learning
Dushyant Rao, Francesco Visin, Andrei Rusu, Razvan Pascanu, Yee Whye Teh, and Raia Hadsell. 2019 · 2019
Earlier work this paper cites.
Latent ordinary differential equations for irregularly-sampled time series
Yulia Rubanova, Ricky TQ Chen, and David K Duvenaud. 2019 · 2019
Earlier work this paper cites.
Multi-task Self-Supervised Learning for Human Activity Detection
Aaqib Saeed, Tanir Ozcelebi, and Johan Lukkien. 2019 · 2019
Cited alongside, same era.
Videobert: A joint model for video and language representation learning. In Proceedings of the IEEE/CVF International Conference on Computer Vision . 7464–7473
Chen Sun, Austin Myers, Carl Vondrick, Kevin Murphy, and Cordelia Schmid. 2019 · 2019
Cited alongside, same era.
Ranked list loss for deep metric learning. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition
Xinshao Wang, Yang Hua, Elyor Kodirov, Guosheng Hu, Romain Garnier, and Neil M Robertson. 2019 · 2019
Cited alongside, same era.
Self-Supervised MultiModal Versatile Networks
Jean-Baptiste Alayrac, Adria Recasens, Rosalia Schneider, Relja Arandjelovic, Jason Ramapuram, Jeffrey De Fauw, Lucas Smaira, Sander Dieleman, and Andrew Zisserman. 2020 · 2020
Cited alongside, same era.
Self-Supervised Learning by Cross-Modal Audio-Video Clustering. In Advances in Neural Information Processing Systems (NeurIPS)
Humam Alwassel, Dhruv Mahajan, Bruno Korbar, Lorenzo Torresani, Bernard Ghanem, and Du Tran. 2020 · 2020
Multimodal clustering networks for self-supervised learning from unlabeled videos. In ICCV
Brian Chen, Andrew Rouditchenko, Kevin Duarte, Hilde Kuehne, Samuel Thomas, Angie Boggust, Rameswar Panda, Brian Kingsbury, Rogerio Feris, David Harwath, et al · 2021
Later among the works it cites.
Exploring simple siamese representation learning. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition . 15750–15758
Xinlei Chen and Kaiming He. 2021 · 2021
Later among the works it cites.
Audio albert: A lite bert for self-supervised learning of audio representation. In 2021 IEEE Spoken Language Technology Workshop (SLT) . IEEE, 344–350
Po-Han Chi, Pei-Hung Chung, Tsung-Han Wu, Chun-Cheng Hsieh, Yen-Hao Chen, Shang-Wen Li, and Hung-yi Lee. 2021 · 2021
Later among the works it cites.
Time Series Change Point Detection with Self-Supervised Contrastive Predictive Coding. In Proceedings of The Web Conference 2021 (WWW ’21) . Association for Computing Machinery
Shohreh Deldari, Daniel V. Smith, Hao Xue, and Flora D. Salim. 2021 · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Labelling unlabelled videos from scratch with multi-modal self-supervision
Yuki Asano, Mandela Patrick, Christian Rupprecht, and Andrea Vedaldi. 2020 · 2020
Cited alongside, same era.
vq-wav2vec: Self-Supervised Learning of Discrete Speech Representations. In 8th International Conference on Learning Representations, ICLR 2020, Addis Ababa, Ethiopia, April 26-30, 2020
Alexei Baevski, Steffen Schneider, and Michael Auli. 2020a · 2020
Cited alongside, same era.
On the Benefits of Early Fusion in Multimodal Representation Learning
George Barnum, Sabera Talukder, and Yisong Yue. 2020 · 2020
Cited alongside, same era.
Learning with Fenchel-Young losses
Mathieu Blondel, André FT Martins, and Vlad Niculae. 2020 · 2020
Cited alongside, same era.
Unsupervised learning of visual features by contrasting cluster assignments. In Advances in Neural Information Processing Systems (NeurIPS)
Mathilde Caron, Ishan Misra, Julien Mairal, Priya Goyal, Piotr Bojanowski, and Armand Joulin. 2020 · 2020
Cited alongside, same era.
A simple framework for contrastive learning of visual representations. In International conference on machine learning . PMLR, 1597–1607
Ting Chen, Simon Kornblith, Mohammad Norouzi, and Geoffrey Hinton. 2020 · 2020
Cited alongside, same era.
Subject-aware contrastive learning for biosignals
Joseph Y Cheng, Hanlin Goh, Kaan Dogrusoz, Oncel Tuzel, and Erdrin Azemi. 2020 · 2020
Cited alongside, same era.
Virtex: Learning visual representations from textual annotations. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition . 11162–11173
Karan Desai and Justin Johnson. 2021 · 2021
Later among the works it cites.
Time-Series Representation Learning via Temporal and Contextual Contrasting. In Proceedings of the Thirtieth International Joint Conference on Artificial Intelligence, IJCAI-21
Emadeldeen Eldele, Mohamed Ragab, Zhenghua Chen, Min Wu, Chee Keong Kwoh, Xiaoli Li, and Cuntai Guan. 2021 · 2021
Later among the works it cites.
Self-Supervised Representation Learning: Introduction, Advances and Challenges
Linus Ericsson, Henry Gouk, Chen Change Loy, and Timothy M Hospedales. 2021 · 2021
Later among the works it cites.
Whitening for Self-Supervised Representation Learning. In Proceedings of the 38th International Conference on Machine Learning (Proceedings of Machine Learning Research, Vol. 139) . PMLR, 3015–3024
Aleksandr Ermolov, Aliaksandr Siarohin, Enver Sangineto, and Nicu Sebe. 2021 · 2021
Later among the works it cites.
Learning contextual tag embeddings for cross-modal alignment of audio and tags. In ICASSP 2021-2021 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) . IEEE, 596–600
Xavier Favory, Konstantinos Drossos, Tuomas Virtanen, and Xavier Serra. 2021 · 2021
Later among the works it cites.
DeCLUTR: Deep Contrastive Learning for Unsupervised Textual Representations. In Proc. of the Annual Meeting of the Assoc. for Computational Linguistics and the International Joint Conf. on Natural Lang. Processing, ACL/IJCNLP
John M. Giorgi, Osvald Nitski, Bo Wang, and Gary D. Bader. 2021 · 2021
Later among the works it cites.
AudioCLIP: Extending CLIP to Image, Text and Audio
Andrey Guzhov, Federico Raue, Jörn Hees, and Andreas Dengel. 2021 · 2021
Later among the works it cites.
Contrastive predictive coding for human activity recognition
Harish Haresamudram, Irfan Essa, and Thomas Plötz. 2021 · 2021
Later among the works it cites.
HuBERT: Self-Supervised Speech Representation Learning by Masked Prediction of Hidden Units
Wei-Ning Hsu, Benjamin Bolte, Yao-Hung Hubert Tsai, Kushal Lakhotia, Ruslan Salakhutdinov, and Abdelrahman Mohamed. 2021 · 2021
Later among the works it cites.
On feature decorrelation in self-supervised learning. In Proceedings of the IEEE/CVF International Conference on Computer Vision . 9598–9608
Tianyu Hua, Wenxiao Wang, Zihui Xue, Sucheng Ren, Yue Wang, and Hang Zhao. 2021 · 2021
Later among the works it cites.
Self-supervised feature learning by cross-modality and cross-view correspondences. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition . 1581–1591
Longlong Jing, Ling Zhang, and Yingli Tian. 2021 · 2021
Later among the works it cites.
Self-supervised Pre-training and Contrastive Representation Learning for Multiple-choice Video QA
Seonhoon Kim, Seohyeong Jeong, Eunbyul Kim, Inho Kang, and Nojun Kwak. 2021a · 2021
Later among the works it cites.
Unsupervised balanced covariance learning for visual-inertial sensor fusion
Youngji Kim, Sungho Yoon, Sujung Kim, and Ayoung Kim. 2021b · 2021
Later among the works it cites.
Self-supervised Learning: The Dark Matter of Intelligence
Yan LeCun and Ishan Misra. 2021 · 2021
Later among the works it cites.
Prototypical Contrastive Learning of Unsupervised Representations. In International Conference on Learning Representations
Junnan Li, Pan Zhou, Caiming Xiong, and Steven Hoi. 2021 · 2021
Later among the works it cites.
Contrastive clustering. In 2021 AAAI
Yunfan Li, Peng Hu, Zitao Liu, Dezhong Peng, Joey Tianyi Zhou, and Xi Peng. [n.d.] · 2021
Later among the works it cites.
Tera: Self-supervised learning of transformer encoder representation for speech
Andy T Liu, Shang-Wen Li, and Hung-yi Lee. 2021a · 2021
Later among the works it cites.
Self-supervised Learning: Generative or Contrastive
Xiao Liu, Fanjin Zhang, Zhenyu Hou, Li Mian, Zhaoyu Wang, Jing Zhang, and Jie Tang. 2021b · 2021
Later among the works it cites.
Audio-visual instance discrimination with cross-modal agreement. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition . 12475–12486
Pedro Morgado, Nuno Vasconcelos, and Ishan Misra. 2021 · 2021
Later among the works it cites.
Temporal Predictive Coding For Model-Based Planning In Latent Space. In International Conference on Machine Learning
Tung Nguyen, Rui Shu, Tuan Pham, Hung Bui, and Stefano Ermon. 2021 · 2021
Later among the works it cites.
Spatiotemporal contrastive video representation learning. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition . 6964–6974
Rui Qian, Tianjian Meng, Boqing Gong, Ming-Hsuan Yang, Huisheng Wang, Serge Belongie, and Yin Cui. 2021 · 2021
Later among the works it cites.
Learning transferable visual models from natural language supervision. In International Conference on Machine Learning . PMLR
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, et al · 2021
Later among the works it cites.
Contrastive Learning with Hard Negative Samples. In International Conference on Learning Representations
Joshua David Robinson, Ching-Yao Chuang, Suvrit Sra, and Stefanie Jegelka. 2021 · 2021
Later among the works it cites.
Contrastive learning of general-purpose audio representations. In ICASSP 2021-2021 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) . IEEE, 3875–3879
Aaqib Saeed, David Grangier, and Neil Zeghidour. 2021a · 2021
Later among the works it cites.
Sense and Learn: Self-supervision for omnipresent sensors
Aaqib Saeed, Victor Ungureanu, and Beat Gfeller. 2021c · 2021
Later among the works it cites.
Data-Efficient Reinforcement Learning with Self-Predictive Representations. In 9th International Conference on Learning Representations, ICLR 2021, Virtual Event
Max Schwarzer, Ankesh Anand, Rishab Goel, R. Devon Hjelm, Aaron C. Courville, and Philip Bachman. 2021 · 2021
Later among the works it cites.
Relating by Contrasting: A Data-efficient Framework for Multimodal Generative Models. In International Conference on Learning Representations
Yuge Shi, Brooks Paige, Philip Torr, and Siddharth N. 2021 · 2021
Later among the works it cites.
DABS: a Domain-Agnostic Benchmark for Self-Supervised Learning. In Thirty-fifth Conference on Neural Information Processing Systems Datasets and Benchmarks Track (Round 1)
Alex Tamkin, Vincent Liu, Rongfei Lu, Daniel Fein, Colin Schultz, and Noah Goodman. 2021 · 2021
Later among the works it cites.
Divide and contrast: Self-supervised learning from uncurated data. In Proceedings of the IEEE/CVF International Conference on Computer Vision . 10063–10074
Yonglong Tian, Olivier J Henaff, and Aäron van den Oord. 2021 · 2021
Later among the works it cites.
Unsupervised Representation Learning for Time Series with Temporal Neighborhood Coding. In International Conference on Learning Representations
Sana Tonekaboni, Danny Eytan, and Anna Goldenberg. 2021 · 2021
Later among the works it cites.
Understanding the Behaviour of Contrastive Loss. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) . 2495–2504
Feng Wang and Huaping Liu. 2021 · 2021
Later among the works it cites.
Contrastive Separative Coding for Self-Supervised Representation Learning. In ICASSP 2021-2021 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) . IEEE, 3865–3869
Jun Wang, Max W Y Lam, Dan Su, and Dong Yu. 2021a · 2021
Later among the works it cites.
Multimodal Self-Supervised Learning of General Audio Representations
Luyu Wang, Pauline Luc, Adria Recasens, Jean-Baptiste Alayrac, and Aaron van den Oord. 2021b · 2021
Later among the works it cites.
Time Series Data Augmentation for Deep Learning: A Survey. In Proceedings of the Thirtieth International Joint Conference on Artificial Intelligence, IJCAI-21 , Zhi-Hua Zhou (Ed.). International Joint Conferences on Artificial Intelligence Organization, 4653–4660
Qingsong Wen, Liang Sun, Fan Yang, Xiaomin Song, Jingkun Gao, Xue Wang, and Huan Xu. 2021 · 2021
Later among the works it cites.
Self-Supervised Learning for Sleep Stage Classification with Predictive and Discriminative Contrastive Coding. In International Conf. on Acoustics, Speech and Signal Processing (ICASSP)
Qinfeng Xiao, Jing Wang, Jianan Ye, Hongjun Zhang, Yuyan Bu, Yiqiong Zhang, and Hao Wu. 2021 · 2021
Later among the works it cites.
A Survey on Deep Semi-supervised Learning
Xiangli Yang, Zixing Song, Irwin King, and Zenglin Xu. 2021b · 2021
Later among the works it cites.
Multimodal Contrastive Training for Visual Representation Learning. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition . 6995–7004
Xin Yuan, Zhe Lin, Jason Kuen, Jianming Zhang, Yilin Wang, Michael Maire, Ajinkya Kale, and Baldo Faieta. 2021 · 2021
Later among the works it cites.
TS2Vec: Towards Universal Representation of Time Series
Zhihan Yue, Yujing Wang, Juanyong Duan, Tianmeng Yang, Congrui Huang, Yunhai Tong, and Bixiong Xu. 2021 · 2021
Later among the works it cites.
Pretext Tasks selection for multitask self-supervised speech representation learning
Salah Zaiem, Titouan Parcollet, and Slim Essid. 2021 · 2021
Later among the works it cites.
Barlow Twins: Self-Supervised Learning via Redundancy Reduction. In Proceedings of the 38th International Conference on Machine Learning, ICML , Marina Meila and Tong Zhang (Eds.), Vol. 139. PMLR
Jure Zbontar, Li Jing, Ishan Misra, Yann LeCun, and Stéphane Deny. 2021 · 2021
Later among the works it cites.
A Transformer-Based Framework for Multivariate Time Series Representation Learning. In Proceedings of the 27th ACM SIGKDD Conference on Knowledge Discovery and Data Mining
George Zerveas, Srideepika Jayaraman, Dhaval Patel, Anuradha Bhamidipaty, and Carsten Eickhoff. 2021 · 2021
Later among the works it cites.
SelfVIO: Self-supervised deep monocular Visual–Inertial Odometry and depth estimation
Yasin Almalioglu, Mehmet Turan, Muhamad Risqi U. Saputra, Pedro P.B. de Gusmão, Andrew Markham, and Niki Trigoni. 2022 · 2022
Closest in time.
Data2vec: A general framework for self-supervised learning in speech, vision and language
Alexei Baevski, Wei-Ning Hsu, Qiantong Xu, Arun Babu, Jiatao Gu, and Michael Auli. 2022 · 2022
Closest in time.
VICReg: Variance-Invariance-Covariance Regularization for Self-Supervised Learning. In International Conference on Learning Representations
Adrien Bardes, Jean Ponce, and Yann LeCun. 2022 · 2022
Closest in time.
Generative adversarial networks for spatio-temporal data: A survey
Nan Gao, Hao Xue, Wei Shao, Sichen Zhao, Kyle Kai Qin, Arian Prabowo, Mohammad Saiedur Rahaman, and Flora D Salim. 2022 · 2022
Closest in time.
Boosting contrastive self-supervised learning with false negative cancellation. In Proceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision . 2785–2795
Tri Huynh, Simon Kornblith, Matthew R Walter, Michael Maire, and Maryam Khademi. 2022 · 2022
Closest in time.
ColloSSL: Collaborative Self-Supervised Learning for Human Activity Recognition
Yash Jain, Chi Ian Tang, Chulhong Min, Fahim Kawsar, and Akhil Mathur. 2022 · 2022
Closest in time.
Cross-Modal Common Representation Learning with Triplet Loss Functions
Felix Ott, David Rügamer, Lucas Heublein, Bernd Bischl, and Christopher Mutschler. 2022 · 2022
Closest in time.
Self-supervised Video Representation Learning with Cross-Stream Prototypical Contrasting. In Proceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision . 108–118
Martine Toering, Ioannis Gatopoulos, Maarten Stol, and Vincent Tao Hu. 2022 · 2022
Closest in time.