Don't miss the Mismatch: Investigating the Objective Function Mismatch for Unsupervised Representation Learning
arXiv:2009.02383 · doi:10.1007/s00521-022-07031-9
Abstract
Finding general evaluation metrics for unsupervised representation learning techniques is a challenging open research question, which recently has become more and more necessary due to the increasing interest in unsupervised methods. Even though these methods promise beneficial representation characteristics, most approaches currently suffer from the objective function mismatch. This mismatch states that the performance on a desired target task can decrease when the unsupervised pretext task is learned too long - especially when both tasks are ill-posed. In this work, we build upon the widely used linear evaluation protocol and define new general evaluation metrics to quantitatively capture the objective function mismatch and the more generic metrics mismatch. We discuss the usability and stability of our protocols on a variety of pretext and target tasks and study mismatches in a wide range of experiments. Thereby we disclose dependencies of the objective function mismatch across several pretext and target tasks with respect to the pretext model's representation size, target model complexity, pretext and target augmentations as well as pretext and target task types. In our experiments, we find that the objective function mismatch reduces performance by ~0.1-5.0% for Cifar10, Cifar100 and PCam in many setups, and up to ~25-59% in extreme cases for the 3dshapes dataset.
13 pages, 6 figures, Published in Neural Computing and Applications
References in corpus (18)
- A Simple Framework for Contrastive Learning of Visual Representations
- Improved Baselines with Momentum Contrastive Learning
- Learning deep representations by mutual information estimation and maximization
- Challenging Common Assumptions in the Unsupervised Learning of Disentangled Representations
- Prototypical Contrastive Learning of Unsupervised Representations
- Large Scale Adversarial Representation Learning
- Unsupervised Learning via Meta-Learning
- A Large-scale Study of Representation Learning with the Visual Task Adaptation Benchmark
- A critical analysis of self-supervision, or what we can learn from a single image
- Evaluation Metrics for Unsupervised Learning Algorithms
- An analysis on the use of autoencoders for representation learning: fundamentals, learning task case studies, explainability and challenges
- AET vs. AED: Unsupervised Representation Learning by Auto-Encoding Transformations rather than Data
- On Mutual Information in Contrastive Learning for Visual Representations
- Self-Supervised Relational Reasoning for Representation Learning
- Unsupervised Deep Metric Learning via Auxiliary Rotation Loss
- CSNNs: Unsupervised, Backpropagation-free Convolutional Neural Networks for Representation Learning
- Guided Generative Adversarial Neural Network for Representation Learning and High Fidelity Audio Generation using Fewer Labelled Audio Data
- Instance Separation Emerges from Inpainting