vix.ing · top · new · best · stats · spec

Towards a Hypothesis on Visual Transformation based Self-Supervision

2019/11/24 by Dipan K. Pal, Pal, Dipan K., Sreena Nallamothu +3
Biochemistry, Genetics and Molecular Biology · Computer Science · Mathematics · #Cell Image Analysis Techniques #Computer Vision and Pattern Recognition (cs.CV) #Data Visualization and Analytics #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #cs.CV #cs.LG #stat.ML

paper · pdf · doi:10.48550/arxiv.1911.10594

Draft

openalex publication_date 2019/11/24 · arxiv created 2020/02/14 · arxiv updated 2020/02/18 · openalex created_date 2025/10/10 · openalex updated_date 2026/07/28

Abstract

We propose the first qualitative hypothesis characterizing the behavior of visual transformation based self-supervision, called the VTSS hypothesis. Given a dataset upon which a self-supervised task is performed while predicting instantiations of a transformation, the hypothesis states that if the predicted instantiations of the transformations are already present in the dataset, then the representation learned will be less useful. The hypothesis was derived by observing a key constraint in the application of self-supervision using a particular transformation. This constraint, which we term the transformation conflict for this paper, forces a network learn degenerative features thereby reducing the usefulness of the representation. The VTSS hypothesis helps us identify transformations that have the potential to be effective as a self-supervision task. Further, it helps to generally predict whether a particular transformation based self-supervision technique would be effective or not for a particular dataset. We provide extensive evaluations on CIFAR 10, CIFAR 100, SVHN and FMNIST confirming the hypothesis and the trends it predicts. We also propose novel cost-effective self-supervision techniques based on translation and scale, which when combined with rotation outperforms all transformations applied individually. Overall, this paper aims to shed light on the phenomenon of visual transformation based self-supervision.

Citations

Related