Don't miss the Mismatch: Investigating the Objective Function Mismatch for Unsupervised Representation Learning

Stuhr, Bonifaz; Brauer, Jürgen

doi:10.1007/s00521-022-07031-9

Computer Science > Computer Vision and Pattern Recognition

arXiv:2009.02383 (cs)

[Submitted on 4 Sep 2020 (v1), last revised 28 Feb 2022 (this version, v2)]

Title:Don't miss the Mismatch: Investigating the Objective Function Mismatch for Unsupervised Representation Learning

Authors:Bonifaz Stuhr, Jürgen Brauer

View PDF

Abstract:Finding general evaluation metrics for unsupervised representation learning techniques is a challenging open research question, which recently has become more and more necessary due to the increasing interest in unsupervised methods. Even though these methods promise beneficial representation characteristics, most approaches currently suffer from the objective function mismatch. This mismatch states that the performance on a desired target task can decrease when the unsupervised pretext task is learned too long - especially when both tasks are ill-posed. In this work, we build upon the widely used linear evaluation protocol and define new general evaluation metrics to quantitatively capture the objective function mismatch and the more generic metrics mismatch. We discuss the usability and stability of our protocols on a variety of pretext and target tasks and study mismatches in a wide range of experiments. Thereby we disclose dependencies of the objective function mismatch across several pretext and target tasks with respect to the pretext model's representation size, target model complexity, pretext and target augmentations as well as pretext and target task types. In our experiments, we find that the objective function mismatch reduces performance by ~0.1-5.0% for Cifar10, Cifar100 and PCam in many setups, and up to ~25-59% in extreme cases for the 3dshapes dataset.

Comments:	13 pages, 6 figures, Published in Neural Computing and Applications
Subjects:	Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI); Machine Learning (cs.LG)
ACM classes:	I.4; I.5; I.2
Cite as:	arXiv:2009.02383 [cs.CV]
	(or arXiv:2009.02383v2 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2009.02383
Journal reference:	Neural Computing and Applications (2022)
Related DOI:	https://doi.org/10.1007/s00521-022-07031-9

Submission history

From: Bonifaz Stuhr [view email]
[v1] Fri, 4 Sep 2020 20:21:17 UTC (4,960 KB)
[v2] Mon, 28 Feb 2022 20:15:42 UTC (1,188 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:Don't miss the Mismatch: Investigating the Objective Function Mismatch for Unsupervised Representation Learning

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Don't miss the Mismatch: Investigating the Objective Function Mismatch for Unsupervised Representation Learning

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators