Self-supervised learning with automatic data augmentation for enhancing representation

Citations

WEB OF SCIENCE

4
Citations

SCOPUS

4

초록

Self-supervised learning has become an increasingly popular method for learning effective representations from unlabeled data. One prominent approach in self-supervised learning is contrastive learning, which trains models to distinguish between similar and dissimilar sample pairs by pulling similar pairs closer and pushing dissimilar pairs farther apart. The key to the success of contrastive learning lies in the quality of the data augmentation, which increases the diversity of the data and helps the model learn more powerful and generalizable representations. While many studies have emphasized the importance of data augmentation, however, most of them rely on human-crafted augmentation strategies. In this paper, we propose a novel method, Self Augmentation on Contrastive Learning with Clustering (SACL), searching for the optimal data augmentation policy automatically using Bayesian optimization and clustering. The proposed approach overcomes the limitations of relying on domain knowledge and avoids the high costs associated with manually designing data augmentation rules. It automatically captures informative and useful features within the data by exploring augmentation policies. We demonstrate that the proposed method surpasses existing approaches that rely on manually designed augmentation rules. Our experiments show SACL outperforms manual strategies, achieving a performance improvement of 1.68% and 1.57% over MoCo v2 on the CIFAR10 and SVHN datasets, respectively. © 2024 Elsevier B.V.

키워드

Auto augmentationClusteringContrastive learningSelf-supervised learning
제목
Self-supervised learning with automatic data augmentation for enhancing representation
저자
Park, ChanjongKim, Eunwoo
DOI
10.1016/j.patrec.2024.06.012
발행일
2024-08
유형
Article
저널명
Pattern Recognition Letters
184
페이지
133 ~ 139