arXiv · 2412.00770
Explorations in Self-Supervised Learning: Dataset Composition Testing for Object Classification
Abstract
This paper investigates the impact of sampling and pretraining using datasets with different image characteristics on the performance of self-supervised learning (SSL) models for object classification. To do this, we sample two apartment datasets from the Omnidata platform based on modality, luminosity, image size, and camera field of view and use them to pretrain a SimCLR model. The encodings generated from the pretrained model are then transferred to a supervised Resnet-50 model for object classification. Through A/B testing, we find that depth pretrained models are more effective on low resolution images, while RGB pretrained models perform better on higher resolution images. We also discover that increasing the luminosity of training images can improve the performance of models on low resolution images without negatively affecting their performance on higher resolution images.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Raynor Kirkson E. Chavez, Kyle Gabriel M. Reynoso. 2024-12-01. Explorations in Self-Supervised Learning: Dataset Composition Testing for Object Classification. https://arxiv.org/abs/2412.00770
Cite the original work for its findings. Save a collection to share your selection of sources.