Reducing Training Data Using Pre-Trained Foundation Models: A Case Study on Traffic Sign Segmentation Using the Segment Anything Model

Sofia Henninger; Maximilian Kellner; Benedikt Rombach; Alexander Reiterer

doi:10.3390/jimaging10090220

Journal of Imaging (Sep 2024)

Reducing Training Data Using Pre-Trained Foundation Models: A Case Study on Traffic Sign Segmentation Using the Segment Anything Model

Sofia Henninger,
Maximilian Kellner,
Benedikt Rombach,
Alexander Reiterer

Affiliations

Sofia Henninger: Fraunhofer Institute for Physical Measurement Techniques IPM, 79110 Freiburg, Germany
Maximilian Kellner: Fraunhofer Institute for Physical Measurement Techniques IPM, 79110 Freiburg, Germany
Benedikt Rombach: Fraunhofer Institute for Physical Measurement Techniques IPM, 79110 Freiburg, Germany
Alexander Reiterer: Fraunhofer Institute for Physical Measurement Techniques IPM, 79110 Freiburg, Germany

DOI: https://doi.org/10.3390/jimaging10090220
Journal volume & issue: Vol. 10, no. 9
p. 220

Abstract

Read online

The utilization of robust, pre-trained foundation models enables simple adaptation to specific ongoing tasks. In particular, the recently developed Segment Anything Model (SAM) has demonstrated impressive results in the context of semantic segmentation. Recognizing that data collection is generally time-consuming and costly, this research aims to determine whether the use of these foundation models can reduce the need for training data. To assess the models’ behavior under conditions of reduced training data, five test datasets for semantic segmentation will be utilized. This study will concentrate on traffic sign segmentation to analyze the results in comparison to Mask R-CNN: the field’s leading model. The findings indicate that SAM does not surpass the leading model for this specific task, regardless of the quantity of training data. Nevertheless, a knowledge-distilled student architecture derived from SAM exhibits no reduction in accuracy when trained on data that have been reduced by 95%.

Published in Journal of Imaging

ISSN: 2313-433X (Online)
Publisher: MDPI AG
Country of publisher: Switzerland
LCC subjects: Technology: Photography; Medicine: Medicine (General): Computer applications to medicine. Medical informatics; Science: Mathematics: Instruments and machines: Electronic computers. Computer science
Website: http://www.mdpi.com/journal/jimaging

About the journal

Abstract

Keywords