Skip to main navigation Skip to search Skip to main content

CRAX: Parameter-Efficient Fine-Tuning of SAM2 for Interactive Crack Annotation

Research output: Chapter in Book or Conference ProceedingsConference Proceedings with Oral Presentationpeer-review

Abstract

We study how to adapt the Segment Anything Model 2 (SAM2) for interactive segmentation of thin, cracklike structures. Rather than building a fully automatic crack detector, we focus on the annotator’s perspective: how many clicks are needed to obtain boundary- and topology-faithful masks. To this end, we curate CRAX, a multi-domain corpus of 35 datasets covering surface cracks, retinal vessels, and plant roots/mycelium, together with leakage-controlled leave-one-domain-out and 5-fold leave-one-dataset-out splits tailored to interactive evaluation. On top of SAM2, we systematically compare three fine-tuning strategies: decoderonly, Low-Rank-Adaptation+decoder, and full encoder+decoder. Using a conservative click simulator and a topology-aware metric suite, we show that full fine-tuning substantially improves performance on challenging thin-structure domains, almost doubling boundary IoU and nearly halving Hausdorff distance in the hardest cross-domain setting, while consistently impro ving cross-dataset generalization within cracks. Most importantly, fine-tuned models reach—and often surpass—the baseline’s 9-click quality after only one to two clicks, reducing annotation effort and making SAM2 more effective for large-scale crack and crack-like labeling.
Original languageEnglish
Title of host publicationProceedings of the 21st International Conference on Computer Vision Theory and Applications - Volume 1: VISAPP
EditorsAntonino Furnari, Petia Radeva
Pages689-700
Number of pages12
Volume1
DOIs
Publication statusPublished - 2026
Event21st International Conference on Computer Vision Theory and Applications - Barceló Marbella hotel, Marbella, Spain
Duration: 9 Mar 202611 Mar 2026
Conference number: 21
https://visapp.scitevents.org/Home.aspx

Conference

Conference21st International Conference on Computer Vision Theory and Applications
Abbreviated titleVISAPP 2026
Country/TerritorySpain
CityMarbella
Period9/03/2611/03/26
Internet address

Research Field

  • High-Performance Vision Systems

Keywords

  • Crack Segmentation
  • Interactive Segmentation
  • Segment Anything Model
  • Parameter-Efficient Fine-Tuning
  • LoRA
  • Thin-Structure Segmentation
  • Structural Health Monitoring
  • Annotation Efficiency

Web of Science subject categories (JCR Impact Factors)

  • Computer Science, Artificial Intelligence

Fingerprint

Dive into the research topics of 'CRAX: Parameter-Efficient Fine-Tuning of SAM2 for Interactive Crack Annotation'. Together they form a unique fingerprint.

Cite this