GAMT: A Geometry-Aware, Multi-view, Training-Free Segmentation Framework for Foundation Models in Medical Imaging

Citations

SCOPUS

0

초록

Medical image segmentation is a critical task in clinical diagnostics and biomedical research. While deep learning has significantly advanced the field, most existing methods rely on task-specific models that require extensive manual annotations for training or adaptation. Vision foundation models, such as the Segment Anything Model (SAM), offer a promising alternative with their universal segmentation capabilities. However, their application to 3D medical imaging remains limited, especially in zero-shot scenarios involving previously unseen anatomical structures. In this work, we introduce GAMT, a zero-shot, training-free framework that repurposes powerful 2D foundation segmentation models (e.g., SAM, SAM-Med2D) for universal 3D biomedical image segmentation. To bridge the dimensionality gap, GAMT performs slice-wise inference along three orthogonal anatomical planes (axial, coronal, and sagittal) and subsequently fuses the predictions to construct a coherent 3D segmentation mask. Crucially, without any model training or fine-tuning, this framework achieves average Dice Similarity Coefficient (DSC) and Normalized Surface Dice (NSD) scores of 0.487 and 0.477, respectively—without requiring model training or fine-tuning. Our code and results are publicly available at https://github.com/SpatialAILab/GAMT.

키워드

foundation-modelMedical image segmentationtraining-freeClinical researchDeep learningDiagnosisFoundationsMedical image processing
제목
GAMT: A Geometry-Aware, Multi-view, Training-Free Segmentation Framework for Foundation Models in Medical Imaging
저자
Jo, SunChoi, AhjinHong, Je Hyeong
DOI
10.1007/978-3-032-23496-4_3
발행일
2026-00
유형
Conference paper
저널명
Lecture Notes in Computer Science
16447 LNCS
페이지
36 ~ 50