Paper Reading AI Learner

SPENet: Self-guided Prototype Enhancement Network for Few-shot Medical Image Segmentation

2025-09-03 03:59:27
Chao Fan, Xibin Jia, Anqi Xiao, Hongyuan Yu, Zhenghan Yang, Dawei Yang, Hui Xu, Yan Huang, Liang Wang

Abstract

Few-Shot Medical Image Segmentation (FSMIS) aims to segment novel classes of medical objects using only a few labeled images. Prototype-based methods have made significant progress in addressing FSMIS. However, they typically generate a single global prototype for the support image to match with the query image, overlooking intra-class variations. To address this issue, we propose a Self-guided Prototype Enhancement Network (SPENet). Specifically, we introduce a Multi-level Prototype Generation (MPG) module, which enables multi-granularity measurement between the support and query images by simultaneously generating a global prototype and an adaptive number of local prototypes. Additionally, we observe that not all local prototypes in the support image are beneficial for matching, especially when there are substantial discrepancies between the support and query images. To alleviate this issue, we propose a Query-guided Local Prototype Enhancement (QLPE) module, which adaptively refines support prototypes by incorporating guidance from the query image, thus mitigating the negative effects of such discrepancies. Extensive experiments on three public medical datasets demonstrate that SPENet outperforms existing state-of-the-art methods, achieving superior performance.

Abstract (translated)

少样本医学图像分割(FSMIS)的目标是仅使用少量标注图像来对新型的医学对象进行分割。基于原型的方法在解决FSMIS问题上取得了显著进展,但它们通常为支持图像生成单一全局原型与查询图像匹配,从而忽略了类内的变化。为了应对这一挑战,我们提出了一种自我引导的原型增强网络(SPENet)。具体来说,我们引入了一个多层次原型生成(MPG)模块,该模块通过同时生成全局原型和自适应数量的局部原型,使得支持图像和查询图像之间可以进行多粒度测量。此外,我们观察到并非支持图像中的所有局部原型都对匹配有帮助,特别是在支持图像与查询图像之间存在显著差异的情况下更是如此。为缓解这一问题,我们提出了一种基于查询引导的局部原型增强(QLPE)模块,该模块通过从查询图像中获取指导信息来自适应地精炼支持原型,从而减轻这些差异所带来的负面影响。在三个公开医学数据集上的广泛实验表明,SPENet优于现有的最先进方法,在性能上取得了更好的结果。

URL

https://arxiv.org/abs/2509.02993

PDF

https://arxiv.org/pdf/2509.02993.pdf


Tags
3D Action Action_Localization Action_Recognition Activity Adversarial Agent Attention Autonomous Bert Boundary_Detection Caption Chat Classification CNN Compressive_Sensing Contour Contrastive_Learning Deep_Learning Denoising Detection Dialog Diffusion Drone Dynamic_Memory_Network Edge_Detection Embedding Embodied Emotion Enhancement Face Face_Detection Face_Recognition Facial_Landmark Few-Shot Gait_Recognition GAN Gaze_Estimation Gesture Gradient_Descent Handwriting Human_Parsing Image_Caption Image_Classification Image_Compression Image_Enhancement Image_Generation Image_Matting Image_Retrieval Inference Inpainting Intelligent_Chip Knowledge Knowledge_Graph Language_Model LLM Matching Medical Memory_Networks Multi_Modal Multi_Task NAS NMT Object_Detection Object_Tracking OCR Ontology Optical_Character Optical_Flow Optimization Person_Re-identification Point_Cloud Portrait_Generation Pose Pose_Estimation Prediction QA Quantitative Quantitative_Finance Quantization Re-identification Recognition Recommendation Reconstruction Regularization Reinforcement_Learning Relation Relation_Extraction Represenation Represenation_Learning Restoration Review RNN Robot Salient Scene_Classification Scene_Generation Scene_Parsing Scene_Text Segmentation Self-Supervised Semantic_Instance_Segmentation Semantic_Segmentation Semi_Global Semi_Supervised Sence_graph Sentiment Sentiment_Classification Sketch SLAM Sparse Speech Speech_Recognition Style_Transfer Summarization Super_Resolution Surveillance Survey Text_Classification Text_Generation Time_Series Tracking Transfer_Learning Transformer Unsupervised Video_Caption Video_Classification Video_Indexing Video_Prediction Video_Retrieval Visual_Relation VQA Weakly_Supervised Zero-Shot