arXiv · 2505.07511
MAIS: Memory-Attention for Interactive Segmentation
Abstract
Interactive medical segmentation reduces annotation effort by refining predictions through user feedback. Vision Transformer (ViT)-based models, such as the Segment Anything Model (SAM), achieve state-of-the-art performance using user clicks and prior masks as prompts. However, existing methods treat interactions as independent events, leading to redundant corrections and limited refinement gains. We address this by introducing MAIS, a Memory-Attention mechanism for Interactive Segmentation that stores past user inputs and segmentation states, enabling temporal context integration. Our approach enhances ViT-based segmentation across diverse imaging modalities, achieving more efficient and accurate refinements.
Explore related subjects
Keep this discovery
Mauricio Orbes-Arteaga, Oeslle Lucena, Sabastien Ourselin, M. Jorge Cardoso. 2025-05-12. MAIS: Memory-Attention for Interactive Segmentation. https://arxiv.org/abs/2505.07511
Cite the original work for its findings. Save a collection to share your selection of sources.