Skip to content
Preprint

IMVS: Interactive Medical Volume Segmentation with Test-Time Adaptation - A New Method for Annotating Radiology Datasets

Sep 2026 · 0 citations · 25 references
Computer Science

TL;DR

IMVS is presented, a human-in-the-loop annotation framework that composes three components into a closed loop rather than a new segmentation primitive: a lightweight 2D Slice Mask Adapter fine-tuned online from user scribbles, a frozen Volume Mask Tracker (VMT) that propagates corrected masks across adjacent slices, and a soft teacher--student alignment that limits forgetting.

Abstract

Annotating large radiology datasets is bottlenecked by the manual effort of delineating structures slice-by-slice in 3D volumes. Interactive methods reduce this effort but stay interaction-inefficient: slice-wise methods (including many foundation models) ignore inter-slice continuity, while 3D and video-based methods propagate a prompt with a \emph{fixed} propagator that never adapts to the target volume, so it drifts on low-contrast or pathological structures and must be re-prompted. We present IMVS, a human-in-the-loop annotation framework that composes three components into a closed loop rather than a new segmentation primitive: a lightweight 2D Slice Mask Adapter (SMA) fine-tuned online from user scribbles, a frozen Volume Mask Tracker (VMT) that propagates corrected masks across adjacent slices, and a soft teacher--student alignment that limits forgetting. The SMA is backbone-agnostic (UNet++, DeepLabV3, TransUNet). Across 8 public CT/MRI datasets, IMVS matches strong interactive baselines in quality while sharply cutting annotation effort: $14.4\times$ faster than a proficient copy-based manual workflow ($22.3\times$ over naive manual), $4.6\times$ over slice-wise and $1.9\times$ over 3D interactive methods. MedSAM2 and ScribblePrompt stay competitive or stronger on well-delineated organs; IMVS's advantage is largest on challenging targets and on interaction efficiency. Source code and Demo Video: https://github.com/AbhilakshSinghReen/imvs.

View source

Similar papers

Preprint Sep 2026

Confidence-Aware Teacher-Student Distillation for 3D Medical Segmentation

Medical image segmentation models typically rely on large amounts of densely annotated volumetric data, limiting their scalability across tasks and imaging modalities. This work addresses the challenge of predicting entire 3D anatomical structures from extreme annotation sparsity. An annotation-efficient student-teache...

Georgios Triantafyllou, D. Iakovidis · 0 citations
Preprint Sep 2026

SAMI3D-DW: Interactive Segmentation of Any 3D Medical Images

Interactive segmentation of 3D medical images supports quantitative analysis of anatomical structures and disease while allowing users to specify and refine their targets. Despite substantial progress by nnInteractive and VISTA3D, reliable segmentation across diverse clinical targets remains challenging, particularly f...

Ping-Qiu Gong, Shi-Yuan Su, Fandong Zhang et al. · 0 citations
Preprint Sep 2026

From Few-Shot Segmentation to Clinician-in-the-Loop Medical Image Analysis

This work reframe FSMIS as a three-layer sequential decision problem, state six hypotheses with an explicit dependency order, and proposes a minimal pilot that can falsify the foundational self-assessment claim before a clinician study.

Yazhou Zhu · 0 citations
Conference Open access Sep 2026

Slice-aware MedSAM Adaptation with cross-slice consistency for 2.5D multi-organ CT segmentation

Abdominal CT multi-organ segmentation is important for computer-aided diagnosis and quantitative analysis, but it remains challenging due to large variations in organ size, shape, and boundary clarity. Existing 2D methods lack inter-slice contextual modeling, while 3D methods require high computational and memory costs...

Qing-Yang Guan, Ke-Wen Qin, Wei-Hai Pang et al. · 0 citations
Aug 2026

Dynamic iterative coarse-to-fine prompt learning based on SAM for precise esophageal cancer gross target volume segmentation

Accurate segmentation of the gross tumor volume from computed tomography images is a core step in the development of precise radiotherapy planning for esophageal cancer, which directly affects treatment efficacy and normal tissue protection. Recently, foundation models represented by the segment anything model (SAM) ha...

Yuxuan Yao, Hong-Fei Sun, Cheng-Wei Chen et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.