Skip to content
Preprint

Lethe: How Hard Is It to Forget? A Benchmark for Federated Unlearning in Medical Imaging

Aug 2026 · 0 citations · 50 references
Computer Science

TL;DR

Lethe is presented, a benchmark for federated unlearning in medical imaging, which evaluates twelve methods across eight task families, from classification and segmentation to denoising, cross-modality synthesis, and vision-language question answering, at three forgetting granularities and against a retrained gold standard on utility, privacy, and cost.

Abstract

Federated learning enables medical-imaging models to be trained across hospitals, and privacy law, most explicitly the GDPR ``right to be forgotten'', turns removing a hospital's, a class's, or a patient's influence from such a model into a federated unlearning problem. This need is most acute in medicine, where patients withdraw consent and hospitals leave collaborations. Yet nearly all unlearning evidence comes from natural images, whose heterogeneity and task structure differ sharply from clinical data, so it is unclear whether existing methods transfer, and no shared protocol covers clinical data. We present Lethe, a benchmark for federated unlearning in medical imaging. It evaluates twelve methods across eight task families, from classification and segmentation to denoising, cross-modality synthesis, and vision-language question answering, at three forgetting granularities and against a retrained gold standard on utility, privacy, and cost. The central result is that what separates methods is the difficulty of the forgetting request, not the method itself. The easy removals that dominate the literature leave the methods that preserve utility indistinguishable, while only hard ones separate them. More striking, on the many medical tasks that generalize across sites, forgetting a client barely changes task performance, leaving residual membership as the signal that must be erased.

View source

Similar papers

#artificial intelligence Preprint Sep 2026

OmniMed-FL: A Robust Multimodal Federated Learning Framework for Clinical Diagnosis

Simultaneous assessment of medical imaging and patient records is often required in clinical diagnosis. However, standard machine learning algorithms cannot analyze these data types together. Meanwhile, compliance with HIPAA and GDPR can constrain centralized aggregation of sensitive patient data. This leaves a crucial...

Ayush Debnath, Ruelia Saha, Sudip Misra · 0 citations
Open access Aug 2026

Regulatory-orientedDeep federated learning framework for multi-hospital medical imaging: privacy-preserving, explainable, and generalizable diagnosis

The proposed FL framework provides a privacy-preserving, explainable, and computationally efficient solution for collaborative AI in medical imaging by combining adaptive federated learning, secure privacy mechanisms, and explainable AI techniques, demonstrating strong potential for deployment in multi-hospital clinica...

Chandra Shakher Tyagi, Partheeban Nagappan, T. R · 0 citations
#artificial intelligence Preprint Sep 2026

Federated LoRA Adaptation of BiomedCLIP Across Four International Chest X-Ray Cohorts

Federated LoRA adaptation improves shared-class AUC on all four cohorts over the unadapted BiomedCLIP backbone over the unadapted BiomedCLIP backbone, showing that the gains come from federated adaptation rather than from the pretrained model's zero-shot ability.

Sanjaya Poudel, Nirajan Kunwor, Manish Dhakal et al. · 0 citations
Conference Aug 2026

Aggregation Algorithm Selection for Non-IID Federated Learning in Healthcare: Analysis of Convergence, Privacy, and Communication Tradeoffs

With Federated Learning (FL), hospitals are able to train models together without moving patient data, which is important due to security laws such as HIPAA and GDPR that make most deployments legally prohibit training models with patient data in a central location. The larger issue is statistical. Hospitals likely hav...

Zubair Raja Ahamed Zaffarullah, Venkata M. Gudala, J. M. Shanthi et al. · 0 citations
Sep 2026

MEDAL: Sequential adapter learning for privacy-preserving multicenter clinical language models.

MedAL is a scalable method for fine-tuning LLMs across many health systems without sharing patient-level data, enabling high-performance local models for reasoning over clinical notes and may also be useful for training multimodal healthcare AI models.

Ahmed Bakr, A. Garcia-Agundez, Travis Atkison et al. · 0 citations
#machine learning Review Aug 2026

CoMedBench: A Multi-Source Benchmark of Synthetic Medical Data Fidelity and Downstream Utility

CoMedBench is introduced, a reproducible benchmark that evaluates a family of generators under a common clinical-validity framework and one shared training and evaluation engine, spanning static tabular and temporal downstream tasks on established critical-care datasets.

Akanta Das, Farhad Al-Amin Dipto, M. S. Anto et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.