Skip to content
Preprint

Hide&Seek: Learning to Explain in an End-to-End Differentiable Network

Aug 2026 · 0 citations · 44 references
Mathematics Computer Science

TL;DR

This paper presents Hide&Seek, an end-to-end differentiable model for instance-wise feature selection and prediction under a single objective without information leakage that outperforms existing state-of-the-art models across a range of experiments and is fast to train.

Abstract

Instance-wise feature selection is a valuable tool for interpreting labeled data and the predictions of black-box models. In contrast to global feature selection techniques, instance-wise methods dynamically identify important features for each instance. A growing number of methods learn a selector, which identifies important features, and a predictor, which uses these to make predictions. However, these pioneering methods face challenges including information leakage and lack of differentiability, which can slow training. In this paper, we present Hide&Seek, an end-to-end differentiable model for instance-wise feature selection. We jointly learn feature selection and prediction under a single objective without information leakage. Hide&Seek outperforms existing state-of-the-art models across a range of experiments and is fast to train. We achieve this by reformulating feature removal as a differentiable operation where instead of discretely removing features, we replace a proportion of each feature. Training is further stabilized via a parsimony-weight annealing framework.

View source

Similar papers

Conference Open access Sep 2026

Learning Local Feature Masks with Variational Information Bottleneck

Instance-wise feature selection (IWFS) identifies informative features for each instance, improving generalization by discarding irrelevant information and enhancing interpretability through personalized explanations. Most IWFS methods adopt a selector--predictor architecture, where a selector generates instance-specif...

Lu Sun, Jun Sakuma · 0 citations
Preprint Aug 2026

Black-Box Knowledge Transfer across Distinct Feature Sets

This work proposes a two-step neural network procedure, estimating the transferable component from abundant unlabeled feature pairs that bridge the two input spaces and the non-transferable component from limited labels, and derives prediction risk bounds that improve on those of a non-transfer alternative when the non...

Oh-Ran Kwon, Daeyoung Ham · 0 citations
#machine learning Preprint Sep 2026

Where Decoder Cosine Similarity Fails for SAE Feature Flow Discovery

Foundation models are increasingly adapted through fine-tuning, model editing, and alignment procedures while retaining previously acquired capabilities. Understanding the internal computations that support these adaptations is therefore becoming increasingly important for continual model evolution. Sparse autoencoders...

Hendrik Droste, C. M. Adriano, Kathrin Korte et al. · 0 citations
#artificial intelligence Preprint Sep 2026

A Study of Hidden-State Optimization Order in Predictive Coding Networks

A boundary-first inference schedule that partitions a model into chunks, first coordinates hidden states at chunk boundaries, and then refines representations within each chunk is proposed, which instantiate in predictive coding networks (PCNs), a local-learning framework in which hidden activities and prediction error...

Xue-Yuan Li, Danilo Vasconcellos Vargas · 0 citations
Conference Aug 2026

Feature Norm Normalization: A Plug-and-Play Module for Boosting Model Accuracy in Machine Unlearning

Driven by stringent privacy regulations and data deletion requirements, machine unlearning has emerged as a critical field focused on selectively removing the influence of specific data from the pre-trained model. In this paper, we focus on achieving the unlearning objective while maintaining better model accuracy. Fir...

Cheng-Hao Yang, Xue Yang, Xiaohu Tang · 0 citations
Open access Jul 2026

Comprehend, Divide, and Conquer: Feature Subspace Exploration via Multi-Agent Hierarchical Reinforcement Learning

Feature selection aims to preprocess the target dataset, find an optimal and most streamlined feature subset, and enhance the downstream machine learning task. Among filter, wrapper, and embedded-based approaches, the reinforcement learning (RL)-based subspace exploration strategy provides a novel objective optimizatio...

Weiliang Zhang, Xiaohan Huang, Ziyue Qiao et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.