Skip to content

Weight-Space Mixture-of-Experts for Implicit Neural Representation Classification

Jul 2026 · arXiv.org · Vol abs/2607.29463 · 0 citations · 46 references
Computer Science

TL;DR

This work proposes a hierarchical Mixture-of-Experts (HMoE) Transformer that processes INR weights using conditional computation aligned with the structure of the underlying implicit network and develops weight-space attribution and pruning methods that identify parameters most relevant for classification.

Abstract

Implicit Neural Representations (INRs) encode signals as the weights of a coordinate-based neural network and have recently been proposed as an alternative domain for downstream learning. While promising, classification directly in weight space remains challenging due to the high dimensionality and complex structure of INR parameters. Furthermore, the way discriminative information is distributed across INR weights remains poorly understood. We propose a hierarchical Mixture-of-Experts (HMoE) Transformer that processes INR weights using conditional computation aligned with the structure of the underlying implicit network. Coupled with a meta-learning framework that shapes INR parameters for downstream tasks, our model achieves state-of-the-art accuracy across standard benchmarks, ranging from low-resolution datasets to high-resolution ImageNet-1K. To gain insight into how INRs encode discriminative information, we develop weight-space attribution and pruning methods that identify parameters most relevant for classification. These analyses reveal how class-specific structure emerges within INR layers and support the suitability of MoE architectures for weight-space learning. Our approach advances both the performance and interpretability of weight-space classifiers.

View source

Similar papers

Preprint Aug 2026

A Unified Backbone--Expert Framework with Relation-Token and Residual--Classifier Interfaces for Automatic Modulation Recognition

A unified backbone-expert framework with a common convolutional state-space backbone and two specialized interfaces for automatic modulation recognition, confirming the benefit of expert-interface decoupling over one-size-fits-all architectures.

Zhixiang Deng, Houbiao Li, Z. Cui · 0 citations
Open access Aug 2026

Embedded One-Class Classification for Deep Neural Network Representations

An Embedded One-Class Classification (EOCC) framework for monitoring task-informed neural network representations andComparisons with depth-based, density-based, covariance-based, isolation-based, and end-to-end deep one-class methods show that EOCC is competitive and frequently achieves low Type II error while maintai...

Edgard M. Maboudou-Tchao, P. S. Senaratne, Randyll Pandohie et al. · 0 citations
Preprint Aug 2026

Sparse Prototype Code Underlies Classification and Prediction Across Modalities

Neural representations have become a central tool for studying the internal mechanisms of modern AI models, yet their complex high-dimensional structure makes them difficult to interpret. We show that classification tasks give rise to a universal representational geometry, shared across state-of-the-art models in visio...

Yehonatan Avidan, Daniel D. Lee, H. Sompolinsky · 0 citations
Sep 2026

Adaptive Multi-Stage Feature-View Fusion via Deep Graph Representation for Clustering.

Deep neural networks primarily aim to enhance the representation ability of high-level semantic features by deepening the network to reinforce critical features. However, this often leads to the under-utilization or discarding of certain low-level information, resulting in suboptimal performance. Due to differences in...

Rui Zhang, Yue-Long Cheng, Jing-Fan Yang et al. · 0 citations
Preprint Aug 2026

MLLM-Routed Heterogeneous Ensembles for Robust Cross-Dataset Image Classification

ARMDIL is an ensemble that uses a multimodal large language model (MLLM) agent to dynamically route each image to the most suitable vision backbone, drastically improves adaptability by allowing new information to be integrated via simple prompt modifications, while enhancing interpretability through natural language r...

Daniel A. Perkins, J. Squires, Janou Milligan et al. · 0 citations
Open access 2026

HIFN-Transformer: Learnable Information-Theoretic Parameters for Interpretable Deep Classification

HIFN-T is presented, a framework extending the Variational Information Bottleneck through four jointly learnable per-layer parameters: information retention, entropy budget, magnitude scaling, and global information gates that generalizes standard VIB as a special case and characterize the role of the entropy budget as...

Mohammed Tawfik · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.