This work proposes a hierarchical Mixture-of-Experts (HMoE) Transformer that processes INR weights using conditional computation aligned with the structure of the underlying implicit network and develops weight-space attribution and pruning methods that identify parameters most relevant for classification.
Abstract
Implicit Neural Representations (INRs) encode signals as the weights of a coordinate-based neural network and have recently been proposed as an alternative domain for downstream learning. While promising, classification directly in weight space remains challenging due to the high dimensionality and complex structure of INR parameters. Furthermore, the way discriminative information is distributed across INR weights remains poorly understood. We propose a hierarchical Mixture-of-Experts (HMoE) Transformer that processes INR weights using conditional computation aligned with the structure of the underlying implicit network. Coupled with a meta-learning framework that shapes INR parameters for downstream tasks, our model achieves state-of-the-art accuracy across standard benchmarks, ranging from low-resolution datasets to high-resolution ImageNet-1K. To gain insight into how INRs encode discriminative information, we develop weight-space attribution and pruning methods that identify parameters most relevant for classification. These analyses reveal how class-specific structure emerges within INR layers and support the suitability of MoE architectures for weight-space learning. Our approach advances both the performance and interpretability of weight-space classifiers.
A unified backbone-expert framework with a common convolutional state-space backbone and two specialized interfaces for automatic modulation recognition, confirming the benefit of expert-interface decoupling over one-size-fits-all architectures.
An Embedded One-Class Classification (EOCC) framework for monitoring task-informed neural network representations andComparisons with depth-based, density-based, covariance-based, isolation-based, and end-to-end deep one-class methods show that EOCC is competitive and frequently achieves low Type II error while maintai...
Edgard M. Maboudou-Tchao, P. S. Senaratne, Randyll Pandohie et al.· Mathematics· 0 citations
Neural representations have become a central tool for studying the internal mechanisms of modern AI models, yet their complex high-dimensional structure makes them difficult to interpret. We show that classification tasks give rise to a universal representational geometry, shared across state-of-the-art models in visio...
Yehonatan Avidan, Daniel D. Lee, H. Sompolinsky· 0 citations
Deep neural networks primarily aim to enhance the representation ability of high-level semantic features by deepening the network to reinforce critical features. However, this often leads to the under-utilization or discarding of certain low-level information, resulting in suboptimal performance. Due to differences in...
Rui Zhang, Yue-Long Cheng, Jing-Fan Yang et al.· IEEE Transactions on Image P...· 0 citations
ARMDIL is an ensemble that uses a multimodal large language model (MLLM) agent to dynamically route each image to the most suitable vision backbone, drastically improves adaptability by allowing new information to be integrated via simple prompt modifications, while enhancing interpretability through natural language r...
Daniel A. Perkins, J. Squires, Janou Milligan et al.· 0 citations
HIFN-T is presented, a framework extending the Variational Information Bottleneck through four jointly learnable per-layer parameters: information retention, entropy budget, magnitude scaling, and global information gates that generalizes standard VIB as a special case and characterize the role of the entropy budget as...
Mohammed Tawfik· IEEE Access· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.