Skip to content

TabSOM: A tabular-to-image encoding method based on self-organizing maps

Aug 2026 · 1 citation · 27 references
Computer Science

TL;DR

TabSOM is proposed, a tabular-to-image encoding built on the Self-Organizing Map, which provides a spatial layout in which every input feature occupies a fixed canvas position derived from its component plane via collision-free Hungarian assignment and a graph that captures pairwise feature relationships derived from the SOM component planes.

Abstract

Tabular-to-image methods have emerged as novel approaches to leverage the high predictive performance of convolutional neural networks and vision transformers. They convert tabular data into image representations, mapping each feature at a fixed pixel location derived from a dimensionality-reduction method (e.g., t-SNE, UMAP, PCA). However, they encode only the marginal value of each feature and discard information about feature relationships. We propose TabSOM, a tabular-to-image encoding built on the Self-Organizing Map (SOM), which provides: (i) a spatial layout in which every input feature occupies a fixed canvas position derived from its component plane via collision-free Hungarian assignment; and (ii) a graph that captures pairwise feature relationships derived from the SOM component planes. The resulting image stacks two multi-scale node channels: one encodes feature values at fixed scales, while the other encodes pairwise feature interactions as spatial connections between related features. Two SOM-derived interpretability approaches are introduced: a prototype-inspired partial dependence plot and a class--separation importance score. Benchmarked against twelve existing tabular-to-image methods across public binary-classification datasets, TabSOM ranks first or second on every dataset and achieves the lowest variance of any method evaluated. Interpretability obtained with TabSOM was validated against Random Forest, XGBoost, and SHAP, the class-separation score shows reasonable agreement with established baselines on the top-ranked features while capturing complementary structural information from input data. These results demonstrate that TabSOM provides an effective and interpretable approach for applying deep learning architectures to tabular data, bridging the performance--interpretability gap in this domain.

View source

Similar papers

#machine learning Preprint Aug 2026

VG-TIE: An interpretable tabular-to-image encoding method based on visibility graphs

Experiments show that VG-TIE is competitive with other tabular-to-image methods while providing interpretability on feature importance and ranking similar to intrinsic interpretable methods, and highlights the potential of the proposed image-based transformation to provide an effective framework that expands the use of...

David Chushig-Muzo, L. López-Ramos, Ángeles Rodríguez de Cara et al. · 0 citations
Preprint Aug 2026

Test-Time Augmentation for Tabular-to-Image Classifiers under Distribution Shifts

Results indicate that TTA improves OOD performance, with composite and photometric strategies providing the best trade-off between robustness and variance, in contrast to frequency-domain transformations that alter the encoder's feature-to-intensity mapping consistently degrade performance.

Malena Loza, Felipe Grijalva, Eva Milara et al. · 0 citations
Sep 2026

Information-Theoretic Analysis of Positional Encoding Strategies in Vision Transformers: A Comparative Study of Four Approaches.

Vision Transformers (ViTs) rely on positional encoding (PE) because self-attention has no native notion of token order or image-grid location, yet the information-theoretic properties of different PE strategies and their downstream consequences for model behaviour remain insufficiently characterised. We present a syste...

D. Bandur, M. Bandur, B. Jakšić · 0 citations

: A Self-Attention

Sudharshan Banakar, K. Chandrashekhar, M.tech · 50 citations · ⚡1
Preprint Sep 2026

GraLoD: Graphics-Inspired Continuous Level-of-Detail Learning for Image Restoration

GraLoD, a plug-and-play framework that treats restoration scale as a spatially varying and stage-dependent continuous variable, is proposed and minimal-sufficient footprint calibration (MSFC) together with structure-aware regularization (SAR) is introduced to encourage restoration-effective and spatially coherent LOD a...

Hu Gao, Li-Zhuang Ma, Yulong Chen · 0 citations
Preprint Sep 2026

Structure-Guided Masked Autoencoders for Ultra-High Resolution Scientific Image Understanding

Self-supervised pre-training with Vision Transformers, including Masked Autoencoders (MAE), is difficult to apply to gigapixel scientific images. Random masking is poorly matched to the structured, multi-scale morphology of scientific data, while uniform tokenization produces prohibitively long sequences that make $O(N...

En-Zhi Zhang, Du Wu, Rui Zhong et al. · 0 citations

Related blog posts

Microsoft Research Blog Jul 13, 2026

Verifying Rust cryptography in SymCrypt, from standards to code

Cryptographic code supports vital protections in modern computing systems. Learn how a new method helps verify code as developers write it while preserving speed and adaptability as it gets implemented and evolves. The post Verifying Rust cryptography in SymCrypt, from standards to code appeared first on Microsoft Research.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.