Skip to content

TabPack: Efficient Hyperparameter Ensembles for Tabular Deep Learning

Jul 2026 · arXiv.org · Vol abs/2607.05380 · 0 citations · 36 references
Computer Science

TL;DR

This work introduces TabPack, an efficient MLP ensemble with strong out-of-the-box performance and reduced reliance on traditional tuning, and specifies ranges from which to sample MLP hyperparameter rather than exact hyperparameter values, which naturally demands less precision for good performance.

Abstract

In deep learning for tabular data, efficient ensembles of multilayer perceptrons (MLPs) have recently emerged as effective and practical architectures. Existing methods of this kind use the same hyperparameters for all underlying MLPs, which requires hyperparameter tuning for achieving the best performance. In this work, we introduce TabPack, an efficient MLP ensemble with strong out-of-the-box performance and reduced reliance on traditional tuning. In a single run, TabPack samples and trains many MLPs with different hyperparameters efficiently in parallel and selects ensemble members on the fly during training. Thus, TabPack only requires specifying ranges from which to sample MLP hyperparameter rather than exact hyperparameter values, which naturally demands less precision for good performance. In experiments on medium-to-large public datasets, TabPack with default settings performs on par with extensively tuned prior methods, thus substantially reducing effort and compute resources needed to achieve competitive results on tabular tasks. Notably, running the default TabPack configuration on a modern MacBook took less time than tuning some baselines on an industry-grade GPU.

View source

Similar papers

Preprint Aug 2026

TabDPT-Turbo: Efficient In-Context Learning for Tabular Prediction

This work adopts an alternate approach, sticking with row-based attention while incorporating long context pre-training to eliminate the need for retrieval in TabDPT-Turbo, a model that provides comparable default performance to TabDPT v1.1 on TabArena-Lite, CC18, and CTR23, at orders of magnitude faster.

Rasa Hosseinzadeh, Alex Labach, Zexin Xue et al. · 4 citations
#machine learning Preprint Sep 2026

Solving In-Table Prediction Problems by Deep Neural Networks with Performance Evaluation Using Synthetic Data

This work investigates whether NNs can predict the values of arbitrarily selected columns in a given table based on the remaining known columns, and concludes that the attention-based structure outperforms the other two networks, when a sufficiently large number of training examples is available and a relatively large...

Xiao Zhao, Daniela Oelke · 0 citations
Preprint Aug 2026

Localized TabICLv2: Scaling Tabular In-Context Learning through k-NN

Localized TabICLv2 introduces a method that reduces the inference cost of TabICLv2 by retrieving only the k nearest training neighbours for each test point, measured by similarity in the model's Stage 2 row-representation space, rather than using the full training context.

Beimnet Bekele Guta · 0 citations
#machine learning Preprint Sep 2026

Agentic Search Spaces for Tabular Machine Learning

This study suggests that LLM agents can provide practical value for tabular ML by expanding the design space, and suggests that the two strongest agentic ensembles surpass the best AutoGluon ensemble of conventional models.

Renat Sergazinov, Artem Chistyakov, Sergey Pankevich et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.