Skip to content

Distributional Split Criteria for Random Forests: Extensions, Shrinkage, and the Robustness of Mean Splitting

Jul 2026 · arXiv.org · Vol abs/2607.23721 · 0 citations · 4 references
Computer Science Mathematics

TL;DR

A family of distributional criteria inside a single honest-forest implementation of isotropic random-Fourier-feature maximum mean discrepancy, an anisotropic diagonal-bandwidth variant, an adaptive per-split frequency-selection variant, and a non-kernel sliced-Wasserstein criterion are implemented.

Abstract

Distributional random forests replace mean-based CART splitting with criteria that compare the full conditional response distribution in candidate children. We implement and systematically study a family of such criteria inside a single honest-forest implementation: isotropic random-Fourier-feature maximum mean discrepancy (MMD), an anisotropic diagonal-bandwidth variant, an adaptive per-split frequency-selection variant, and a non-kernel sliced-Wasserstein criterion, together with post-hoc kernel-mean shrinkage of the forest weights. Using paired-seed comparisons across synthetic quantile mechanisms, real univariate benchmarks, a California-housing subsample curve, and multivariate synthetic and real responses, we characterize where each extension pays. Three findings recur. First, among distributional criteria ordinary isotropic MMD is already close to best in class: the anisotropic, adaptive-frequency, and sliced-Wasserstein extensions, and post-hoc shrinkage, do not systematically improve on it. Second, on scalar tabular regression mean-based CART splitting remains the robust default and wins many cells. Third, multivariate responses are the regime where distributional splitting clearly earns its keep, most sharply on a pure-dependence copula where the energy score separates the criteria even though marginal CRPS does not. The evidence supports a simple allocation story: distributional splitting helps only when non-location structure is both present and estimable; otherwise it dilutes split-selection power away from the mean. All criteria, the honest forest, and the paired-comparison harness are implemented in the open-source \texttt{drforest} library, whose Rust-backed split search makes broad criterion sweeps inexpensive.

View source

Similar papers

Preprint Sep 2026

Median-based Splitting Rules for Causal Trees and Forests

Heavy-tailed and skewed outcomes are common in the randomized experiments and observational studies used to estimate heterogeneous treatment effects, yet the mean-squared-error criterion that guides splitting in honest causal trees is sensitive to the extreme values they generate. Building on the causal forest framewor...

Lennard Maßmann, Karolina Gliszczyńska-Schroeder · 0 citations
#machine learning Preprint Sep 2026

Improving the Predictive Performance of Bootstrap Aggregating by Dirichlet Resampling

We revisit Breiman's observation that reducing inter-tree correlation without weakening individual trees can improve random forests. Building on this principle, we introduce two variants: Dirichlet-Multinomial Bagging Random Forest (DM) and Dirichlet-Weighted Random Forest (DW). Both modulate sample reweighting via a c...

Quoc V. Le, Joon Suk Park · 0 citations
Preprint Sep 2026

Sliced $L^p$ Distributional Balancing

A popular class of causal inference methods addresses confounding through weighting, which reweights treated and control groups to balance their covariate distributions without using outcome information, thereby preserving a design-based perspective. In this paper, we propose sliced $L^p$ distributional balancing (SLDB...

Hao-Ran Zhang, Guanhua Chen, Chan Park · 0 citations
Preprint Sep 2026

Selective Inference for CART with Binary Outcomes

Finite-sample conditional tests of a common success probability within a parent selected by deterministic Gini CART, which separate information loss from computational limitations and exhibit a selected fiber disconnected under single-label swaps.

Tomoshige Nakamura · 0 citations
Open access Sep 2026

The revival of bagged trees: hierarchical shrinkage as a regularization tool

We present methodological refinements and new insights into hierarchical shrinkage (HS), a post-hoc regularization technique for decision trees that shrinks node predictions toward ancestral means. By contrasting HS with the implicit regularization from feature sub-sampling in random forests (RFs), we clarify when bagg...

Markus Loecher, A. Gevaert, B. Pfeifer et al. · 0 citations
Preprint Aug 2026

Handling Missing Data in Probabilistic Regression Trees

Probabilistic Regression Trees (PRTrees) are a smooth and consistent alternative to classical regression trees, producing continuous predictions through probabilistic split assignments. This paper extends the PRTree framework to accommodate missing predictor values directly during tree construction, eliminating the nee...

T. S. Prass, A. Neimaier, G. Pumi · 1 citation

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.