Skip to content

Conditioning Tree-Based Diffusions and Flows for Probabilistic Tabular Regression

Jul 2026 · arXiv.org · Vol abs/2607.28864 · 0 citations · 23 references
Mathematics Computer Science

TL;DR

DiffGBM exposes the score-side recipe---residualization, EDM-style preconditioning, log-sigma time sampling, noise-level features, loss weighting, and histogram resolution---as jointly tunable axes over a shared LightGBM surface rather than one frozen bundle.

Abstract

Tree-based diffusion models fit flexible conditional predictive distributions for tabular regression without a neural density estimator, but they inherit their design defaults---noising path, parameterization, training distribution, features, sampler---from the neural setting. We show these defaults are the binding constraint: what a gradient-boosted ensemble actually solves is a supervised regression problem whose conditioning they determine. We present DiffGBM, which makes them explicit along two axes. First, a Gaussian-path flow-matching trainer for $p(y \mid x)$ that learns a velocity field directly and recovers the score algebraically, admitting few-step deterministic ODE sampling. Second, we expose the score-side recipe---residualization, EDM-style preconditioning, log-sigma time sampling, noise-level features, loss weighting, and histogram resolution---as jointly tunable axes over a shared LightGBM surface rather than one frozen bundle. This \emph{score-flex} space represents the published recipe as a special case; across eleven tabular benchmarks under fold-0 tuning, folds-1--5 evaluation, and a matched 40-trial budget and sampler, the selected configurations beat that baseline on \emph{every} dataset (paired Wilcoxon $11/0$, $p<10^{-3}$), with the best aggregate CRPS skill (0.725 vs.\ 0.699) of any row. The two rows are complementary: score-flex buys accuracy with a stochastic sampler and is the slowest row, while flow matching is the cheapest sampler ($5.2\times$ faster than the published baseline) and the best-calibrated DiffGBM row. Tuned non-diffusion baselines still win individual datasets, and stochastic ($\varepsilon>0$) flow samplers do not Pareto-dominate the deterministic corner.

View source

Similar papers

Preprint Aug 2026

A Plug-in Interpretation of Conditioning in Score-Based Diffusion Models

A conditioning mechanism for diffusion models based on multi-speed joint diffusion of the target and the condition learns an unconditional joint score network and enforces conditioning at inference via a plug-in correction term, and derives explicit conditional reverse-time SDEs and approximate probability-flow ODEs.

Libo Chen, Souvik Ghosh, Teo Deveney et al. · 0 citations
Jul 2026

Learning Sampling Parameters for Diffusion Models

This work introduces LeSAMP, a framework for learning prompt-conditioned, timestep-varying sampling parameters, and suggests that learned sampling-parameter policies provide a complementary approach to existing post-training methods for improving diffusion model outputs.

Arisrei Lim, Yossi Gandelsman · 0 citations
Jul 2026

Inference-Time Scaling of Diffusion Models via Progressive Seed Pruning

Progressive Seed Pruning consistently improves reward-guided selection and achieves higher GenEval scores (automated) and better human evaluation on prompt-alignment than best-of-$N$, importance-sampling, and tree-search baselines at matched compute.

Rogério Guimarães, Pietro Perona · 0 citations
Preprint Aug 2026

Designing Reinforcement Learning for Diffusion Models: A Unified Path-Space View

The empirical gap between these method families is identified as a variance-reduction effect rather than a difference in RL principle, and a multi-sample KDE value-gradient estimator that reuses rollout groups, together with scale-bounded weight families that retain stable existing recipes while excluding singular ones...

Yi-Xian Xu, Yuanrui Zhang, Shengjie Luo et al. · 2 citations
#machine learning Preprint Aug 2026

Scaling Reinforcement Learning for Diffusion Models via Velocity Matching

This work proposes reward-based velocity matching (RVM), a simple trajectory-free update that acts directly on the velocity field and provides a general framework that recovers recent fine-tuning methods, including RAM and DiffusionNFT, as special cases.

Jaemoo Choi, Wei Guo, Yuchen Zhu et al. · 1 citation

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.