Skip to content
Conference

ProReGen: Progressive Residual Generation under Attribute Correlations

2026 · International Conference on Learning Representations · Vol 2026, pp. 6667 - 6695 · 0 citations
Medicine

TL;DR

ProReGen is presented, a progressive residual generation approach inspired by the classical Robinson’s transformation, to partial out from an image attribute x2 its component mx1 that is predictable by other image attributes x1, and the residual γ=x2-mx1 that is not.

Abstract

Attribute correlations in the training data will compromise the ability of a deep generative model (DGM) to synthesize images with under-represented attribute combinations (i.e., minority samples). Existing approaches mitigate this by data re-sampling to remove attribute correlations seen by the DGM, using a classifier to provide pseudo-supervision on generated counterfactual samples, or incorporating inductive bias to explicitly decompose the generation into independent submechanisms. We present ProReGen, a progressive residual generation approach inspired by the classical Robinson’s transformation, to partial out from an image attribute x2 its component mx1 that is predictable by other image attributes x1, and the residual γ=x2-mx1 that is not. This simplifies the problem of learning a DGM gx1,x2 conditioned on correlated inputs, to learning g~x1,γ conditioned on orthogonal inputs. It further allows us to progressively learn g~ by first shifting the burden to abundant majority samples to learn g~x1,γ=0, and then expanding it with additional layers gres to resolve its difference to g~x1,γ using residual attribute γ on limited minority samples. On three benchmark datasets with varying strengths of attribute correlation and one dataset with natural attribute correlation, we demonstrate that ProReGen—with input orthogonalization and progressive residual learning—improved the correctness of minority generations compared to existing strategies.

View source

Similar papers

#machine learning Preprint Sep 2026

Data Unlearning via Inverse Distillation

Multi-step matching models, including flow and diffusion models, produce high-quality outputs but incur substantial inference costs and may reproduce unwanted components of their training datasets. We introduce Inverse Distillation Unlearning (IDU), a unified framework that simultaneously distills a teacher multi-step...

Aleksei Leonov, N. Kornilov, Zhen-He Zhang et al. · 0 citations

with Pretrained Generative Model for Self-Supervised Learning

GenView is presented, a controllable framework that augments the diversity of positive views leveraging the power of pretrained generative models while preserving semantics, and an adaptive view generation method that dynamically adjusts the noise level in sampling to ensure the preservation of essential semantic meani...

Xiao-Jie Li, Yibo Yang, Xiang-Tai Li et al. · 0 citations
Preprint Sep 2026

PAPT++: Risk-Aware Adversarial Tuning and Generation for Single Domain Generalization

Single domain generalization (SDG) aims to learn a model from one labeled source domain that generalizes to unseen target domains. A common strategy is to enrich the source distribution with augmented or generated samples, and recent text-to-image (T2I) diffusion models provide a strong generative prior for this purpos...

Zhi-Peng Xu, De Cheng, Xinyang Jiang et al. · 0 citations
Open access 2026

MSTabVAE: Multi-Step Latent Conditional Variational Autoencoder for Imbalanced Tabular Data Synthesis

Recent advances in artificial intelligence have expanded its applications in the financial domain, particularly in fraud detection, a critical task for preventing losses for both customers and institutions. However, fraud detection is challenging due to severe class imbalance, which significantly degrades detection per...

Min-Ji Kang, Hyeryung Jang · 0 citations
#machine learning Preprint Sep 2026

Comparing Latent Concept Formation in State Space Models and Transformers via Sparse Autoencoders

The quadratic scaling of Transformer self-attention has driven the adoption of sub-quadratic Selective State Space Models (SSMs) like Mamba, which compress past context into a fixed-size recurrent hidden state. This strict informational bottleneck raises a foundational question for mechanistic interpretability: do SSMs...

Rithin Nagaraj, Rupa Laalasa Oruganti, Prerna Subhashchandra Kunder et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.