Skip to content

FairDiffuseVQVAE: Sampling-Time Fairness in Tabular Diffusion via Conditional Refinement of Vector-Quantized Latents

Jul 2026 · arXiv.org · Vol abs/2607.28945 · 0 citations · 56 references
Computer Science

TL;DR

FairDiffuseVQVAE is introduced, a two-stage architecture that decouples fidelity from fairness that achieves the highest mean Demographic Parity Ratio and Equalized Odds Ratio and attains the lowest mean pair-wise correlation error of any published method.

Abstract

Synthetic tabular data is increasingly used in privacy-preserving data sharing, data augmentation, and to mitigate downstream classifier bias. State-of-the-art tabular diffusion models such as TabDDPM and TabSyn achieve excellent distributional fidelity but offer no mechanism for fairness; conversely, fairness-aware tabular generators (DECAF, FairTGAN, FairTabDDPM) impose explicit fairness penalties at training time, yielding modest fairness gains at substantial cost to either sample quality or downstream utility. We introduce FairDiffuseVQVAE, a two-stage architecture that decouples fidelity from fairness: a vector-quantized autoencoder with a row-level discriminator (Stage~1, no fairness terms) is followed by a DiffuseVAE-style continuous diffusion refiner that conditions on both the Stage-1 reconstruction and the protected attribute via classifier-free guidance (Stage~2). Fairness emerges as a property of the sampling distribution -- uniform sampling of the protected attribute at inference time enforces demographic parity by construction, rather than from competing loss terms. On the Adult, Bank and COMPAS datasets, FairDiffuseVQVAE achieves the highest mean Demographic Parity Ratio ($0.702$, $+47\%$ over FairTabDDPM) and Equalized Odds Ratio ($0.686$, $+100\%$). It also attains the lowest mean pair-wise correlation error ($0.034$) of any published method, while explicitly trading $\sim$$15$ AUC points for these fairness gains.

View source

Similar papers

#machine learning Preprint Sep 2026

Efficient Fairness Auditing Across Guidance Scales in Text-to-Image Diffusion Models via Causal Abstraction

Fairness auditing of text-to-image diffusion models often requires generating large numbers of images across sampling configurations, making comprehensive evaluation computationally expensive. We propose a causal-abstraction-based audit instrument for efficiently evaluating fairness under interventions on the classifie...

Nabila Tasfiha Rahman, Rajatsubhra Chakraborty, De-Peng Xu et al. · 0 citations
Preprint Aug 2026

Fairness-Aware Mixture-of-Experts via Subgroup Reweighting and Gate Regularization

This work identifies routing-induced bias, a failure mode in which subgroup imbalance drives the gating network to route subgroups onto a few experts, and proposes an end-to-end Mixture-of-Experts (MoE) framework that corrects it and improves fairness while maintaining competitive predictive performance.

Sunhee Hwang · 0 citations

Fair Graph Learning Needs Expressiveness: Rethinking Fairness from the Spectral Perspective

It is formally proved that GNN architectures lacking spectral expressiveness impose strict constraints on the representation space, so that harmful linear correlations between sensitive attributes and target prediction logits are preserved whenever a low-expressive backbone is paired with a debiasing operator acting wi...

Ming-Qi Yang, Zhao-Yu Liu · 0 citations
Preprint Aug 2026

Fairness-Aware Test-Time Prompt Tuning

FairTPT is developed, a novel fairness-aware episodic TTA method that jointly minimizes target marginal entropy while maximizing spurious marginal entropy through soft-prompt tuning and establishes a foundation for robust TTA, which is essential for achieving fairness in practice.

Yoann L. Launay, Parameswaran Kamalaruban, Tom Kempton et al. · 0 citations
Preprint Aug 2026

FairReL: Deepfake Detection using Fairness-Aware Representation Learning

Although recent deepfake detectors achieve high overall accuracy, their errors remain unevenly distributed across demographic subgroups, with real faces from certain groups more often misclassified as fake. Existing fairness-aware detectors typically regularise the entire feature representation, without identifying or...

Xiaoman Lu, Jiaqi Li, Shuntian Zheng et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.