Skip to content

Author

Matthias Grossglauser

1 paper indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#machine learning Preprint Sep 2026

Uncertainty-Normalized Margins for Direct Preference Optimization

Direct preference optimization (DPO) models binary preferences through a Bradley-Terry model with a common noise scale, without explicitly accounting for preference strength or prompt-dependent uncertainty from human feedback. We introduce uncertainty-normalized margin DPO (UNM-DPO), which combines strength-dependent m...

Sadegh Khorasani, P. Mikkola, Matthias Grossglauser · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.