#natural language process...
Feb 2026
Small Reward Models via Backward Inference
FLIP (FLipped Inference for Prompt reconstruction), a reference-free and rubric-free reward modeling approach that reformulates reward modeling through backward inference that enables reliable reward modeling in downscaled regimes where judgment methods fail, is proposed.
Yike Wang, Faeze Brahman, Shangbin Feng et al.
· arXiv.org · 3 citations