Preprint
Aug 2026
Rubric Dropout: A Simple Way to Mitigate Reward Hacking in Rubric-as-Reward RL
This work proposes Rubric Dropout, a one-line fix borrowed from neuron dropout that randomly drops a subset of the rubric's criteria before computing the reward, so the policy never optimizes the same rubric twice.
Minglai Yang, Xinyu Guo, Utkarsh Tyagi et al.
· 0 citations