Distributional LLM-as-a-Judge
This work proposes a novel training framework that explicitly aligns the LLM-generated judgment distribution with human evaluation distributions, and incorporates adversarial training to ensure a robust alignment with this true distribution, rather than overfitting to its imperfect approximation.