Review
Aug 2026
ConRub-Med: Reinforcement Learning with Consensus Rubrics for Open-Ended Medical Question Answering
ConRub-Med is introduced to preserve useful distinctions as rubric feedback moves from construction to policy optimization, and ranks first on six of nine benchmarks and achieves the highest medical and generalization averages.
Taojie Zhu, Yuan Xia, Tao Sun et al.
· 0 citations