Preprint
Aug 2026
RISE-RL: Rubric-Informed Selective Exploration for Open-Ended Reinforcement Learning
Results indicate that selective internalization through reward filtering and policy support shaping is effective for open-ended reinforcement learning.
Jinkun Hou, Zhuo Liu, Huimin Ren et al.
· 0 citations