Preprint
Aug 2026
From Solver Feedback to Faithful Plans: Multi-Role Reinforcement Learning for Symbolic Planning
Results show that organizing solver feedback into generation, verification, and repair roles enables more scalable and faithful annotation-free symbolic planning.
Chenghao Zhang, Yikai Mao, Shan Liu et al.
· 0 citations