Learning-Guided Task Refinement for Multi-UAV Swarm Coordination
Multi-UAV swarm coordination requires high-level task refinement across heterogeneous game-like operation segments with different controllers, risks, and time-dependent rewards. Existing hand-written refinement rules are difficult to tune when an intermediate action has no direct reward but changes downstream losses an...