Robust Multi-Agent Reinforcement Learning for Small UAS Separation Assurance under GPS Degradation and Spoofing
This work derives a closed-form expression for this adversarial perturbation, bypassing the iterative inner optimization of adversarial training entirely and enabling linear-time evaluation in the state dimension, and shows that this expression approximates the exact minimizer of the value function over the modeled uncertainty set with second-order accuracy.