Learning High-Risk High-Precision Motion Control
This work proposes and evaluates State-Conditioned Shooting (SCOOT), a novel DRL algorithm that builds on advantage-weighted regression (AWR) with three key modifications, and showcases the features’ performance in learning physically-based billiard shots demonstrating high action precision and discovering multiple sho...