Multi-UAV swarm coordination requires high-level task refinement across heterogeneous game-like operation segments with different controllers, risks, and time-dependent rewards. Existing hand-written refinement rules are difficult to tune when an intermediate action has no direct reward but changes downstream losses an...
Chen Wang, Yin-Xuan Huang, Lai-Long Luo et al.· 2026 12th International Conf...· 0 citations
Orthogonal Decomposition for Social Recommendation (ODSR) is proposed, an embedding-space framework that orthogonally decomposes the aggregated social message into an aligned component and an orthogonal deviation, and learns a dimension-wise vector gate to regulate the deviation under ranking supervision.
Rongfeng Guo, Yinxuan Huang, Wei Chen et al.· Proceedings of the 32nd ACM...· 0 citations
OODA-Tool, a typed closed-loop policy designed to mitigate state preservation from action realization, consistently improves task success across model sizes, with larger gains on smaller models and on tasks whose actions depend strongly on information accumulated across turns and prior tool results.
Rongfeng Guo, Yin-Xuan Huang, Yusen Wu et al.· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.