Skip to content

Author

Hao Zhang

We have 2 of 93 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Sep 2026

Trajectory-Constrained Human-Guided Reinforcement Learning for Autonomous Driving.

In this article, we propose trajectory-constrained human-guided reinforcement learning (TCHug-RL), a new framework that 1) produces guidance at the full-trajectory level to suit the planning horizon of autonomous driving and 2) provides a formal criterion for deciding when and how to inject human input into the policy-update loop. First, we propose a trajectory-level similarity of human guidance, which is defined such that the satisfaction of trajectory-level similarity implies the satisfaction of single-step similarity. Consequently, trajectory-level similarity is strictly stronger than the single-step one. Then, the trajectory-level similarity is modeled as a constraint to guide the subsequent policy optimization. In this way, not only is a clear criterion for human guidance provided, but the influence of such guidance on the policy is also transparent and interpretable. Finally, we leverage the Lagrangian method to solve the constrained optimization problem and provide a strong duality analysis for it in the unparameterized policy space. Furthermore, a practical algorithm implementation of TCHug-RL and experiments are provided, demonstrating that TCHug-RL effectively leverages human guidance and achieves improvements of 22.1% in learning efficiency and 18.7% in overall performance in autonomous driving tasks, compared to the state-of-the-art methods.

Li-Fei Dai, Yiqun Liu, Hao Zhang et al. · 0 citations
Open access Jul 2026

Attention-based Spatio-Temporal Graph Convolutional Networks-enhanced deep reinforcement learning for adaptive traffic signal control in urban traffic networks.

A novel adaptive traffic signal control framework by integrating Attention-based Spatio-Temporal Graph Convolutional Networks (ASTGCN) with Multi-Agent Deep Deterministic Policy Gradient (MADDPG) is proposed, providing a scalable and data-driven solution for intelligent traffic signal control in urban traffic networks, supporting the development of smart mobility systems.

Jing Wang, Xiaopeng Wang, Yang Mo et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.