Aug 2026· IEEE Transactions on Cybernetics· Vol PP· 0 citations
Medicine
TL;DR
A novel deep reinforcement learning framework that enables safe and efficient navigation in such communication-free, signal-free, and lane-free intersection environments while meeting stringent safety requirements for practical deployment is proposed.
Abstract
Autonomous intersection without reliance on lane markings, traffic signals, or intervehicle communication remains a critical challenge for decentralized autonomous vehicles (DAVs). In this study, we propose a novel deep reinforcement learning (DRL) framework that enables safe and efficient navigation in such communication-free, signal-free, and lane-free intersection environments. Built upon a continuous proximal policy optimization (CPPO) foundation, our method, CPPO-RA-CL, integrates curriculum learning and a reference-action-guided loss to control AVs' acceleration and steering angle under complex multi-AV scenarios. To enhance perception, we design a multimodal data fusion architecture that combines visual and sensor-based inputs in an ego-centric coordinate system (ECCS). Furthermore, we introduce a rule-embedded hybrid policy to ensure long-term deployment safety by combining learned and fallback control. Extensive simulations demonstrate that our approach achieves robust navigation across diverse geometries, including four-way and three-way unstructured intersections. In particular, long-horizon evaluations over 100million km of cumulative vehicle travel report only 36collisions, corresponding to safety levels consistent with real-world statistics. This work highlights the feasibility of scalable, decentralized AV control using DRL without external coordination or infrastructure support while meeting stringent safety requirements for practical deployment.
Unmanned surface and underwater vehicles face challenges in autonomously learning cooperative encirclement for high-value targets under partial observability, intermittent communication, and collision risks. This paper proposes a heterogeneous multi-agent reinforcement learning framework with safe decisionmaking. The f...
In mixed traffic, decision-making for autonomous vehicles (AVs) confronts three interrelated challenges. First, physics-based priors incorporated into reinforcement learning (RL) models fail to capture latent interactive vehicle intentions and diverse driver behaviors, limiting the proactive reasoning capabilities. Sec...
Jie Fang, Wei Zheng, Mengyun Xu et al.· 0 citations
This paper presents an observation-only autonomy framework for Unmanned Underwater Vehicles (UUVs) navigation in dynamic underwater environments that integrates persistent occupancy mapping, global clearance-aware planning, and risk-aware local control. The proposed pipeline constructs occupancy maps solely from onboar...
This work proposes a modified Multi-Agent Twin-Delayed Deep Deterministic Policy Gradient (M-MATD3) algorithm, specifically designed to mitigate common issues such as overestimation bias and high variance observed in standard MATD3.
The proposed framework improves the safety and crossing efficiency of autonomous vehicle decision-making at unsignalized intersections and introduces a composite prioritized replay mechanism into the Twin Delayed Deep Deterministic Policy Gradient algorithm.
Shufeng Wang, Yu-Hang Wang, Yongxin Lei et al.· Machines· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.