Skip to content
Open access

Deep Reinforcement Learning for Communication-Free Distributed Control of Autonomous Vehicles in Unstructured Intersection.

Aug 2026 · IEEE Transactions on Cybernetics · Vol PP · 0 citations
Medicine

TL;DR

A novel deep reinforcement learning framework that enables safe and efficient navigation in such communication-free, signal-free, and lane-free intersection environments while meeting stringent safety requirements for practical deployment is proposed.

Abstract

Autonomous intersection without reliance on lane markings, traffic signals, or intervehicle communication remains a critical challenge for decentralized autonomous vehicles (DAVs). In this study, we propose a novel deep reinforcement learning (DRL) framework that enables safe and efficient navigation in such communication-free, signal-free, and lane-free intersection environments. Built upon a continuous proximal policy optimization (CPPO) foundation, our method, CPPO-RA-CL, integrates curriculum learning and a reference-action-guided loss to control AVs' acceleration and steering angle under complex multi-AV scenarios. To enhance perception, we design a multimodal data fusion architecture that combines visual and sensor-based inputs in an ego-centric coordinate system (ECCS). Furthermore, we introduce a rule-embedded hybrid policy to ensure long-term deployment safety by combining learned and fallback control. Extensive simulations demonstrate that our approach achieves robust navigation across diverse geometries, including four-way and three-way unstructured intersections. In particular, long-horizon evaluations over 100million km of cumulative vehicle travel report only 36collisions, corresponding to safety levels consistent with real-world statistics. This work highlights the feasibility of scalable, decentralized AV control using DRL without external coordination or infrastructure support while meeting stringent safety requirements for practical deployment.

Read PDF

Similar papers

Conference Aug 2026

Heterogeneous Multi-Agent Autonomous Learning and Safe Cooperative Decision-Making

Unmanned surface and underwater vehicles face challenges in autonomously learning cooperative encirclement for high-value targets under partial observability, intermittent communication, and collision risks. This paper proposes a heterogeneous multi-agent reinforcement learning framework with safe decisionmaking. The f...

Jiang-Li Cao, Chao Liu, Guo-Ping Zhang · 0 citations
Preprint Aug 2026

Knowledge-Data-Dual-Driven Reinforcement Learning for Autonomous Vehicle Control in Mixed Traffic

In mixed traffic, decision-making for autonomous vehicles (AVs) confronts three interrelated challenges. First, physics-based priors incorporated into reinforcement learning (RL) models fail to capture latent interactive vehicle intentions and diverse driver behaviors, limiting the proactive reasoning capabilities. Sec...

Jie Fang, Wei Zheng, Mengyun Xu et al. · 0 citations
Preprint Aug 2026

Unified Planning-Learning Framework for Robust UUV Navigation Under Partial Observability

This paper presents an observation-only autonomy framework for Unmanned Underwater Vehicles (UUVs) navigation in dynamic underwater environments that integrates persistent occupancy mapping, global clearance-aware planning, and risk-aware local control. The proposed pipeline constructs occupancy maps solely from onboar...

M. E. Deowan, Eleni Kelasidi · 0 citations
Aug 2026

Deep reinforcement learning–based safe path planning for leader–follower robots

This work proposes a modified Multi-Agent Twin-Delayed Deep Deterministic Policy Gradient (M-MATD3) algorithm, specifically designed to mitigate common issues such as overestimation bias and high variance observed in standard MATD3.

Ehsan Kazemi Tameh, Mohammadreza Estarki, Saeed Khodaygan · 0 citations
Open access Aug 2026

Trajectory Prediction-Aided Deep Reinforcement Learning for Autonomous Vehicle Decision-Making at Unsignalized Intersections

The proposed framework improves the safety and crossing efficiency of autonomous vehicle decision-making at unsignalized intersections and introduces a composite prioritized replay mechanism into the Twin Delayed Deep Deterministic Policy Gradient algorithm.

Shufeng Wang, Yu-Hang Wang, Yongxin Lei et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.