Skip to content
Open access

Deep reinforcement learning for portfolio allocation: a comparative study on five Vietnamese equities

Jul 2026 · Asian Journal of Economics and Banking · 0 citations · 19 references

Abstract

This study applies deep reinforcement learning (DRL) to multi-asset portfolio optimization in the Vietnamese stock market, aiming to evaluate the performance and stability of different DRL algorithms under emerging market conditions. Seven algorithms – A2C, proximal policy optimization (PPO), deep deterministic policy gradient, twin delayed deep deterministic policy gradient (TD3), soft actor-critic (SAC), truncated quantile critics (TQC) and RecurrentPPO – are trained on daily data from January 2018 to December 2024 and evaluated out-of-sample from January 2023 through September 2025 using five liquid equities (SBT.VN, BID.VN, CTG.VN, HPG.VN and VCB.VN). The state representation includes technical indicators (RSI, MACD, SMA, EMA, Bollinger Bands and OBV), as well as risk features such as rolling covariance and a turbulence index. PPO achieves the highest annual return (0.1666) and demonstrates the most stable performance. TD3 delivers comparable cumulative growth with higher variability. RecurrentPPO attains the highest Sharpe ratio (1.0461), highlighting the importance of temporal modeling. SAC and TQC produce more conservative but stable outcomes. This study does not propose a new DRL algorithm; instead, it provides a controlled and reproducible benchmarking framework for comparing multiple DRL models under identical conditions in an emerging market setting.

Read PDF

Similar papers

Jul 2026

SciPhy Reinforcement Learning for Portfolio Optimization

The results demonstrate that the proposed framework successfully translates known signal quality into a robust, multi-period, and cost-aware allocation mechanism with strictly controlled volatility and turnover.

I. Halperin, A. Itkin · 0 citations
2026

Fixed Deep Residual Networks with Multi-Objective Proximal Policy Optimization for Adaptive Trading

Adaptive trading systems must balance return generation with downside control while processing noisy, non stationary market data. This study evaluates a preference conditioned trading framework that combines a randomly initialized, permanently frozen deep residual network (DRN) with a single multi-objective proximal po...

Xi-Chen Song · 0 citations
Jul 2026

CLaC@FinMMEval 2026 Task 3: Sentiment-Augmented Deep Reinforcement Learning for Active Trading - An Alpha-Reward Approach

This paper presents our system for Task 3 of the CLEF 2026 FinMMEval Lab, which requires daily long, flat, or short trading decisions for Bitcoin (BTC) and Tesla (TSLA) using news and historical market data. We formulate the problem as a discrete-action Markov Decision Process and compare four deep reinforcement learni...

Andrei Neagu, Eeham Khan, Leila Kosseim · 0 citations
Review Open access Aug 2026

A survey on LLM-enhanced reinforcement learning in financial markets

A three-paradigm taxonomy (feature-based, auxiliary-based, and policy-based) based on the functional role of LLMs within the RL pipeline is proposed, which provides superior scalability and stability, though often at the expense of representational depth.

Ghusoon Hadi al-Aldaffaie, Alireza Taheri, Amirfarhad Farhadi et al. · 0 citations
Review Open access Aug 2026

Mapping the Methodological Bifurcation of Quantitative Portfolio Optimization: A PRISMA-Compliant Systematic Review with BERTopic–SPECTER Analysis (2003–2025)

A rank-weighted similarity analysis, designed to neutralise the c-TF-IDF collinearity artefact, shows that deep reinforcement learning is the most isolated paradigm.

Gharmili Meryem, Boudri Imane, Alj Abdelkamel · 0 citations
Aug 2026

Quantum-Inspired Portfolio Optimization Using Reinforcement Learning for Dynamic Stock Allocation

A new Quantum-Inspired Portfolio Optimization (QIPO-RL) model is presented that combines quantum-inspired search techniques, an adaptive RL agent, and asset weights to create a framework for a Reinforcement Learning (RL) based dynamic stock allocation algorithm.

Kishore Kumar Sambangi · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.