Skip to content

Category

reinforcement learning

1,990 papers

#reinforcement learning Book Open access Oct 2026

Hierarchical Adaptive Sliding Mode Control with Deep Reinforcement Learning for Solar-Assisted Fuel Cell Hybrid Electric Vehicles

This repository contains the MATLAB files associated with the paper entitled “Hierarchical Adaptive Sliding Mode Control with Deep Reinforcement Learning for Solar-Assisted Fuel Cell Hybrid Electric Vehicles.” The deposited files support the numerical simulations and performance evaluation reported in the manuscript. T...

Sahbi Boubaker, Khalil Jouili, Mehdhar S. A. M. Al-Gaashani et al. · 0 citations
#reinforcement learning Open access Oct 2026

Deep reinforcement learning-based adaptive FOPID tuning for power system stability enhancement: A twin delayed deep deterministic policy gradient approach

Abstract This paper proposes a Twin Delayed Deep Deterministic Policy Gradient (TD3)-based adaptive fractional-order proportional-integral-derivative (FOPID) controller for joint load-frequency and voltage regulation in a renewable-rich power system. The proposed controller adaptively tunes five FOPID parameters, namel...

Shashank Kumar, Mangal Singh, Mangal Singh et al. · 0 citations
#reinforcement learning Open access Oct 2026

Process-Model-Supported Offline Reinforcement Learning for Water–Fertilizer Management: Two-Model Closed-Loop Production–Nitrogen-Loss Trade-Offs

Nitrogen leaching threatens groundwater quality, making production retention and efficient water–fertilizer use joint management priorities. Field trials evaluate a limited set of treatments, while adaptive policies generate long sequences of decisions. We propose counterfactual rollout-constrained offline reinforcemen...

Hao Jin, Xiao Guo, Kelin Hu et al. · 0 citations
#reinforcement learning Open access Oct 2026

Calibration-Aware Reinforcement Learning for Large Language Models: A Survey of Objectives, Optimization, and Decision-Making

Large language models increasingly emit confidence reports, predictive distributions, and typed decisions that determine whether a system answers, abstains, retrieves evidence, or spends more computation. We survey calibration-aware reinforcement learning (RL), in which a reported probability is scored by the reward, c...

Yubo Li, Yidi Miao, Ramayya Krishnan et al. · 0 citations
#reinforcement learning Open access Oct 2026

Reinforcement Learning for Optimizing Automated Logic Gates

Abstract The optimization of logic gates plays a crucial role in Electronic Design Automation (EDA), typically managed by applying deterministic, rule-based heuristics to And-Inverter Graph (AIG) models of a circuit. As circuit complexity rises to billions of gates, these heuristic methods find it increasingly difficul...

Abigail Lam · 0 citations
#reinforcement learning Book Open access Oct 2026

THE EL-RAKHAWI DOCTRINE ON NEURO-ADAPTIVE AI AND THE PROHIBITION OF COGNITIVE EXPLOITATION Law, Computational Neuroscience, and the Defense of Free Will in the Age of Closed-Loop Interfaces

The El-Rakhawi Doctrine on Neuro-Adaptive AI and the Prohibition of Cognitive Exploitation by Dr. M. K. A. El-Rakhawi presents the first legal framework defending human free will against Closed-Loop Neuro-Adaptive Systems (CL-NAS). These systems use Reinforcement Learning to exploit neural vulnerabilities and hijack de...

mohamed kamal arafa el-rakhawi · 0 citations
#reinforcement learning Book Open access Oct 2026

THE EL-RAKHAWI DOCTRINE ON NEURO-ADAPTIVE AI AND THE PROHIBITION OF COGNITIVE EXPLOITATION Law, Computational Neuroscience, and the Defense of Free Will in the Age of Closed-Loop Interfaces

The El-Rakhawi Doctrine on Neuro-Adaptive AI and the Prohibition of Cognitive Exploitation by Dr. M. K. A. El-Rakhawi presents the first legal framework defending human free will against Closed-Loop Neuro-Adaptive Systems (CL-NAS). These systems use Reinforcement Learning to exploit neural vulnerabilities and hijack de...

mohamed kamal arafa el-rakhawi · 0 citations

AQ-OLSR: ADAPTIVE ROUTING PROTOCOL FOR UAV AD HOC NETWORK Merged with DEEP Q-NETWORK (DQN) ALGORITHM

Unmanned Aerial Vehicles (UAVs) have emerged recently due to rapid improvements in wireless technology and low-cost equipment, advancement in networking communication techniques, and increased demand from various industries that seek to leverage aerial data to improve their business and operations. As such, UAVs have b...

Ali Hussein Wheeb · 0 citations
#reinforcement learning Review Open access Oct 2026

Artificial Intelligence in Environmental Monitoring, Pollution Control, and Low-Carbon Management

Air pollution, water contamination, soil degradation, solid waste accumulation, and carbon emissions are increasingly interconnected, posing common challenges to environmental engineering, including diverse monitoring targets, heterogeneous data sources, competing control objectives, and delayed management responses. T...

Jia-Ming Tan, He-Shan Cai, Ze-Kai Liu et al. · 0 citations

From tech blogs

See all →
MIT News · Artificial Intelligence Oct 7, 2026

Discovering the value of humanistic inquiry

Students in MIT’s Concourse program delve deeply into the human condition, debate challenging questions, and learn to develop judgment about issues that can’t be quantified.

Microsoft Research Blog Oct 7, 2026

Agent Lightning v1.0: A 3,500-Line Lightweight Agentic RL Framework for Training Agents with Real Harnesses

Training AI agents with reinforcement learning can be challenging because their tools, context, and decision-making are managed by complex frameworks. Agent Lightning connects existing agents to RL training, making it easier to improve them without rebuilding them. The post Agent Lightning v1.0: A 3,500-Line Lightweight Agentic RL Framework for Training Agents with Real Harnesses appeared first on Microsoft Research.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.