Skip to content

Category

reinforcement learning

1,990 papers

#reinforcement learning Open access Oct 2026

Quantum Electronics Explained: Quantum Systems, Devices & Advanced Technologies

Quantum Electronics Explained: Quantum Systems, Devices & Advanced Technologies is a publication-grade Open Educational Resource (OER) module covering the physical principles, macroscopic quantum mechanics, circuit quantum electrodynamics (cQED), and cryogenic microwave control governing quantum hardware. Serving as an...

Prep4Uni.Online · 0 citations
#reinforcement learning Open access Oct 2026

NeuroSymbolic-RLNet: a neuro-symbolic reinforcement learning framework for transparent sequential decision-making

The integration of symbolic reasoning with deep reinforcement learning presents a promising paradigm for achieving transparent and interpretable sequential decision-making in complex environments. This work introduces NeuroSymbolic-RLNet, a novel hybrid framework that combines symbolic state transition graphs with neur...

Mohammed Abdullah Alsuwaiket · 0 citations
#reinforcement learning Open access Oct 2026

Architectural Exploration of Reinforcement Learning and Graph Neural Networks for Gate-Level Logic Synthesis in Modern EDA

This review-style study examines how reinforcement learning agents and graph neural networks are being integrated into gate-level logic synthesis for electronic design automation. It discusses how GNN-derived structural embeddings support pre-layout estimation of signal probability and switching activity, how RL agents...

Zandro Guinialope · 0 citations
#reinforcement learning Open access Oct 2026

Security-Aware Adaptive Computation Offloading in Mobile Edge Computing Using Reinforcement Learning and Deep Q-Networks

Mobile Edge Computing (MEC) enables resource-constrained mobile devices to offload computation-intensive tasks to nearby edge servers. Existing computation offloading approaches primarily optimise latency, energy consumption, or resource allocation, but often do not consider security constraints and multi-user queue st...

Brindeshwar Sharma · 0 citations
#reinforcement learning Open access Oct 2026

Automated Digital Logic Circuit Optimization Using Intelligent Systems

ABSTRACT In view of the development of semiconductor manufacturing technology to the sub-3nm nodes, traditional Electronic Design Automation (EDA) approaches face challenges when navigating through the huge search space associated with multi-objective optimization in the realm of Power, Performance, and Area (PPA). Cla...

Ma. Cheena Iry Pajares · 0 citations
#reinforcement learning Open access Oct 2026

Runtime Adaptive Behavioral Direction Control in Autoregressive Transformers via Closed-Loop Activation Telemetry

Large language models encode behavioral constraints directly within their latent representations. Existing approaches to modifying these behaviors rely on parameter-level adjustments (Supervised Fine-Tuning, Reinforcement Learning from Human Feedback, or permanent weight abliteration). These irreversibly alter model we...

Stefan Beierle · 0 citations
#reinforcement learning Open access Oct 2026

Project TALOS: Tactical Agentic Literature Orchestration System

Project TALOS is an autonomous research intelligence platform powered by deep reinforcement learning (DDDQN), multi-tier LLM orchestration, and the Grey Wolf Optimizer (GWO). It conducts end-to-end scientific literature discovery and evaluation across 18 academic APIs.

Christos Smarlamakis, Efstratios Georgopoulos · 0 citations

A Hybrid Model‐Data‐Driven Scheduling Strategy for Vehicle‐to‐Grid Interaction Based on Virtual Power Plant Coordination

Electric vehicles (EVs) have experienced vigorous development in recent years. However, their large‐scale integration into the power grid presents challenges related to the ‘curse of dimensionality’ and uncertainties, making it difficult to balance rapid grid demand response with the interests of EV users. To overcome...

Lu Chen, Xiaona Lv, Jinhu Fang et al. · 0 citations
#reinforcement learning Open access Oct 2026

Automated Digital Logic Circuit Optimization Using Intelligent Systems

ABSTRACT In view of the development of semiconductor manufacturing technology to the sub-3nm nodes, traditional Electronic Design Automation (EDA) approaches face challenges when navigating through the huge search space associated with multi-objective optimization in the realm of Power, Performance, and Area (PPA). Cla...

Ma. Cheena Iry Pajares · 0 citations
#reinforcement learning Open access Oct 2026

Between Determinism and Chance: From Llull's Machine to Generative AI Models

In 1937, Jorge Luis Borges looked back some 650 years at Ramon Llull‘s machinic ars inveniendi. This is one of five essays that looks at contemporary learning machines through the lens of Borges‘s characteristically engimatic historical note and literary reflection. It examines the epistemological tension between deter...

Daria Sergeevna Bylieva · 0 citations
#reinforcement learning Book Open access Oct 2026

EL-RAKHAWI DOCTRINE OF TOPOLOGICAL MORAL GRAVITY AND ONTOLOGICAL REINFORCEMENT LEARNING ENGINEERING THE MATHEMATICAL CONSCIENCE OF AUTONOMOUS AGENTS IN HIGH-DIMENSIONAL DECISION SPACES

EL-RAKHAWI DOCTRINE OF TOPOLOGICAL MORAL GRAVITY: EXECUTIVE SUMMARY This treatise establishes Topological Moral Gravity, reframing AI ethics from external rules to the fundamental geometry of the decision space. We prove standard gradient descent operates in a morally flat Euclidean space. Our solution: bending the dec...

mohamed kamal arafa el-rakhawi · 0 citations
#reinforcement learning Book Open access Oct 2026

EL-RAKHAWI DOCTRINE OF TOPOLOGICAL MORAL GRAVITY AND ONTOLOGICAL REINFORCEMENT LEARNING ENGINEERING THE MATHEMATICAL CONSCIENCE OF AUTONOMOUS AGENTS IN HIGH-DIMENSIONAL DECISION SPACES

EL-RAKHAWI DOCTRINE OF TOPOLOGICAL MORAL GRAVITY: EXECUTIVE SUMMARY This treatise establishes Topological Moral Gravity, reframing AI ethics from external rules to the fundamental geometry of the decision space. We prove standard gradient descent operates in a morally flat Euclidean space. Our solution: bending the dec...

mohamed kamal arafa el-rakhawi · 0 citations

From tech blogs

See all →
MIT News · Artificial Intelligence Oct 7, 2026

Discovering the value of humanistic inquiry

Students in MIT’s Concourse program delve deeply into the human condition, debate challenging questions, and learn to develop judgment about issues that can’t be quantified.

Microsoft Research Blog Oct 7, 2026

Agent Lightning v1.0: A 3,500-Line Lightweight Agentic RL Framework for Training Agents with Real Harnesses

Training AI agents with reinforcement learning can be challenging because their tools, context, and decision-making are managed by complex frameworks. Agent Lightning connects existing agents to RL training, making it easier to improve them without rebuilding them. The post Agent Lightning v1.0: A 3,500-Line Lightweight Agentic RL Framework for Training Agents with Real Harnesses appeared first on Microsoft Research.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.