Skip to content

Category

reinforcement learning

1,990 papers

#reinforcement learning Open access Oct 2026

Project TALOS: Tactical Agentic Literature Orchestration System

Project TALOS is an autonomous research intelligence platform powered by deep reinforcement learning (DDDQN), multi-tier LLM orchestration, and the Grey Wolf Optimizer (GWO). It conducts end-to-end scientific literature discovery and evaluation across 18 academic APIs.

Christos Smarlamakis, Efstratios Georgopoulos · 0 citations

Building an Error Management Climate in Construction Projects

An error management climate (EMC) promotes open discussion, analysis, and constructive handling of errors, fostering learning and innovation. However, how to develop such a climate in construction organizations remains poorly understood. This paper examines practical mechanisms for cultivating EMC in project environm...

Peter E. D. Love · 0 citations
#reinforcement learning Open access Oct 2026

Project TALOS: Tactical Agentic Literature Orchestration System

Project TALOS is an autonomous research intelligence platform powered by deep reinforcement learning (DDDQN), multi-tier LLM orchestration, and the Grey Wolf Optimizer (GWO). It conducts end-to-end scientific literature discovery and evaluation across 18 academic APIs.

Christos Smarlamakis, Efstratios Georgopoulos · 0 citations
#reinforcement learning Open access Oct 2026

Within Reach: Actions as Archetypes

An action is an archetype: substantiated structure for how a thing is done, laid down by what doing it has resolved or produced. This paper states that commitment formally within the Language of Stress, a value-primitive theory of consciousness, and follows what it does. It is the treatment that the constraint map for...

Joshua Pace · 0 citations
#reinforcement learning Open access Oct 2026

Dynamic Difficulty Adjustment in Video Games: A Review of Approaches, Applications and Open Challenges

This paper is a narrative literature review examining Dynamic Difficulty Adjustment (DDA) in video games. It covers the psychological foundations of DDA (Flow Theory and Self-Determination Theory), the main technical approaches used to implement it including rule-based methods, player modeling, reinforcement learning,...

Shabnam Ali Ahmed Khan · 0 citations
#reinforcement learning Open access Oct 2026

Project TALOS: Tactical Agentic Literature Orchestration System

Project TALOS is an autonomous research intelligence platform powered by deep reinforcement learning (DDDQN), multi-tier LLM orchestration, and the Grey Wolf Optimizer (GWO). It conducts end-to-end scientific literature discovery and evaluation across 18 academic APIs.

Christos Smarlamakis, Efstratios Georgopoulos · 0 citations
#reinforcement learning Open access Oct 2026

Project TALOS: Tactical Agentic Literature Orchestration System

Project TALOS is an autonomous research intelligence platform powered by deep reinforcement learning (DDDQN), multi-tier LLM orchestration, and the Grey Wolf Optimizer (GWO). It conducts end-to-end scientific literature discovery and evaluation across 18 academic APIs.

Christos Smarlamakis, Efstratios Georgopoulos · 0 citations
#reinforcement learning Open access Oct 2026

CTAM-RL: Continual Threat-Aware Memory for Continual Reinforcement Learning in Autonomous Cyber Defense

{ "upload_type": "software", "title": "CTAM-RL: Continual Threat-Aware Memory for Continual Reinforcement Learning in Autonomous Cyber Defense (v1.1.0)", "description": "Code, extracted attack profiles, and per-seed results for the manuscript 'A Continual Reinforcement Learning Framework for Autonomous Cyber Defense Un...

Kimia Memarpour, Bahar Memarpour, Kimia Shirini et al. · 0 citations
#reinforcement learning Open access Oct 2026

zannunakiz/Q1_Research_DQN-autonomous-vehicles: Q1 DQN_AV V1.0.0

Safety-Aware DQN Variants for Lightweight Sensor-Based Collision Avoidance in Autonomous Vehicles This repository contains the official research software, experiment code, and datasets supporting the manuscript "Safety-Aware DQN Variants for Lightweight Sensor-Based Collision Avoidance in Autonomous Vehicles." Paper In...

Richky Abednego · 0 citations

From tech blogs

See all →
MIT News · Artificial Intelligence Oct 7, 2026

Discovering the value of humanistic inquiry

Students in MIT’s Concourse program delve deeply into the human condition, debate challenging questions, and learn to develop judgment about issues that can’t be quantified.

Microsoft Research Blog Oct 7, 2026

Agent Lightning v1.0: A 3,500-Line Lightweight Agentic RL Framework for Training Agents with Real Harnesses

Training AI agents with reinforcement learning can be challenging because their tools, context, and decision-making are managed by complex frameworks. Agent Lightning connects existing agents to RL training, making it easier to improve them without rebuilding them. The post Agent Lightning v1.0: A 3,500-Line Lightweight Agentic RL Framework for Training Agents with Real Harnesses appeared first on Microsoft Research.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.