Skip to content

Category

reinforcement learning

1,990 papers

Self-Programming Artificial Intelligence: Autonomous Learning and Evolutionary Algorithms

This paper presents a hybrid framework for self-programming artificial intelligence (AI), integrating reinforcement learning (RL), genetic programming (GP), neural architecture search (NAS), and meta-learning to enable autonomous code evolution with minimal human intervention. The proposed system is designed to optimiz...

Henry James Neenaalebari · 0 citations
#reinforcement learning Editorial Open access Oct 2026

Editorial: Advancing sustainability and resilience in agri-food supply chains through multi-criteria decision-making methods

Agri-food supply chains (AFSCs) are found at the heart of many of the major challenges faced today, such as climate variability, geopolitical instability, digital disruption, and increased customer expectations for healthy, sustainably produced, yet affordable foods. In particular, smalland medium-sized enterprises (SM...

Behzad Mosalla Nezhad, Jabir Arif, Guoyu Zhao et al. · 0 citations
#reinforcement learning Open access Oct 2026

Bio-inspired exploration of butterfly wing morphology guides deep reinforcement learning optimization of robotic flight

How wing morphology shapes flight performance is a central question in functional biology and bio-inspired robotics, yet these relationships are difficult to isolate in living organisms. Here, we use a robotic butterfly with tunable wing geometry to systematically characterize morphology–performance relationships by co...

Haifeng Huang, Ze Chen, Lung‐Jieh Yang et al. · 0 citations
#reinforcement learning Open access Oct 2026

Generative Artificial Intelligence-Driven Enterprise Knowledge Base

As enterprises generate vast amounts of heterogeneous knowledge, existing knowledge management systems face challenges in providing contextually relevant, accurate, and domain-specific answers. To address these issues, the authors propose KQGen, a unified framework that integrates knowledge graph embeddings with genera...

邹志新, Jiahui Zheng, Xiyin Zheng et al. · 0 citations
#reinforcement learning Open access Oct 2026

Delay-aware multi-agent reinforcement learning for mixed-autonomy platoon control

It is recognized that controlling mixed-autonomy platoons comprising connected and automated vehicles (CAVs) and human-driven vehicles (HDVs) can enhance traffic flow. While multi-agent reinforcement learning (MARL) is a promising real-time control paradigm, existing MARL-based platoon controllers rarely account for co...

Jingyuan Zhou, Zhicheng Wang, Kaidi Yang · 0 citations
#reinforcement learning Open access Oct 2026

Intelligent planning and scheduling system for supply chain logistics routes based on reinforcement learning

The optimization of logistics routing and scheduling in supply chains is a critical challenge, particularly in dynamic environments characterized by fluctuating demand, traffic conditions, and stringent time-capacity constraints. Traditional optimization methods and heuristic-based approaches often struggle to adapt an...

Qiong Liu · 0 citations

From tech blogs

See all →
MIT News · Artificial Intelligence Oct 7, 2026

Discovering the value of humanistic inquiry

Students in MIT’s Concourse program delve deeply into the human condition, debate challenging questions, and learn to develop judgment about issues that can’t be quantified.

Microsoft Research Blog Oct 7, 2026

Agent Lightning v1.0: A 3,500-Line Lightweight Agentic RL Framework for Training Agents with Real Harnesses

Training AI agents with reinforcement learning can be challenging because their tools, context, and decision-making are managed by complex frameworks. Agent Lightning connects existing agents to RL training, making it easier to improve them without rebuilding them. The post Agent Lightning v1.0: A 3,500-Line Lightweight Agentic RL Framework for Training Agents with Real Harnesses appeared first on Microsoft Research.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.