Skip to content

Category

reinforcement learning

1,901 papers

#reinforcement learning Open access Jan 2027

A single safe policy for fast charging control of multi-chemistry and multi-size battery packs using contextual reinforcement learning

Fast charging of lithium-ion battery (LIB) packs is constrained by thermal runaway risk, cell state-of-charge (SOC) imbalance, and accelerated capacity fade, challenges compounded by spatial temperature gradients and the diversity of cathode chemistries and configurations across electric vehicles and stationary storage...

Saumya Karan, P. Bharadwaj · 0 citations
#reinforcement learning Open access Oct 2026

Increased monetary incentives enhance goal-directed reinforcement learning but fail to suppress automatic learning processes

Adaptive reinforcement learning requires assigning outcomes to the features of actions that causally determine them. Yet humans also learn reward associations with action features that are explicitly known to be outcome-irrelevant, and these associations can bias subsequent choices. Whether such maladaptive learning is...

Ido Ben-Artzi, Noham Wolpe, Nitzan Shahar · 0 citations

Representative view selection and hard view learning for multiview 3D model classification

View-based 3D model classification benefits from mature 2D visual backbones, but dense multi-view rendering usually contains many repeated observations. This paper therefore focuses on two practical questions: how to keep a compact set of representative views, and how to make the classifier learn more from difficult vi...

Xiaopeng Li · 0 citations
#reinforcement learning Open access Oct 2026

Machine learning-based energy management and bidirectional power control for battery-fed electric vehicle traction drive

Abstract Efficient energy management in a battery-fed electric vehicle (EV) traction system requires coordinated control of propulsion power, battery operating conditions, Direct Current (DC)-link voltage, and regenerative braking under rapidly changing driving conditions. This study develops a machine learning (ML)-ba...

Velappagari Sekhar, Syed Suraya, R. Dharmaprakash et al. · 0 citations

Reinforcement and Deep Learning Approaches to Dynamic Cache Optimization

With explosive data-driven application growth and increasing complexity of contemporary computing environments, conventional static cache management methods fall short more and more. This chapter discusses how reinforcement learning (RL) and deep learning (DL) models are disrupting cache optimization by making caching...

Patel Smit Vasant Kumar, Kruti Dataram, Uma Shankar et al. · 0 citations
#reinforcement learning Open access Oct 2026

Calibrated Decisions Are Not Calibrated Probabilities: An Exact-Target Audit of Jev and Three Open Decision Models

Decision models answer typed questions with probabilities instead of text, and their main selling point is that those probabilities can be trusted. TypeSafe says its Jev model, trained with an unpublished method called Reinforcement Learning for Calibrated Decisions (RLCD), returns "epistemically honest" probabilities....

Mohit Shankar Velu · 0 citations

TARL-GBS: research on a two-level agent reinforcement learning algorithm for government-driven blockchain sharding

Government blockchain exhibits distinct characteristics including diverse transaction types, stringent permission hierarchies, and complex cross-departmental approval processes. Existing sharding scheduling mechanisms often lead to issues such as unauthorized downgrading of high-access transactions and missing approval...

Xude Zhou · 0 citations
#reinforcement learning Open access Oct 2026

YouTube Dominates IMO 2025/2026 Search; No Math Content Found — E8 Intelligence Research

FINDING: The search results are dominated by YouTube titles and metadata for IMO 2025/2026 problems, with no actual mathematical content extracted. The only substantive item is an arXiv paper on robotic garment folding (LeHome Challenge 2026), which is unrelated to olympiad mathematics. | MATH: No equations, constants,...

Andrew Stewart Caldin · 0 citations
#reinforcement learning Open access Oct 2026

Successor representation supports structural learning of syllable sequences

Humans efficiently learn the temporal structure of speech, yet the underlying cognitive mechanisms remain unclear. Recent research in visuospatial sequential memory has proposed the successor representation (SR), important in reinforcement learning, which encodes multi-step transitional relationships. To test whether S...

Jiali Liu, Yixiang Wang, Jiayu Feng et al. · 0 citations
#reinforcement learning Open access Oct 2026

A Concise Framework for AI-Driven Blockchain: Integrating DRL Consensus and GNN Security Auditing

This research introduces a novel, high-performance hybrid framework merging Deep Reinforcement Learning (DRL) for dynamic consensus optimization with Graph Neural Networks (GNN) for advanced smart contract security auditing. Traditional blockchain architectures frequently struggle with balancing scalability and securit...

Annu Anuj Sharma · 0 citations

From tech blogs

See all →
MIT News · Artificial Intelligence Oct 7, 2026

Discovering the value of humanistic inquiry

Students in MIT’s Concourse program delve deeply into the human condition, debate challenging questions, and learn to develop judgment about issues that can’t be quantified.

Microsoft Research Blog Oct 7, 2026

Agent Lightning v1.0: A 3,500-Line Lightweight Agentic RL Framework for Training Agents with Real Harnesses

Training AI agents with reinforcement learning can be challenging because their tools, context, and decision-making are managed by complex frameworks. Agent Lightning connects existing agents to RL training, making it easier to improve them without rebuilding them. The post Agent Lightning v1.0: A 3,500-Line Lightweight Agentic RL Framework for Training Agents with Real Harnesses appeared first on Microsoft Research.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.