Aug 2026· Proceedings of the VLDB Endowment· Vol 19, pp. 3928-3940· 0 citations· 38 references
TL;DR
DBAgent is presented, an autonomous agent for Huawei Cloud Data Warehouse Service (DWS) integrated with Autopilot (DWS's production monitoring, alerting, and auto-remediation service) that consumes DWS telemetry and Autopilot alerts and produces evidence-grounded reports with low hallucination on DWS benchmark.
Abstract
Database Operations and Maintenance (O&M) is a critical but complex and labor-intensive task. Recent LLM-based assistants promise to lower the barrier by reading manuals/tickets and exploring diagnostic search trees. However, existing LLM-based solutions fall short due to fundamental limitations in learning from expert demonstration. Such design fails to internalize domain dynamics (how interventions change plans, resources, etc.), and struggles under workload and statistics drift. To address this, we present DBAgent, an autonomous agent for Huawei Cloud Data Warehouse Service (DWS) integrated with Autopilot (DWS's production monitoring, alerting, and auto-remediation service). DBAgent consumes DWS telemetry (e.g., KPIs and execution plans) and Autopilot alerts to diagnose and remediate incidents in production clusters. DBAgent emulates an expert's iterative
Think-Act-Observe
problem-solving loop with a policy trained via reinforcement learning (RL). It couples dynamic tool use for information gathering, a multimodal perception module for database-native signals, and an RL-based reasoning engine that plans, verifies, and generates evidence-grounded remediation recommendations. Extensive experiments show that DBAgent handles a broad range of complex O&M tasks. It surpasses the strongest baseline by +23% success rate and produces evidence-grounded reports with low hallucination (~5%) on DWS benchmark.
HxAgent is introduced, an iterative LLM-based planning agent with a proactive correction strategy that achieves 97.4% Exact-Match accuracy on MiniWoB++, comparable to the best baselines without human demonstrations and surpassing the recent WALT by 10.5%.
Tuong Nguyen, Duy Cao, Viet Nguyen et al.· 0 citations
Long-horizon autonomous research tasks such as machine learning engineering require systems to make interdependent decisions under a limited budget. Existing LLM-based agents typically organize candidate-solution improvement through tree, graph, or chain structures, meaning that the search process determines how inform...
Shaokang Fu, Yulong Tao, Linbo Jin et al.· 1 citation
PILOT (Proactive Insight Learner for Online Tree-Experiments), an LLM-agent framework that organizes three roles within a constrained control loop where deterministic services enforce all safety, statistical, and permission boundaries, is presented.
Jiuning Lin, Ruiquan Lan, Xiaodong Zhu et al.· 0 citations
LLMs enable multi-agent systems (MAS) to tackle complex tasks, but manually designing agent roles, prompts, and communication structures requires substantial expertise and effort. This motivates learning policies that construct query-specific MAS from execution reward. Existing approaches typically train these policies...
Bei-Cheng Xu, Bo-Wen Fan, Wei Qian et al.· 0 citations
Recent gains in language model capability have come more from data than from architecture. Frontier labs and data companies produce verifiable agentic tasks, which supervised finetuning and reinforcement learning then turn into capability.This production line still rests on human labour and on human-in-the-loop collabo...
Haotian Luo, Hao-Yu Wang, Ze-Yu Qin et al.· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.