Evidence-in-the-Loop: Trace-Driven Optimization for Customer-Service LLM Agents
The paper contributes three reusable deployment patterns: hybrid RAG evidence construction, multi-channel retrieval and reranking produce auditable FAQ candidates, and trace-driven RAG and reranker improvement, where reranker fine-tuning is evaluated not only for in-domain gain but also for forgetting risk.