MODEL AND AGENT EVALUATION SYSTEM IN A MULTI-AGENT INTELLIGENT COMPETITIVE INTELLIGENCE PLATFORM: ARCHITECTURE AND METHODOLOGY
The article addresses the problem of evaluating the quality of large language models (LLMs) and intelligent agents in multi-agent competitive intelligence automation platforms. The aim is to develop the architecture of a two-level evaluation system featuring an Evaluator Agent with reasoning chain tracing, a control da...