MODEL AND AGENT EVALUATION SYSTEM IN A MULTI-AGENT INTELLIGENT COMPETITIVE INTELLIGENCE PLATFORM: ARCHITECTURE AND METHODOLOGY
The aim is to develop the architecture of a two-level evaluation system featuring an Evaluator Agent with reasoning chain tracing, a control dataset of 30+ use cases, and a validation mechanism on samples of real user queries.