Skip to main content

Overview

AccuracyEval uses an LLM judge to score agent responses against expected answers on a 0.0–1.0 scale.

Quick Start

Install @agentium/core@4.0.0, @agentium/eval@4.0.0, and the optional openai SDK. Set OPENAI_API_KEY; this example makes live provider calls. For a no-key runnable project, start with the quality-gate recipe.

Configuration