Skip to main content
Configure evaluations to verify that your agents provide the results that you expect. Create test cases with prompts and expected responses, run evaluations to generate agent responses, then review results to identify where your agents can improve. Use the Cognite CLI for evaluation cases you keep in Git and run from the terminal or CI. You can also run evaluations with Evaluate agents in Cognite Data Fusion (CDF). These paths use different workflows and artifacts; don’t treat them as interchangeable.

Before you start

You must create an agent with the Cognite CLI or in the Agent builder.

Choose how to evaluate

Evaluate with the Cognite CLI or Evaluate agents in CDF.
Last modified on August 21, 2026