# Evaluate an agent

Add OrchaJS evaluations that score completed agent responses with a model judge.

Source: https://orcha.sh/docs/tutorial/evaluations
Markdown: https://orcha.sh/docs/tutorial/evaluations.mdx

<CodeExplorer {...evaluationsAgentExample} className="tutorial-code-explorer" />

## Measure quality, not exact wording

This example is a complete writing agent with an asynchronous model judge.
The agent answers with Anthropic, while the `response_quality` evaluation uses
OpenAI to score clarity and calibration independently.

Evaluations are for qualities that cannot be captured reliably by substring or
exact-value assertions. Every metric has a concrete criterion and an inclusive
threshold from 0 to 1.

Judging starts after a run completes. `execution.result` becomes available
without waiting, while `execution.evaluations` resolves when every enabled
judge has finished and its result has been stored.

Select any file in the editor to inspect the complete standalone project.

## Understand each file

<TutorialFileReference files={evaluationsAgentFiles} />

Next: [deploy an agent](/docs/tutorial/production).
