Encyclopedia Evalica / Evaluation / Alignment

Alignment illustration

Alignment

/uh'leyen.muhnt/The process of calibrating an LLM judge so its scores match human judgment. Alignment typically involves iterating on rubrics and prompts until agreement is high. (noun)

After our alignment sessions, the LLM judge scores matched human ratings much more closely.

Customer example

Navan treated alignment as an iterative process, tuning its eval prompt like an ML classifier until the evaluator reached >0.9 macro F1 and using reasoning to debug failures. Read more

Related Evaluation terms

From the docs

Get started with Evals

Braintrust is the observability platform for agents in production. By actively applying intelligence to agent traces and automatically surfacing critical patterns, Braintrust helps teams at Notion, Stripe, Box, OpenAI, and Cloudflare ship quality agents at scale.

Start building