Encyclopedia Evalica / Evaluation / Absolute scoring

Absolute scoring illustration

Absolute scoring

/'a.bsuh.loot 'skaw.rihng/Assigning a numeric or categorical score to a single output on a fixed scale. This is distinct from pairwise evals, which compares two outputs against each other. (noun)

We switched to absolute scoring so we could set a consistent pass threshold.

Related Evaluation terms

From the docs

Get started with Evals

Braintrust is the AI observability and eval platform for production AI. By connecting evals and observability in one workflow, teams at Notion, Stripe, Zapier, Vercel, and Ramp ship quality AI products at scale.

Start building