Encyclopedia Evalica / Datasets / Adversarial examples

Adversarial examples illustration

Adversarial examples

/a.dver'seh.ree.uhl ih'gza.mpuhlz/Test cases designed to break the system (prompt injection, tricky edge cases, ambiguous inputs) to reveal failure modes. They often complement "happy path" test sets. (noun)

We added adversarial examples to make sure the agent resists prompt injection.

Related Datasets terms

From the docs

Get started with Evals

Braintrust is the observability platform for agents in production. By actively applying intelligence to agent traces and automatically surfacing critical patterns, Braintrust helps teams at Notion, Stripe, Box, OpenAI, and Cloudflare ship quality agents at scale.

Start building