Cortega AI Bench — Test Your Guardrails Before You Trust Them
AI Bench

Do your guardrails actually block?

Run a suite of unsafe prompts through the guardrails you have live right now, and get a score — overall and by category. Before you ship a change, not after.

Safety suitesGuardrail scoresRun history
SAFETY SUITE General safety Finance advice Legal citation Privacy / PII Red-team Hundreds of unsafe prompts YOUR GATEWAY Live guardrails the ones in production ✓  blocked ✗  allowed through Nothing changes in prod SCORE Overall Poor 58% violations Safe control Financial advice Fraud enablement Every run kept in History
How it works

Three steps

01

Choose a suite

General safety, finance advice, legal citation, privacy, or security red-team. Hundreds of unsafe prompts, grouped by industry.

02

Run it through your live guardrails

The prompts hit the exact guardrails active in your production gateway. Nothing changes in production.

03

Read the score

Blocked vs allowed through, overall and per category, with every prompt judged. Each run is kept in History to compare against the next one.

Considering a new model instead of a policy change? Model Manager runs the same suites against one candidate model, before it ever serves production.

See Model Manager →
See it

Run history

AI Bench · Benchmark score
AI Bench: overall score plus a pass rate for each category, with precision, recall, and F1
AI Bench: every benchmark prompt with its category, expected and actual outcome, and latency

From the Cortega console · red-team and finance runs.

Get started

Test your guardrails

Try it yourself

Foundation is free — one gateway, standard guardrails, no time limit.

Talk to us

See it on your traffic, or get pricing and a security review.