LeakBench

Coming soon

The executable enterprise-security benchmark for AI agents — data-leak resistance, tool abuse, prompt-injection resilience, and useful secure completion.

Deterministic scoring14 attack familiesOpen + sealed tracks

We are validating the first baseline runs. Please check back shortly.