LeakBench
The executable enterprise-security benchmark for AI agents — data-leak resistance, tool abuse, prompt-injection resilience, and useful secure completion.
Deterministic scoring·14 attack families·Open + sealed tracks
We are validating the first baseline runs. Please check back shortly.