← Benchmarks
B3 AI Security Benchmark
B3 AI Security Benchmark evaluates agent-backbone robustness against contextual prompt-injection attacks collected in Lakera's Agent Breaker Challenge.
id b3-ai-security-benchmark · max 1 · 1 models reported
| # | Model | Score |
|---|---|---|
| 1 | Mistral Large 4 Mistral AI | 0.93 |
