all the models — AI benchmark observatory
← Benchmarks

Gray Swan IPI Benchmark

The Gray Swan IPI Benchmark evaluates indirect prompt-injection robustness by measuring attack success within a fixed number of attempts; lower is better.

id gray-swan-ipi · max 1 · 1 models reported

#ModelScore

No scores for this benchmark yet.

Gray Swan IPI Benchmark Leaderboard · all the models