all the models — AI benchmark observatory
← Benchmarks

IF

Instruction-Following Evaluation (IFEval) benchmark for large language models, focusing on verifiable instructions with 25 types of instructions and around 500 prompts containing one or more verifiable constraints

id if · max 1 · 2 models reported

#ModelScore

No scores for this benchmark yet.

IF Leaderboard · all the models