all the models — AI benchmark observatory
← Benchmarks

MM IF-Eval

A challenging multimodal instruction-following benchmark that includes both compose-level constraints for output responses and perception-level constraints tied to input images, with comprehensive evaluation pipeline.

id mm-if-eval · max 1 · 2 models reported

#ModelScore

No scores for this benchmark yet.

MM IF-Eval Leaderboard · all the models