← Benchmarks
MM IF-Eval
A challenging multimodal instruction-following benchmark that includes both compose-level constraints for output responses and perception-level constraints tied to input images, with comprehensive evaluation pipeline.
id mm-if-eval · max 1 · 2 models reported
| # | Model | Score |
|---|
No scores for this benchmark yet.
