all the models — AI benchmark observatory
← Benchmarks

MIABench

MIABench evaluates multimodal instruction alignment and following capabilities.

id miabench · max 100 · 1 models reported

#ModelScore
1Qwen3 VL 235B A22B Thinking
Alibaba Cloud / Qwen Team · open
0.93