all the models — AI benchmark observatory
← Benchmarks

MME-RealWorld

A comprehensive evaluation benchmark for Multimodal Large Language Models featuring over 13,366 high-resolution images and 29,429 question-answer pairs across 43 subtasks and 5 real-world scenarios. The largest manually annotated multimodal benchmark to date, designed to test MLLMs on challenging high-resolution real-world scenarios.

id mme-realworld · max 1 · 1 models reported

#ModelScore

No scores for this benchmark yet.

MME-RealWorld Leaderboard · all the models