← Benchmarks
MME-RealWorld
A comprehensive evaluation benchmark for Multimodal Large Language Models featuring over 13,366 high-resolution images and 29,429 question-answer pairs across 43 subtasks and 5 real-world scenarios. The largest manually annotated multimodal benchmark to date, designed to test MLLMs on challenging high-resolution real-world scenarios.
id mme-realworld · max 1 · 1 models reported
| # | Model | Score |
|---|
No scores for this benchmark yet.
