← Benchmarks
MMVet
MM-Vet is an evaluation benchmark that examines large multimodal models on complicated multimodal tasks requiring integrated capabilities. It assesses six core vision-language capabilities: recognition, knowledge, spatial awareness, language generation, OCR, and math through questions that require one or more of these capabilities.
id mmvet · max 1 · 2 models reported
| # | Model | Score |
|---|
No scores for this benchmark yet.
