all the models — AI benchmark observatory
← Benchmarks

MMVet

MM-Vet is an evaluation benchmark that examines large multimodal models on complicated multimodal tasks requiring integrated capabilities. It assesses six core vision-language capabilities: recognition, knowledge, spatial awareness, language generation, OCR, and math through questions that require one or more of these capabilities.

id mmvet · max 1 · 2 models reported

#ModelScore

No scores for this benchmark yet.

MMVet Leaderboard · all the models