all the models — AI benchmark observatory
← Benchmarks

MMVetGPT4Turbo

MM-Vet evaluation using GPT-4 Turbo for scoring. This variant of MM-Vet examines large multimodal models on complicated multimodal tasks requiring integrated capabilities across six core vision-language abilities: recognition, knowledge, spatial awareness, language generation, OCR, and math.

id mmvetgpt4turbo · max 1 · 1 models reported

#ModelScore

No scores for this benchmark yet.