all the models — AI benchmark observatory
← Models

Qwen2.5 VL 72B Instruct

Alibaba Cloud / Qwen Team · open weight · qwen2.5-vl-72b

Qwen2.5-VL is the new flagship vision-language model of Qwen, significantly improved from Qwen2-VL. It excels at recognizing objects, analyzing text/charts/layouts in images, acting as a visual agent, understanding long videos (over 1 hour) with event pinpointing, performing visual localization (bounding boxes/points), and generating structured outputs from documents.

DocVQAAndroid ContChartQAOCRBenchAI2DMMBench

Benchmark scores

Pricing

  • No provider pricing.

AA metrics

No Artificial Analysis link yet.

Arena Elo

  • No Arena snapshot linked.