all the models — AI benchmark observatory
← Benchmarks

TreeBench

TreeBench evaluates visual grounded reasoning, requiring models to localize and reason about fine-grained visual details.

id treebench · max 1 · 3 models reported

#ModelScore

No scores for this benchmark yet.