all the models — AI benchmark observatory
← Benchmarks

ZClawBench

ZClawBench evaluates Claw-style agent task execution quality, measuring a model's ability to autonomously complete complex multi-step coding tasks in real-world environments.

id zclawbench · max 1 · 4 models reported

#ModelScore

No scores for this benchmark yet.

ZClawBench Leaderboard · all the models