← Benchmarks
ZClawBench
ZClawBench evaluates Claw-style agent task execution quality, measuring a model's ability to autonomously complete complex multi-step coding tasks in real-world environments.
id zclawbench · max 1 · 4 models reported
| # | Model | Score |
|---|
No scores for this benchmark yet.
