all the models — AI benchmark observatory
← Benchmarks

WildClawBench

WildClawBench is an agentic coding benchmark from InternLM/Claw-Eval that reports overall model performance on real-world tool-using development tasks.

id wildclawbench · max 1 · 5 models reported

#ModelScore

No scores for this benchmark yet.