← Benchmarks
WildClawBench
WildClawBench is an agentic coding benchmark from InternLM/Claw-Eval that reports overall model performance on real-world tool-using development tasks.
id wildclawbench · max 1 · 5 models reported
| # | Model | Score |
|---|
No scores for this benchmark yet.
