all the models — AI benchmark observatory
← Benchmarks

CL-bench

CL-bench is an open-source benchmark with its own data and rubrics for evaluating models on coding and agentic tasks, scored using a setup fully aligned with the official procedure.

id cl-bench · max 1 · 3 models reported

#ModelScore

No scores for this benchmark yet.