all the models — AI benchmark observatory
← Benchmarks

OJBench

OJBench is a competition-level code benchmark designed to assess the competitive-level code reasoning abilities of large language models. It comprises 232 programming competition problems from NOI and ICPC, categorized into Easy, Medium, and Hard difficulty levels. The benchmark evaluates models' ability to solve complex competitive programming challenges using Python and C++.

id ojbench · max 1 · 9 models reported

#ModelScore

No scores for this benchmark yet.