all the models — AI benchmark observatory
← Benchmarks

OneMillion Bench

OneMillion Bench evaluates AI agents on high-economic-value tasks that require sustained, reliable execution across long-horizon real-world workflows.

id onemillion-bench · max 1 · 4 models reported

#ModelScore

No scores for this benchmark yet.

OneMillion Bench Leaderboard · all the models