← Benchmarks
OneMillion Bench
OneMillion Bench evaluates AI agents on high-economic-value tasks that require sustained, reliable execution across long-horizon real-world workflows.
id onemillion-bench · max 1 · 4 models reported
| # | Model | Score |
|---|
No scores for this benchmark yet.
