all the models — AI benchmark observatory
← Benchmarks

SWE-Atlas

SWE-Atlas is a software engineering benchmark focused on debugging, evaluating a model's ability to localize and fix bugs in real-world codebases.

id swe-atlas · max 1 · 2 models reported

#ModelScore

No scores for this benchmark yet.

SWE-Atlas Leaderboard · all the models