all the models — AI benchmark observatory
← Benchmarks

SWE Atlas - Test Writing

SWE Atlas - Test Writing evaluates a model's ability to author meaningful tests for real-world software projects, measuring how well agents can understand code and produce correct, useful test coverage.

id swe-atlas-test-writing · max 1 · 1 models reported

#ModelScore

No scores for this benchmark yet.

SWE Atlas - Test Writing Leaderboard · all the models