all the models — AI benchmark observatory
← Benchmarks

DSBench-Hard

DSBench-Hard is DeepSeek's internal test set of difficult coding-agent problems.

id dsbench-hard · max 1 · 3 models reported

#ModelScore

No scores for this benchmark yet.