all the models — AI benchmark observatory
← Benchmarks

DRACO

DRACO is a deep research benchmark that evaluates an agent's ability to gather, synthesize, and reason over information to answer complex research questions. Scores are based on official rubrics per question, with the final score being the average across all questions.

id draco · max 1 · 3 models reported

#ModelScore
1Hy4 preview
Tencent · open
0.77
2MiniMax M3
MiniMax · open
0.73
DRACO Leaderboard · all the models