all the models — AI benchmark observatory
← Benchmarks

RULER 64k

RULER 64k evaluates the official 13-task RULER v1 suite at a 65536-token context budget.

id ruler-64k · max 1 · 4 models reported

#ModelScore
1MiniCPM-SALA
OpenBMB · open
0.93
RULER 64k Leaderboard · all the models