all the models — AI benchmark observatory
← Benchmarks

RULER 128k

RULER 128k evaluates the official 13-task RULER v1 suite at a 131072-token context budget.

id ruler-128k · max 1 · 4 models reported

#ModelScore

No scores for this benchmark yet.