all the models — AI benchmark observatory
← Benchmarks

MRCR 1M

MRCR 1M is a variant of the Multi-Round Coreference Resolution benchmark designed for testing extremely long context capabilities with approximately 1 million tokens. It evaluates models' ability to maintain reasoning and attention across ultra-long conversations.

id mrcr-1m · max 1 · 4 models reported

#ModelScore

No scores for this benchmark yet.

MRCR 1M Leaderboard · all the models