all the models — AI benchmark observatory
← Benchmarks

MRCR v2

MRCR v2 (Multi-Round Coreference Resolution version 2) is an enhanced version of the synthetic long-context reasoning task. It extends the original MRCR framework with improved evaluation criteria and additional complexity for testing models' ability to maintain attention and reasoning across extended contexts.

id mrcr-v2 · max 1 · 3 models reported

#ModelScore

No scores for this benchmark yet.