all the models — AI benchmark observatory
← Benchmarks

NoLiMa 128K

NoLiMa evaluated at a 131072-token context length. Tests latent associative reasoning in long contexts with minimal lexical overlap between questions and needles.

id nolima-128k · max 1 · 1 models reported

#ModelScore

No scores for this benchmark yet.