all the models — AI benchmark observatory
← Benchmarks

LOCA-Bench (256k)

LOCA-Bench is a long-context agentic benchmark. The 256k variant evaluates agents using the official ReAct mode with an environment description length of 256k tokens, measuring how well models reason and act over very long contexts.

id loca-bench-256k · max 1 · 1 models reported

#ModelScore

No scores for this benchmark yet.