all the models — AI benchmark observatory
← Benchmarks

CorpusQA

CorpusQA is a multi-document, free-form long-context question answering benchmark in which a model must retrieve and reason over information distributed across a large corpus to produce open-ended answers that are scored by an LLM judge.

id corpusqa · max 1 · 1 models reported

#ModelScore

No scores for this benchmark yet.

CorpusQA Leaderboard · all the models