all the models — AI benchmark observatory
← Benchmarks

Legal Agent Benchmark

The Legal Agent Benchmark (LAB) is Harvey's open-source benchmark for evaluating AI agents on complex, long-horizon legal work. Tasks are scored under an all-pass standard against expert-curated rubrics, where a task passes only if every required rubric criterion (facts, conclusions, citations, structure, and analytical moves) passes.

id legal-agent-benchmark · max 1 · 14 models reported

#ModelScore

No scores for this benchmark yet.

Legal Agent Benchmark Leaderboard · all the models