← Benchmarks
Labwork Bench
Labwork Bench is an OpenAI evaluation of whether models can connect perturbations to outcomes in real wet-lab protocols for troubleshooting and optimization tasks.
id labworkbench · max 1 · 1 models reported
| # | Model | Score |
|---|
No scores for this benchmark yet.
