all the models — AI benchmark observatory
← Benchmarks

Labwork Bench

Labwork Bench is an OpenAI evaluation of whether models can connect perturbations to outcomes in real wet-lab protocols for troubleshooting and optimization tasks.

id labworkbench · max 1 · 1 models reported

#ModelScore

No scores for this benchmark yet.