all the models — AI benchmark observatory
← Benchmarks

Gorilla Benchmark API Bench

APIBench, a comprehensive dataset of over 11,000 instruction-API pairs from HuggingFace, TorchHub, and TensorHub APIs for evaluating language models' ability to generate accurate API calls.

id gorilla-benchmark-api-bench · max 1 · 3 models reported

#ModelScore

No scores for this benchmark yet.