← Benchmarks
Gorilla Benchmark API Bench
APIBench, a comprehensive dataset of over 11,000 instruction-API pairs from HuggingFace, TorchHub, and TensorHub APIs for evaluating language models' ability to generate accurate API calls.
id gorilla-benchmark-api-bench · max 1 · 3 models reported
| # | Model | Score |
|---|
No scores for this benchmark yet.
