all the models — AI benchmark observatory
← Benchmarks

VoiceBench Avg

VoiceBench is the first benchmark designed to provide a multi-faceted evaluation of LLM-based voice assistants, evaluating capabilities including general knowledge, instruction-following, reasoning, and safety using both synthetic and real spoken instruction data with diverse speaker characteristics and environmental conditions.

id voicebench-avg · max 1 · 3 models reported

#ModelScore

No scores for this benchmark yet.