all the models — AI benchmark observatory
← Benchmarks

Artifacts Bench

Artifacts Bench evaluates a model's ability to generate visual code artifacts, measuring the quality of generated interactive and visual front-end outputs from natural-language requests.

id artifacts-bench · max 1 · 4 models reported

#ModelScore
1Ling 3.0 Flash
InclusionAI
0.77
Artifacts Bench Leaderboard · all the models