all the models — AI benchmark observatory
← Benchmarks

Global PIQA

Global PIQA is a multilingual commonsense reasoning benchmark that evaluates physical interaction knowledge across 100 languages and cultures. It tests AI systems' understanding of physical world knowledge in diverse cultural contexts through multiple choice questions about everyday situations requiring physical commonsense.

id global-piqa · max 1 · 15 models reported

#ModelScore
1Gemini 3 Pro
Google
0.93
2Gemini 3 Flash
Google
0.93
Global PIQA Leaderboard · all the models