all the models — AI benchmark observatory
← Benchmarks

CommonSenseQA

CommonSenseQA is a multiple-choice question answering dataset that requires different types of commonsense knowledge to predict correct answers. It contains 12,102 questions with one correct answer and four distractors, designed to test semantic reasoning and conceptual relationships. Questions are created based on ConceptNet concepts and require prior world knowledge for accurate reasoning.

id commonsenseqa · max 1 · 1 models reported

#ModelScore

No scores for this benchmark yet.