← Benchmarks
TOMATO
TOMATO (Temporal Reasoning Multimodal Evaluation) assesses multimodal models on motion and temporal perception in video, testing understanding of actions, motion, and changes over time.
id tomato · max 1 · 4 models reported
| # | Model | Score |
|---|
No scores for this benchmark yet.
