← Benchmarks
RefCOCOg
RefCOCOg is a referring expression comprehension benchmark that evaluates spatial grounding in images. Given a natural language expression describing an object, the model must localize the correct region, evaluated by accuracy at a 0.5 IoU threshold. It features longer, more descriptive expressions than RefCOCO and RefCOCO+.
id refcocog · max 1 · 1 models reported
| # | Model | Score |
|---|---|---|
| 1 | Nova 2 Omni Amazon | 0.86 |
