RefCOCO average

A referring-expression grounding benchmark averaged across RefCOCO variants to test whether a model can localize described objects correctly.

ModelsModels with a result.6
Top resultBest result on this test.N/A
Top-3 spreadPoints from first to third.N/A
YearYear of release.2026

Results

6 results
#ModelResultThe score from the source.
Qwen3.6-27B92.5%