← back to model results
ground truth
Animated Concepts #3
csssource ↗
model outputs
Gemini 3 Flash Preview →
A 0.77T 0.32
Qwen3-VL-8B-Instruct →
A 0.63T 0.22
GPT-5.4 →
A 0.89T 0.22
Claude Sonnet 4.6 →
A 0.81T 0.27
LLaMA 4 Scout →
A 0.44T 0.21
1<span class="load5"></span>
2