AMIGO: Agentic Multi-Image Grounding Oracle Benchmark
arXiv:2603.28662v2 Announce Type: replace-cross Abstract: Agentic vision-language models increasingly act through extended interactions, but most evaluations still focus on single-image, single-turn correctness. We introduce \textbf{AMIGO} (\textbf{A}gentic \t
arXiv cs.AI··Updated just now·28 sightings