It's Not What the Image Shows: Irrelevant Context Destabilises VLM Judges Without Informing Them
arXiv:2609.37863v1 Announce Type: cross Abstract: Vision-language models (VLMs) are increasingly used in place of human annotators, making it important that substitutability tests reflect the model rather than incidental evaluation conditions. We…
Read the full story at arXiv cs.AI ↗
ImpactNotable 31/100
Why it mattersRule-based estimate: event keywords (+4); trust 6/10.
RegionsGlobal
Published1 h ago (Wed, 30 Sep 2026 04:00:00 GMT)
RetrievedWed, 30 Sep 2026 04:00:56 GMT via rss
ClassifiedWed, 30 Sep 2026 04:01:19 GMT by heuristic
AuthorNagham Omar, Mahmoud Jabarin, Kinan Ibraheem, Lotem Peled-Cohen