04:00
2026-07-22
arxiv.org
artificial-intelligence
Attributes Should Come from Images, Not Class Names: Distribution-Conditioned Attribute Selection for Vision-Language Models
A new study from arXiv (2607.18695v1) finds that descriptors generated by large language models (LLMs) for zero-shot classification carry little visual evidence, collapsing ImageNet accuracy from 59.5โฆ