Stop Guessing! Images Know Their Own Attributes Best! This research proposes selecting attributes for vision-language models directly from the target images themselves, rather than relying on large language models (LLMs) to generate descriptors based only on class names. Previous methods generated a