04:00
2026-07-13
arxiv.org
computer-vision
C-GAP: Class-Aware and Online Prompting Improves Vision-Language Models on Imbalanced Classes
A new framework called C-GAP (Caption-Guided Augmentation and Prompting) improves minority-class detection in vision-language models by up to 53% without retraining or additional annotations, accordinβ¦