Co-occurring keywords
Papers
Beyond Words: Augmenting Discriminative Richness via Diffusions in Unsupervised Prompt Learning
CVPR 2025
FINECAPTION: Compositional Image Captioning Focusing on Wherever You Want at Any Granularity
CVPR 2025
SlideChat: A Large Vision-Language Assistant for Whole-Slide Pathology Image Understanding
CVPR 2025
VERA: Explainable Video Anomaly Detection via Verbalized Learning of Vision-Language Models
CVPR 2025
Exploring the Potential of Large Vision-Language Models for Unsupervised Text-Based Person Retrieval
AAAI 2025
Leveraging Large Vision-Language Model as User Intent-Aware Encoder for Composed Image Retrieval
AAAI 2025