Co-occurring keywords
Papers
Gaze-Language Alignment for Zero-Shot Prediction of Visual Search Targets from Human Gaze Scanpaths
ICCV 2025
Bringing CLIP to the Clinic: Dynamic Soft Labels and Negation-Aware Learning for Medical Analysis
CVPR 2025
CAPSTONE: Composable Attribute‐Prompted Scene Translation for Zero‐Shot Vision–Language Reasoning
EMNLP 2025
LMOD: A Large Multimodal Ophthalmology Dataset and Benchmark for Large Vision-Language Models
NAACL 2025
D-CoDe: Scaling Image-Pretrained VLMs to Video via Dynamic Compression and Question Decomposition
EMNLP 2025
Hierarchical Divide-and-Conquer Grouping for Classification Adaptation of Pre-Trained Models
ICCV 2025
Dual-Process Image Generation
ICCV 2025