conftrace
_
Papers
Trends
Conferences
Explore
More
Authors
Topics
Keywords
Insights
Papers
Trends
Conferences
Explore
Authors
Topics
Keywords
Insights
Achievements
Home
›
Keywords
›
vision-language model
vision-language model
2348 papers
Explore in graph
Also known as
VLM
VL MODEL
VLMS
Co-occurring keywords
multimodal learning
(4645)
zero-shot learning
(3650)
vision language model
(767)
contrastive learning
(4032)
large language model
(13587)
visual question answering
(1017)
transfer learning
(5449)
few-shot learning
(3398)
knowledge distillation
(3725)
semantic segmentation
(3186)
Papers
MmAP: Multi-Modal Alignment Prompt for Cross-Domain Multi-Task Learning
AAAI 2024
Leveraging Diffusion Perturbations for Measuring Fairness in Computer Vision
AAAI 2024
Data Adaptive Traceback for Vision-Language Foundation Models in Image Classification
AAAI 2024
An Empirical Study of CLIP for Text-Based Person Search
AAAI 2024
COMMA: Co-articulated Multi-Modal Learning
AAAI 2024
Structure-CLIP: Towards Scene Graph Knowledge to Enhance Multi-Modal Structured Representations
AAAI 2024
VLCounter: Text-Aware Visual Representation for Zero-Shot Object Counting
AAAI 2024
Adaptive Uncertainty-Based Learning for Text-Based Person Retrieval
AAAI 2024
Mining Fine-Grained Image-Text Alignment for Zero-Shot Captioning via Text-Only Training
AAAI 2024
ViLT-CLIP: Video and Language Tuning CLIP with Multimodal Prompt Learning and Scenario-Guided Optimization
AAAI 2024
Towards Learning a Generalist Model for Embodied Navigation
CVPR 2024
Enhancing Vision-Language Pre-training with Rich Supervisions
CVPR 2024
Holistic Features are almost Sufficient for Text-to-Video Retrieval
CVPR 2024
Beyond Text: Frozen Large Language Models in Visual Signal Comprehension
CVPR 2024
Emergent Open-Vocabulary Semantic Segmentation from Off-the-shelf Vision-Language Models
CVPR 2024
Eyes Wide Shut? Exploring the Visual Shortcomings of Multimodal LLMs
CVPR 2024
Do Vision and Language Encoders Represent the World Similarly?
CVPR 2024
CADTalk: An Algorithm and Benchmark for Semantic Commenting of CAD Programs
CVPR 2024
Discovering and Mitigating Visual Biases through Keyword Explanation
CVPR 2024
Contrasting Intra-Modal and Ranking Cross-Modal Hard Negatives to Enhance Visio-Linguistic Compositional Understanding
CVPR 2024
AHIVE: Anatomy-aware Hierarchical Vision Encoding for Interactive Radiology Report Retrieval
CVPR 2024
Exploring Regional Clues in CLIP for Zero-Shot Semantic Segmentation
CVPR 2024
SkyScript: A Large and Semantically Diverse Vision-Language Dataset for Remote Sensing
AAAI 2024
Detecting and Preventing Hallucinations in Large Vision Language Models
AAAI 2024
Multimodal Ensembling for Zero-Shot Image Classification
AAAI 2024
<
1
…
64
65
66
…
94
>