conftrace
_
Papers
Trends
Conferences
Explore
More
Authors
Topics
Keywords
Insights
Papers
Trends
Conferences
Explore
Authors
Topics
Keywords
Insights
Achievements
Home
›
Keywords
›
vision-language model
vision-language model
2348 papers
Explore in graph
Also known as
VLM
VL MODEL
VLMS
Co-occurring keywords
multimodal learning
(4645)
zero-shot learning
(3650)
vision language model
(767)
contrastive learning
(4032)
large language model
(13587)
visual question answering
(1017)
transfer learning
(5449)
few-shot learning
(3398)
knowledge distillation
(3725)
semantic segmentation
(3186)
Papers
SeTAR: Out-of-Distribution Detection with Selective Low-Rank Approximation
NIPS 2024
Aggregate-and-Adapt Natural Language Prompts for Downstream Generalization of CLIP
NIPS 2024
VLG-CBM: Training Concept Bottleneck Models with Vision-Language Guidance
NIPS 2024
SpatialPIN: Enhancing Spatial Reasoning Capabilities of Vision-Language Models through Prompting and Interacting 3D Priors
NIPS 2024
JourneyBench: A Challenging One-Stop Vision-Language Understanding Benchmark of Generated Images
NIPS 2024
Conjugated Semantic Pool Improves OOD Detection with Pre-trained Vision-Language Models
NIPS 2024
Self-Calibrated Tuning of Vision-Language Models for Out-of-Distribution Detection
NIPS 2024
Magnet: We Never Know How Text-to-Image Diffusion Models Work, Until We Learn How Vision-Language Models Function
NIPS 2024
Free Lunch in Pathology Foundation Model: Task-specific Model Adaptation with Concept-Guided Feature Enhancement
NIPS 2024
Relationship Prompt Learning is Enough for Open-Vocabulary Semantic Segmentation
NIPS 2024
Training-Free Open-Ended Object Detection and Segmentation via Attention as Prompts
NIPS 2024
Boosting Vision-Language Models with Transduction
NIPS 2024
Enhancing Domain Adaptation through Prompt Gradient Alignment
NIPS 2024
ChatTracker: Enhancing Visual Tracking Performance via Chatting with Multimodal Large Language Model
NIPS 2024
DiPEx: Dispersing Prompt Expansion for Class-Agnostic Object Detection
NIPS 2024
Multimodal Task Vectors Enable Many-Shot Multimodal In-Context Learning
NIPS 2024
Revisiting Few-Shot Object Detection with Vision-Language Models
NIPS 2024
Towards Open-Vocabulary Semantic Segmentation Without Semantic Labels
NIPS 2024
CALVIN: Improved Contextual Video Captioning via Instruction Tuning
NIPS 2024
WATT: Weight Average Test Time Adaptation of CLIP
NIPS 2024
Interpreting and Analysing CLIP's Zero-Shot Image Classification via Mutual Knowledge
NIPS 2024
Measuring Dejavu Memorization Efficiently
NIPS 2024
TabPedia: Towards Comprehensive Visual Table Understanding with Concept Synergy
NIPS 2024
TOPA: Extending Large Language Models for Video Understanding via Text-Only Pre-Alignment
NIPS 2024
Few-Shot Adversarial Prompt Learning on Vision-Language Models
NIPS 2024
<
1
…
53
54
55
…
94
>