SOTAVerified

cross-modal alignment

Papers

Showing 221230 of 342 papers

TitleStatusHype
LangBridge: Interpreting Image as a Combination of Language Embeddings0
Linguistic Query-Guided Mask Generation for Referring Image Segmentation0
Learning Better Visual Representations for Weakly-Supervised Object Detection Using Natural Language Supervision0
Learning by Hallucinating: Vision-Language Pre-training with Weak Supervision0
Learning Joint Embedding with Modality Alignments for Cross-Modal Retrieval of Recipes and Food Images0
Learning Multi-Modal Nonlinear Embeddings: Performance Bounds and an Algorithm0
Learning to Localize Actions in Instructional Videos with LLM-Based Multi-Pathway Text-Video Alignment0
Let Me Finish My Sentence: Video Temporal Grounding with Holistic Text Understanding0
Prototype-guided Cross-modal Completion and Alignment for Incomplete Text-based Person Re-identification0
RAC3: Retrieval-Augmented Corner Case Comprehension for Autonomous Driving with Vision-Language Models0
Show:102550
← PrevPage 23 of 35Next →

No leaderboard results yet.