SOTAVerified

Text Matching

Matching a target text to a source text based on their meaning.

Papers

Showing 101–125 of 364 papers

TitleStatusHype
Scaling Computer-Use Grounding via User Interface Decomposition and Synthesis—0
Descriptive Image-Text Matching with Graded Contextual Similarity—0
Compositional Image-Text Matching and Retrieval by Grounding EntitiesCode0
LGD: Leveraging Generative Descriptions for Zero-Shot Referring Image Segmentation—0
Instruction-augmented Multimodal Alignment for Image-Text and Element Matching—0
Dependency Structure Augmented Contextual Scoping Framework for Multimodal Aspect-Based Sentiment Analysis—0
PRECTR: A Synergistic Framework for Integrating Personalized Search Relevance Matching and CTR Prediction—0
CodeReviewQA: The Code Review Comprehension Assessment for Large Language Models—0
MedUnifier: Unifying Vision-and-Language Pre-training on Medical Data with Vision Generation Task using Discrete Visual Representations—0
Object-centric Binding in Contrastive Language-Image Pretraining—0
CGI: Identifying Conditional Generative Models with Example Images—0
MASS: Overcoming Language Bias in Image-Text Matching—0
Learning Textual Prompts for Open-World Semi-Supervised Learning—0
Multi-Head Attention Driven Dynamic Visual-Semantic Embedding for Enhanced Image-Text Matching—0
A Concept-Centric Approach to Multi-Modality Learning—0
You Only Submit One Image to Find the Most Suitable Generative Model—0
ViUniT: Visual Unit Tests for More Robust Visual Programming—0
Vulnerability of Text-Matching in ML/AI Conference Reviewer Assignments to CollusionsCode0
Automatic Prompt Generation and Grounding Object Detection for Zero-Shot Image Anomaly Detection—0
VLM-HOI: Vision Language Models for Interpretable Human-Object Interaction Analysis—0
EntityCLIP: Entity-Centric Image-Text Matching via Multimodal Attentive Contrastive Learning—0
Bridging the Modality Gap: Dimension Information Alignment and Sparse Spatial Constraint for Image-Text Matching—0
Storyboard guided Alignment for Fine-grained Video Action Recognition—0
Enhance Graph Alignment for Large Language Models—0
DISCO: A Hierarchical Disentangled Cognitive Diagnosis Framework for Interpretable Job RecommendationCode0
Show:102550
← PrevPage 5 of 15Next →

No leaderboard results yet.