SOTAVerified

Text Retrieval

Text Retrieval is the task of finding the most text result (such as an answer, paragraph, or passage) given a query (which could be a question, keywords, or any relevant text)

Papers

Showing 351–400 of 671 papers

TitleStatusHype
Exploring Combinations of Ontological Features and Keywords for Text Retrieval—0
Extracting Molecular Properties from Natural Language with Multimodal Contrastive Learning—0
FecTek: Enhancing Term Weight in Lexicon-Based Retrieval with Feature Context and Term-level Knowledge—0
CaReBench: A Fine-Grained Benchmark for Video Captioning and Retrieval—0
Fine-tuning Multimodal Transformers on Edge: A Parallel Split Learning Approach—0
FLAP: Fast Language-Audio Pre-training—0
FocalLens: Instruction Tuning Enables Zero-Shot Conditional Image Representations—0
Free-ATM: Exploring Unsupervised Learning on Diffusion-Generated Images with Free Attention Masks—0
Free-Form Multi-Modal Multimedia Retrieval (4MR)—0
GAFNet: A Global Fourier Self Attention Based Novel Network for multi-modal downstream tasks—0
GazBy: Gaze-Based BERT Model to Incorporate Human Attention in Neural Information Retrieval—0
Generalizing Multimodal Pre-training into Multilingual via Language Acquisition—0
Generative Negative Text Replay for Continual Vision-Language Pretraining—0
Global–Local Information Soft-Alignment for Cross-Modal Remote-Sensing Image–Text Retrieval—0
Harvest Video Foundation Models via Efficient Post-Pretraining—0
Hashtag Processing for Enhanced Clustering of Tweets—0
HaVTR: Improving Video-Text Retrieval Through Augmentation Using Large Foundation Models—0
Unified Generative & Dense Retrieval for Query Rewriting in Sponsored Search—0
HENASY: Learning to Assemble Scene-Entities for Egocentric Video-Language Model—0
HGAN: Hierarchical Graph Alignment Network for Image-Text Retrieval—0
Hierarchical Gumbel Attention Network for Text-based Person Search—0
HiT: Hierarchical Transformer with Momentum Contrast for Video-Text Retrieval—0
HiVLP: Hierarchical Interactive Video-Language Pre-Training—0
HiVLP: Hierarchical Vision-Language Pre-Training for Fast Image-Text Retrieval—0
How to Make Cross Encoder a Good Teacher for Efficient Image-Text Retrieval?—0
How Vital is the Jurisprudential Relevance: Law Article Intervened Legal Case Retrieval and Matching—0
Hybrid Retrieval and Multi-stage Text Ranking Solution at TREC 2022 Deep Learning Track—0
Hypernymization of named entity-rich captions for grounding-based multi-modal pretraining—0
i-Code Studio: A Configurable and Composable Framework for Integrative AI—0
IG Captioner: Information Gain Captioners are Strong Zero-shot Classifiers—0
ImageBERT: Cross-modal Pre-training with Large-scale Weak-supervised Image-Text Data—0
Image-text Retrieval: A Survey on Recent Research and Development—0
Image-Text Retrieval with Binary and Continuous Label Supervision—0
Improving Adversarial Transferability of Vision-Language Pre-training Models through Collaborative Multimodal Interaction—0
Exploring Train and Test-Time Augmentations for Audio-Language Learning—0
Improving embedding with contrastive fine-tuning on small datasets with expert-augmented scores—0
Improving General Text Embedding Model: Tackling Task Conflict and Data Imbalance through Model Merging—0
Improving Medical Visual Representation Learning with Pathological-level Cross-Modal Alignment and Correlation Exploration—0
Improving Retrieval for RAG based Question Answering Models on Financial Documents—0
In-Batch Negatives for Knowledge Distillation with Tightly-Coupled Teachers for Dense Retrieval—0
Is Cross-modal Information Retrieval Possible without Training?—0
jina-clip-v2: Multilingual Multimodal Embeddings for Text and Images—0
Jina CLIP: Your CLIP Model Is Also Your Text Retriever—0
Killing two birds with one stone: Can an audio captioning system also be used for audio-text retrieval?—0
Knowledge-grounded Adaptation Strategy for Vision-language Models: Building Unique Case-set for Screening Mammograms for Residents Training—0
Knowledge Transfer Across Modalities with Natural Language Supervision—0
Label Embedding using Hierarchical Structure of Labels for Twitter Classification—0
Label Smoothing for Text Mining—0
LaMP: Language-Motion Pretraining for Motion Generation, Retrieval, and Captioning—0
LaT: Latent Translation with Cycle-Consistency for Video-Text Retrieval—0
Show:102550
← PrevPage 8 of 14Next →

No leaderboard results yet.