SOTAVerified

Semantic Retrieval

Papers

Showing 1–50 of 86 papers

TitleStatusHype
Engineering RAG Systems for Real-World Applications: Design, Development, and Evaluation—0
Semantic-enhanced Modality-asymmetric Retrieval for Online E-commerce Search—0
MERIT: Multilingual Semantic Retrieval with Interleaved Multi-Condition Query—0
Scene Detection Policies and Keyframe Extraction Strategies for Large-Scale Video Analysis—0
Deep Retrieval at CheckThat! 2025: Identifying Scientific Papers from Implicit Social Media Mentions via Hybrid Retrieval and Re-Ranking—0
HDLxGraph: Bridging Large Language Models and HDL Repositories via HDL Graph DatabasesCode0
Optimizing Retrieval Augmented Generation for Object Constraint Language—0
Ultra Lowrate Image Compression with Semantic Residual Coding and Compression-aware Diffusion—0
An Open-Source Dual-Loss Embedding Model for Semantic Retrieval in Higher Education—0
RAG-MCP: Mitigating Prompt Bloat in LLM Tool Selection via Retrieval-Augmented Generation—0
Enhancing LLM Language Adaption through Cross-lingual In-Context Pre-training—0
Cross-Document Cross-Lingual NLI via RST-Enhanced Graph Fusion and Interpretability Prediction—0
Semantic Retrieval Augmented Contrastive Learning for Sequential Recommendation—0
Role of Databases in GenAI Applications—0
Toward Agentic AI: Generative Information Retrieval Inspired Intelligent Communications and Networking—0
GNN-Coder: Boosting Semantic Code Retrieval with Combined GNNs and Transformer—0
AnDB: Breaking Boundaries with an AI-Native Database for Universal Semantic AnalysisCode0
Spatial-RAG: Spatial Retrieval Augmented Generation for Real-World Geospatial Reasoning Questions—0
Multimodal semantic retrieval for product searchCode0
SCBench: A KV Cache-Centric Analysis of Long-Context MethodsCode5
1-800-SHARED-TASKS at RegNLP: Lexical Reranking of Semantic Retrieval (LeSeR) for Regulatory Question Answering—0
Semantic Retrieval at Walmart—0
Improving Tool Retrieval by Leveraging Large Language Models for Query Generation—0
AmazonQAC: A Large-Scale, Naturalistic Query Autocomplete Dataset—0
Athena: Retrieval-augmented Legal Judgment Prediction with Large Language Models—0
PAR: Prompt-Aware Token Reduction Method for Efficient Large Multimodal Models—0
How to Make LLMs Strong Node Classifiers?—0
LLM Agents as 6G Orchestrator: A Paradigm for Task-Oriented Physical-Layer Automation—0
Learning Spatially-Aware Language and Audio Embeddings—0
Enhancing Relevance of Embedding-based Retrieval at Walmart—0
APTNESS: Incorporating Appraisal Theory and Emotion Support Strategies for Empathetic Response GenerationCode0
GLARE: Low Light Image Enhancement via Generative Latent Feature based Codebook RetrievalCode2
UQE: A Query Engine for Unstructured Databases—0
MINERS: Multilingual Language Models as Semantic RetrieversCode1
Renal digital pathology visual knowledge search platform based on language large model and book knowledge—0
Towards an In-Depth Comprehension of Case Relevance for Better Legal Retrieval—0
Improving Cross-lingual Representation for Semantic Retrieval with Code-switching—0
LLM Based Multi-Agent Generation of Semi-structured Documents from Semantic Templates in the Public Administration DomainCode1
Pattern retrieval of traffic congestion using graph-based associations of traffic domain-specific features—0
M4LE: A Multi-Ability Multi-Range Multi-Task Multi-Domain Long-Context Evaluation Benchmark for Large Language ModelsCode1
If the Sources Could Talk: Evaluating Large Language Models for Research Assistance in HistoryCode0
Dual-Stream Knowledge-Preserving Hashing for Unsupervised Video Retrieval—0
Sentence Embedding Models for Ancient Greek Using Multilingual Knowledge DistillationCode1
MedCPT: Contrastive Pre-trained Transformers with Large-scale PubMed Search Logs for Zero-shot Biomedical Information RetrievalCode2
Unified Embedding Based Personalized Retrieval in Etsy Search—0
Simultaneous or Sequential Training? How Speech Representations Cooperate in a Multi-Task Self-Supervised Learning System—0
Event-Centric Query Expansion in Web Search—0
Surface-Based Retrieval Reduces Perplexity of Retrieval-Augmented Language ModelsCode0
Measuring and Mitigating Constraint Violations of In-Context Learning for Utterance-to-API Semantic Parsing—0
Description-Based Text Similarity—0
Show:102550
← PrevPage 1 of 2Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1Human baselineSoft-F10.84—Unverified
2k-NN with sentence n-grams, GPT-2 embeddings, fICASoft-F10.51—Unverified
3DBTW, GPT-1 embeddings, fICASoft-F10.51—Unverified
4LSA baselineSoft-F10.39—Unverified
5Universal Sentence EncoderSoft-F10.38—Unverified
6Sentence BERTSoft-F10.31—Unverified