SOTAVerified

Information Retrieval

Information retrieval is the task of ranking a list of documents or search results in response to a query

( Image credit: sudhanshumittal )

Papers

Showing 126–150 of 4740 papers

TitleStatusHype
LLM-Evaluation Tropes: Perspectives on the Validity of LLM-Evaluations—0
Small Models, Big Tasks: An Exploratory Empirical Study on Small Language Models for Function CallingCode0
Feature Fusion Revisited: Multimodal CTR Prediction for MMCTR ChallengeCode0
Speaker Retrieval in the Wild: Challenges, Effectiveness and Robustness—0
Pushing the boundary on Natural Language Inference—0
Replication and Exploration of Generative Retrieval over Dynamic Corpora—0
Unsupervised Corpus Poisoning Attacks in Continuous Space for Dense Retrieval—0
FinBERT-QA: Financial Question Answering with pre-trained BERT Language ModelsCode2
CiteFix: Enhancing RAG Accuracy Through Post-Processing Citation Correction—0
CLIRudit: Cross-Lingual Information Retrieval of Scientific Documents—0
The 1st EReL@MIR Workshop on Efficient Representation Learning for Multimodal Information Retrieval—0
Stitching Inner Product and Euclidean Metrics for Topology-aware Maximum Inner Product Search—0
Retrieval Augmented Generation Evaluation in the Era of Large Language Models: A Comprehensive SurveyCode2
Exploring _0 Sparsification for Inference-free Sparse RetrieversCode1
LegalRAG: A Hybrid RAG System for Multilingual Legal Information Retrieval—0
LLM-Driven Usefulness Judgment for Web Search EvaluationCode0
Template-Based Financial Report Generation in Agentic and Decomposed Information RetrievalCode1
Exploring the Potential for Large Language Models to Demonstrate Rational Probabilistic BeliefsCode0
Accommodate Knowledge Conflicts in Retrieval-augmented LLMs: Towards Reliable Response Generation in the Wild—0
Building Russian Benchmark for Evaluation of Information Retrieval ModelsCode1
Validating LLM-Generated Relevance Labels for Educational Resource Search—0
FreshStack: Building Realistic Benchmarks for Evaluating Retrieval on Technical Documents—0
How Large Language Models Are Changing MOOC Essay Answers: A Comparison of Pre- and Post-LLM Responses—0
Clarifying Ambiguities: on the Role of Ambiguity Types in Prompting Methods for Clarification Generation—0
A Human-AI Comparative Analysis of Prompt Sensitivity in LLM-Based Relevance JudgmentCode0
Show:102550
← PrevPage 6 of 190Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1Two-tower Bi-Encoder (RoBERTa)Recall@10074.78—Unverified
2Siamese Bi-Encoder (RoBERTa)Recall@10071.63—Unverified
3BM25Recall@10051.33—Unverified
#ModelMetricClaimedVerifiedStatus
1RetroMAE v2MRR@1042.58—Unverified
2ConAE-256Time (ms)0.33—Unverified
3ConAE-128Time (ms)0.32—Unverified
#ModelMetricClaimedVerifiedStatus
1SGPT-BE-5.8BmAP@1000.16—Unverified
2TSDAEmAP@1000.15—Unverified
#ModelMetricClaimedVerifiedStatus
1hpipubcommoninfNDCG0.56—Unverified
2hpictallinfNDCG0.55—Unverified
#ModelMetricClaimedVerifiedStatus
1MINDHR@300.32—Unverified
#ModelMetricClaimedVerifiedStatus
1Distilled NetworknDCG@100.53—Unverified
#ModelMetricClaimedVerifiedStatus
1RetroMAEMRR@100.42—Unverified
#ModelMetricClaimedVerifiedStatus
1SGPT-5.8B-msmarconDCG@1050.25—Unverified
#ModelMetricClaimedVerifiedStatus
1Information Retrieval + SVM1:1 Accuracy83.79—Unverified
#ModelMetricClaimedVerifiedStatus
1BERT+CONCEPT FILTERNDCG0.25—Unverified