SOTAVerified

MTEB Benchmark

Papers

Showing 1–13 of 13 papers

TitleStatusHype
NeoBERT: A Next-Generation BERTCode2
KaLM-Embedding: Superior Training Data Brings A Stronger Embedding ModelCode2
Jina Embeddings 2: 8192-Token General-Purpose Text Embeddings for Long DocumentsCode1
C-Pack: Packed Resources For General Chinese EmbeddingsCode1
GATE: General Arabic Text Embedding for Enhanced Semantic Textual Similarity with Matryoshka Representation Learning and Hybrid Loss Training—0
Optimization of embeddings storage for RAG systems using quantization and dimensionality reduction techniques—0
FaMTEB: Massive Text Embedding Benchmark in Persian Language—0
GenEOL: Harnessing the Generative Power of LLMs for Training-Free Sentence EmbeddingsCode0
Contextual Document Embeddings—0
jina-embeddings-v3: Multilingual Embeddings With Task LoRA—0
A Bi-metric Framework for Fast Similarity SearchCode0
Recent advances in text embedding: A Comprehensive Review of Top-Performing Methods on the MTEB Benchmark—0
Text Embeddings by Weakly-Supervised Contrastive Pre-training—0
Show:102550

No leaderboard results yet.