SOTAVerified

Sentence Completion

Papers

Showing 1–25 of 91 papers

TitleStatusHype
Evaluating Gender Bias in Large Language Models—0
KatzBot: Revolutionizing Academic Chatbot for Enhanced CommunicationCode0
BiasAlert: A Plug-and-play Tool for Social Bias Detection in LLMs—0
Ranking LLMs by compression—0
Mixture-of-Subspaces in Low-Rank AdaptationCode0
Enhancing Bangla Language Next Word Prediction and Sentence Completion through Extended RNN with Bi-LSTM Model On N-gram Language—0
MixLoRA: Enhancing Large Language Models Fine-Tuning with LoRA-based Mixture of ExpertsCode3
Language Model Sentence Completion with a Parser-Driven Rhetorical Control MethodCode0
Parameter-Efficient Sparsity Crafting from Dense to Mixture-of-Experts for Instruction Tuning on General TasksCode2
Illuminating the Black Box: A Psychometric Investigation into the Multifaceted Nature of Large Language Models—0
LLM in a flash: Efficient Large Language Model Inference with Limited Memory—0
Mamba: Linear-Time Sequence Modeling with Selective State SpacesCode6
The Falcon Series of Open Language Models—0
mahaNLP: A Marathi Natural Language Processing Library—0
BTRec: BERT-Based Trajectory Recommendation for Personalized ToursCode0
Sheared LLaMA: Accelerating Language Model Pre-training via Structured PruningCode2
Mistral 7BCode6
Investigating Subtler Biases in LLMs: Ageism, Beauty, Institutional, and Nationality Bias in Generative ModelsCode0
Exploiting Language Models as a Source of Knowledge for Cognitive Agents—0
I-WAS: a Data Augmentation Method with GPT-2 for Simile Detection—0
Llama 2: Open Foundation and Fine-Tuned Chat ModelsCode8
Stay on topic with Classifier-Free Guidance—0
ScoNe: Benchmarking Negation Reasoning in Language Models With Fine-Tuning and In-Context LearningCode0
The CoT Collection: Improving Zero-shot and Few-shot Learning of Language Models via Chain-of-Thought Fine-TuningCode2
PaLM 2 Technical Report—0
Show:102550
← PrevPage 1 of 4Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1CompassMTL 567M with TailorAccuracy96.1—Unverified
2CompassMTL 567MAccuracy95.6—Unverified
3DeBERTa-Large 304M (classification-based)Accuracy95.6—Unverified
4GPT-4 (10-shot)Accuracy95.3—Unverified
5LLaMA3+MoSLoRAAccuracy95—Unverified
6LLaMA-2 13B + MixLoRAAccuracy94.7—Unverified
7DeBERTa-Large 304MAccuracy94.7—Unverified
8Unicorn 11B (fine-tuned)Accuracy93.9—Unverified
9LLaMA-3 8B + MixLoRAAccuracy93.3—Unverified
10LLaMA-2 7B + MixLoRAAccuracy93.1—Unverified