SOTAVerified

Common Sense Reasoning

Common sense reasoning tasks are intended to require the model to go beyond pattern recognition. Instead, the model should use "common sense" or world knowledge to make inferences.

Papers

Showing 110 of 939 papers

TitleStatusHype
Comparing Apples to Oranges: A Dataset & Analysis of LLM Humour Understanding from Traditional Puns to Topical Jokes0
LoSiA: Efficient High-Rank Fine-Tuning via Subnet Localization and OptimizationCode0
CheckManual: A New Challenge and Benchmark for Manual-based Appliance Manipulation0
EditInspector: A Benchmark for Evaluation of Text-Guided Image Edits0
Prime the search: Using large language models for guiding geometric task and motion planning by warm-starting tree searchCode0
AmbiK: Dataset of Ambiguous Tasks in Kitchen EnvironmentCode0
ATLAS: Learning to Optimally Memorize the Context at Test Time0
Spatial Knowledge Graph-Guided Multimodal Synthesis0
CaseEdit: Enhancing Localized Commonsense Reasoning via Null-Space Constrained Knowledge Editing in Small Parameter Language Models0
Align-GRAG: Reasoning-Guided Dual Alignment for Graph Retrieval-Augmented Generation0
Show:102550
← PrevPage 1 of 94Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1Human BenchmarkAverage F10.93Unverified
2Golden TransformerAverage F10.92Unverified
3YaLM 1.0B few-shotAverage F10.86Unverified
4ruT5-large-finetuneAverage F10.81Unverified
5ruT5-base-finetuneAverage F10.79Unverified
6ruBert-base finetuneAverage F10.74Unverified
7ruRoberta-large finetuneAverage F10.73Unverified
8ruBert-large finetuneAverage F10.68Unverified
9RuGPT3XL few-shotAverage F10.67Unverified
10MT5 LargeAverage F10.57Unverified