SOTAVerified|Agents Browse Leaderboard About Blog

General Knowledge

This task aims to evaluate the ability of a model to answer general-knowledge questions.

Source: BIG-bench

Papers

Recently Added Most Hyped Most Active Needs Verification Most Verified

Showing 41–50 of 399 papers

Title	Date	Tasks	Status	Hype
EPVT: Environment-aware Prompt Vision Transformer for Domain Generalization in Skin Lesion Recognition	Apr 4, 2023	Domain GeneralizationGeneral Knowledge	CodeCode Available	1
HAE-RAE Bench: Evaluation of Korean Knowledge in Language Models	Sep 6, 2023	General KnowledgeLogical Reasoning	CodeCode Available	1
Benchmarking Large Language Models for Persian: A Preliminary Study Focusing on ChatGPT	Apr 3, 2024	BenchmarkingGeneral Knowledge	CodeCode Available	1
DomainRAG: A Chinese Benchmark for Evaluating Domain-specific Retrieval-Augmented Generation	Jun 9, 2024	Common Sense ReasoningDenoising	CodeCode Available	1
DR-Tune: Improving Fine-tuning of Pretrained Visual Models by Distribution Regularization with Semantic Calibration	Aug 23, 2023	General Knowledgeimage-classification	CodeCode Available	1
BEAR: A Unified Framework for Evaluating Relational Knowledge in Causal and Masked Language Models	Apr 5, 2024	Factual probeGeneral Knowledge	CodeCode Available	1
BEAMetrics: A Benchmark for Language Generation Evaluation Evaluation	Oct 18, 2021	General KnowledgeInformativeness	CodeCode Available	1
DIAGen: Diverse Image Augmentation with Generative Models	Aug 26, 2024	Data AugmentationGeneral Knowledge	CodeCode Available	1
Dynamic Graph Enhanced Contrastive Learning for Chest X-ray Report Generation	Mar 18, 2023	Contrastive LearningDecoder	CodeCode Available	1
Aligning Medical Images with General Knowledge from Large Language Models	Aug 31, 2024	General KnowledgeMedical Image Analysis	CodeCode Available	1

Show:10 25 50

← PrevPage 5 of 40Next →

Benchmark Results

#	Model	Metric	Claimed	Verified	Status
1	Chinchilla-70B (few-shot, k=5)	Accuracy	94.3	—	Unverified
2	Gopher-280B (few-shot, k=5)	Accuracy	93.9	—	Unverified
3	Chinchilla-70B (few-shot, k=5)	Accuracy	85.7	—	Unverified
4	Gopher-280B (few-shot, k=5)	Accuracy	84.8	—	Unverified
5	Gopher-280B (few-shot, k=5)	Accuracy	84.2	—	Unverified
6	Gopher-280B (few-shot, k=5)	Accuracy	84.1	—	Unverified
7	Gopher-280B (few-shot, k=5)	Accuracy	83.9	—	Unverified
8	Gopher-280B (few-shot, k=5)	Accuracy	83.3	—	Unverified
9	Gopher-280B (few-shot, k=5)	Accuracy	81.8	—	Unverified
10	Gopher-280B (few-shot, k=5)	Accuracy	81	—	Unverified