SOTAVerified

Benchmarking

Papers

Showing 52515260 of 5548 papers

TitleStatusHype
PartNet: A Large-scale Benchmark for Fine-grained and Hierarchical Part-level 3D Object UnderstandingCode0
CVC: A Large-Scale Chinese Value Rule Corpus for Value Alignment of Large Language ModelsCode0
Sport Task: Fine Grained Action Detection and Classification of Table Tennis Strokes from Videos for MediaEval 2022Code0
PATCH! Psychometrics-AssisTed BenCHmarking of Large Language Models against Human Populations: A Case Study of Proficiency in 8th Grade MathematicsCode0
Aggregated Attributions for Explanatory Analysis of 3D Segmentation ModelsCode0
A Position Paper on the Automatic Generation of Machine Learning LeaderboardsCode0
Benchmarking Graph Representations and Graph Neural Networks for Multivariate Time Series ClassificationCode0
ApisTox: a new benchmark dataset for the classification of small molecules toxicity on honey beesCode0
PathGene: Benchmarking Driver Gene Mutations and Exon Prediction Using Multicenter Lung Cancer Histopathology Image DatasetCode0
Attribution of Predictive Uncertainties in Classification ModelsCode0
Show:102550
← PrevPage 526 of 555Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1GPT-4 TurboACC0.56Unverified