SOTAVerified

Benchmarking

Papers

Showing 52015210 of 5548 papers

TitleStatusHype
2017 Robotic Instrument Segmentation ChallengeCode0
AI Fairness 360: An Extensible Toolkit for Detecting, Understanding, and Mitigating Unwanted Algorithmic BiasCode0
Benchmarking Intersectional Biases in NLPCode0
Benchmarking Commercial Intent Detection Services with Practice-Driven EvaluationsCode0
Towards Fair and Privacy-Preserving Federated Deep ModelsCode0
SPDEBench: An Extensive Benchmark for Learning Regular and Singular Stochastic PDEsCode0
Deep Neural Network Benchmarks for Selective ClassificationCode0
Abstraction Alignment: Comparing Model-Learned and Human-Encoded Conceptual RelationshipsCode0
Arabic Speech Recognition by End-to-End, Modular Systems and HumanCode0
Benchmarking Image Perturbations for Testing Automated Driving Assistance SystemsCode0
Show:102550
← PrevPage 521 of 555Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1GPT-4 TurboACC0.56Unverified