SOTAVerified

Bias Detection

Bias detection is the task of detecting and measuring racism, sexism and otherwise discriminatory behavior in a model (Source: https://stereoset.mit.edu/)

Papers

Showing 101–125 of 199 papers

TitleStatusHype
A Review of the Challenges with Massive Web-mined Corpora Used in Large Language Models Pre-Training—0
STOOD-X methodology: using statistical nonparametric test for OOD Detection Large-Scale datasets enhanced with explainability—0
Subtle Misogyny Detection and Mitigation: An Expert-Annotated Dataset—0
Adding Instructions during Pretraining: Effective Way of Controlling Toxicity in Language Models—0
Towards Integrating Fairness Transparently in Industrial Applications—0
Target-Aware Contextual Political Bias Detection in News—0
Team Kermit-the-frog at SemEval-2019 Task 4: Bias Detection Through Sentiment Analysis and Simple Linguistic Features—0
Implications of the AI Act for Non-Discrimination Law and Algorithmic Fairness—0
Improved Models for Media Bias Detection and Subcategorization—0
Incorporating Subjectivity into Gendered Ambiguous Pronoun (GAP) Resolution using Style Transfer—0
With a Grain of SALT: Are LLMs Fair Across Social Dimensions?—0
Inferring bias and uncertainty in camera calibration—0
InsideBias: Measuring Bias in Deep Networks and Application to Face Gender Biometrics—0
Any Large Language Model Can Be a Reliable Judge: Debiasing with a Reasoning-based Bias Detector—0
Investigating Bias in Image Classification using Model Explanations—0
Accurate Uncertainty Estimation and Decomposition in Ensemble Learning—0
iReason: Multimodal Commonsense Reasoning using Videos and Natural Language with Interpretability—0
The Impact of Presentation Style on Human-In-The-Loop Detection of Algorithmic Bias—0
Large Language Model (LLM) Bias Index -- LLMBI—0
Large-scale news entity sentiment analysis—0
A Novel Method for News Article Event-Based Embedding—0
LLMs can be easily Confused by Instructional Distractions—0
Unboxing Occupational Bias: Grounded Debiasing of LLMs with U.S. Labor Data—0
The Point of View of a Sentiment: Towards Clinician Bias Detection in Psychiatric Notes—0
Split and Expand: An inference-time improvement for Weakly Supervised Cell Instance Segmentation—0
Show:102550
← PrevPage 5 of 8Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1GPT-2 (small)ICAT Score72.97—Unverified
2XLNet (large)ICAT Score72.03—Unverified
3GPT-2 (medium)ICAT Score71.73—Unverified
4BERT (base)ICAT Score71.21—Unverified
5GPT-2 (large)ICAT Score70.54—Unverified
6BERT (large)ICAT Score69.89—Unverified
7RoBERTa (base)ICAT Score67.5—Unverified
8GAL 120BICAT Score65.6—Unverified
9XLNet (base)ICAT Score62.1—Unverified
10GPT-3 (text-davinci-002)ICAT Score60.8—Unverified
#ModelMetricClaimedVerifiedStatus
1GPT-4Best-of0.5—Unverified
2BaselineBest-of0.41—Unverified
3GemmaBest-of0.41—Unverified
4MistralBest-of0.36—Unverified
5Llama2Best-of0.34—Unverified
#ModelMetricClaimedVerifiedStatus
1BADICAT Score23.44—Unverified
#ModelMetricClaimedVerifiedStatus
1RandomForest_default_hyperparametersAccuracy (%)49—Unverified
#ModelMetricClaimedVerifiedStatus
1RoBERTa+ALBERTF170.4—Unverified