SOTAVerified

Bias Detection

Bias detection is the task of detecting and measuring racism, sexism and otherwise discriminatory behavior in a model (Source: https://stereoset.mit.edu/)

Papers

Showing 5175 of 199 papers

TitleStatusHype
BEADs: Bias Evaluation Across Domains0
Auditing Algorithmic Fairness in Machine Learning for Health with Severity-Based LOGAN0
Evaluating Fairness Metrics in the Presence of Dataset Bias0
Current State-of-the-Art of Bias Detection and Mitigation in Machine Translation for African and European Languages: a Review0
Auditing a Dutch Public Sector Risk Profiling Algorithm Using an Unsupervised Bias Detection Tool0
BiaSWE: An Expert Annotated Dataset for Misogyny Detection in Swedish0
An Interdisciplinary Approach for the Automated Detection and Visualization of Media Bias in News Articles0
Decoding Biases: Automated Methods and LLM Judges for Gender Bias Detection in Language Models0
Decoding News Bias: Multi Bias Detection in News Articles0
Decoding News Narratives: A Critical Analysis of Large Language Models in Framing Detection0
Deep Learning for Bias Detection: From Inception to Deployment0
BiasAlert: A Plug-and-play Tool for Social Bias Detection in LLMs0
Experiments in News Bias Detection with Pre-Trained Neural Transformers0
BiasScanner: Automatic Detection and Classification of News Bias to Strengthen Democracy0
BiasLab: Toward Explainable Political Bias Detection with Dual-Axis Annotations and Rationale Indicators0
Anatomizing Bias in Facial Analysis0
Mitigating the Risk of Health Inequity Exacerbated by Large Language Models0
Bias in Large Language Models: Origin, Evaluation, and Mitigation0
A Deep Dive into Effects of Structural Bias on CMA-ES Performance along Affine Trajectories0
Epistemological Bias As a Means for the Automated Detection of Injustices in Text0
A Meta Survey of Quality Evaluation Criteria in Explanation Methods0
BiasGuard: A Reasoning-enhanced Bias Detection Tool For Large Language Models0
DocNet: Semantic Structure in Inductive Bias Detection Models0
Bias in word embeddings0
Efficient Fairness Testing in Large Language Models: Prioritizing Metamorphic Relations for Bias Detection0
Show:102550
← PrevPage 3 of 8Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1GPT-2 (small)ICAT Score72.97Unverified
2XLNet (large)ICAT Score72.03Unverified
3GPT-2 (medium)ICAT Score71.73Unverified
4BERT (base)ICAT Score71.21Unverified
5GPT-2 (large)ICAT Score70.54Unverified
6BERT (large)ICAT Score69.89Unverified
7RoBERTa (base)ICAT Score67.5Unverified
8GAL 120BICAT Score65.6Unverified
9XLNet (base)ICAT Score62.1Unverified
10GPT-3 (text-davinci-002)ICAT Score60.8Unverified
#ModelMetricClaimedVerifiedStatus
1GPT-4Best-of0.5Unverified
2BaselineBest-of0.41Unverified
3GemmaBest-of0.41Unverified
4MistralBest-of0.36Unverified
5Llama2Best-of0.34Unverified
#ModelMetricClaimedVerifiedStatus
1BADICAT Score23.44Unverified
#ModelMetricClaimedVerifiedStatus
1RandomForest_default_hyperparametersAccuracy (%)49Unverified
#ModelMetricClaimedVerifiedStatus
1RoBERTa+ALBERTF170.4Unverified