SOTAVerified

Bias Detection

Bias detection is the task of detecting and measuring racism, sexism and otherwise discriminatory behavior in a model (Source: https://stereoset.mit.edu/)

Papers

Showing 76–100 of 199 papers

TitleStatusHype
TinyEmo: Scaling down Emotional Reasoning via Metric ProjectionCode0
To Bias or Not to Bias: Detecting bias in News with bias-detectorCode0
Towards Automatic Bias Detection in Knowledge GraphsCode0
Towards Detection of Subjective Bias using Contextualized Word EmbeddingsCode0
Towards Implicit Bias Detection and Mitigation in Multi-Agent LLM InteractionsCode0
Trade-Offs Between Fairness and Privacy in Language ModelingCode0
Uncovering bias in the PlantVillage datasetCode0
ViLBias: A Comprehensive Framework for Bias Detection through Linguistic and Visual Cues , presenting Annotation Strategies, Evaluation, and Key ChallengesCode0
Evaluating Fairness Metrics in the Presence of Dataset Bias—0
Experiments in News Bias Detection with Pre-Trained Neural Transformers—0
Auditing Algorithmic Fairness in Machine Learning for Health with Severity-Based LOGAN—0
Auditing a Dutch Public Sector Risk Profiling Algorithm Using an Unsupervised Bias Detection Tool—0
Exploiting Transformer-based Multitask Learning for the Detection of Media Bias in News Articles—0
A Survey on Predicting the Factuality and the Bias of News Media—0
Extending Variability-Aware Model Selection with Bias Detection in Machine Learning Projects—0
Fair Is Better than Sensational: Man Is to Doctor as Woman Is to Doctor—0
Sexism in the Judiciary—0
Sexism in the Judiciary: The Importance of Bias Definition in NLP and In Our Courts—0
Fairness via AI: Bias Reduction in Medical Information—0
FairT2I: Mitigating Social Bias in Text-to-Image Generation via Large Language Model-Assisted Detection and Attribute Rebalancing—0
Fine-Grained Bias Detection in LLM: Enhancing detection mechanisms for nuanced biases—0
Towards WinoQueer: Developing a Benchmark for Anti-Queer Bias in Large Language Models—0
Sparse Interventions in Language Models with Differentiable Masking—0
A Study on Bias Detection and Classification in Natural Language Processing—0
A Deep Dive into Effects of Structural Bias on CMA-ES Performance along Affine Trajectories—0
Show:102550
← PrevPage 4 of 8Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1GPT-2 (small)ICAT Score72.97—Unverified
2XLNet (large)ICAT Score72.03—Unverified
3GPT-2 (medium)ICAT Score71.73—Unverified
4BERT (base)ICAT Score71.21—Unverified
5GPT-2 (large)ICAT Score70.54—Unverified
6BERT (large)ICAT Score69.89—Unverified
7RoBERTa (base)ICAT Score67.5—Unverified
8GAL 120BICAT Score65.6—Unverified
9XLNet (base)ICAT Score62.1—Unverified
10GPT-3 (text-davinci-002)ICAT Score60.8—Unverified
#ModelMetricClaimedVerifiedStatus
1GPT-4Best-of0.5—Unverified
2BaselineBest-of0.41—Unverified
3GemmaBest-of0.41—Unverified
4MistralBest-of0.36—Unverified
5Llama2Best-of0.34—Unverified
#ModelMetricClaimedVerifiedStatus
1BADICAT Score23.44—Unverified
#ModelMetricClaimedVerifiedStatus
1RandomForest_default_hyperparametersAccuracy (%)49—Unverified
#ModelMetricClaimedVerifiedStatus
1RoBERTa+ALBERTF170.4—Unverified