SOTAVerified

Hate Speech Detection

Hate speech detection is the task of detecting if communication such as text, audio, and so on contains hatred and or encourages violence towards a person or a group of people. This is usually based on prejudice against 'protected characteristics' such as their ethnicity, gender, sexual orientation, religion, age et al. Some example benchmarks are ETHOS and HateXplain. Models can be evaluated with metrics like the F-score or F-measure.

Papers

Showing 251–275 of 507 papers

TitleStatusHype
SWE2: SubWord Enriched and Significant Word Emphasized Framework for Hate Speech Detection—0
Systematic Offensive Stereotyping (SOS) Bias in Language Models—0
Tâches Auxiliaires Multilingues pour le Transfert de Modèles de Détection de Discours Haineux (Multilingual Auxiliary Tasks for Zero-Shot Cross-Lingual Transfer of Hate Speech Detection)—0
Who Speaks Matters: Analysing the Influence of the Speaker's Ethnicity on Hate Classification—0
The binary trio at SemEval-2019 Task 5: Multitarget Hate Speech Detection in Tweets—0
The Effects of User Features on Twitter Hate Speech Detection—0
The Impact of Persona-based Political Perspectives on Hateful Content Detection—0
Measuring Catastrophic Forgetting in Cross-Lingual Transfer Paradigms: Exploring Tuning Strategies—0
The Risk of Racial Bias in Hate Speech Detection—0
The Role of Context in Detecting the Target of Hate Speech—0
Thesis Distillation: Investigating The Impact of Bias in NLP Models on Hate Speech Detection—0
To BAN or not to BAN: Bayesian Attention Networks for Reliable Hate Speech Detection—0
ToKen: Task Decomposition and Knowledge Infusion for Few-Shot Hate Speech Detection—0
Towards A Multi-agent System for Online Hate Speech Detection—0
Towards Argument Mining for Social Good: A Survey—0
Towards Code-switched Classification Exploiting Constituent Language Resources—0
Towards countering hate speech against journalists on social media—0
Towards Fairness Assessment of Dutch Hate Speech Detection—0
Towards generalisable hate speech detection: a review on obstacles and solutions—0
Towards hate speech detection in low-resource languages: Comparing ASR to acoustic word embeddings on Wolof and Swahili—0
Towards High-Fidelity Synthetic Multi-platform Social Media Datasets via Large Language Models—0
ToxSyn-PT: A Large-Scale Synthetic Dataset for Hate Speech Detection in Portuguese—0
Transferring Knowledge via Neighborhood-Aware Optimal Transport for Low-Resource Hate Speech Detection—0
Trustworthy Hate Speech Detection Through Visual Augmentation—0
TuEval at SemEval-2019 Task 5: LSTM Approach to Hate Speech Detection in English and Spanish—0
Show:102550
← PrevPage 11 of 21Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1BiLSTM + static BEF1-score0.8—Unverified
2BERTF1-score0.79—Unverified
3BiLSTM+Attention+FTF1-score0.77—Unverified
4OPT-175B (few-shot)F1-score0.76—Unverified
5CNN+Attention+FT+GVF1-score0.74—Unverified
6OPT-175B (one-shot)F1-score0.71—Unverified
7OPT-175B (zero-shot)F1-score0.67—Unverified
8SVMF1-score0.66—Unverified
9Random ForestsF1-score0.64—Unverified
10Davinci (zero-shot)F1-score0.63—Unverified
#ModelMetricClaimedVerifiedStatus
1BERT-MRPAUROC0.86—Unverified
2BERT-RPAUROC0.85—Unverified
3BERT-HateXplain [LIME]AUROC0.85—Unverified
4BERT-HateXplain [Attn]AUROC0.85—Unverified
5BERT [Attn]AUROC0.84—Unverified
6BiRNN-HateXplain [Attn]AUROC0.81—Unverified
7BiRNN-Attn [Attn]AUROC0.8—Unverified
8CNN-GRU [LIME]AUROC0.79—Unverified
9BiRNN [LIME]AUROC0.77—Unverified
10XG-HSI-BERTAccuracy0.75—Unverified
#ModelMetricClaimedVerifiedStatus
1MLARAMHamming Loss0.29—Unverified
2MLkNNHamming Loss0.16—Unverified
3Binary RelevanceHamming Loss0.14—Unverified
4Neural Classifier ChainsHamming Loss0.13—Unverified
5Neural Binary RelevanceHamming Loss0.11—Unverified
#ModelMetricClaimedVerifiedStatus
1Mozafari et al., 2019AAA50.94—Unverified
2SVMAAA46.51—Unverified
3Kennedy et al., 2020AAA45.5—Unverified
#ModelMetricClaimedVerifiedStatus
1HateBERTMacro F10.74—Unverified
2BERTMacro F10.72—Unverified
#ModelMetricClaimedVerifiedStatus
1mBertAccuracy0.83—Unverified
2Logistic RegressionAccuracy0.7—Unverified
#ModelMetricClaimedVerifiedStatus
1HXP + CLAP + CLIPTEST F1 (macro)0.85—Unverified
2BERT + ViT + MFCCTEST F1 (macro)0.79—Unverified
#ModelMetricClaimedVerifiedStatus
1HateBERTMacro F10.49—Unverified
2BERTMacro F10.48—Unverified
#ModelMetricClaimedVerifiedStatus
1HateBERTMacro F10.81—Unverified
2BERTMacro F10.8—Unverified
#ModelMetricClaimedVerifiedStatus
1Multilingual BERTF1-score0.75—Unverified
2AutoMLF1-score0.74—Unverified
#ModelMetricClaimedVerifiedStatus
1AOM mBERTF10.85—Unverified
#ModelMetricClaimedVerifiedStatus
1BaselineF10.7—Unverified
#ModelMetricClaimedVerifiedStatus
1RoBERTa-large-STMacro F180.7—Unverified
#ModelMetricClaimedVerifiedStatus
1Baseline BERT (task A)F10.77—Unverified