SOTAVerified

Abuse Detection

Abuse detection is the task of identifying abusive behaviors, such as hate speech, offensive language, sexism and racism, in utterances from social media platforms (Source: https://arxiv.org/abs/1802.00385).

Papers

Showing 1–50 of 73 papers

TitleStatusHype
Entropy-based Attention Regularization Frees Unintended Bias Mitigation from ListsCode1
ConvAbuse: Data, Analysis, and Benchmarks for Nuanced Abuse Detection in Conversational AICode1
AbuseAnalyzer: Abuse Detection, Severity and Target Prediction for Gab PostsCode1
KUISAIL at SemEval-2020 Task 12: BERT-CNN for Offensive Speech Identification in Social MediaCode1
Intersectional Bias in Hate Speech and Abusive Language DatasetsCode1
Multimodal Meme Dataset (MultiOFF) for Identifying Offensive Content in Image and TextCode1
Kungfupanda at SemEval-2020 Task 12: BERT-Based Multi-Task Learning for Offensive Language DetectionCode1
Comparative Studies of Detecting Abusive Language on TwitterCode1
One-step and Two-step Classification for Abusive Language Detection on TwitterCode1
Creating and Evaluating Code-Mixed Nepali-English and Telugu-English Datasets for Abusive Language Detection Using Traditional and Deep Learning Models—0
Predictive Response Optimization: Using Reinforcement Learning to Fight Online Social Network Abuse—0
A survey of textual cyber abuse detection using cutting-edge language models and large language models—0
HP-BERT: A framework for longitudinal study of Hinduphobia on social media via LLMsCode0
Towards Cross-Lingual Audio Abuse Detection in Low-Resource Settings with Few-Shot LearningCode0
DetoxBench: Benchmarking Large Language Models for Multitask Fraud & Abuse Detection—0
CoLLAB: A Collaborative Approach for Multilingual Abuse Detection—0
Breaking the Silence Detecting and Mitigating Gendered Abuse in Hindi, Tamil, and Indian English Online SpacesCode0
Overview of the 2023 ICON Shared Task on Gendered Abuse Detection in Indic Languages—0
Voucher Abuse Detection with Prompt-based Fine-tuning on Graph Neural Networks—0
Detection of Children Abuse by Voice and Audio Classification by Short-Time Fourier Transform Machine Learning implemented on Nvidia Edge GPU device—0
TCAB: A Large-Scale Text Classification Attack BenchmarkCode0
Machine Generated Text: A Comprehensive Survey of Threat Models and Detection Methods—0
Explainable Abuse Detection as Intent Classification and Slot FillingCode0
Adversarial Robustness for Tabular Data through Cost and Utility Awareness—0
Enriching Abusive Language Detection with Community Context—0
DE-ABUSE@TamilNLP-ACL 2022: Transliteration as Data Augmentation for Abuse Detection in Tamil—0
Darkness can not drive out darkness: Investigating Bias in Hate SpeechDetection Models—0
Improving Generalizability in Implicitly Abusive Language Detection with Concept Activation VectorsCode0
Multilingual and Multimodal Abuse Detection—0
The Online Behaviour of the Algerian Abusers in Social Media Networks—0
Abuse and Fraud Detection in Streaming Services Using Heuristic-Aware Machine Learning—0
ADIMA: Abuse Detection In Multilingual AudioCode0
Identifying Adversarial Attacks on Text Classifiers—0
Toxicity Detection for Indic Multilingual Social Media Content—0
Improving Generalizability in Implicitly Abusive Language Detection with Concept Activation Vectors—0
What Models Know About Their Attackers: Deriving Attacker Information From Latent Representations—0
A Large-Scale English Multi-Label Twitter Dataset for Cyberbullying and Online Abuse Detection—0
AAA: Fair Evaluation for Abuse Detection Systems WantedCode0
Generalisability of Topic Models in Cross-corpora Abusive Language Detection—0
Modeling Users and Online Communities for Abuse Detection: A Position on Ethics and Explainability—0
Confronting Abusive Language Online: A Survey from the Ethical and Human Rights Perspective—0
UoB at SemEval-2020 Task 12: Boosting BERT with Corpus Level Information—0
Joint Modelling of Emotion and Abusive Language Detection—0
Evaluating Performance of an Adult Pornography Classifier for Child Sexual Abuse Detection—0
LIIR at SemEval-2020 Task 12: A Cross-Lingual Augmentation Approach for Multilingual Offensive Language Identification—0
WAC: A Corpus of Wikipedia Conversations for Online Abuse DetectionCode0
Stereotypical Bias Removal for Hate Speech Detection Task using Knowledge-based Generalizations—0
HateMonitors: Language Agnostic Abuse Detection in Social MediaCode0
Tackling Online Abuse: A Survey of Automated Abuse Detection Methods—0
Challenges and frontiers in abusive content detectionCode0
Show:102550
← PrevPage 1 of 2Next →

No leaderboard results yet.