SOTAVerified

Binary Classification

Papers

Showing 1–50 of 2574 papers

TitleStatusHype
An Automated Classifier of Harmful Brain Activities for Clinical Usage Based on a Vision-Inspired Pre-trained Framework—0
DDL: A Dataset for Interpretable Deepfake Detection and Localization in Real-World Scenarios—0
Inverse Scene Text RemovalCode0
Divide, Specialize, and Route: A New Approach to Efficient Ensemble Learning—0
Private Model Personalization Revisited—0
Exploring Strategies for Personalized Radiation Therapy Part I Unlocking Response-Related Tumor Subregions with Class Activation Mapping—0
I Know Which LLM Wrote Your Code Last Summer: LLM generated Code Stylometry for Authorship Attribution—0
Universal Rates of ERM for Agnostic Learning—0
Detecting immune cells with label-free two-photon autofluorescence and deep learning—0
LLM-Powered Intent-Based Categorization of Phishing Emails—0
On the existence of consistent adversarial attacks in high-dimensional linear classification—0
Optimization of bi-directional gated loop cell based on multi-head attention mechanism for SSD health state classification model—0
Adversarial Surrogate Risk Bounds for Binary Classification—0
DEAL: Disentangling Transformer Head Activations for LLM Steering—0
DIsoN: Decentralized Isolation Networks for Out-of-Distribution Detection in Medical Imaging—0
EDINET-Bench: Evaluating LLMs on Complex Financial Tasks using Japanese Financial StatementsCode1
Zeroth-Order Optimization Finds Flat Minima—0
CHANCERY: Evaluating Corporate Governance Reasoning Capabilities in Language Models—0
Does Prompt Design Impact Quality of Data Imputation by LLMs?—0
Quantum Ensembling Methods for Healthcare and Life Science—0
Trade-offs in Data Memorization via Strong Data Processing Inequalities—0
Localized Forest Fire Risk Prediction: A Department-Aware Approach for Operational Decision Support—0
LGAR: Zero-Shot LLM-Guided Neural Ranking for Abstract Screening in Systematic Literature ReviewsCode0
PatchDEMUX: A Certifiably Robust Framework for Multi-label Classifiers Against Adversarial PatchesCode0
Hidden Persuasion: Detecting Manipulative Narratives on Social Media During the 2022 Russian Invasion of Ukraine—0
The Rich and the Simple: On the Implicit Bias of Adam and SGD—0
Individualised Counterfactual Examples Using Conformal Prediction Intervals—0
AbsoluteNet: A Deep Learning Neural Network to Classify Cerebral Hemodynamic Responses of Auditory Processing—0
Leveraging Cascaded Binary Classification and Multimodal Fusion for Dementia Detection through Spontaneous Speech—0
Token-level Accept or Reject: A Micro Alignment Approach for Large Language ModelsCode0
Revolutionizing Wildfire Detection with Convolutional Neural Networks: A VGG16 Model Approach—0
Rhapsody: A Dataset for Highlight Detection in PodcastsCode0
A Smart Healthcare System for Monkeypox Skin Lesion Detection and Tracking—0
Mechanical in-sensor computing: a programmable meta-sensor for structural damage classification without external electronic power—0
Debate-to-Detect: Reformulating Misinformation Detection as a Real-World Debate with Large Language Models—0
Anatomy-Guided Multitask Learning for MRI-Based Classification of Placenta Accreta Spectrum and its Subtypes—0
Attention with Trained Embeddings Provably Selects Important Tokens—0
Self-Classification Enhancement and Correction for Weakly Supervised Object Detection—0
Self-Boost via Optimal Retraining: An Analysis via Approximate Message Passing—0
Adaptive Estimation and Learning under Temporal Distribution Shift—0
XDementNET: An Explainable Attention Based Deep Convolutional Network to Detect Alzheimer Progression from MRI dataCode0
CSAGC-IDS: A Dual-Module Deep Learning Network Intrusion Detection Model for Complex and Imbalanced Data—0
Safety Alignment Can Be Not Superficial With Explicit Safety Signals—0
BusterX: MLLM-Powered AI-Generated Video Forgery Detection and ExplanationCode1
VenusX: Unlocking Fine-Grained Functional Understanding of ProteinsCode1
SoftPQ: Robust Instance Segmentation Evaluation via Soft Matching and Tunable ThresholdsCode0
Deepfake Forensic Analysis: Source Dataset Attribution and Legal Implications of Synthetic Media Manipulation—0
Comparing LLM Text Annotation Skills: A Study on Human Rights Violations in Social Media DataCode0
Zero-Shot Multi-modal Large Language Model v.s. Supervised Deep Learning: A Comparative Study on CT-Based Intracranial Hemorrhage SubtypingCode0
Crowd Scene Analysis using Deep Learning Techniques—0
Show:102550
← PrevPage 1 of 52Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1Trompt + OpenAI embeddingAUROC0.98—Unverified
2LightGBM + OpenAI embeddingAUROC0.97—Unverified
3FTTransformer + RoBERTa fintuneAUROC0.96—Unverified
4LightGBM + RoBERTa embeddingAUROC0.95—Unverified
5FTTransformer + RoBERTa embeddingAUROC0.94—Unverified
6ResNet + RoBERTa embeddingAUROC0.93—Unverified
7ResNet + OpenAI embeddingAUROC0.92—Unverified
8FTTransformer + OpenAI embeddingAUROC0.91—Unverified
#ModelMetricClaimedVerifiedStatus
1Trompt + OpenAI embeddingAUROC0.81—Unverified
2Multimodal-Net All-TextAUROC0.8—Unverified
3ResNet + RoBERTa finetuneAUROC0.79—Unverified
4LightGBM + RoBERTa embeddingAUROC0.77—Unverified
#ModelMetricClaimedVerifiedStatus
1Attention based CNNF1 score0.93—Unverified
#ModelMetricClaimedVerifiedStatus
1XGBoostF1-Score98.79—Unverified