SOTAVerified

Topic Classification

Papers

Showing 1–50 of 186 papers

TitleStatusHype
Prototypical Verbalizer for Prompt-based Few-shot TuningCode4
Language Through a Prism: A Spectral Approach for Multiscale Language RepresentationsCode1
2kenize: Tying Subword Sequences for Chinese Script ConversionCode1
SynthesizRR: Generating Diverse Datasets with Retrieval AugmentationCode1
DocSCAN: Unsupervised Text Classification via Learning from NeighborsCode1
Hierarchical Multi-Label Classification of Scientific DocumentsCode1
Hierarchical Transformers for Long Document ClassificationCode1
MultiEURLEX -- A multi-lingual and multi-label legal document classification dataset for zero-shot cross-lingual transferCode1
SIB-200: A Simple, Inclusive, and Big Evaluation Dataset for Topic Classification in 200+ Languages and DialectsCode1
Cross-Lingual Adaptation using Structural Correspondence LearningCode1
Polyglot Prompt: Multilingual Multitask PrompTrainingCode1
GrEmLIn: A Repository of Green Baseline Embeddings for 87 Low-Resource Languages Injected with Multilingual Graph KnowledgeCode1
HUE: Pretrained Model and Dataset for Understanding Hanja Documents of Ancient KoreaCode1
Explaining NLP Models via Minimal Contrastive Editing (MiCE)Code1
L3Cube-IndicNews: News-based Short Text and Long Document Classification Datasets in Indic LanguagesCode1
KLUE: Korean Language Understanding EvaluationCode1
LexC-Gen: Generating Data for Extremely Low-Resource Languages with Large Language Models and Bilingual LexiconsCode1
Mind Your Outliers! Investigating the Negative Impact of Outliers on Active Learning for Visual Question AnsweringCode1
Newswire: A Large-Scale Structured Database of a Century of Historical NewsCode1
Revisiting LSTM Networks for Semi-Supervised Text Classification via Mixed Objective FunctionCode1
TEMPERA: Test-Time Prompting via Reinforcement LearningCode1
Zero-Shot Text Classification via Self-Supervised TuningCode1
MultiEURLEX - A multi-lingual and multi-label legal document classification dataset for zero-shot cross-lingual transferCode1
Adapting Pre-trained Language Models to African Languages via Multilingual Adaptive Fine-TuningCode1
Entailment as Few-Shot LearnerCode1
MasakhaNEWS: News Topic Classification for African languagesCode1
Label Semantic Aware Pre-training for Few-shot Text ClassificationCode1
In-Context Learning with Iterative Demonstration SelectionCode1
Baselines and Bigrams: Simple, Good Sentiment and Topic Classification—0
Analysis of Policy Agendas: Lessons Learned from Automatic Topic Classification of Croatian Political Texts—0
AWS CORD-19 Search: A Neural Search Engine for COVID-19 Literature—0
Evaluating Pixel Language Models on Non-Standardized Languages—0
Expanding the Text Classification Toolbox with Cross-Lingual Embeddings—0
Few-Shot Cross-Lingual Transfer for Prompting Large Language Models in Low-Resource Languages—0
A Multilingual Bag-of-Entities Model for Zero-Shot Cross-Lingual Text Classification—0
Attention-Enhancing Backdoor Attacks Against BERT-based Models—0
A Multilingual Bag-of-Entities Model for Zero-Shot Cross-Lingual Text Classification—0
A Statistical Theory of Contrastive Learning via Approximate Sufficient Statistics—0
Embracing Error to Enable Rapid Crowdsourcing—0
Estimating Confidence of Predictions of Individual Classifiers and TheirEnsembles for the Genre Classification Task—0
From Measurement Instruments to Data: Leveraging Theory-Driven Synthetic Training Data for Classifying Social Constructs—0
Cross-Lingual Classification of Topics in Political Texts—0
Assessing In-context Learning and Fine-tuning for Topic Classification of German Web Data—0
Co-Training for Topic Classification of Scholarly Data—0
A Soft Contrastive Learning-based Prompt Model for Few-shot Sentiment Analysis—0
CTM - A Model for Large-Scale Multi-View Tweet Topic Classification—0
CTM -- A Model for Large-Scale Multi-View Tweet Topic Classification—0
CTM - A Model for Large-Scale Multi-View Tweet Topic Classification—0
Data Sets: Word Embeddings Learned from Tweets and General Data—0
American Stories: A Large-Scale Structured Text Dataset of Historical U.S. Newspapers—0
Show:102550
← PrevPage 1 of 4Next →

No leaderboard results yet.