SOTAVerified

Masked Language Modeling

Papers

Showing 201–250 of 475 papers

TitleStatusHype
Embracing Ambiguity: Improving Similarity-oriented Tasks with Contextual Synonym Knowledge—0
Emerging Cross-lingual Structure in Pretrained Language Models—0
Emerging Property of Masked Token for Effective Pre-training—0
Enabling Autoregressive Models to Fill In Masked Tokens—0
Enhancing BERT-Based Visual Question Answering through Keyword-Driven Sentence Selection—0
Enhancing Domain-Specific Encoder Models with LLM-Generated Data: How to Leverage Ontologies, and How to Do Without Them—0
ESALE: Enhancing Code-Summary Alignment Learning for Source Code Summarization—0
Exploring Multi-Modal Contextual Knowledge for Open-Vocabulary Object Detection—0
Exposing the Implicit Energy Networks behind Masked Language Models via Metropolis--Hastings—0
Extrapolating Multilingual Understanding Models as Multilingual Generators—0
FARM: Functional Group-Aware Representations for Small Molecules—0
How Useful is Continued Pre-Training for Generative Unsupervised Domain Adaptation?—0
Foundation Posteriors for Approximate Probabilistic Inference—0
Gender-tuning: Empowering Fine-tuning for Debiasing Pre-trained Language Models—0
General Framework for Reversible Data Hiding in Texts Based on Masked Language Modeling—0
Generating multiple-choice questions for medical question answering with distractors and cue-masking—0
Generative Prompt Tuning for Relation Classification—0
GeoRecon: Graph-Level Representation Learning for 3D Molecules via Reconstruction-Based Pretraining—0
Global memory transformer for processing long documents—0
Go-tuning: Improving Zero-shot Learning Abilities of Smaller Language Models—0
GPTs at Factify 2022: Prompt Aided Fact-Verification—0
GraphCodeBERT: Pre-training Code Representations with Data Flow—0
HAD: Hybrid Architecture Distillation Outperforms Teacher in Genomic Sequence Modeling—0
HCDIR: End-to-end Hate Context Detection, and Intensity Reduction model for online comments—0
HOP+: History-enhanced and Order-aware Pre-training for Vision-and-Language Navigation—0
How does the pre-training objective affect what large language models learn about linguistic properties?—0
Image as a Foreign Language: BEiT Pretraining for Vision and Vision-Language Tasks—0
ImageBERT: Cross-modal Pre-training with Large-scale Weak-supervised Image-Text Data—0
Image BERT Pre-training with Online Tokenizer—0
Improving BERT with Hybrid Pooling Network and Drop Mask—0
Improving Low-Resource Morphological Inflection via Self-Supervised Objectives—0
Improving the Reusability of Pre-trained Language Models in Real-world Applications—0
In-Context Learning can distort the relationship between sequence likelihoods and biological fitness—0
Investigating Masking-based Data Generation in Language Models—0
"Is Whole Word Masking Always Better for Chinese BERT?": Probing on Chinese Grammatical Error Correction—0
"Is Whole Word Masking Always Better for Chinese BERT?": Probing on Chinese Grammatical Error Correction—0
“Is Whole Word Masking Always Better for Chinese BERT?”: Probing on Chinese Grammatical Error Correction—0
Iterative Mask Filling: An Effective Text Augmentation Method Using Masked Language Modeling—0
Joint unsupervised and supervised learning for context-aware language identification—0
Joint Unsupervised and Supervised Training for Multilingual ASR—0
KECP: Knowledge Enhanced Contrastive Prompting for Few-shot Extractive Question Answering—0
Knowing Where to Focus: Attention-Guided Alignment for Text-based Person Search—0
Knowledgeable Prompt-tuning: Incorporating Knowledge into Prompt Verbalizer for Text Classification—0
Knowledge Distillation vs. Pretraining from Scratch under a Fixed (Computation) Budget—0
KUL@SMM4H’22: Template Augmented Adaptive Pre-training for Tweet Classification—0
LakotaBERT: A Transformer-based Model for Low Resource Lakota Language—0
LAnoBERT: System Log Anomaly Detection based on BERT Masked Language Model—0
Larger-Scale Transformers for Multilingual Masked Language Modeling—0
LayoutMask: Enhance Text-Layout Interaction in Multi-modal Pre-training for Document Understanding—0
Enhancing Continual Learning with Global Prototypes: Counteracting Negative Representation Drift—0
Show:102550
← PrevPage 5 of 10Next →

No leaderboard results yet.