SOTAVerified

Text Generation

Text Generation is the task of generating text with the goal of appearing indistinguishable to human-written text. This task is more formally known as "natural language generation" in the literature.

Text generation can be addressed with Markov processes or deep generative models like LSTMs. Recently, some of the most advanced methods for text generation include BART, GPT and other GAN-based approaches. Text generation systems are evaluated either through human ratings or automatic evaluation metrics like METEOR, ROUGE, and BLEU.

Further readings:

( Image credit: Adversarial Ranking for Language Generation )

Papers

Showing 24012450 of 5335 papers

TitleStatusHype
Evaluating Computational Language Models with Scaling Properties of Natural Language0
Evaluating Compact LLMs for Zero-Shot Iberian Language Tasks on End-User Devices0
CDM: Combining Extraction and Generation for Definition Modeling0
ARGUS: Hallucination and Omission Evaluation in Video-LLMs0
A Glimpse in ChatGPT Capabilities and its impact for AI research0
Adapting Pre-trained Generative Models for Extractive Question Answering0
Absformer: Transformer-based Model for Unsupervised Multi-Document Abstractive Summarization0
Evaluating an Automata Approach to Query Containment0
Evaluating a Dynamic Time Warping Based Scoring Algorithm for Facial Expressions in ASL Animations0
CCPrefix: Counterfactual Contrastive Prefix-Tuning for Many-Class Classification0
Eval all, trust a few, do wrong to none: Comparing sentence generation models0
Argument Summarization and its Evaluation in the Era of Large Language Models0
Ev2R: Evaluating Evidence Retrieval in Automated Fact-Checking0
CBAG: Conditional Biomedical Abstract Generation0
Estimating Subjective Crowd-Evaluations as an Additional Objective to Improve Natural Language Generation0
CAVALRY-V: A Large-Scale Generator Framework for Adversarial Attacks on Video MLLMs0
Argument linking in LTAG: A constraint-based implementation with XMG0
Aggregation methods for efficient collocation detection0
Error Norm Truncation: Robust Training in the Presence of Data Noise for Text Generation Models0
Error-Correcting Neural Sequence Prediction0
Error-Correcting Codes For Approximate Neural Sequence Prediction0
CAT-Gen: Improving Robustness in NLP Models via Controlled Adversarial Text Generation0
Adapting Neural Single-Document Summarization Model for Abstractive Multi-Document Summarization: A Pilot Study0
Equi-Tuning: Group Equivariant Fine-Tuning of Pretrained Models0
EoRA: Training-free Compensation for Compressed LLM with Eigenspace Low-Rank Approximation0
ENTRUST: Argument Reframing with Language Models and Entailment0
Category-Driven Content Selection0
Are You Robert or RoBERTa? Deceiving Online Authorship Attribution Models Using Neural Text Generators0
Entropy-UID: A Method for Optimizing Information Density0
Entropy optimized semi-supervised decomposed vector-quantized variational autoencoder model based on transfer learning for multiclass text classification and generation0
Categorical SDEs with Simplex Diffusion0
Entropy-Guided Watermarking for LLMs: A Test-Time Framework for Robust and Traceable Text Generation0
Entity-Based Semantic Adequacy for Data-to-Text Generation0
Cat and Mouse -- Can Fake Text Generation Outpace Detector Systems?0
A Review on Large Language Models for Visual Analytics0
Entity-based De-noising Modeling for Controllable Dialogue Summarization0
Entertainment chatbot for the digital inclusion of elderly people without abstraction capabilities0
Entity-to-Text based Data Augmentation for various Named Entity Recognition Tasks0
CaseSummarizer: A System for Automated Summarization of Legal Texts0
A Review of Digital Learning Environments for Teaching Natural Language Processing in K-12 Education0
Ensembles and Cocktails: Robust Finetuning for Natural Language Generation0
Ensemble Learning for Large Language Models in Text and Code Generation: A Survey0
Case Relation Transformer: A Crossmodal Language Generation Model for Fetching Instructions0
Are the Tools up to the Task? an Evaluation of Commercial Dialog Tools in Developing Conversational Enterprise-grade Dialog Systems0
A Generative Model of Vector Space Semantics0
Adapting LLMs for Efficient Context Processing through Soft Prompt Compression0
Enriching and Controlling Global Semantics for Text Summarization0
Enhancing Visual Reliance in Text Generation: A Bayesian Perspective on Mitigating Hallucination in Large Vision-Language Models0
Enhancing Variational Autoencoders with Mutual Information Neural Estimation for Text Generation0
Enhancing Trust in Large Language Models with Uncertainty-Aware Fine-Tuning0
Show:102550
← PrevPage 49 of 107Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1T5B BaselineBLEU48.74Unverified
2FactT5BBLEU48.37Unverified
3JointGT BaselineBLEU47.51Unverified
4FactJointGTBLEU47.39Unverified
5Control Prefixes (T5-large)METEOR0.41Unverified
6T5METEOR0.12Unverified
7BARTMETEOR0.11Unverified
#ModelMetricClaimedVerifiedStatus
1LeakGANBLEU-20.95Unverified
2partGANBLEU-20.91Unverified
3RankGANBLEU-20.85Unverified
4RelGAN (100)BLEU-20.85Unverified
5SeqGANBLEU-20.83Unverified
#ModelMetricClaimedVerifiedStatus
1LeakGANBLEU-20.96Unverified
2PPOGANBLEU-20.91Unverified
3RelGANBLEU-20.88Unverified
4SeqGANBLEU-20.86Unverified
5RankGANBLEU-20.78Unverified
#ModelMetricClaimedVerifiedStatus
1UniCRSDistinct-30.65Unverified
2CRFRDistinct-30.52Unverified
3KGSFDistinct-30.43Unverified
4C2CRSDistinct-30.33Unverified
5KBRDDistinct-30.3Unverified
#ModelMetricClaimedVerifiedStatus
1UniLMCIDEr14.92Unverified
2BART (TextBox 2.0)CIDEr12.98Unverified
3BARTMETEOR0.3Unverified
4T5METEOR0.29Unverified
#ModelMetricClaimedVerifiedStatus
1Beam search + A*esque (beam)BLEU-134.4Unverified
2Beam search + A*esque (sample)BLEU-134.4Unverified
3Beam search + A*esque (greedy)BLEU-134.3Unverified
4Beam searchBLEU-133.7Unverified
#ModelMetricClaimedVerifiedStatus
1RankGANBLEU-20.81Unverified
2SeqGANBLEU-20.74Unverified
3LeakGANBLEU-20.46Unverified
#ModelMetricClaimedVerifiedStatus
1TGen++METEOR0.17Unverified
2TGenMETEOR0.15Unverified
3TGen+METEOR0.15Unverified
#ModelMetricClaimedVerifiedStatus
1GPT2-124Meval_loss3.12Unverified
2GPT2-81M-LOOPeval_loss3.11Unverified
3GPT2-Hermiteeval_loss2.91Unverified
#ModelMetricClaimedVerifiedStatus
1LLaMA-65B+CFG (zero-shot)Accuracy96.6Unverified
2LLaMA-30B+CFG (zero-shot)Accuracy96.4Unverified
3LLaMA-13B+CFG (zero-shot)Accuracy95.1Unverified
#ModelMetricClaimedVerifiedStatus
1CNN-VAENLL332.1Unverified
2SA-VAENLL327.5Unverified
3Aggressive VAENLL326.7Unverified
#ModelMetricClaimedVerifiedStatus
1BART (TextBox 2.0)BLEU-410.2Unverified
#ModelMetricClaimedVerifiedStatus
1STWGAN-GPBLEU-30.62Unverified
#ModelMetricClaimedVerifiedStatus
1PALMROUGE-L41.41Unverified
#ModelMetricClaimedVerifiedStatus
1BART (TextBox 2.0)ROUGE-L64.34Unverified
#ModelMetricClaimedVerifiedStatus
1AEM+AttentionBLEU-114.17Unverified
#ModelMetricClaimedVerifiedStatus
1GPT-4ASR65.1Unverified
#ModelMetricClaimedVerifiedStatus
1BART (TextBox 2.0)ROUGE-L42.96Unverified
#ModelMetricClaimedVerifiedStatus
1Graph2SeqBLEU22Unverified
#ModelMetricClaimedVerifiedStatus
1WGANGP + DGflowJS-40.19Unverified