SOTAVerified

Text Generation

Text Generation is the task of generating text with the goal of appearing indistinguishable to human-written text. This task is more formally known as "natural language generation" in the literature.

Text generation can be addressed with Markov processes or deep generative models like LSTMs. Recently, some of the most advanced methods for text generation include BART, GPT and other GAN-based approaches. Text generation systems are evaluated either through human ratings or automatic evaluation metrics like METEOR, ROUGE, and BLEU.

Further readings:

( Image credit: Adversarial Ranking for Language Generation )

Papers

Showing 201250 of 5335 papers

TitleStatusHype
Building Cooperative Embodied Agents Modularly with Large Language ModelsCode2
Fine-Grained Human Feedback Gives Better Rewards for Language Model TrainingCode2
From Pixels to Graphs: Open-Vocabulary Scene Graph Generation with Vision-Language ModelsCode2
Authorship Obfuscation in Multilingual Machine-Generated Text DetectionCode2
Inseq: An Interpretability Toolkit for Sequence Generation ModelsCode2
Retrieval Augmented Generation Evaluation in the Era of Large Language Models: A Comprehensive SurveyCode2
Expressive Text-to-Image Generation with Rich TextCode2
Evolutionary Computation in the Era of Large Language Model: Survey and RoadmapCode2
FActScore: Fine-grained Atomic Evaluation of Factual Precision in Long Form Text GenerationCode2
Evaluating Morphological Compositional Generalization in Large Language ModelsCode2
Shadowcast: Stealthy Data Poisoning Attacks Against Vision-Language ModelsCode2
Balancing LoRA Performance and Efficiency with Simple Shard SharingCode2
Efficient Minimum Bayes Risk Decoding using Low-Rank Matrix Completion AlgorithmsCode2
eVAE: Evolutionary Variational AutoencoderCode2
AutoPatent: A Multi-Agent Framework for Automatic Patent GenerationCode2
Steering Large Language Models between Code Execution and Textual ReasoningCode2
Symphony Generation with Permutation Invariant Language ModelCode2
Ecco: An Open Source Library for the Explainability of Transformer Language ModelsCode2
DriVLMe: Enhancing LLM-based Autonomous Driving Agents with Embodied and Social ExperiencesCode2
ECG-Chat: A Large ECG-Language Model for Cardiac Disease DiagnosisCode2
TeenyTinyLlama: open-source tiny language models trained in Brazilian PortugueseCode2
Texar: A Modularized, Versatile, and Extensible Toolkit for Text GenerationCode2
DiffusionBERT: Improving Generative Masked Language Models with Diffusion ModelsCode2
DiffusionPen: Towards Controlling the Style of Handwritten Text GenerationCode2
CoNT: Contrastive Neural Text GenerationCode2
The Language Interpretability Tool: Extensible, Interactive Visualizations and Analysis for NLP ModelsCode2
DiffuSeq: Sequence to Sequence Text Generation with Diffusion ModelsCode2
DRAGIN: Dynamic Retrieval Augmented Generation based on the Information Needs of Large Language ModelsCode2
Towards a Unified Multi-Dimensional Evaluator for Text GenerationCode2
IDGenRec: LLM-RecSys Alignment with Textual ID LearningCode2
Few-Shot Text Generation with Pattern-Exploiting TrainingCode2
Intelligent Artistic Typography: A Comprehensive Review of Artistic Text Design and GenerationCode2
BERTScore: Evaluating Text Generation with BERTCode1
BERTGEN: Multi-task Generation through BERTCode1
Dependency-based Mixture Language ModelsCode1
Defending Against Backdoor Attacks in Natural Language GenerationCode1
BERTScore is Unfair: On Social Bias in Language Model-Based Metrics for Text GenerationCode1
Defending Against Unforeseen Failure Modes with Latent Adversarial TrainingCode1
DeTiME: Diffusion-Enhanced Topic Modeling using Encoder-decoder based LLMCode1
Adaptive Markup Language Generation for Contextually-Grounded Visual Document UnderstandingCode1
Benchmarking Large Language Models on Controllable Generation under Diversified InstructionsCode1
Adaptive Machine Translation with Large Language ModelsCode1
Deep Graph Convolutional Encoders for Structured Data to Text GenerationCode1
Data-to-text Generation with Macro PlanningCode1
BenchCLAMP: A Benchmark for Evaluating Language Models on Syntactic and Semantic ParsingCode1
A Knowledge-Enhanced Pretraining Model for Commonsense Story GenerationCode1
A Langevin-like Sampler for Discrete DistributionsCode1
Data-to-text Generation with Variational Sequential PlanningCode1
Data-to-Text Bilingual GenerationCode1
Data-to-Text Generation with Iterative Text EditingCode1
Show:102550
← PrevPage 5 of 107Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1T5B BaselineBLEU48.74Unverified
2FactT5BBLEU48.37Unverified
3JointGT BaselineBLEU47.51Unverified
4FactJointGTBLEU47.39Unverified
5Control Prefixes (T5-large)METEOR0.41Unverified
6T5METEOR0.12Unverified
7BARTMETEOR0.11Unverified
#ModelMetricClaimedVerifiedStatus
1LeakGANBLEU-20.95Unverified
2partGANBLEU-20.91Unverified
3RankGANBLEU-20.85Unverified
4RelGAN (100)BLEU-20.85Unverified
5SeqGANBLEU-20.83Unverified
#ModelMetricClaimedVerifiedStatus
1LeakGANBLEU-20.96Unverified
2PPOGANBLEU-20.91Unverified
3RelGANBLEU-20.88Unverified
4SeqGANBLEU-20.86Unverified
5RankGANBLEU-20.78Unverified
#ModelMetricClaimedVerifiedStatus
1UniCRSDistinct-30.65Unverified
2CRFRDistinct-30.52Unverified
3KGSFDistinct-30.43Unverified
4C2CRSDistinct-30.33Unverified
5KBRDDistinct-30.3Unverified
#ModelMetricClaimedVerifiedStatus
1UniLMCIDEr14.92Unverified
2BART (TextBox 2.0)CIDEr12.98Unverified
3BARTMETEOR0.3Unverified
4T5METEOR0.29Unverified
#ModelMetricClaimedVerifiedStatus
1Beam search + A*esque (beam)BLEU-134.4Unverified
2Beam search + A*esque (sample)BLEU-134.4Unverified
3Beam search + A*esque (greedy)BLEU-134.3Unverified
4Beam searchBLEU-133.7Unverified
#ModelMetricClaimedVerifiedStatus
1RankGANBLEU-20.81Unverified
2SeqGANBLEU-20.74Unverified
3LeakGANBLEU-20.46Unverified
#ModelMetricClaimedVerifiedStatus
1TGen++METEOR0.17Unverified
2TGenMETEOR0.15Unverified
3TGen+METEOR0.15Unverified
#ModelMetricClaimedVerifiedStatus
1GPT2-124Meval_loss3.12Unverified
2GPT2-81M-LOOPeval_loss3.11Unverified
3GPT2-Hermiteeval_loss2.91Unverified
#ModelMetricClaimedVerifiedStatus
1LLaMA-65B+CFG (zero-shot)Accuracy96.6Unverified
2LLaMA-30B+CFG (zero-shot)Accuracy96.4Unverified
3LLaMA-13B+CFG (zero-shot)Accuracy95.1Unverified
#ModelMetricClaimedVerifiedStatus
1CNN-VAENLL332.1Unverified
2SA-VAENLL327.5Unverified
3Aggressive VAENLL326.7Unverified
#ModelMetricClaimedVerifiedStatus
1BART (TextBox 2.0)BLEU-410.2Unverified
#ModelMetricClaimedVerifiedStatus
1STWGAN-GPBLEU-30.62Unverified
#ModelMetricClaimedVerifiedStatus
1PALMROUGE-L41.41Unverified
#ModelMetricClaimedVerifiedStatus
1BART (TextBox 2.0)ROUGE-L64.34Unverified
#ModelMetricClaimedVerifiedStatus
1AEM+AttentionBLEU-114.17Unverified
#ModelMetricClaimedVerifiedStatus
1GPT-4ASR65.1Unverified
#ModelMetricClaimedVerifiedStatus
1BART (TextBox 2.0)ROUGE-L42.96Unverified
#ModelMetricClaimedVerifiedStatus
1Graph2SeqBLEU22Unverified
#ModelMetricClaimedVerifiedStatus
1WGANGP + DGflowJS-40.19Unverified