SOTAVerified

Text Generation

Text Generation is the task of generating text with the goal of appearing indistinguishable to human-written text. This task is more formally known as "natural language generation" in the literature.

Text generation can be addressed with Markov processes or deep generative models like LSTMs. Recently, some of the most advanced methods for text generation include BART, GPT and other GAN-based approaches. Text generation systems are evaluated either through human ratings or automatic evaluation metrics like METEOR, ROUGE, and BLEU.

Further readings:

( Image credit: Adversarial Ranking for Language Generation )

Papers

Showing 16011650 of 5335 papers

TitleStatusHype
Text as Image: Learning Transferable Adapter for Multi-Label Classification0
Prompt Highlighter: Interactive Control for Multi-Modal LLMsCode1
Efficient Large Language Models: A SurveyCode3
Measuring Misogyny in Natural Language Generation: Preliminary Results from a Case Study on two Reddit Communities0
Mitigating Open-Vocabulary Caption HallucinationsCode1
Breast Ultrasound Report Generation using LangChain0
Compositional Generalization for Data-to-Text Generation0
E4SRec: An Elegant Effective Efficient Extensible Solution of Large Language Models for Sequential RecommendationCode1
Prompt Optimization via Adversarial In-Context LearningCode1
A Survey on Large Language Model (LLM) Security and Privacy: The Good, the Bad, and the Ugly0
Mitigating Fine-Grained Hallucination by Fine-Tuning Large Vision-Language Models with Caption RewritesCode1
Hot PATE: Private Aggregation of Distributions for Diverse Task0
NLEBench+NorGLM: A Comprehensive Empirical Analysis and Benchmark Dataset for Generative Language Models in NorwegianCode0
Unsupervised Approach to Evaluate Sentence-Level Fluency: Do We Really Need Reference?Code0
TextGenSHAP: Scalable Post-hoc Explanations in Text Generation with Long Documents0
Bootstrapping Interactive Image-Text Alignment for Remote Sensing Image CaptioningCode1
FFT: Towards Harmlessness Evaluation and Analysis for LLMs with Factuality, Fairness, ToxicityCode0
LMRL Gym: Benchmarks for Multi-Turn Reinforcement Learning with Language ModelsCode1
LayerCollapse: Adaptive compression of neural networks0
MM-Narrator: Narrating Long-form Videos with Multimodal In-Context Learning0
Reinforcement Replaces Supervision: Query focused Summarization using Deep Reinforcement LearningCode0
A Benchmark for Evaluating Machine Translation Metrics on Dialects Without Standard OrthographyCode1
Ascle: A Python Natural Language Processing Toolkit for Medical Text GenerationCode1
Safe-CLIP: Removing NSFW Concepts from Vision-and-Language ModelsCode1
LLMGA: Multimodal Large Language Model based Generation AssistantCode2
Optimizing and Fine-tuning Large Language Model for Urban Renewal0
Data Generation for Post-OCR correction of Cyrillic handwritingCode1
Towards Vision Enhancing LLMs: Empowering Multimodal Knowledge Storage and Sharing in LLMs0
Large Language Models in Law: A Survey0
Local Convergence of Approximate Newton Method for Two Layer Nonlinear Regression0
UHGEval: Benchmarking the Hallucination of Chinese Large Language Models via Unconstrained GenerationCode1
Faster Minimum Bayes Risk Decoding with Confidence-based PruningCode0
Data-to-Text Bilingual GenerationCode1
DP-NMT: Scalable Differentially-Private Machine TranslationCode0
Controlled Text Generation via Language Model ArithmeticCode2
CMed-GPT: Prompt Tuning for Entity-Aware Chinese Medical Dialogue Generation0
Large Language Models as Topological Structure Enhancers for Text-Attributed Graphs0
Rethinking Radiology Report Generation via Causal Inspired Counterfactual Augmentation0
Evaluation Metrics of Language Generation Models for Synthetic Traffic Generation Tasks0
Beyond Turing: A Comparative Analysis of Approaches for Detecting Machine-Generated Text0
Causal ATE Mitigates Unintended Bias in Controlled Text GenerationCode0
A Self-enhancement Approach for Domain-specific Chatbot Training via Knowledge Mining and Digest0
Advancements in Generative AI: A Comprehensive Review of GANs, GPT, Autoencoders, Diffusion Model, and Transformers0
Can Language Model Moderators Improve the Health of Online Discourse?0
Language Generation from Brain RecordingsCode1
PixT3: Pixel-based Table-To-Text GenerationCode0
Generative AI for Hate Speech Detection: Evaluation and Findings0
The Curious Decline of Linguistic Diversity: Training Language Models on Synthetic TextCode0
WatME: Towards Lossless Watermarking Through Lexical Redundancy0
AMRFact: Enhancing Summarization Factuality Evaluation with AMR-Driven Negative Samples GenerationCode0
Show:102550
← PrevPage 33 of 107Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1T5B BaselineBLEU48.74Unverified
2FactT5BBLEU48.37Unverified
3JointGT BaselineBLEU47.51Unverified
4FactJointGTBLEU47.39Unverified
5Control Prefixes (T5-large)METEOR0.41Unverified
6T5METEOR0.12Unverified
7BARTMETEOR0.11Unverified
#ModelMetricClaimedVerifiedStatus
1LeakGANBLEU-20.95Unverified
2partGANBLEU-20.91Unverified
3RankGANBLEU-20.85Unverified
4RelGAN (100)BLEU-20.85Unverified
5SeqGANBLEU-20.83Unverified
#ModelMetricClaimedVerifiedStatus
1LeakGANBLEU-20.96Unverified
2PPOGANBLEU-20.91Unverified
3RelGANBLEU-20.88Unverified
4SeqGANBLEU-20.86Unverified
5RankGANBLEU-20.78Unverified
#ModelMetricClaimedVerifiedStatus
1UniCRSDistinct-30.65Unverified
2CRFRDistinct-30.52Unverified
3KGSFDistinct-30.43Unverified
4C2CRSDistinct-30.33Unverified
5KBRDDistinct-30.3Unverified
#ModelMetricClaimedVerifiedStatus
1UniLMCIDEr14.92Unverified
2BART (TextBox 2.0)CIDEr12.98Unverified
3BARTMETEOR0.3Unverified
4T5METEOR0.29Unverified
#ModelMetricClaimedVerifiedStatus
1Beam search + A*esque (beam)BLEU-134.4Unverified
2Beam search + A*esque (sample)BLEU-134.4Unverified
3Beam search + A*esque (greedy)BLEU-134.3Unverified
4Beam searchBLEU-133.7Unverified
#ModelMetricClaimedVerifiedStatus
1RankGANBLEU-20.81Unverified
2SeqGANBLEU-20.74Unverified
3LeakGANBLEU-20.46Unverified
#ModelMetricClaimedVerifiedStatus
1TGen++METEOR0.17Unverified
2TGenMETEOR0.15Unverified
3TGen+METEOR0.15Unverified
#ModelMetricClaimedVerifiedStatus
1GPT2-124Meval_loss3.12Unverified
2GPT2-81M-LOOPeval_loss3.11Unverified
3GPT2-Hermiteeval_loss2.91Unverified
#ModelMetricClaimedVerifiedStatus
1LLaMA-65B+CFG (zero-shot)Accuracy96.6Unverified
2LLaMA-30B+CFG (zero-shot)Accuracy96.4Unverified
3LLaMA-13B+CFG (zero-shot)Accuracy95.1Unverified
#ModelMetricClaimedVerifiedStatus
1CNN-VAENLL332.1Unverified
2SA-VAENLL327.5Unverified
3Aggressive VAENLL326.7Unverified
#ModelMetricClaimedVerifiedStatus
1BART (TextBox 2.0)BLEU-410.2Unverified
#ModelMetricClaimedVerifiedStatus
1STWGAN-GPBLEU-30.62Unverified
#ModelMetricClaimedVerifiedStatus
1PALMROUGE-L41.41Unverified
#ModelMetricClaimedVerifiedStatus
1BART (TextBox 2.0)ROUGE-L64.34Unverified
#ModelMetricClaimedVerifiedStatus
1AEM+AttentionBLEU-114.17Unverified
#ModelMetricClaimedVerifiedStatus
1GPT-4ASR65.1Unverified
#ModelMetricClaimedVerifiedStatus
1BART (TextBox 2.0)ROUGE-L42.96Unverified
#ModelMetricClaimedVerifiedStatus
1Graph2SeqBLEU22Unverified
#ModelMetricClaimedVerifiedStatus
1WGANGP + DGflowJS-40.19Unverified