SOTAVerified

Text Generation

Text Generation is the task of generating text with the goal of appearing indistinguishable to human-written text. This task is more formally known as "natural language generation" in the literature.

Text generation can be addressed with Markov processes or deep generative models like LSTMs. Recently, some of the most advanced methods for text generation include BART, GPT and other GAN-based approaches. Text generation systems are evaluated either through human ratings or automatic evaluation metrics like METEOR, ROUGE, and BLEU.

Further readings:

( Image credit: Adversarial Ranking for Language Generation )

Papers

Showing 151200 of 5335 papers

TitleStatusHype
LLMGA: Multimodal Large Language Model based Generation AssistantCode2
KoSBi: A Dataset for Mitigating Social Bias Risks Towards Safer Large Language Model ApplicationCode2
Keyformer: KV Cache Reduction through Key Tokens Selection for Efficient Generative InferenceCode2
Language-Driven Representation Learning for RoboticsCode2
In-Context Retrieval-Augmented Language ModelsCode2
Building Cooperative Embodied Agents Modularly with Large Language ModelsCode2
InfiniGen: Efficient Generative Inference of Large Language Models with Dynamic KV Cache ManagementCode2
Language Models Can See: Plugging Visual Controls in Text GenerationCode2
LLM-Inference-Bench: Inference Benchmarking of Large Language Models on AI AcceleratorsCode2
Multimodality for NLP-Centered Applications: Resources, Advances and FrontiersCode2
Improving Factuality and Reasoning in Language Models through Multiagent DebateCode2
In-Context Editing: Learning Knowledge from Self-Induced DistributionsCode2
Inseq: An Interpretability Toolkit for Sequence Generation ModelsCode2
Intelligent Artistic Typography: A Comprehensive Review of Artistic Text Design and GenerationCode2
MiniLLM: Knowledge Distillation of Large Language ModelsCode2
HALC: Object Hallucination Reduction via Adaptive Focal-Contrast DecodingCode2
Hardware-Aware Parallel Prompt Decoding for Memory-Efficient Acceleration of LLM InferenceCode2
Language Models can Self-Lengthen to Generate Long TextsCode2
Harmonizing Visual Text Comprehension and GenerationCode2
Learning Transferable Visual Models From Natural Language SupervisionCode2
GPT-NER: Named Entity Recognition via Large Language ModelsCode2
GlyphControl: Glyph Conditional Control for Visual Text GenerationCode2
GPTScore: Evaluate as You DesireCode2
LlamaFactory: Unified Efficient Fine-Tuning of 100+ Language ModelsCode2
Generative Pretrained Structured Transformers: Unsupervised Syntactic Language Models at ScaleCode2
Grounding Language Models to Images for Multimodal Inputs and OutputsCode2
Fine-Grained Human Feedback Gives Better Rewards for Language Model TrainingCode2
Few-Shot Text Generation with Pattern-Exploiting TrainingCode2
CelebV-Text: A Large-Scale Facial Text-Video DatasetCode2
BatGPT: A Bidirectional Autoregessive Talker from Generative Pre-trained TransformerCode2
BayLing: Bridging Cross-lingual Alignment and Instruction Following through Interactive Translation for Large Language ModelsCode2
Make LoRA Great Again: Boosting LoRA with Adaptive Singular Values and Mixture-of-Experts Optimization AlignmentCode2
FActScore: Fine-grained Atomic Evaluation of Factual Precision in Long Form Text GenerationCode2
Benchmarking Uncertainty Quantification Methods for Large Language Models with LM-PolygraphCode2
Balancing LoRA Performance and Efficiency with Simple Shard SharingCode2
From Pixels to Graphs: Open-Vocabulary Scene Graph Generation with Vision-Language ModelsCode2
A Contrastive Framework for Neural Text GenerationCode2
ChatEval: Towards Better LLM-based Evaluators through Multi-Agent DebateCode2
HelloBench: Evaluating Long Text Generation Capabilities of Large Language ModelsCode2
Get my drift? Catching LLM Task Drift with Activation DeltasCode2
eVAE: Evolutionary Variational AutoencoderCode2
Evaluating Morphological Compositional Generalization in Large Language ModelsCode2
Efficient Minimum Bayes Risk Decoding using Low-Rank Matrix Completion AlgorithmsCode2
Ecco: An Open Source Library for the Explainability of Transformer Language ModelsCode2
ECG-Chat: A Large ECG-Language Model for Cardiac Disease DiagnosisCode2
A Judge-free LLM Open-ended Generation Benchmark Based on the Distributional HypothesisCode2
Evolutionary Computation in the Era of Large Language Model: Survey and RoadmapCode2
OmniMamba: Efficient and Unified Multimodal Understanding and Generation via State Space ModelsCode2
DiffusionBERT: Improving Generative Masked Language Models with Diffusion ModelsCode2
DiffusionPen: Towards Controlling the Style of Handwritten Text GenerationCode2
Show:102550
← PrevPage 4 of 107Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1T5B BaselineBLEU48.74Unverified
2FactT5BBLEU48.37Unverified
3JointGT BaselineBLEU47.51Unverified
4FactJointGTBLEU47.39Unverified
5Control Prefixes (T5-large)METEOR0.41Unverified
6T5METEOR0.12Unverified
7BARTMETEOR0.11Unverified
#ModelMetricClaimedVerifiedStatus
1LeakGANBLEU-20.95Unverified
2partGANBLEU-20.91Unverified
3RankGANBLEU-20.85Unverified
4RelGAN (100)BLEU-20.85Unverified
5SeqGANBLEU-20.83Unverified
#ModelMetricClaimedVerifiedStatus
1LeakGANBLEU-20.96Unverified
2PPOGANBLEU-20.91Unverified
3RelGANBLEU-20.88Unverified
4SeqGANBLEU-20.86Unverified
5RankGANBLEU-20.78Unverified
#ModelMetricClaimedVerifiedStatus
1UniCRSDistinct-30.65Unverified
2CRFRDistinct-30.52Unverified
3KGSFDistinct-30.43Unverified
4C2CRSDistinct-30.33Unverified
5KBRDDistinct-30.3Unverified
#ModelMetricClaimedVerifiedStatus
1UniLMCIDEr14.92Unverified
2BART (TextBox 2.0)CIDEr12.98Unverified
3BARTMETEOR0.3Unverified
4T5METEOR0.29Unverified
#ModelMetricClaimedVerifiedStatus
1Beam search + A*esque (beam)BLEU-134.4Unverified
2Beam search + A*esque (sample)BLEU-134.4Unverified
3Beam search + A*esque (greedy)BLEU-134.3Unverified
4Beam searchBLEU-133.7Unverified
#ModelMetricClaimedVerifiedStatus
1RankGANBLEU-20.81Unverified
2SeqGANBLEU-20.74Unverified
3LeakGANBLEU-20.46Unverified
#ModelMetricClaimedVerifiedStatus
1TGen++METEOR0.17Unverified
2TGenMETEOR0.15Unverified
3TGen+METEOR0.15Unverified
#ModelMetricClaimedVerifiedStatus
1GPT2-124Meval_loss3.12Unverified
2GPT2-81M-LOOPeval_loss3.11Unverified
3GPT2-Hermiteeval_loss2.91Unverified
#ModelMetricClaimedVerifiedStatus
1LLaMA-65B+CFG (zero-shot)Accuracy96.6Unverified
2LLaMA-30B+CFG (zero-shot)Accuracy96.4Unverified
3LLaMA-13B+CFG (zero-shot)Accuracy95.1Unverified
#ModelMetricClaimedVerifiedStatus
1CNN-VAENLL332.1Unverified
2SA-VAENLL327.5Unverified
3Aggressive VAENLL326.7Unverified
#ModelMetricClaimedVerifiedStatus
1BART (TextBox 2.0)BLEU-410.2Unverified
#ModelMetricClaimedVerifiedStatus
1STWGAN-GPBLEU-30.62Unverified
#ModelMetricClaimedVerifiedStatus
1PALMROUGE-L41.41Unverified
#ModelMetricClaimedVerifiedStatus
1BART (TextBox 2.0)ROUGE-L64.34Unverified
#ModelMetricClaimedVerifiedStatus
1AEM+AttentionBLEU-114.17Unverified
#ModelMetricClaimedVerifiedStatus
1GPT-4ASR65.1Unverified
#ModelMetricClaimedVerifiedStatus
1BART (TextBox 2.0)ROUGE-L42.96Unverified
#ModelMetricClaimedVerifiedStatus
1Graph2SeqBLEU22Unverified
#ModelMetricClaimedVerifiedStatus
1WGANGP + DGflowJS-40.19Unverified