SOTAVerified

Text Generation

Text Generation is the task of generating text with the goal of appearing indistinguishable to human-written text. This task is more formally known as "natural language generation" in the literature.

Text generation can be addressed with Markov processes or deep generative models like LSTMs. Recently, some of the most advanced methods for text generation include BART, GPT and other GAN-based approaches. Text generation systems are evaluated either through human ratings or automatic evaluation metrics like METEOR, ROUGE, and BLEU.

Further readings:

( Image credit: Adversarial Ranking for Language Generation )

Papers

Showing 101–150 of 5335 papers

TitleStatusHype
Adaptable Logical Control for Large Language ModelsCode2
Keyformer: KV Cache Reduction through Key Tokens Selection for Efficient Generative InferenceCode2
Balancing LoRA Performance and Efficiency with Simple Shard SharingCode2
Building Cooperative Embodied Agents Modularly with Large Language ModelsCode2
Language Models can Self-Lengthen to Generate Long TextsCode2
Large Language Models Can Learn Temporal ReasoningCode2
Intelligent Artistic Typography: A Comprehensive Review of Artistic Text Design and GenerationCode2
MiniLLM: Knowledge Distillation of Large Language ModelsCode2
KoSBi: A Dataset for Mitigating Social Bias Risks Towards Safer Large Language Model ApplicationCode2
LITA: Language Instructed Temporal-Localization AssistantCode2
In-Context Retrieval-Augmented Language ModelsCode2
InfiniGen: Efficient Generative Inference of Large Language Models with Dynamic KV Cache ManagementCode2
Get my drift? Catching LLM Task Drift with Activation DeltasCode2
AnyText2: Visual Text Generation and Editing With Customizable AttributesCode2
LongForm: Effective Instruction Tuning with Reverse InstructionsCode2
In-Context Editing: Learning Knowledge from Self-Induced DistributionsCode2
Inseq: An Interpretability Toolkit for Sequence Generation ModelsCode2
Make LoRA Great Again: Boosting LoRA with Adaptive Singular Values and Mixture-of-Experts Optimization AlignmentCode2
CelebV-Text: A Large-Scale Facial Text-Video DatasetCode2
mbrs: A Library for Minimum Bayes Risk DecodingCode2
Language-Driven Representation Learning for RoboticsCode2
MemLong: Memory-Augmented Retrieval for Long Text ModelingCode2
MonoFormer: One Transformer for Both Diffusion and AutoregressionCode2
Most Language Models can be Poets too: An AI Writing Assistant and Constrained Text Generation StudioCode2
LLMGA: Multimodal Large Language Model based Generation AssistantCode2
Harmonizing Visual Text Comprehension and GenerationCode2
A Contrastive Framework for Neural Text GenerationCode2
HALC: Object Hallucination Reduction via Adaptive Focal-Contrast DecodingCode2
Benchmarking Uncertainty Quantification Methods for Large Language Models with LM-PolygraphCode2
Hardware-Aware Parallel Prompt Decoding for Memory-Efficient Acceleration of LLM InferenceCode2
HelloBench: Evaluating Long Text Generation Capabilities of Large Language ModelsCode2
GPTScore: Evaluate as You DesireCode2
GPT-NER: Named Entity Recognition via Large Language ModelsCode2
Grounding Language Models to Images for Multimodal Inputs and OutputsCode2
GlyphControl: Glyph Conditional Control for Visual Text GenerationCode2
HyperSteer: Activation Steering at Scale with HypernetworksCode2
BatGPT: A Bidirectional Autoregessive Talker from Generative Pre-trained TransformerCode2
BayLing: Bridging Cross-lingual Alignment and Instruction Following through Interactive Translation for Large Language ModelsCode2
From Pixels to Graphs: Open-Vocabulary Scene Graph Generation with Vision-Language ModelsCode2
Generative Pretrained Structured Transformers: Unsupervised Syntactic Language Models at ScaleCode2
Fine-Grained Human Feedback Gives Better Rewards for Language Model TrainingCode2
Few-Shot Text Generation with Pattern-Exploiting TrainingCode2
Evaluating Morphological Compositional Generalization in Large Language ModelsCode2
A Judge-free LLM Open-ended Generation Benchmark Based on the Distributional HypothesisCode2
eVAE: Evolutionary Variational AutoencoderCode2
Evolutionary Computation in the Era of Large Language Model: Survey and RoadmapCode2
Ecco: An Open Source Library for the Explainability of Transformer Language ModelsCode2
AutoPatent: A Multi-Agent Framework for Automatic Patent GenerationCode2
ECG-Chat: A Large ECG-Language Model for Cardiac Disease DiagnosisCode2
Efficient Minimum Bayes Risk Decoding using Low-Rank Matrix Completion AlgorithmsCode2
Show:102550
← PrevPage 3 of 107Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1T5B BaselineBLEU48.74—Unverified
2FactT5BBLEU48.37—Unverified
3JointGT BaselineBLEU47.51—Unverified
4FactJointGTBLEU47.39—Unverified
5Control Prefixes (T5-large)METEOR0.41—Unverified
6T5METEOR0.12—Unverified
7BARTMETEOR0.11—Unverified
#ModelMetricClaimedVerifiedStatus
1LeakGANBLEU-20.95—Unverified
2partGANBLEU-20.91—Unverified
3RankGANBLEU-20.85—Unverified
4RelGAN (100)BLEU-20.85—Unverified
5SeqGANBLEU-20.83—Unverified
#ModelMetricClaimedVerifiedStatus
1LeakGANBLEU-20.96—Unverified
2PPOGANBLEU-20.91—Unverified
3RelGANBLEU-20.88—Unverified
4SeqGANBLEU-20.86—Unverified
5RankGANBLEU-20.78—Unverified
#ModelMetricClaimedVerifiedStatus
1UniCRSDistinct-30.65—Unverified
2CRFRDistinct-30.52—Unverified
3KGSFDistinct-30.43—Unverified
4C2CRSDistinct-30.33—Unverified
5KBRDDistinct-30.3—Unverified
#ModelMetricClaimedVerifiedStatus
1UniLMCIDEr14.92—Unverified
2BART (TextBox 2.0)CIDEr12.98—Unverified
3BARTMETEOR0.3—Unverified
4T5METEOR0.29—Unverified
#ModelMetricClaimedVerifiedStatus
1Beam search + A*esque (beam)BLEU-134.4—Unverified
2Beam search + A*esque (sample)BLEU-134.4—Unverified
3Beam search + A*esque (greedy)BLEU-134.3—Unverified
4Beam searchBLEU-133.7—Unverified
#ModelMetricClaimedVerifiedStatus
1RankGANBLEU-20.81—Unverified
2SeqGANBLEU-20.74—Unverified
3LeakGANBLEU-20.46—Unverified
#ModelMetricClaimedVerifiedStatus
1TGen++METEOR0.17—Unverified
2TGenMETEOR0.15—Unverified
3TGen+METEOR0.15—Unverified
#ModelMetricClaimedVerifiedStatus
1GPT2-124Meval_loss3.12—Unverified
2GPT2-81M-LOOPeval_loss3.11—Unverified
3GPT2-Hermiteeval_loss2.91—Unverified
#ModelMetricClaimedVerifiedStatus
1LLaMA-65B+CFG (zero-shot)Accuracy96.6—Unverified
2LLaMA-30B+CFG (zero-shot)Accuracy96.4—Unverified
3LLaMA-13B+CFG (zero-shot)Accuracy95.1—Unverified
#ModelMetricClaimedVerifiedStatus
1CNN-VAENLL332.1—Unverified
2SA-VAENLL327.5—Unverified
3Aggressive VAENLL326.7—Unverified
#ModelMetricClaimedVerifiedStatus
1BART (TextBox 2.0)BLEU-410.2—Unverified
#ModelMetricClaimedVerifiedStatus
1STWGAN-GPBLEU-30.62—Unverified
#ModelMetricClaimedVerifiedStatus
1PALMROUGE-L41.41—Unverified
#ModelMetricClaimedVerifiedStatus
1BART (TextBox 2.0)ROUGE-L64.34—Unverified
#ModelMetricClaimedVerifiedStatus
1AEM+AttentionBLEU-114.17—Unverified
#ModelMetricClaimedVerifiedStatus
1GPT-4ASR65.1—Unverified
#ModelMetricClaimedVerifiedStatus
1BART (TextBox 2.0)ROUGE-L42.96—Unverified
#ModelMetricClaimedVerifiedStatus
1Graph2SeqBLEU22—Unverified
#ModelMetricClaimedVerifiedStatus
1WGANGP + DGflowJS-40.19—Unverified