SOTAVerified

Task-Oriented Dialogue Systems

Achieving a pre-defined task through a dialog.

Papers

Showing 1–25 of 308 papers

TitleStatusHype
An Efficient Task-Oriented Dialogue Policy: Evolutionary Reinforcement Learning Injected by Elite Individuals—0
WHEN TO ACT, WHEN TO WAIT: Modeling Structural Trajectories for Intent Triggerability in Task-Oriented DialogueCode1
EnSToM: Enhancing Dialogue Systems with Entropy-Scaled Steering Vectors for Topic MaintenanceCode0
clem:todd: A Framework for the Systematic Benchmarking of LLM-Based Task-Oriented Dialogue System Realisations—0
LANID: LLM-assisted New Intent DiscoveryCode0
Interpretable and Robust Dialogue State Tracking via Natural Language Summarization with LLMs—0
Leveraging Graph Structures and Large Language Models for End-to-End Synthetic Task-Oriented DialoguesCode0
Intent-driven In-context Learning for Few-shot Dialogue State Tracking—0
Towards Automatic Evaluation of Task-Oriented Dialogue Flows—0
Large Language Models as User-Agents for Evaluating Task-Oriented-Dialogue Systems—0
ReSpAct: Harmonizing Reasoning, Speaking, and Acting Towards Building Large Language Model-Based Conversational AI Agents—0
DARD: A Multi-Agent Approach for Task-Oriented Dialog Systems—0
A Pointer Network-based Approach for Joint Extraction and Detection of Multi-Label Multi-Class Intents—0
Pseudo-Label Enhanced Prototypical Contrastive Learning for Uniformed Intent DiscoveryCode0
MediTOD: An English Dialogue Dataset for Medical History Taking with Comprehensive Annotations—0
Intent Detection in the Age of LLMs—0
Diversity-grounded Channel Prototypical Learning for Out-of-Distribution Intent Detection—0
Confidence Estimation for LLM-Based Dialogue State TrackingCode0
Keyword-Aware ASR Error Augmentation for Robust Dialogue State Tracking—0
Infusing Emotions into Task-oriented Dialogue Systems: Understanding, Management, and Generation—0
Unsupervised Extraction of Dialogue Policies from Conversations—0
Making Task-Oriented Dialogue Datasets More Natural by Synthetically Generating Indirect User Requests—0
Towards Spoken Language Understanding via Multi-level Multi-grained Contrastive Learning—0
Unsupervised Mutual Learning of Dialogue Discourse Parsing and Topic SegmentationCode0
Enhancing Dialogue State Tracking Models through LLM-backed User-Agents Simulation—0
Show:102550
← PrevPage 1 of 13Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1T5-3b(UnifiedSKG)Entity F170.07—Unverified
2COMETEntity F163.6—Unverified
3DF-NetEntity F162.7—Unverified
4DF-NetEntity F162.5—Unverified
5GLMPEntity F159.97—Unverified
6TTOSEntity F155.38—Unverified
7KB-retrieverEntity F153.7—Unverified
8DSREntity F151.9—Unverified
9KV Retrieval NetEntity F148—Unverified
10THPNEntity F137.8—Unverified
#ModelMetricClaimedVerifiedStatus
1T5METEOR0.33—Unverified
2BARTMETEOR0.09—Unverified
#ModelMetricClaimedVerifiedStatus
1BART (TextBox 2.0)BLEU-420.17—Unverified