SOTAVerified

AudioCaps

Papers

Showing 4150 of 64 papers

TitleStatusHype
Mitigating Audiovisual Mismatch in Visual-Guide Audio Captioning0
Multiscale Matching Driven by Cross-Modal Similarity Consistency for Audio-Text Retrieval0
VoiceLDM: Text-to-Speech with Environmental Context0
Text-to-Audio Generation Synchronized with Videos0
Quality Over Quantity? LLM-Based Curation for a Data-Efficient Audio-Video Foundation Model0
AC/DC: LLM-based Audio Comprehension via Dialogue Continuation0
Rethinking Transfer and Auxiliary Learning for Improving Audio Captioning Transformer0
Retrieval-Augmented Text-to-Audio Generation0
Unbiased Sliced Wasserstein Kernels for High-Quality Audio Captioning0
Audio-text Retrieval in Context0
Show:102550
← PrevPage 5 of 7Next →

No leaderboard results yet.