SOTAVerified

Sentence

Papers

Showing 47914800 of 10752 papers

TitleStatusHype
Exploiting Semantic Embedding and Visual Feature for Facial Action Unit Detection0
Image Change Captioning by Learning From an Auxiliary Task0
Cascaded Prediction Network via Segment Tree for Temporal Video Grounding0
Sketch, Ground, and Refine: Top-Down Dense Video CaptioningCode0
Structured Multi-Level Interaction Network for Video Moment Localization via Language Query0
Multi-Stage Aggregated Transformer Network for Temporal Language Localization in Videos0
Learning From the Master: Distilling Cross-Modal Advanced Knowledge for Lip Reading0
Towards Bridging Event Captioner and Sentence Localizer for Weakly Supervised Dense Event Captioning0
Multi-Modal Relational Graph for Cross-Modal Video Moment Retrieval0
Transitional Adaptation of Pretrained Models for Visual Storytelling0
Show:102550
← PrevPage 480 of 1076Next →

No leaderboard results yet.