SOTAVerified

Caption Generation

Papers

Showing 176–200 of 310 papers

TitleStatusHype
O2NA: An Object-Oriented Non-Autoregressive Approach for Controllable Video Captioning—0
OBJ2TEXT: Generating Visually Descriptive Language from Object Layouts—0
PathM3: A Multimodal Multi-Task Multiple Instance Learning Framework for Whole Slide Image Classification and Captioning—0
Predicting the Mumble of Wireless Channel with Sequence-to-Sequence Models—0
Relationship-based Neural Baby Talk—0
REST: REtrieve & Self-Train for generative action recognition—0
Rethinking the Form of Latent States in Image Captioning—0
Retrieval-Augmented Multimodal Language Modeling—0
Review Networks for Caption Generation—0
RUC+CMU: System Report for Dense Captioning Events in Videos—0
Scene-based Factored Attention for Image Captioning—0
Scene Graph Generation for Better Image Captioning?—0
Scene Understanding for Autonomous Manipulation with Deep Learning—0
See It All: Contextualized Late Aggregation for 3D Dense Captioning—0
Seq2Mol: Automatic design of de novo molecules conditioned by the target protein sequences through deep neural networks—0
Sequence to Sequence - Video to Text—0
Set Prediction Guided by Semantic Concepts for Diverse Video Captioning—0
Simultaneous Segmentation and Recognition: Towards more accurate Ego Gesture Recognition—0
Skip-Gram − Zipf + Uniform = Vector Additivity—0
SLAM-AAC: Enhancing Audio Captioning with Paraphrasing Augmentation and CLAP-Refine through LLMs—0
Social Media Ready Caption Generation for Brands—0
Soft + Hardwired Attention: An LSTM Framework for Human Trajectory Prediction and Abnormal Event Detection—0
Spatio-Temporal Dynamics and Semantic Attribute Enriched Visual Encoding for Video Captioning—0
Stacked Cross-modal Feature Consolidation Attention Networks for Image Captioning—0
Stack-VS: Stacked Visual-Semantic Attention for Image Caption Generation—0
Show:102550
← PrevPage 8 of 13Next →

No leaderboard results yet.