SOTAVerified

Zero-Shot Video Question Answer

This task present the results of Zeroshot Question Answer results on TGIF-QA dataset for LLM powered Video Conversational Models.

Papers

Showing 76–85 of 85 papers

TitleStatusHype
Long Story Short: Story-level Video Understanding from 20K Short Films—0
0/1 Deep Neural Networks via Block Coordinate Descent—0
MoReVQA: Exploring Modular Reasoning Models for Video Question Answering—0
GPT-4o: Visual perception performance of multimodal large language models in piglet activity understanding—0
GPT-4o System Card—0
ENTER: Event Based Interpretable Reasoning for VideoQA—0
Vista-LLaMA: Reliable Video Narrator via Equal Distance to Visual Tokens—0
DeepStack: Deeply Stacking Visual Tokens is Surprisingly Simple and Effective for LMMs—0
CinePile: A Long Video Question Answering Dataset and Benchmark—0
Zero-Shot Video Question Answering with Procedural Programs—0
Show:102550
← PrevPage 4 of 4Next →

No leaderboard results yet.