SOTAVerified

Video Understanding

A crucial task of Video Understanding is to recognise and localise (in space and time) different actions or events appearing in the video.

Source: Action Detection from a Robot-Car Perspective

Papers

Showing 10311040 of 1149 papers

TitleStatusHype
Prompting Video-Language Foundation Models with Domain-specific Fine-grained Heuristics for Video Question Answering0
PromptonomyViT: Multi-Task Prompt Learning Improves Video Transformers using Synthetic Scene Data0
Prototypes are Balanced Units for Efficient and Effective Partially Relevant Video Retrieval0
PVChat: Personalized Video Chat with One-Shot Learning0
PVC: Progressive Visual Token Compression for Unified Image and Video Processing in Large Vision-Language Models0
PVUW 2025 Challenge Report: Advances in Pixel-level Understanding of Complex Videos in the Wild0
PYSKL: a toolbox for skeleton-based video understanding0
Q-Bench-Video: Benchmarking the Video Quality Understanding of LMMs0
Q-Bench-Video: Benchmark the Video Quality Understanding of LMMs0
Q-Frame: Query-aware Frame Selection and Multi-Resolution Adaptation for Video-LLMs0
Show:102550
← PrevPage 104 of 115Next →

No leaderboard results yet.