SOTAVerified

Video Understanding

A crucial task of Video Understanding is to recognise and localise (in space and time) different actions or events appearing in the video.

Source: Action Detection from a Robot-Car Perspective

Papers

Showing 711720 of 1149 papers

TitleStatusHype
Dual-path Adaptation from Image to Video TransformersCode1
TemporalMaxer: Maximize Temporal Context with only Max Pooling for Temporal Action LocalizationCode1
Localizing Moments in Long Video Via Multimodal GuidanceCode1
Video4MRI: An Empirical Study on Brain Magnetic Resonance Image Analytics with CNN-based Video Classification Frameworks0
MINOTAUR: Multi-task Video Grounding From Multimodal QueriesCode0
AIM: Adapting Image Models for Efficient Video Action RecognitionCode2
Semi-Parametric Video-Grounded Text Generation0
Building Scalable Video Understanding Benchmarks through Sports0
STPrivacy: Spatio-Temporal Privacy-Preserving Action Recognition0
Test of Time: Instilling Video-Language Models with a Sense of TimeCode1
Show:102550
← PrevPage 72 of 115Next →

No leaderboard results yet.