SOTAVerified

Video Understanding

A crucial task of Video Understanding is to recognise and localise (in space and time) different actions or events appearing in the video.

Source: Action Detection from a Robot-Car Perspective

Papers

Showing 741750 of 1149 papers

TitleStatusHype
Dynamic Appearance: A Video Representation for Action Recognition with Joint Training0
Contrastive Masked Autoencoders for Self-Supervised Video HashingCode1
A Unified Model for Video Understanding and Knowledge Embedding with Heterogeneous Knowledge Graph Dataset0
EVEREST: Efficient Masked Video Autoencoder by Removing Redundant Spatiotemporal TokensCode1
Masked Autoencoders for Egocentric Video Understanding @ Ego4D Challenge 2022Code0
InternVideo-Ego4D: A Pack of Champion Solutions to Ego4D ChallengesCode1
UniFormerV2: Spatiotemporal Learning by Arming Image ViTs with Video UniFormerCode2
Exploring State Change Capture of Heterogeneous Backbones @ Ego4D Hands and Objects Challenge 20220
Grounded Video Situation Recognition0
VTC: Improving Video-Text Retrieval with User CommentsCode1
Show:102550
← PrevPage 75 of 115Next →

No leaderboard results yet.