SOTAVerified

Video Understanding

A crucial task of Video Understanding is to recognise and localise (in space and time) different actions or events appearing in the video.

Source: Action Detection from a Robot-Car Perspective

Papers

Showing 831840 of 1149 papers

TitleStatusHype
Concept Graph Neural Networks for Surgical Video Understanding0
Audio Visual Scene-Aware Dialog Generation with Transformer-based Video Representations0
ActionFormer: Localizing Moments of Actions with TransformersCode2
Learning Optical Flow with Adaptive Graph ReasoningCode1
A Coding Framework and Benchmark towards Low-Bitrate Video UnderstandingCode0
A Dataset for Medical Instructional Video Classification and Question AnsweringCode1
Capturing Temporal Information in a Single Frame: Channel Sampling Strategies for Action RecognitionCode0
End-to-end Generative Pretraining for Multimodal Video Captioning0
Multiview Transformers for Video Recognition0
MERLOT Reserve: Neural Script Knowledge through Vision and Language and Sound0
Show:102550
← PrevPage 84 of 115Next →

No leaderboard results yet.