| 1st Place Solution for MeViS Track in CVPR 2024 PVUW Workshop: Motion Expression guided Video Segmentation | Jun 11, 2024 | Referring Video Object SegmentationSegmentation | CodeCode Available | 1 | 5 |
| BST: Badminton Stroke-type Transformer for Skeleton-based Action Recognition in Racket Sports | Feb 28, 2025 | Action RecognitionLine Detection | CodeCode Available | 1 | 5 |
| In-N-Out Generative Learning for Dense Unsupervised Video Segmentation | Mar 29, 2022 | Contrastive LearningSemantic Segmentation | CodeCode Available | 1 | 5 |
| Stochastic positional embeddings improve masked image modeling | Jul 31, 2023 | Language ModellingMasked Language Modeling | CodeCode Available | 1 | 5 |
| SASVi - Segment Any Surgical Video | Feb 12, 2025 | SegmentationVideo Segmentation | CodeCode Available | 1 | 5 |
| Multi-modal Segment Assemblage Network for Ad Video Editing with Importance-Coherence Reward | Sep 25, 2022 | DecoderVideo Editing | CodeCode Available | 1 | 5 |
| Generic Event Boundary Detection: A Benchmark for Event Segmentation | Jan 26, 2021 | Action DetectionBoundary Detection | CodeCode Available | 1 | 5 |
| One-Shot Video Object Segmentation | Nov 16, 2016 | Foreground SegmentationObject | CodeCode Available | 1 | 5 |
| General and Task-Oriented Video Segmentation | Jul 9, 2024 | DisentanglementDiversity | CodeCode Available | 1 | 5 |
| Flow-based Video Segmentation for Human Head and Shoulders | Apr 20, 2021 | DecoderImage Matting | CodeCode Available | 1 | 5 |
| Concatenated Masked Autoencoders as Spatial-Temporal Learner | Nov 2, 2023 | Action RecognitionData Augmentation | CodeCode Available | 1 | 5 |
| A Survey on Deep Learning Technique for Video Segmentation | Jul 2, 2021 | Autonomous DrivingDeep Learning | CodeCode Available | 1 | 5 |
| Multi-Granularity Video Object Segmentation | Dec 2, 2024 | ObjectSegmentation | CodeCode Available | 1 | 5 |
| PanoVOS: Bridging Non-panoramic and Panoramic Views with Transformer for Video Segmentation | Sep 21, 2023 | Autonomous DrivingSegmentation | CodeCode Available | 1 | 5 |
| Betrayed by Attention: A Simple yet Effective Approach for Self-supervised Video Object Segmentation | Nov 29, 2023 | ClusteringObject | CodeCode Available | 1 | 5 |
| AuxAdapt: Stable and Efficient Test-Time Adaptation for Temporally Consistent Video Semantic Segmentation | Oct 24, 2021 | Optical Flow EstimationSegmentation | CodeCode Available | 1 | 5 |
| Adversarial Pixel Restoration as a Pretext Task for Transferable Perturbations | Jul 18, 2022 | object-detectionObject Detection | CodeCode Available | 1 | 5 |
| AutoVisual Fusion Suite: A Comprehensive Evaluation of Image Segmentation and Voice Conversion Tools on HuggingFace Platform | Dec 17, 2023 | Image SegmentationSegmentation | CodeCode Available | 1 | 5 |
| Few-shot Structure-Informed Machinery Part Segmentation with Foundation Models and Graph Neural Networks | Jan 17, 2025 | Few-Shot Semantic SegmentationSegmentation | CodeCode Available | 1 | 5 |
| Global Knowledge Calibration for Fast Open-Vocabulary Segmentation | Mar 16, 2023 | Knowledge DistillationOpen Vocabulary Semantic Segmentation | CodeCode Available | 1 | 5 |
| Physarum Powered Differentiable Linear Programming Layers and Applications | Apr 30, 2020 | Few-Shot LearningMeta-Learning | CodeCode Available | 1 | 5 |
| EPIC-KITCHENS VISOR Benchmark: VIdeo Segmentations and Object Relations | Sep 26, 2022 | ObjectSegmentation | CodeCode Available | 1 | 5 |
| RankSeg: Adaptive Pixel Classification with Image Category Ranking for Segmentation | Mar 8, 2022 | ClassificationInstance Segmentation | CodeCode Available | 1 | 5 |
| Efficient Semantic Video Segmentation with Per-frame Inference | Feb 26, 2020 | Knowledge DistillationOptical Flow Estimation | CodeCode Available | 1 | 5 |
| D2Conv3D: Dynamic Dilated Convolutions for Object Segmentation in Videos | Nov 15, 2021 | Multi-Object Tracking and SegmentationSegmentation | CodeCode Available | 1 | 5 |