SOTAVerified

Object Detection

Papers

Showing 1–50 of 10957 papers

TitleStatusHype
A Real-Time System for Egocentric Hand-Object Interaction Detection in Industrial Domains—0
RS-TinyNet: Stage-wise Feature Fusion Network for Detecting Tiny Objects in Remote Sensing Images—0
Dual LiDAR-Based Traffic Movement Count Estimation at a Signalized Intersection: Deployment, Data Collection, and Preliminary Analysis—0
Decoupled PROB: Decoupled Query Initialization Tasks and Objectness-Class Learning for Open World Object Detection—0
Vision-based Perception for Autonomous Vehicles in Obstacle Avoidance Scenarios—0
Tomato Multi-Angle Multi-Pose Dataset for Fine-Grained Phenotyping—0
ECORE: Energy-Conscious Optimized Routing for Deep Learning Models at the Edge—0
Beyond One Shot, Beyond One Perspective: Cross-View and Long-Horizon Distillation for Better LiDAR RepresentationsCode1
MambaFusion: Height-Fidelity Dense Global Fusion for Multi-modal 3D Object DetectionCode2
Weakly-supervised Contrastive Learning with Quantity Prompts for Moving Infrared Small Target DetectionCode0
Detection of Rail Line Track and Human Beings Near the Track to Avoid Accidents—0
Improve Underwater Object Detection through YOLOv12 Architecture and Physics-informed AugmentationCode1
Seg-R1: Segmentation Can Be Surprisingly Simple with Reinforcement LearningCode2
Towards Reliable Detection of Empty Space: Conditional Marked Point Processes for Object DetectionCode0
DuET: Dual Incremental Object Detection via Exemplar-Free Task Arithmetic—0
A Comprehensive Dataset for Underground Miner Detection in Diverse Scenario—0
LASFNet: A Lightweight Attention-Guided Self-Modulation Feature Fusion Network for Multimodal Object DetectionCode0
ThermalDiffusion: Visual-to-Thermal Image-to-Image Translation for Autonomous Navigation—0
Lightweight Multi-Frame Integration for Robust YOLO Object Detection in Videos—0
TDiR: Transformer based Diffusion for Image Restoration Tasks—0
Feature Hallucination for Self-supervised Action Recognition—0
From Codicology to Code: A Comparative Study of Transformer and YOLO-based Detectors for Layout Analysis in Historical Documents—0
A Survey of Multi-sensor Fusion Perception for Embodied AI: Background, Methods, Challenges and Prospects—0
Unfolding the Past: A Comprehensive Deep Learning Approach to Analyzing Incunabula Pages—0
YOLOv13: Real-Time Object Detection with Hypergraph-Enhanced Adaptive Visual PerceptionCode5
Class Agnostic Instance-level Descriptor for Visual Instance Search—0
Can AI Dream of Unseen Galaxies? Conditional Diffusion Model for Galaxy Morphology AugmentationCode0
Retrospective Memory for Camouflaged Object Detection—0
VisText-Mosquito: A Multimodal Dataset and Benchmark for AI-Based Mosquito Breeding Site Detection and ReasoningCode0
YOLOv11-RGBT: Towards a Comprehensive Single-Stage Multispectral Object Detection FrameworkCode4
Comparison of Two Methods for Stationary Incident Detection Based on Background Image—0
How Real is CARLAs Dynamic Vision Sensor? A Study on the Sim-to-Real Gap in Traffic Object Detection—0
Sparse Convolutional Recurrent Learning for Efficient Event-based Neuromorphic Object Detection—0
UAV Object Detection and Positioning in a Mining Industrial Metaverse with Custom Geo-Referenced Data—0
FindMeIfYouCan: Bringing Open Set metrics to near , far and farther Out-of-Distribution Object Detection—0
Lecture Video Visual Objects (LVVO) Dataset: A Benchmark for Visual Object Detection in Educational VideosCode0
Focusing on Tracks for Online Multi-Object TrackingCode2
MatchPlant: An Open-Source Pipeline for UAV-Based Single-Plant Detection and Data ExtractionCode0
Vision-based Lifting of 2D Object Detections for Automated Driving—0
Teleoperated Driving: a New Challenge for 3D Object Detection in Compressed Point Clouds—0
FSATFusion: Frequency-Spatial Attention Transformer for Infrared and Visible Image FusionCode0
Improving Medical Visual Representation Learning with Pathological-level Cross-Modal Alignment and Correlation Exploration—0
Semantic-decoupled Spatial Partition Guided Point-supervised Oriented Object DetectionCode1
Uncertainty-Masked Bernoulli Diffusion for Camouflaged Object Detection Refinement—0
DySS: Dynamic Queries and State-Space Learning for Efficient 3D Object Detection from Multi-Camera Videos—0
CEM-FBGTinyDet: Context-Enhanced Foreground Balance with Gradient Tuning for tiny Objects—0
WD-DETR: Wavelet Denoising-Enhanced Real-Time Object Detection Transformer for Robot Perception with Event Cameras—0
Data Augmentation For Small Object using Fast AutoAugment—0
Hierarchical Neural Collapse Detection Transformer for Class Incremental Object Detection—0
ADAM: Autonomous Discovery and Annotation Model using LLMs for Context-Aware Annotations—0
Show:102550
← PrevPage 1 of 220Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1Co-DETRbox mAP66—Unverified
2InternImage-H (M3I Pre-training)box mAP65.5—Unverified
3M3I Pre-training (InternImage-H)box mAP65.4—Unverified
4MoCaEbox mAP65.1—Unverified
5Co-DETR (Swin-L)box mAP64.8—Unverified
6Focal-Stable-DINO (Focal-Huge, no TTA)box mAP64.8—Unverified
7EVAbox mAP64.7—Unverified
8Group DETR v2box mAP64.5—Unverified
9FocalNet-H (DINO)box mAP64.4—Unverified
10InternImage-XLbox mAP64.3—Unverified