SOTAVerified

Open Vocabulary Object Detection

Open-vocabulary detection (OVD) aims to generalize beyond the limited number of base classes labeled during the training phase. The goal is to detect novel classes defined by an unbounded (open) vocabulary at inference.

Papers

Showing 126–145 of 145 papers

TitleStatusHype
Weakly Supervised Open-Vocabulary Object Detection—0
Learning Pseudo-Labeler beyond Noun Concepts for Open-Vocabulary Object Detection—0
Spuriosity Rankings for Free: A Simple Framework for Last Layer Retraining Based on Object Detection—0
YOLOv8-Based Visual Detection of Road Hazards: Potholes, Sewer Covers, and Manholes—0
Region-centric Image-Language Pretraining for Open-Vocabulary DetectionCode0
EdaDet: Open-Vocabulary Object Detection Using Early Dense Alignment—0
Contrastive Feature Masking Open-Vocabulary Vision Transformer—0
Exploring Multi-Modal Contextual Knowledge for Open-Vocabulary Object Detection—0
Open-Vocabulary Object Detection via Scene Graph Discovery—0
Scaling Open-Vocabulary Object DetectionCode0
DetCLIPv2: Scalable Open-Vocabulary Object Detection Pre-training via Word-Region Alignment—0
MaMMUT: A Simple Architecture for Joint Learning for MultiModal TasksCode0
Prompt-Guided Transformers for End-to-End Open-Vocabulary Object Detection—0
Open-Vocabulary Object Detection using Pseudo Caption Labels—0
Investigating the Role of Attribute Context in Vision-Language Models for Object Recognition and Detection—0
Open-Vocabulary Object Detection With an Open Corpus—0
Learning to Detect and Segment for Open Vocabulary Object Detection—0
Fine-grained Visual-Text Prompt-Driven Self-Training for Open-Vocabulary Object Detection—0
F-VLM: Open-Vocabulary Object Detection upon Frozen Vision and Language ModelsCode0
Simple Open-Vocabulary Object Detection with Vision TransformersCode0
Show:102550
← PrevPage 6 of 6Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1Cooperative Foundational ModelsAP 0.550.3—Unverified
2DE-ViTAP 0.550—Unverified
3Yolov8-nanoAP 0.547.2—Unverified
4DITOAP 0.546.1—Unverified
5OV-DQUO(RN50x4)AP 0.545.6—Unverified
6LP-OVOD (OWL-ViT Proposals)AP 0.544.9—Unverified
7CLIPSelfAP 0.544.3—Unverified
8CORA+AP 0.543.1—Unverified
9BARONAP 0.542.7—Unverified
10SIA-OVD (RN50x4)AP 0.541.9—Unverified
#ModelMetricClaimedVerifiedStatus
1LaMI-DETRAP novel-LVIS base training43.4—Unverified
2DITOAP novel-LVIS base training40.4—Unverified
3OV-DQUO(ViT-L/14)AP novel-LVIS base training39.3—Unverified
4CoDet (EVA02-L)AP novel-LVIS base training37—Unverified
5CLIPSelfAP novel-LVIS base training34.9—Unverified
6OVMRAP novel-LVIS base training34.4—Unverified
7DE-ViTAP novel-LVIS base training34.3—Unverified
8CFM-ViTAP novel-LVIS base training33.9—Unverified
9CLIM (RN50x64)AP novel-LVIS base training32.3—Unverified
10RO-ViTAP novel-LVIS base training32.1—Unverified
#ModelMetricClaimedVerifiedStatus
1Object-Centric-OVDmask AP5022.3—Unverified
2ViLDmask AP5018.2—Unverified
#ModelMetricClaimedVerifiedStatus
1Object-Centric-OVDmask AP5042.9—Unverified
2Deticmask AP5042.2—Unverified