SOTAVerified

Open Vocabulary Object Detection

Open-vocabulary detection (OVD) aims to generalize beyond the limited number of base classes labeled during the training phase. The goal is to detect novel classes defined by an unbounded (open) vocabulary at inference.

Papers

Showing 101–125 of 145 papers

TitleStatusHype
Open-vocabulary vs. Closed-set: Best Practice for Few-shot Object Detection Considering Text DescribabilityCode0
LUDVIG: Learning-free Uplifting of 2D Visual features to Gaussian Splatting scenes—0
Boosting Open-Vocabulary Object Detection by Handling Background Samples—0
VOVTrack: Exploring the Potentiality in Videos for Open-Vocabulary Object Tracking—0
Search and Detect: Training-Free Long Tail Object Detection via Web-Image Retrieval—0
HA-FGOVD: Highlighting Fine-grained Attributes via Explicit Linear Composition for Open-Vocabulary Object Detection—0
End-to-end Open-vocabulary Video Visual Relationship Detection using Multi-modal Prompting—0
A Lightweight Modular Framework for Low-Cost Open-Vocabulary Object Detection TrainingCode0
On the Potential of Open-Vocabulary Models for Object Detection in Unusual Street Scenes—0
Unconstrained Open Vocabulary Image Classification: Zero-Shot Transfer from Text to Image via CLIP InversionCode0
BACON: Improving Clarity of Image Captions via Bag-of-Concept Graphs—0
V3Det Challenge 2024 on Vast Vocabulary and Open Vocabulary Object Detection: Methods and Results—0
Open-Vocabulary X-ray Prohibited Item Detection via Fine-tuning CLIP—0
Enhanced Object Detection: A Study on Vast Vocabulary Object Detection Track for V3Det Challenge 2024—0
Learning Background Prompts to Discover Implicit Knowledge for Open Vocabulary Object Detection—0
Open-Vocabulary Object Detection via Neighboring Region Attention Alignment—0
Watch Your Step: Optimal Retrieval for Continual Learning at Scale—0
DetCLIPv3: Towards Versatile Generative Open-vocabulary Object Detection—0
Open-Vocabulary Object Detectors: Robustness Challenges under Distribution Shifts—0
Open-Vocabulary Object Detection with Meta Prompt Representation and Instance Contrastive Optimization—0
LLMs Meet VLMs: Boost Open Vocabulary Object Detection with Fine-grained Descriptors—0
LCV2: An Efficient Pretraining-Free Framework for Grounded Visual Question Answering—0
Exploring Region-Word Alignment in Built-in Detector for Open-Vocabulary Object Detection—0
Scene-adaptive and Region-aware Multi-modal Prompt for Open Vocabulary Object Detection—0
Generating Enhanced Negatives for Training Language-Based Object DetectorsCode0
Show:102550
← PrevPage 5 of 6Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1Cooperative Foundational ModelsAP 0.550.3—Unverified
2DE-ViTAP 0.550—Unverified
3Yolov8-nanoAP 0.547.2—Unverified
4DITOAP 0.546.1—Unverified
5OV-DQUO(RN50x4)AP 0.545.6—Unverified
6LP-OVOD (OWL-ViT Proposals)AP 0.544.9—Unverified
7CLIPSelfAP 0.544.3—Unverified
8CORA+AP 0.543.1—Unverified
9BARONAP 0.542.7—Unverified
10SIA-OVD (RN50x4)AP 0.541.9—Unverified
#ModelMetricClaimedVerifiedStatus
1LaMI-DETRAP novel-LVIS base training43.4—Unverified
2DITOAP novel-LVIS base training40.4—Unverified
3OV-DQUO(ViT-L/14)AP novel-LVIS base training39.3—Unverified
4CoDet (EVA02-L)AP novel-LVIS base training37—Unverified
5CLIPSelfAP novel-LVIS base training34.9—Unverified
6OVMRAP novel-LVIS base training34.4—Unverified
7DE-ViTAP novel-LVIS base training34.3—Unverified
8CFM-ViTAP novel-LVIS base training33.9—Unverified
9CLIM (RN50x64)AP novel-LVIS base training32.3—Unverified
10RO-ViTAP novel-LVIS base training32.1—Unverified
#ModelMetricClaimedVerifiedStatus
1Object-Centric-OVDmask AP5022.3—Unverified
2ViLDmask AP5018.2—Unverified
#ModelMetricClaimedVerifiedStatus
1Object-Centric-OVDmask AP5042.9—Unverified
2Deticmask AP5042.2—Unverified