SOTAVerified

Image Retrieval

Image Retrieval is a fundamental and long-standing computer vision task that involves finding images similar to a given query from a large database. It is often considered a form of fine-grained, instance-level classification. The task is integral to image recognition alongside classification and cross-modal retrieval. By leveraging visual similarity and other criteria, image retrieval enables users to efficiently discover relevant images, making it a crucial tool in applications such as search and recommendation.

Extending CLIP for Category-to-image Retrieval in E-commerce

( Image credit: DELF )

Papers

Showing 801850 of 2239 papers

TitleStatusHype
Deep Feature Aggregation and Image Re-ranking with Heat Diffusion for Image RetrievalCode0
Gray Level Co-Occurrence Matrices: Generalisation and Some New FeaturesCode0
A Discriminatively Learned CNN Embedding for Person Re-identificationCode0
HADA: A Graph-based Amalgamation Framework in Image-text RetrievalCode0
Deep Exemplar-based ColorizationCode0
ARNet: Self-Supervised FG-SBIR with Unified Sample Feature Alignment and Multi-Scale Token RecyclingCode0
BiVLC: Extending Vision-Language Compositionality Evaluation with Text-to-Image RetrievalCode0
Deep Discrete Hashing with Self-supervised Pairwise LabelsCode0
DeepDiary: Automatic Caption Generation for Lifelogging Image StreamsCode0
A Hybrid Compact Neural Architecture for Visual Place RecognitionCode0
Deep Constrained Dominant Sets for Person Re-identificationCode0
Gram Barcodes for Histopathology Tissue Texture RetrievalCode0
Deep Class-Wise Hashing: Semantics-Preserving Hashing via Class-wise LossCode0
Deep Cauchy Hashing for Hamming Space RetrievalCode0
Adaptive Fine-Grained Sketch-Based Image RetrievalCode0
Socializing the Semantic Gap: A Comparative Survey on Image Tag Assignment, Refinement and RetrievalCode0
Geolocating Earth Imagery from ISS: Integrating Machine Learning with Astronaut Photography for Enhanced Geographic MappingCode0
Feature Learning based Deep Supervised Hashing with Pairwise LabelsCode0
Binary Generative Adversarial Networks for Image RetrievalCode0
Geolocation Estimation of Photos using a Hierarchical Model and Scene ClassificationCode0
Generative Domain-Migration Hashing for Sketch-to-Image RetrievalCode0
Generating Diverse and Meaningful CaptionsCode0
Decoupling the Role of Data, Attention, and Losses in Multimodal TransformersCode0
Decoupling Semantic Similarity from Spatial Alignment for Neural NetworksCode0
A Neural Divide-and-Conquer Reasoning Framework for Image Retrieval from Linguistically Complex TextCode0
Generalization in Metric Learning: Should the Embedding Layer be the Embedding Layer?Code0
12-in-1: Multi-Task Vision and Language Representation LearningCode0
Retrieving Users' Opinions on Social Media with Multimodal Aspect-Based Sentiment AnalysisCode0
From Selective Deep Convolutional Features to Compact Binary Representations for Image RetrievalCode0
Attribute-Aware Attention Model for Fine-grained Representation LearningCode0
Date Estimation in the Wild of Scanned Historical Photos: An Image Retrieval ApproachCode0
Hardness-Aware Deep Metric LearningCode0
Implicit Differentiable Outlier Detection Enable Robust Deep Multimodal AnalysisCode0
Data-Efficient Ranking Distillation for Image Retrieval0
Bidirectional Retrieval Made Simple0
An Enhanced Object Detection Model for Scene Graph Generation0
Data-Efficient Generalization for Zero-shot Composed Image Retrieval0
Zero-Shot Sketch-Based Image Retrieval with Structure-aware Asymmetric Disentanglement0
An Element-wise Visual-enhanced BiLSTM-CRF Model for Location Name Recognition0
A deep learning pipeline for product recognition on store shelves0
DART^3: Leveraging Distance for Test Time Adaptation in Person Re-Identification0
DarkRank: Accelerating Deep Metric Learning via Cross Sample Similarities Transfer0
DALG: Deep Attentive Local and Global Modeling for Image Retrieval0
Beyond Instance-Level Image Retrieval: Leveraging Captions to Learn a Global Visual Representation for Semantic Retrieval0
An Efficient Image Retrieval Based on Fusion of Low-Level Visual Features0
Curriculum Learning for Data-Efficient Vision-Language Alignment0
Beyond Captioning: Task-Specific Prompting for Improved VLM Performance in Mathematical Reasoning0
An Efficient Framework for Zero-Shot Sketch-Based Image Retrieval0
A deep image retrieval network using Max-m-Min pooling and morphological feature generating residual blocks0
A Convolutional Neural Network-based Patent Image Retrieval Method for Design Ideation0
Show:102550
← PrevPage 17 of 45Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1SuperGlobalmAP80.2Unverified
2AMESmAP80Unverified
3Hypergraph propagation+community selectionmAP73Unverified
4TokenmAP66.57Unverified
5DELG+ α QE reranking+ RRT rerankingmAP64Unverified
6FIRemAP61.2Unverified
7HOWmAP56.9Unverified
8ResNet101+ArcFace GLDv2-train-cleanmAP51.6Unverified
9DELF–HQE+SPmAP50.3Unverified
10HesAff–rSIFT–HQE+SPmAP49.7Unverified
#ModelMetricClaimedVerifiedStatus
1AMESmAP90.7Unverified
2Hypergraph propagation+Community selectionmAP88.4Unverified
3TokenmAP82.28Unverified
4FIRemAP81.8Unverified
5DELG+ α QE reranking + RRT rerankingmAP80.4Unverified
6HOWmAP79.4Unverified
7ResNet101+ArcFace GLDv2-train-cleanmAP74.2Unverified
8DELF–HQE+SPmAP73.4Unverified
9HesAff–rSIFT–HQE+SPmAP71.3Unverified
10DELF–ASMK*+SPmAP67.8Unverified
#ModelMetricClaimedVerifiedStatus
1AMESmAP89.7Unverified
2SuperGlobalmAP86.7Unverified
3Hypergraph propagationmAP83.3Unverified
4TokenmAP78.56Unverified
5DELG+ α QE reranking + RRT rerankingmAP77.7Unverified
6ResNet101+ArcFace GLDv2-train-cleanmAP70.3Unverified
7FIRemAP70Unverified
8DELF–HQE+SPmAP69.3Unverified
9HOWmAP62.4Unverified
10R–R-MACmAP59.4Unverified
#ModelMetricClaimedVerifiedStatus
1AMESmAP94.9Unverified
2Hypergraph propagationmAP92.6Unverified
3TokenmAP89.34Unverified
4DELG+ α QE reranking + RRT rerankingmAP88.5Unverified
5FIRemAP85.3Unverified
6ResNet101+ArcFace GLDv2-train-cleanmAP84.9Unverified
7DELF–HQE+SPmAP84Unverified
8HOWmAP81.6Unverified
9R–R-MACmAP78.9Unverified
10R–GeMmAP77.2Unverified
#ModelMetricClaimedVerifiedStatus
1Swin-T (MosaiCLIP, CC-12M)Recall@1 (HN-Atom, UC)44.5Unverified
2RN-50 (MosaiCLIP, CC-12M)Recall@1 (HN-Atom, UC)44.4Unverified
3MosaiCLIP (YFCC-FT)Recall@1 (HN-Atom, UC)41.5Unverified
4RN-50 (NegCLIP, CC-12M)Recall@1 (HN-Atom, UC)41.4Unverified
5MosaiCLIP (CC-FT)Recall@1 (HN-Atom, UC)40.9Unverified
6Swin-T (NegCLIP, CC-12M)Recall@1 (HN-Atom, UC)39.6Unverified
7CLIP (YFCC-FT)Recall@1 (HN-Atom, UC)39.5Unverified
8ViT-L-14 (LAION400M)Recall@1 (HN-Atom + HN-Comp, SC)39.44Unverified
9NegCLIP (YFCC-FT)Recall@1 (HN-Atom, UC)39Unverified
10CLIP-FT (YFCC-FT)Recall@1 (HN-Atom, UC)38.3Unverified
#ModelMetricClaimedVerifiedStatus
1DQU-CIR(Recall@10+Recall@50)/271.77Unverified
2TMCIR(Recall@10+Recall@50)/266.56Unverified
3SPN4CIR (SPRC)(Recall@10+Recall@50)/266.41Unverified
4SPRC(Recall@10+Recall@50)/264.85Unverified
5Candidate Set Re-ranking(Recall@10+Recall@50)/262.15Unverified
6RUTIR (BLIP B/16)(Recall@10+Recall@50)/261.32Unverified
7CASE(Recall@10+Recall@50)/259.73Unverified
8CaLa(Recall@10+Recall@50)/257.96Unverified
9BLIP4CIR+Bi(Recall@10+Recall@50)/255.4Unverified
10CLIP4Cir (v3)(Recall@10+Recall@50)/255.36Unverified
#ModelMetricClaimedVerifiedStatus
1X-VLM (base)R@186.9Unverified
2RCARR@162.6Unverified
3SGRAFR@158.5Unverified
4VisualSpartaR@157.4Unverified
5LGSGMR@157.4Unverified
6TERAN MrSwR@156.5Unverified
7TERAN Symm.R@155.7Unverified
8VSRNR@154.7Unverified
9CAMPR@151.5Unverified
10SCAN i-tR@144Unverified
#ModelMetricClaimedVerifiedStatus
1TMCIR(Recall@5+Recall_subset@1)/283.46Unverified
2SPN4CIR (SPRC)(Recall@5+Recall_subset@1)/282.69Unverified
3SPRC2(Recall@5+Recall_subset@1)/282.66Unverified
4SPRC(Recall@5+Recall_subset@1)/281.39Unverified
5Candidate Set Re-ranking(Recall@5+Recall_subset@1)/280.9Unverified
6CaLa(Recall@5+Recall_subset@1)/278.74Unverified
7CASE (Pre-trained on LaSCo.Ca)(Recall@5+Recall_subset@1)/278.25Unverified
8CASE(Recall@5+Recall_subset@1)/277.5Unverified
9VISTA (base)(Recall@5+Recall_subset@1)/275.9Unverified
10MMRet-MLLM(Recall@5+Recall_subset@1)/275.7Unverified
#ModelMetricClaimedVerifiedStatus
1Unicom+ViT-L@336pxR@191.2Unverified
2ROADMAP (DeiT-B)R@186Unverified
3CGD (SG/GS)R@184.2Unverified
4ROADMAP (ResNet-50)R@183.1Unverified
5ProxyNCA++R@181.4Unverified
6PNP LossR@181.1Unverified
7Cross-Batch MemoryR@180.6Unverified
8Smooth-APR@180.1Unverified
9NormSoftmax2048 (ResNet-50)R@179.5Unverified
10EPSHN512R@178.3Unverified
#ModelMetricClaimedVerifiedStatus
1InternVL-G-FTR@185.9Unverified
2InternVL-C-FTR@185.2Unverified
3R2D2 (ViT-L/14)R@184.4Unverified
4CN-CLIP (ViT-L/14@336px)R@184.4Unverified
5CN-CLIP (ViT-H/14)R@183.8Unverified
6CN-CLIP (ViT-L/14)R@182.7Unverified
7CN-CLIP (ViT-B/16)R@179.1Unverified
8R2D2 (ViT-B)R@178.3Unverified
9Wukong (ViT-L/14)R@177.4Unverified
10Wukong (ViT-B/32)R@167.6Unverified
#ModelMetricClaimedVerifiedStatus
1Offline DiffusionMAP96.2Unverified
2CNN+IME layerMAP92Unverified
3DELF+FT+ATT+DIR+QEMAP90Unverified
4DIR+QE*MAP89Unverified