SOTAVerified

zero-shot-classification

Papers

Showing 151200 of 422 papers

TitleStatusHype
Geodesic Multi-Modal Mixup for Robust Fine-TuningCode0
Generative Diffusion Model Bootstraps Zero-shot Classification of Fetal Ultrasound Images In Underrepresented African PopulationsCode0
Automatic Report Generation for Histopathology images using pre-trained Vision TransformersCode0
Real-Time Cell Sorting with Scalable In Situ FPGA-Accelerated Deep LearningCode0
Automated Medical Report Generation for ECG Data: Bridging Medical Text and Signal Processing with Deep LearningCode0
Perturb and Recover: Fine-tuning for Effective Backdoor Removal from CLIPCode0
From Unimodal to Multimodal: Scaling up Projectors to Align ModalitiesCode0
A Unified Debiasing Approach for Vision-Language Models across Modalities and TasksCode0
Forget NLI, Use a Dictionary: Zero-Shot Topic Classification for Low-Resource Languages with Application to LuxembourgishCode0
Optimizing CLIP Models for Image Retrieval with Maintained Joint-Embedding AlignmentCode0
On the effectiveness of Large Language Models in the mechanical design domainCode0
Self-supervised Multi-modal Training from Uncurated Image and Reports Enables Zero-shot Oversight Artificial Intelligence in RadiologyCode0
OFF-CLIP: Improving Normal Detection Confidence in Radiology CLIP with Simple Off-Diagonal Term Auto-AdjustmentCode0
On the use of Silver Standard Data for Zero-shot Classification Tasks in Information ExtractionCode0
OverPrompt: Enhancing ChatGPT through Efficient In-Context LearningCode0
NECOMIMI: Neural-Cognitive Multimodal EEG-informed Image Generation with Diffusion ModelsCode0
Non-Contrastive Learning Meets Language-Image Pre-TrainingCode0
Fine-Grained Zero-Shot Learning with DNA as Side InformationCode0
Multimodal Remote Sensing Scene Classification Using VLMs and Dual-Cross Attention NetworksCode0
Online Zero-Shot Classification with CLIPCode0
Multi-level Cross-modal Feature Alignment via Contrastive Learning towards Zero-shot Classification of Remote Sensing Image ScenesCode0
CLaMP: Contrastive Language-Music Pre-training for Cross-Modal Symbolic Music Information RetrievalCode0
Evaluation of Output Embeddings for Fine-Grained Image ClassificationCode0
ModalChorus: Visual Probing and Alignment of Multi-modal Embeddings via Modal Fusion MapCode0
CLAMP: A Contrastive Language And Molecule Pre-training NetworkCode0
Evaluating the Fairness of Discriminative Foundation Models in Computer VisionCode0
Estimating Uncertainty in Multimodal Foundation Models using Public Internet DataCode0
Mitigating Word Bias in Zero-shot Prompt-based ClassifiersCode0
Enhancing Visual Classification using Comparative DescriptorsCode0
AmorLIP: Efficient Language-Image Pretraining via AmortizationCode0
A Statistical Theory of Contrastive Pre-training and Multimodal Generative AICode0
LiT Tuned Models for Efficient Species DetectionCode0
LLM Chain Ensembles for Scalable and Accurate Data AnnotationCode0
Linear Representations of Sentiment in Large Language ModelsCode0
CapS-Adapter: Caption-based MultiModal Adapter in Zero-Shot ClassificationCode0
Learning Deep Representations of Fine-grained Visual DescriptionsCode0
Learning Portrait Style RepresentationsCode0
Gradient Matching Generative Networks for Zero-Shot LearningCode0
LAION-5B: An open large-scale dataset for training next generation image-text modelsCode0
Can Graph Neural Networks Learn Language with Extremely Weak Text Supervision?Code0
Large Language Models versus Classical Machine Learning: Performance in COVID-19 Mortality Prediction Using High-Dimensional Tabular DataCode0
Lex2Sent: A bagging approach to unsupervised sentiment analysisCode0
Investigating the Emergent Audio Classification Ability of ASR Foundation ModelsCode0
Direct side information learning for zero-shot regressionCode0
Connecting NeRFs, Images, and TextCode0
DINOv2 Meets Text: A Unified Framework for Image- and Pixel-Level Vision-Language AlignmentCode0
Boosting Visual-Language Models by Exploiting Hard SamplesCode0
Improving Zero-Shot Detection of Low Prevalence Chest Pathologies using Domain Pre-trained Language ModelsCode0
KPL: Training-Free Medical Knowledge Mining of Vision-Language ModelsCode0
MoRE: Multi-Modal Contrastive Pre-training with Transformers on X-Rays, ECGs, and Diagnostic ReportCode0
Show:102550
← PrevPage 4 of 9Next →

No leaderboard results yet.