SOTAVerified

Visual Question Answering

MLLM Leaderboard

Papers

Showing 15761600 of 2177 papers

TitleStatusHype
Bilaterally Slimmable Transformer for Elastic and Efficient Visual Question AnsweringCode0
Towards Escaping from Language Bias and OCR Error: Semantics-Centered Text Visual Question Answering0
Can you even tell left from right? Presenting a new challenge for VQA0
CLIP Models are Few-shot Learners: Empirical Studies on VQA and Visual Entailment0
Enabling Multimodal Generation on CLIP via Vision-Language Knowledge Distillation0
Barlow constrained optimization for Visual Question AnsweringCode0
Modeling Coreference Relations in Visual Dialog0
Dynamic Key-value Memory Enhanced Multi-step Graph Reasoning for Knowledge-based Visual Question AnsweringCode0
Recent, rapid advancement in visual question answering architecture: a review0
On Modality Bias Recognition and ReductionCode0
Joint Answering and Explanation for Visual Commonsense ReasoningCode0
Measuring CLEVRness: Blackbox testing of Visual Reasoning Models0
OG-SGG: Ontology-Guided Scene Graph Generation. A Case Study in Transfer Learning for Telepresence RoboticsCode0
Privacy Preserving Visual Question Answering0
Delving Deeper into Cross-lingual Visual Question AnsweringCode0
An experimental study of the vision-bottleneck in VQA0
Can Open Domain Question Answering Systems Answer Visual Knowledge Questions?0
NEWSKVQA: Knowledge-Aware News Video Question Answering0
OFA: Unifying Architectures, Tasks, and Modalities Through a Simple Sequence-to-Sequence Learning FrameworkCode0
Grounding Answers for Visual Questions Asked by Visually Impaired PeopleCode0
Compositionality as Lexical SymmetryCode0
Transformer Module Networks for Systematic Generalization in Visual Question AnsweringCode0
Learning to Compose Diversified Prompts for Image Emotion Classification0
MGA-VQA: Multi-Granularity Alignment for Visual Question Answering0
SA-VQA: Structured Alignment of Visual and Semantic Representations for Visual Question Answering0
Show:102550
← PrevPage 64 of 88Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1MMCTAgent (GPT-4 + GPT-4V)GPT-4 score74.24Unverified
2Qwen2-VL-72BGPT-4 score74Unverified
3InternVL2.5-78BGPT-4 score72.3Unverified
4GPT-4o +text rationale +IoTGPT-4 score72.2Unverified
5Lyra-ProGPT-4 score71.4Unverified
6GLM-4V-PlusGPT-4 score71.1Unverified
7Phantom-7BGPT-4 score70.8Unverified
8InternVL2.5-38BGPT-4 score68.8Unverified
9InternVL2-26B (SGP, token ratio 64%)GPT-4 score65.6Unverified
10Baichuan-Omni (7B)GPT-4 score65.4Unverified