SOTAVerified

Visual Question Answering

MLLM Leaderboard

Papers

Showing 901925 of 2177 papers

TitleStatusHype
Interpretable Visual Question Answering via Reasoning Supervision0
Interpretable Visual Reasoning via Probabilistic Formulation under Natural Supervision0
Bilinear Graph Networks for Visual Question Answering0
Analysis of Visual Question Answering Algorithms with attention model0
Inverse Visual Question Answering: A New Benchmark and VQA Diagnosis Tool0
Inverse Visual Question Answering with Multi-Level Attentions0
Learning What Makes a Difference from Counterfactual Examples and Gradient Supervision0
Graph Neural Networks in Vision-Language Image Understanding: A Survey0
A Unified Framework for Multilingual and Code-Mixed Visual Question Answering0
Graph-based Heuristic Search for Module Selection Procedure in Neural Module Network0
Learning Visual Knowledge Memory Networks for Visual Question Answering0
Lego: Learning to Disentangle and Invert Personalized Concepts Beyond Object Appearance in Text-to-Image Diffusion Models0
ISAAQ - Mastering Textbook Questions with Pre-trained Transformers and Bottom-Up and Top-Down Attention0
Is GPT-3 all you need for Visual Question Answering in Cultural Heritage?0
GRAM: Global Reasoning for Multi-Page VQA0
GRADE: Quantifying Sample Diversity in Text-to-Image Models0
It Takes Two to Tango: Towards Theory of AI's Mind0
iVQA: Inverse Visual Question Answering0
Jaeger: A Concatenation-Based Multi-Transformer VQA Model0
Learning to Select Question-Relevant Relations for Visual Question Answering0
Learning to Specialize with Knowledge Distillation for Visual Question Answering0
AMXFP4: Taming Activation Outliers with Asymmetric Microscaling Floating-Point for 4-bit LLM Inference0
GPT-4V Explorations: Mining Autonomous Driving0
Learning to Recognize the Unseen Visual Predicates0
LEGO-Puzzles: How Good Are MLLMs at Multi-Step Spatial Reasoning?0
Show:102550
← PrevPage 37 of 88Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1MMCTAgent (GPT-4 + GPT-4V)GPT-4 score74.24Unverified
2Qwen2-VL-72BGPT-4 score74Unverified
3InternVL2.5-78BGPT-4 score72.3Unverified
4GPT-4o +text rationale +IoTGPT-4 score72.2Unverified
5Lyra-ProGPT-4 score71.4Unverified
6GLM-4V-PlusGPT-4 score71.1Unverified
7Phantom-7BGPT-4 score70.8Unverified
8InternVL2.5-38BGPT-4 score68.8Unverified
9InternVL2-26B (SGP, token ratio 64%)GPT-4 score65.6Unverified
10Baichuan-Omni (7B)GPT-4 score65.4Unverified