SOTAVerified

TextVQA

Papers

Showing 31–40 of 47 papers

TitleStatusHype
EE-MLLM: A Data-Efficient and Compute-Efficient Multimodal Large Language Model—0
Enhancing Instruction-Following Capability of Visual-Language Models by Reducing Image Redundancy—0
EvoMoE: Expert Evolution in Mixture of Experts for Multimodal Large Language Models—0
Exploring Sparse Spatial Relation in Graph Inference for Text-Based VQA—0
FlexAttention for Efficient High-Resolution Vision-Language Models—0
Graph Relation Transformer: Incorporating pairwise object features into the Transformer architecture—0
HyViLM: Enhancing Fine-Grained Recognition with a Hybrid Encoder for Vision-Language Models—0
Locate Then Generate: Bridging Vision and Language with Bounding Box for Scene-Text VQA—0
Making the V in Text-VQA Matter—0
Multiple-Question Multiple-Answer Text-VQA—0
Show:102550
← PrevPage 4 of 5Next →

No leaderboard results yet.