SOTAVerified

document understanding

Document understanding involves document classification, layout analysis, information extraction, and DocQA.

Papers

Showing 151200 of 309 papers

TitleStatusHype
StructFormer: Document Structure-based Masked Attention and its Impact on Language Model Pre-Training0
Survey on Question Answering over Visually Rich Documents: Methods, Challenges, and Trends0
SynthDoc: Bilingual Documents Synthesis for Visual Document Understanding0
Table-Of-Contents generation on contemporary documents0
Table Structure Extraction with Bi-directional Gated Recurrent Unit Networks0
Test-Time Adaptation for Visual Document Understanding0
DuReader_vis: A Chinese Dataset for Open-domain Document Visual Question Answering0
The Hidden Structure -- Improving Legal Document Understanding Through Explicit Text Formatting0
The Law of Large Documents: Understanding the Structure of Legal Contracts Using Visual Cues0
The MERIT Dataset: Modelling and Efficiently Rendering Interpretable Transcripts0
TokenSelect: Efficient Long-Context Inference and Length Extrapolation for LLMs via Dynamic Token-Level KV Cache Selection0
Towards Complex Document Understanding By Discrete Reasoning0
Towards Efficient Resume Understanding: A Multi-Granularity Multi-Modal Pre-Training Approach0
Towards Natural Language-Based Document Image Retrieval: New Dataset and Benchmark0
GlobalDoc: A Cross-Modal Vision-Language Framework for Real-World Document Image Retrieval and Classification0
Transformer-based Approach for Document Understanding0
TRIE: End-to-End Text Reading and Information Extraction for Document Understanding0
Two to Five Truths in Non-Negative Matrix Factorization0
Understanding Long Documents with Different Position-Aware Attentions0
UniDoc: Unified Pretraining Framework for Document Understanding0
Unified Pretraining Framework for Document Understanding0
Unimodal and Multimodal Representation Training for Relation Extraction0
ViRED: Prediction of Visual Relations in Engineering Drawings0
Vision Grid Transformer for Document Layout Analysis0
WebFormer: The Web-page Transformer for Structure Information Extraction0
"What is the value of templates?" Rethinking Document Information Extraction Datasets for LLMs0
What Makes a Good Dataset for Symbol Description Reading?0
WikiMixQA: A Multimodal Benchmark for Question Answering over Tables and Charts0
Workshop on Document Intelligence Understanding0
XFUND: A Benchmark Dataset for Multilingual Visually Rich Form Understanding0
Deep Learning based Visually Rich Document Content Understanding: A Survey0
Zero-Shot Prompting and Few-Shot Fine-Tuning: Revisiting Document Image Classification Using Large Language Models0
WildDoc: How Far Are We from Achieving Comprehensive and Robust Document Understanding in the Wild?0
VRDU: A Benchmark for Visually-rich Document Understanding0
Acronym Identification and Disambiguation Shared Tasks for Scientific Document Understanding0
A LayoutLMv3-Based Model for Enhanced Relation Extraction in Visually-Rich Documents0
A Multi-Modal Multilingual Benchmark for Document Image Classification0
Arctic-TILT. Business Document Understanding at Sub-Billion Scale0
A Retrospective Recount of Computer Architecture Research with a Data-Driven Study of Over Four Decades of ISCA Publications0
A Simple yet Effective Layout Token in Large Language Models for Document Understanding0
Assessing Generative AI value in a public sector context: evidence from a field experiment0
A Survey and Approach to Chart Classification0
A Survey on MLLM-based Visually Rich Document Understanding: Methods, Challenges, and Emerging Trends0
A Survey on Vietnamese Document Analysis and Recognition: Challenges and Future Directions0
AT-BERT: Adversarial Training BERT for Acronym Identification Winning Solution for SDU@AAAI-210
A Token-level Text Image Foundation Model for Document Understanding0
Attention-Based Graph Neural Network with Global Context Awareness for Document Understanding0
Attention Where It Matters: Rethinking Visual Document Understanding with Selective Region Concentration0
A User-Centered Concept Mining System for Query and Document Understanding at Tencent0
Auto-encodeurs pour la compr\'ehension de documents parl\'es (Auto-encoders for Spoken Document Understanding)0
Show:102550
← PrevPage 4 of 7Next →

No leaderboard results yet.