SOTAVerified

document understanding

Document understanding involves document classification, layout analysis, information extraction, and DocQA.

Papers

Showing 76100 of 309 papers

TitleStatusHype
Automatic Knowledge Extraction with Human Interface0
Decontextualization: Making Sentences Stand-Alone0
Automated Parsing of Engineering Drawings for Structured Information Extraction Using a Fine-tuned Document Understanding Transformer0
DAViD: Domain Adaptive Visually-Rich Document Understanding with Synthetic Insights0
DavarOCR: A Toolbox for OCR and Multi-Modal Document Understanding0
Arctic-TILT. Business Document Understanding at Sub-Billion Scale0
Extract with Order for Coherent Multi-Document Summarization0
Auto-encodeurs pour la compr\'ehension de documents parl\'es (Auto-encoders for Spoken Document Understanding)0
A User-Centered Concept Mining System for Query and Document Understanding at Tencent0
CREPE: Coordinate-Aware End-to-End Document Parser0
WildDoc: How Far Are We from Achieving Comprehensive and Robust Document Understanding in the Wild?0
ClueWeb22: 10 Billion Web Documents with Visual and Semantic Information0
Attention Where It Matters: Rethinking Visual Document Understanding with Selective Region Concentration0
DrVideo: Document Retrieval Based Long Video Understanding0
Attention-Based Graph Neural Network with Global Context Awareness for Document Understanding0
Acronym Identification and Disambiguation Shared Tasks for Scientific Document Understanding0
ERNIE-mmLayout: Multi-grained MultiModal Transformer for Document Understanding0
Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling0
Fast-StrucTexT: An Efficient Hourglass Transformer with Modality-guided Dynamic Token Merge for Document Understanding0
A Multi-Modal Multilingual Benchmark for Document Image Classification0
DONUT-hole: DONUT Sparsification by Harnessing Knowledge and Optimizing Learning Efficiency0
A LayoutLMv3-Based Model for Enhanced Relation Extraction in Visually-Rich Documents0
DUBLIN -- Document Understanding By Language-Image Network0
Efficient End-to-End Visual Document Understanding with Rationale Distillation0
DOGE: Towards Versatile Visual Document Grounding and Referring0
Show:102550
← PrevPage 4 of 13Next →

No leaderboard results yet.