SOTAVerified

Document Layout Analysis

"Document Layout Analysis is performed to determine physical structure of a document, that is, to determine document components. These document components can consist of single connected components-regions [...] of pixels that are adjacent to form single regions [...] , or group of text lines. A text line is a group of characters, symbols, and words that are adjacent, “relatively close” to each other and through which a straight line can be drawn (usually with horizontal or vertical orientation)." L. O'Gorman, "The document spectrum for page layout analysis," in IEEE Transactions on Pattern Analysis and Machine Intelligence, vol. 15, no. 11, pp. 1162-1173, Nov. 1993.

Image credit: PubLayNet: largest dataset ever for document layout analysis

Papers

Showing 2650 of 99 papers

TitleStatusHype
appjsonify: An Academic Paper PDF-to-JSON Conversion ToolkitCode1
Training data-efficient image transformers & distillation through attentionCode1
DANIEL: A fast Document Attention Network for Information Extraction and Labelling of handwritten documentsCode1
DocXChain: A Powerful Open-Source Toolchain for Document Parsing and BeyondCode0
Vision Grid Transformer for Document Layout AnalysisCode0
ICDAR 2021 Competition on Historical Map SegmentationCode0
SFDLA: Source-Free Document Layout AnalysisCode0
Text Role Classification in Scientific Charts Using Multimodal TransformersCode0
dhSegment: A generic deep-learning approach for document segmentationCode0
LayoutLMv3: Pre-training for Document AI with Unified Text and Image MaskingCode0
BaDLAD: A Large Multi-Domain Bengali Document Layout Analysis DatasetCode0
VSR: A Unified Framework for Document Layout Analysis combining Vision, Semantics and RelationsCode0
A Graphical Approach to Document Layout AnalysisCode0
Multimodal weighted graph representation for information extraction from visually rich documents.Code0
Multi-Task Handwritten Document Layout AnalysisCode0
M^6Doc: A Large-Scale Multi-Format, Multi-Type, Multi-Layout, Multi-Language, Multi-Annotation Category Dataset for Modern Document Layout AnalysisCode0
Class-Agnostic Region-of-Interest Matching in Document ImagesCode0
DCQA: Document-Level Chart Question Answering towards Complex Reasoning and Common-Sense UnderstandingCode0
LayoutLMv2: Multi-modal Pre-training for Visually-Rich Document UnderstandingCode0
Information Extraction from Visually Rich Documents Using Directed Weighted Graph Neural NetworkCode0
PdfTable: A Unified Toolkit for Deep Learning-Based Table ExtractionCode0
Document Layout Annotation: Database and Benchmark in the Domain of Public AffairsCode0
LayoutReader: Pre-training of Text and Layout for Reading Order DetectionCode0
Vision-Based Layout Detection from Scientific Literature using Recurrent Convolutional Neural Networks0
Visual Detection with Context for Document Layout Analysis0
Show:102550
← PrevPage 2 of 4Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1CDeC-NetTable0.98Unverified
2VGTOverall0.96Unverified
3TRDLUOverall0.96Unverified
4VSROverall0.96Unverified
5DETROverall0.96Unverified
6LayoutLMv3-BOverall0.95Unverified
7DiT-LOverall0.95Unverified
8DoPTAOverall0.95Unverified
9UDocOverall0.94Unverified
10ResNext-101-32×8dOverall0.94Unverified
#ModelMetricClaimedVerifiedStatus
1CV-GroupClass Average IoU83.4Unverified
2CNKIClass Average IoU77.8Unverified
3VAI-OCRClass Average IoU70.7Unverified
4DeepLabV3+Class Average IoU66.5Unverified
5L3i++Class Average IoU (Few-shot setting)61.1Unverified
#ModelMetricClaimedVerifiedStatus
1DoPTA mAP70.72Unverified
2DocLayout-YOLO mAP70.3Unverified
3VGT mAP68.8Unverified
#ModelMetricClaimedVerifiedStatus
1Faster_RCNNOverall0.96Unverified
2fglihaiOverall0.96Unverified
3Faster-RCNNOverall0.95Unverified
#ModelMetricClaimedVerifiedStatus
1fglihaiOverall0.92Unverified
2USYD NLP_CS29-2Overall0.92Unverified
3Faster-RCNNOverall0.91Unverified
#ModelMetricClaimedVerifiedStatus
1VisualWordGridFAR28.7Unverified