SOTAVerified

A two-stage approach for table extraction in invoices

2022-10-10Unverified0· sign in to hype

Thomas Saout, Frédéric Lardeux, Frédéric Saubion

Unverified — Be the first to reproduce this paper.

Reproduce

Abstract

The automated analysis of administrative documents is an important field in document recognition that is studied for decades. Invoices are key documents among these huge amounts of documents available in companies and public services. Invoices contain most of the time data that are presented in tables that should be clearly identified to extract suitable information. In this paper, we propose an approach that combines an image processing based estimation of the shape of the tables with a graph-based representation of the document, which is used to identify complex tables precisely. We propose an experimental evaluation using a real case application.

Tasks

Reproductions