deepdoctection

deepdoctection

Document AI framework for layout and extraction

Description

Extracting tables and paragraphs from scans and PDFs with plain OCR yields unstructured text and broken tables. deepdoctection is a Python document AI framework that chains layout detection, table recognition, OCR and language models into one pipeline with structured output.

Version 1.0 is PyTorch-only, supports fine-tuned models like LayoutLM and LiLT from Hugging Face, and is split into smaller packages.

Features



Layout analysis:Titles, paragraphs, figures and tables.

Table recognition:Rebuilds rows and columns.

OCR integration:Several OCR engines.

Model ecosystem:Document models from Hugging Face.