Skip to content
#

scanned-documents

Here are 48 public repositories matching this topic...

Dedoc is a library (service) for automate documents parsing and bringing to a uniform format. It automatically extracts content, logical structure, tables, and meta information from textual electronic documents. (Parse document; Document content extraction; Logical structure extraction; PDF parser; Scanned document parser; DOCX parser; HTML parser

  • Updated Nov 19, 2024
  • Python

Efficient Text Localization Algorithm, Image Inversion Detection of Scanned Documents & Language Identification based on Shape Context and Traditional Computer Vision.

  • Updated Dec 18, 2021
  • Python

Improve this page

Add a description, image, and links to the scanned-documents topic page so that developers can more easily learn about it.

Curate this topic

Add this topic to your repo

To associate your repository with the scanned-documents topic, visit your repo's landing page and select "manage topics."

Learn more