OCR

Page 2 of 2

  1. Marker: Open-Source PDF to Markdown Conversion with Deep Learning

    PDF documents remain one of the most common formats for knowledge distribution, yet they are among the most difficult to process programmatically...

    AI
  2. LayoutParser: Unified Open-Source Toolkit for Document Image Analysis

    If you have ever tried to extract structured information from a scanned PDF, a historical newspaper archive, or a stack of invoices, you know the...

    Open Source
  3. GOT-OCR2.0: General OCR Theory Towards OCR-2.0 with Unified End-to-End Model

    Optical Character Recognition has been a solved problem for decades -- for clean scanned documents with straightforward text. But the real world of...

    AI