OCR
Page 2 of 2
- AI
Marker: Open-Source PDF to Markdown Conversion with Deep Learning
PDF documents remain one of the most common formats for knowledge distribution, yet they are among the most difficult to process programmatically...
- Open Source
LayoutParser: Unified Open-Source Toolkit for Document Image Analysis
If you have ever tried to extract structured information from a scanned PDF, a historical newspaper archive, or a stack of invoices, you know the...
- AI
GOT-OCR2.0: General OCR Theory Towards OCR-2.0 with Unified End-to-End Model
Optical Character Recognition has been a solved problem for decades -- for clean scanned documents with straightforward text. But the real world of...