OCR
Page 1 of 2
- AI
Surya: Open-Source Multilingual OCR and Document Understanding
Optical Character Recognition is one of the oldest applications of computer vision, but traditional OCR engines have struggled to keep pace with...
- Open Source
RapidLayout: Open-Source Document Layout Analysis for Chinese and English
Document layout analysis is the critical first step in any document understanding pipeline. Before OCR can extract text, before tables can be parsed...
- Open Source
PDF-Extract-Kit: Comprehensive PDF Content Extraction Toolkit
PDFs remain the most common format for document exchange, but extracting structured content from them is notoriously difficult. PDF-Extract-Kit...
- AI
PaddleOCR: Baidu's Ultra-Lightweight OCR Toolkit with 80+ Language Support
PaddleOCR is Baidu's industrial-grade, ultra-lightweight optical character recognition (OCR) toolkit built on the PaddlePaddle deep learning...
- AI
olmOCR: AI2's Open-Source PDF-to-Markdown Toolkit for LLM Training Data
Converting PDFs to clean, machine-readable text at scale is one of the foundational challenges in LLM dataset preparation. Traditional PDF parsers...
- AI
MinerU: Open-Source PDF Document Parsing and Data Extraction
PDF is the universal format for document distribution, but it is arguably the worst format for data extraction. PDFs store visual layouts —...