OCR

Page 1 of 2

  1. Surya: Open-Source Multilingual OCR and Document Understanding

    Optical Character Recognition is one of the oldest applications of computer vision, but traditional OCR engines have struggled to keep pace with...

    AI
  2. RapidLayout: Open-Source Document Layout Analysis for Chinese and English

    Document layout analysis is the critical first step in any document understanding pipeline. Before OCR can extract text, before tables can be parsed...

    Open Source
  3. PDF-Extract-Kit: Comprehensive PDF Content Extraction Toolkit

    PDFs remain the most common format for document exchange, but extracting structured content from them is notoriously difficult. PDF-Extract-Kit...

    Open Source
  4. PaddleOCR: Baidu's Ultra-Lightweight OCR Toolkit with 80+ Language Support

    PaddleOCR is Baidu's industrial-grade, ultra-lightweight optical character recognition (OCR) toolkit built on the PaddlePaddle deep learning...

    AI
  5. olmOCR: AI2's Open-Source PDF-to-Markdown Toolkit for LLM Training Data

    Converting PDFs to clean, machine-readable text at scale is one of the foundational challenges in LLM dataset preparation. Traditional PDF parsers...

    AI
  6. MinerU: Open-Source PDF Document Parsing and Data Extraction

    PDF is the universal format for document distribution, but it is arguably the worst format for data extraction. PDFs store visual layouts —...

    AI