Open Source OCR Projects
Discover 18 open source OCR repositories from GitHub, each with an analysis of what it does, key features, use cases and alternatives. OCR projects here are most often combined with Data Extraction, Python and PDF. Last updated October 3, 2026.
18 repositories · updated October 3, 2026

docling-api: Convert Documents to Markdown Through an API
docling-api is a self-hostable FastAPI service that converts documents and images into Markdown using Docling. It suits teams that need synchronous or queued batch processing, with CPU and GPU deployment options.

papermerge: Organize and Search Scanned Documents
Papermerge is a web-based document management system for organizing scanned archives. It uses OCR and full-text search to make documents easier to find, but this repository is archived and development has moved to papermerge-core.

xberg: Extract Text and Structure from Documents
Xberg is a Rust-based document intelligence engine that extracts text, tables, metadata, and structured data from many file types. Use it as a library, CLI, REST API, or MCP server, with bindings for multiple languages.

marker: Convert Documents into Structured Text
Marker converts PDFs and other documents into Markdown, JSON, HTML, or chunks, preserving structure such as tables, equations, and images. It suits developers building document-processing workflows who can run its local models and inference backend.

Ollama-OCR: Extract Text from Images and PDFs with Vision Models
Ollama-OCR uses vision-language models served by Ollama to extract text and structured content from images and PDFs. It offers a Python package for single or batch processing and a Streamlit interface for interactive use.

text-extract-api: Extract Text and Data from Documents
A self-hostable API that turns PDFs, images, and Office files into Markdown or structured JSON using OCR and Ollama models. It suits teams that need document processing and PII removal with control over where files are processed.