Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages.
-
Updated
Jul 22, 2026 - Python
Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages.
天枢 - 企业级 AI ��站式数据预处理平台 | PDF/Office转Markdown | 支持MCP协议AI助手集成 | Vue3+FastAPI全栈方案 | 文档解析 | 多模态信息提取
Local-first document parsing workbench for five OCR models: PDF/Office to Markdown with WebUI, model switching, CLI and Apple Silicon support.
Recognition of license plate numbers, in any format, by automatic detection with Yolov8, pipeline of filters and paddleocr as OCR
A self-hosted PDF OCR API that converts scanned documents to markdown. Powered by PaddleOCR-VL, runs on GPU via Docker.
Native Swift + MLX port of Paddle0CR-VL, a 0.9B doc parsing VLM for OCR, tables, formulas, and charts on Apple Silicon
Advanced PDF processing, OCR, and AI vision analysis nodes for ComfyUI. Extract images from PDFs, perform multilingual OCR with Surya, detect objects with Florence-2, and analyze document layouts.
Local-first web UI for PaddleOCR-VL: OCR PDFs and images with Docker, live logs, previews, and ZIP exports.
Minimal script to run PaddleOCR-VL on a PDF and output a MD file with images
Production-oriented PDF OCR parsing platform for large batch, distributed control/worker workflows with shared storage and OpenAI-compatible model services.
Bob 终极 OCR:支持本地隐私模型 + 大模型API云端 OCR 高精度,一键切换。快、免费、隐私好。识别能力更强,可选更多模型
An open-source alternative to Mathpix Snip Tool. Instantly extract LaTeX, math equations, and text from images and screenshots.
零依赖的百度智能云 PaddleOCR-VL 文档解析 API 评估 Demo
Convert user-supplied LS-DYNA Keyword and Theory Manual PDFs into structured, traceable Markdown documents.
PaddleVLM Editor is a local, Docker-based application that extracts structured content from PDFs and images using PaddleOCR-VL's Vision-Language Model, then lets you review and edit the results in a rich-text block editor — all without sending data to any cloud service.
One-command OmniDocBench v1.6 evaluation setup on AMD Windows (ROCm/HIP). Model-agnostic: swap any document parsing model via adapters. PaddleOCR-VL-1.6 as reference. CDM formula scoring included.
OCR-DFlash is a research runner for accelerating PDF document parsing with PaddleOCR-VL, PP-DocLayoutV3, PDF native text drafts, and bounded LLMA-style token verification.
PaddleOCR 기반 OCR 검수·문서 변환·게시 워크플로우
Python data pipeline for arXiv metadata: SQLAlchemy + Alembic schema, PostgreSQL storage, PDF download tracking, and optional PaddleOCR processing.
PaddleOCR-VL Fastapi Inference Server with vLLM backend for Deployment on Nvidia GPUs
To associate your repository with the paddleocr-vl topic, visit your repo's landing page and select "manage topics."