💡PaddleOCR
PaddlePaddle/PaddleOCR · homepage ↗
Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages.
⭐ Hugely popular: 88k stars, gaining about 451 a week
View on GitHub ↗repo profile
momentum
durability
bus factor = how many people it takes to cover more than half the commits (6 months). 1 is a solo project; higher means the work is spread across a team. top-author share is the single busiest author's slice of those commits.
since we covered it
why it's a big deal
- Recognizes scene text, printed text, handwritten notes, tables and formulas - all in one toolkit.
- Supports end-to-end pipelines from image/PDF ingestion to structured output (JSON, Markdown).
- Enables deployment on CPUs, GPUs, mobile devices and on-premises - suitable for enterprise use.
under the hood
- Licensed under Apache-2.0. GitHub.
- Built on PaddlePaddle and written in Python with high-performance modules for inference, layout parsing and vision-language tasks.
- Core models include PP-OCRv5 (multilingual text recognition), PP-StructureV3 (complex document layout parsing) and PP-ChatOCRv4 (key information extraction + LLM integration).
our take from PR#21, 2025-11-12
star history
- PR#21 64k 2025-11-12
- now 88k + 24k since first covered
curve is sampled from GitHub's star history, plus our own daily readings since we covered it; the dashed stretch is before we first covered it, the solid line since. figures at coverage are the numbers we printed then (approx.), current count is live.
understory
Better known than its recent output, coasting a little on attention.
- output, commits & releases
- clout, star velocity
output = commits & releases; clout = star velocity, both 0 to 100 monthly indices; the gap where output runs above clout is the understory. The understory →
covered in
-
Industry-grade OCR & document AI toolkit
similar projects
compare these →- ⚡ Unsloth
Local UI to run and train LLMs and diffusion models, including Qwen3.8, Kimi K3, MiniMax-H3, Gemma 4, DeepSeek-V4, FLUX and more.
72k ACTIVE - 🎙️ Real-Time-Voice-Cloning
Clone a voice in 5 seconds to generate arbitrary speech in real-time
60k ACTIVE - 🎙️ VibeVoice
Open-Source Frontier Voice AI
53k ACTIVE
replaces
alternatives to AWS Textract →- AWS Textract
AWS
closed source - Google Document AI
Google
closed source - ABBYY closed source
comments
Sign in with GitHub to add your blip on PaddleOCR.