Skip to content
repository radar
Voice, Vision & Multimodal

💡PaddleOCR

ACTIVE STEADY

PaddlePaddle/PaddleOCR · homepage ↗

Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages.

⭐ Hugely popular: 88k stars, gaining about 451 a week

View on GitHub ↗

repo profile

vintage 2020 6 years old
delivery library
language Python
license Apache-2.0

momentum

total stars 88k
stars added last week +451
weekly downloads 582k /wk · PyPI
used by 7k repos & packages
commits / week 1 cooling · -33% vs prior mo
issues closed 85% ~19d to close, median
contributors 295
release cadence monthly
last activity 24d ago

durability

backing company-owned Baidu
openness permissive
security score 5/10 OpenSSF Scorecard · 2026-08-10
bus factor 2 concentrated
top-author share 39% 6 mo

bus factor = how many people it takes to cover more than half the commits (6 months). 1 is a solo project; higher means the work is spread across a team. top-author share is the single busiest author's slice of those commits.

since we covered it

monthly average + 3k/mo (+4%/mo) · + 24k total since PR#21

why it's a big deal

  • Recognizes scene text, printed text, handwritten notes, tables and formulas - all in one toolkit.
  • Supports end-to-end pipelines from image/PDF ingestion to structured output (JSON, Markdown).
  • Enables deployment on CPUs, GPUs, mobile devices and on-premises - suitable for enterprise use.

under the hood

  • Licensed under Apache-2.0. GitHub.
  • Built on PaddlePaddle and written in Python with high-performance modules for inference, layout parsing and vision-language tasks.
  • Core models include PP-OCRv5 (multilingual text recognition), PP-StructureV3 (complex document layout parsing) and PP-ChatOCRv4 (key information extraction + LLM integration).

our take from PR#21, 2025-11-12

star history

PR#21 · 64k88k now May 2020Aug 2026
  1. PR#21 64k 2025-11-12
  2. now 88k + 24k since first covered

curve is sampled from GitHub's star history, plus our own daily readings since we covered it; the dashed stretch is before we first covered it, the solid line since. figures at coverage are the numbers we printed then (approx.), current count is live.

understory

Better known than its recent output, coasting a little on attention.

-28 understory score output 51 · clout 79
Aug 2025 Jul 2026
  • output, commits & releases
  • clout, star velocity

output = commits & releases; clout = star velocity, both 0 to 100 monthly indices; the gap where output runs above clout is the understory. The understory →

covered in

  • PR#21 2025-11-12 above the radar

    Industry-grade OCR & document AI toolkit

similar projects

compare these →
  • Unsloth

    Local UI to run and train LLMs and diffusion models, including Qwen3.8, Kimi K3, MiniMax-H3, Gemma 4, DeepSeek-V4, FLUX and more.

    72k ACTIVE
  • 🎙️ Real-Time-Voice-Cloning

    Clone a voice in 5 seconds to generate arbitrary speech in real-time

    60k ACTIVE
  • 🎙️ VibeVoice

    Open-Source Frontier Voice AI

    53k ACTIVE
  • AWS Textract

    AWS

    closed source
  • Google Document AI

    Google

    closed source
  • ABBYY
    closed source

comments

Sign in with GitHub to add your blip on PaddleOCR.

loading comments…