🐱KittenTTS
State-of-the-art TTS model under 25MB 😻
⭐ Very popular: 15k stars, gaining about 90 a week
View on GitHub ↗repo profile
momentum
durability
bus factor = how many people it takes to cover more than half the commits (6 months). 1 is a solo project; higher means the work is spread across a team. top-author share is the single busiest author's slice of those commits.
since we covered it
why it's a big deal
- Local TTS removes a major dependency for fully self-hosted agents.
- Small runtime footprint makes voice interfaces practical on commodity hardware.
- Fits naturally into agent workflows that need fast voice output without cloud latency.
under the hood
- Optimized neural TTS pipeline designed for efficient local inference.
- Lightweight model design aimed at low-resource environments.
- Simple integration path for agent frameworks and automation scripts.
our take from PR#28, 2026-02-25
star history
- PR#28 11k 2026-02-25
- now 15k + 4k since first covered
curve is sampled from GitHub's star history, plus our own daily readings since we covered it; the dashed stretch is before we first covered it, the solid line since. figures at coverage are the numbers we printed then (approx.), current count is live.
understory
Better known than its recent output, coasting a little on attention.
- output, commits & releases
- clout, star velocity
output = commits & releases; clout = star velocity, both 0 to 100 monthly indices; the gap where output runs above clout is the understory. The understory →
covered in
-
Lightweight text-to-speech for local AI workflows
similar projects
compare these →- 💡 PaddleOCR
5.7× the stars
Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages.
88k ACTIVE - ⚡ Unsloth
4.7× the stars
Local UI to run and train LLMs and diffusion models, including Qwen3.8, Kimi K3, MiniMax-H3, Gemma 4, DeepSeek-V4, FLUX and more.
72k ACTIVE - 🎙️ Real-Time-Voice-Cloning
3.9× the stars
Clone a voice in 5 seconds to generate arbitrary speech in real-time
60k ACTIVE
comments
Sign in with GitHub to add your blip on KittenTTS.