🎙️Real-Time-Voice-Cloning
CorentinJ/Real-Time-Voice-Cloning
Clone a voice in 5 seconds to generate arbitrary speech in real-time
⭐ Hugely popular: 60k stars, gaining about 15 a week
View on GitHub ↗repo profile
momentum
durability
since we covered it
why it's a big deal
- Democratizes advanced voice synthesis with a fast, minimal setup.
- Useful for prototyping conversational agents, accessibility tools, and creative projects.
- A go-to reference implementation for many follow-up research and commercial systems.
under the hood
- Based on SV2TTS: speaker encoder (generalized embeddings), Tacotron2-like synthesizer, and WaveRNN vocoder.
- Pretrained models available for immediate inference.
- Supports training on your own dataset for custom speaker adaptation.
our take from PR#17, 2025-09-17
star history
- PR#17 56 2025-09-17
- now 60k + 60k since first covered
curve is sampled from GitHub's star history, plus our own daily readings since we covered it; the dashed stretch is before we first covered it, the solid line since. figures at coverage are the numbers we printed then (approx.), current count is live.
understory
Better known than its recent output, coasting a little on attention.
- output, commits & releases
- clout, star velocity
output = commits & releases; clout = star velocity, both 0 to 100 monthly indices; the gap where output runs above clout is the understory. The understory →
covered in
-
Instant Voice Cloning
similar projects
compare these →- 👁️ supervision
We write your reusable computer vision tools. 💜
49k ACTIVE - ⚡ Unsloth
Local UI to run and train LLMs and diffusion models, including Qwen3.8, Kimi K3, MiniMax-H3, Gemma 4, DeepSeek-V4, FLUX and more.
72k ACTIVE - 🎧 ebook2audiobook
leaner, 20k stars
Generate audiobooks from e-books, voice cloning & 1158+ languages!
20k ACTIVE
comments
Sign in with GitHub to add your blip on Real-Time-Voice-Cloning.