Skip to content
repository radar
Voice, Vision & Multimodal

🎙️Real-Time-Voice-Cloning

ACTIVE BREAKOUT

CorentinJ/Real-Time-Voice-Cloning

Clone a voice in 5 seconds to generate arbitrary speech in real-time

⭐ Hugely popular: 60k stars, gaining about 15 a week

View on GitHub ↗

repo profile

vintage 2019 7 years old
language Python
license Other

momentum

total stars 60k
stars added last week +15
HN peak 94 pts 11mo ago
commits / week 0 steady
issues closed 61% ~37d to close, median
contributors 20
release cadence n/a
last activity 5mo ago

durability

backing community / independent
security score 3.3/10 OpenSSF Scorecard · 2026-08-10

since we covered it

monthly average + 5k/mo (+9659%/mo) · + 60k total since PR#17

why it's a big deal

  • Democratizes advanced voice synthesis with a fast, minimal setup.
  • Useful for prototyping conversational agents, accessibility tools, and creative projects.
  • A go-to reference implementation for many follow-up research and commercial systems.

under the hood

  • Based on SV2TTS: speaker encoder (generalized embeddings), Tacotron2-like synthesizer, and WaveRNN vocoder.
  • Pretrained models available for immediate inference.
  • Supports training on your own dataset for custom speaker adaptation.

our take from PR#17, 2025-09-17

star history

PR#17 · 5660k now May 2019Aug 2026
  1. PR#17 56 2025-09-17
  2. now 60k + 60k since first covered

curve is sampled from GitHub's star history, plus our own daily readings since we covered it; the dashed stretch is before we first covered it, the solid line since. figures at coverage are the numbers we printed then (approx.), current count is live.

understory

Better known than its recent output, coasting a little on attention.

-77 understory score output 0 · clout 77
Aug 2025 Jul 2026
  • output, commits & releases
  • clout, star velocity

output = commits & releases; clout = star velocity, both 0 to 100 monthly indices; the gap where output runs above clout is the understory. The understory →

covered in

  • PR#17 2025-09-17 on the radar

    Instant Voice Cloning

similar projects

compare these →
  • 👁️ supervision

    We write your reusable computer vision tools. 💜

    49k ACTIVE
  • Unsloth

    Local UI to run and train LLMs and diffusion models, including Qwen3.8, Kimi K3, MiniMax-H3, Gemma 4, DeepSeek-V4, FLUX and more.

    72k ACTIVE
  • 🎧 ebook2audiobook

    leaner, 20k stars

    Generate audiobooks from e-books, voice cloning & 1158+ languages!

    20k ACTIVE

comments

Sign in with GitHub to add your blip on Real-Time-Voice-Cloning.

loading comments…