TESIGN
SAVED

EN · KO

Auto-generated — not yet edited · only the numbers are verified

  1. TESIGN / RADAR
  2. REPOSITORY CARD
  3. jamiepine/voicebox

jamiepine/voicebox

The open-source AI voice studio. Clone, dictate, create.

★ 53.3k▲ +191 / 7dDaily star increases over 14 days. Hollow bars mean no data. The last day may be partial.2026-09-03 no data2026-09-04 +142026-09-05 +222026-09-06 +162026-09-07 no data2026-09-08 no data2026-09-09 +22026-09-10 +102026-09-11 +182026-09-12 +82026-09-13 +92026-09-14 +92026-09-15 +962026-09-16 +41largest one-day increase in the 14 days+96last 14 days · stars per dayhollow = no data

RANKS All-time #124 · Rising 30d #80

AT A GLANCE

LANGUAGE
TypeScript
LICENSE
MIT
USAGE
Use, change and redistribute, commercially too. Keep the notice.
ACTIVITY
last commit 39 days ago ()
TOPICS
  • ai
  • cuda
  • mlx
  • qwen3-tts
  • qwen3-tts-ui
  • voice-ai
  • voice-clone
  • whisper
HOMEPAGE
https://voicebox.sh
REPOSITORY
GitHub ↗
CATEGORY
AI

A summary, not legal advice.

README EXCERPT

Voicebox The open-source AI voice studio. Clone any voice. Generate speech. Dictate into any app. Talk to agents in voices you own. The full voice I/O stack, running locally on your machine. voicebox.sh • Docs • Download • Features • API • Troubleshooting Click the image above to watch the demo video on voicebox.sh What is Voicebox? Voicebox is a local-first AI voice studio — a free and open-source alternative to ElevenLabs and WisprFlow in one app. Clone voices from a few seconds of audio, generate speech in 23 languages across 7 TTS engines, dictate into any text field with a global hotkey, and give any MCP-aware AI agent a voice of your choosing. The two cloud incumbents sit on opposite halves of the voice I/O loop — ElevenLabs on output, WisprFlow on input. Voicebox does both, bridges them with a bundled local LLM for refinement and per-profile personas, and runs the whole thing on your machine. - Complete privacy — models, voice data, and captures never leave your machine - 7 TTS engines — Qwen3-TTS, Qwen CustomVoice, LuxTTS, Chatterbox Multilingual, Chatterbox Turbo, HumeAI TADA, and Kokoro - Voice cloning and preset voices — zero-shot cloning from a reference sample, or 50+…

The opening of the GitHub README as stored, at most 1,200 characters. Markdown is not rendered.

TIMELINE

  1. Repository created
  2. Last push
  3. FIRST SEEN BY TESIGN ◌ BACK CATALOG

AI chooses the lists under the owner’s delegation. No human review is running in September 2026. Total stars, increases, cross-source signals, last updates and licences are shown as evidence. How ranks work →

If this repository gets editorial text (why, build, who, start, caveat) it becomes an edited entry. Until then the page shows only stored GitHub metadata and numbers.