AI analysis grounded in the code graph — computed facts, not vibes · 2026-09-08T03:03:19Z
VoiceStudio is a locally-run, open-source audio production suite covering voice cloning, voice design, video dubbing, dictation, transcription and audiobook creation. It bundles 16 TTS engines and 11 ASR engines with a 646-language catalogue, ships as a desktop app (Tauri/Rust front end, Python backend) for macOS, Windows and Linux, and exposes local REST/SSE/WebSocket APIs plus an OpenAI-compatible /v1/audio/speech endpoint and an MCP server for agent integration. It targets developers and creators wanting an ElevenLabs-style workflow without cloud accounts, API keys or usage metering.
The 832-star single-day gain lines up with an active release cadence (v0.4.0 → v0.5.1 across roughly six weeks, with v0.5.2 prep visible in commits) and a steady stream of user-facing feature commits — synced lyrics playback, speech-to-speech voice changing, karaoke caption burn-in, project-level dub casting. The README's "no account, API key, subscription, or usage meter" positioning against ElevenLabs is a plausible growth driver, though CodeHub's graph facts can't independently confirm causation between any specific release and the star spike.
What changed recently, how it's actually built (from the code graph), and whether you should care. Free account — no card, no spam.