All 122 Dependencies
Every package VoiceStudio depends on, ranked by repo health score.
VoiceStudio is a desktop app and local backend for working with voices on your own machine. You can clone a voice from a short reference recording, design a new one from a description, dub videos with timed speech, dictate into any app, transcribe audio, and produce audiobooks, in 646 languages. Its default engine is OmniVoice, and you can switch to others from a model catalogue.
It runs as an Electron desktop app for macOS, Windows and Linux, with a Python FastAPI service behind it, and a Docker image is available too. Models download on first use from Hugging Face, and the compute device (CUDA, Apple MPS or CPU) is detected automatically. Remote GPU workers are optional.
For developers and agents, it exposes a local HTTP API and an MCP server that can generate speech, clone voices, transcribe audio and list voices. Everything is local by default, with analytics off unless you consent.