mirror of
https://github.com/jamiepine/voicebox.git
synced 2026-09-16 13:20:39 -07:00
* docs: audit mdx docs against multi-engine backend and refresh stale content Rewrote developer-facing docs that predated the TTSBackend Protocol / ModelConfig registry refactor (architecture, tts-generation, model-management, transcription). Updated user-facing docs to reflect all seven shipped engines (Qwen, Qwen CustomVoice, LuxTTS, Chatterbox, Chatterbox Turbo, TADA, Kokoro) instead of the outdated "5 engines" claim. Also fixes: - Stale app identifier (com.voicebox.app → sh.voicebox.app) - CUDA backend update flow (now two-archive split, not N-way chunks) - Whisper model list (removed tiny, added turbo) - Broken /development/ and /guides/ route links - Stale just commands and install steps (missing --no-deps chatterbox/tada) - Removed ASCII art diagrams from README and stories.mdx - History Generation schema sync with DB model Co-Authored-By: Claude Opus 4.7 (1M context) <[email protected]> * docs: add DeepWiki badge to README Co-Authored-By: Claude Opus 4.7 (1M context) <[email protected]> * docs: address PR review feedback - architecture.mdx: fix backends/ file list (remove nonexistent qwen_backend.py, rename tada_backend.py → hume_backend.py) - model-management.mdx: Kokoro language count 9 → 8 (matches ModelConfig) - model-management.mdx: ProgressManager path services/ → utils/ - tts-generation.mdx: ModelConfig example uses field(default_factory=...) — mutable default would raise at runtime - tts-generation.mdx: "1080p samples" → "on CUDA" (1080p is video, not audio) - PROJECT_STATUS.md: replace ASCII architecture diagram with prose (matches no-ASCII-art rule) Co-Authored-By: Claude Opus 4.7 (1M context) <[email protected]> * fix(app): guard against undefined engine in FloatingGenerateBox preset check form.getValues('engine') returns string | undefined; Set<string>.has() rejects undefined under strict mode. Added a truthy guard before the preset lookup. Co-Authored-By: Claude Opus 4.7 (1M context) <[email protected]> --------- Co-authored-by: Claude Opus 4.7 (1M context) <[email protected]>
38 lines
2.2 KiB
Plaintext
38 lines
2.2 KiB
Plaintext
---
|
|
title: "Voicebox Documentation"
|
|
description: "Voicebox is a local-first voice cloning studio -- a free and open-source alternative to ElevenLabs."
|
|
---
|
|
|
|
Voicebox is a **local-first voice cloning studio** -- a free and open-source alternative to ElevenLabs. Clone voices from a few seconds of audio, generate speech in 23 languages across 7 TTS engines, apply post-processing effects, and compose multi-voice projects with a timeline editor.
|
|
|
|

|
|
|
|
- **Complete privacy** -- models and voice data stay on your machine
|
|
- **7 TTS engines** -- Qwen3-TTS, Qwen CustomVoice, LuxTTS, Chatterbox Multilingual, Chatterbox Turbo, HumeAI TADA, and Kokoro
|
|
- **Cloning and preset voices** -- zero-shot cloning from a reference sample, or 50+ curated preset voices via Kokoro and Qwen CustomVoice
|
|
- **23 languages** -- from English to Arabic, Japanese, Hindi, Swahili, and more
|
|
- **Post-processing effects** -- pitch shift, reverb, delay, chorus, compression, and filters
|
|
- **Expressive speech** -- paralinguistic tags like `[laugh]`, `[sigh]`, `[gasp]` via Chatterbox Turbo; natural-language delivery control via Qwen CustomVoice
|
|
- **Unlimited length** -- auto-chunking with crossfade for scripts, articles, and chapters
|
|
- **Stories editor** -- multi-track timeline for conversations, podcasts, and narratives
|
|
- **API-first** -- REST API for integrating voice synthesis into your own projects
|
|
- **Native performance** -- built with Tauri (Rust), not Electron
|
|
- **Runs everywhere** -- macOS (MLX/Metal), Windows (CUDA), Linux, AMD ROCm, Intel Arc, Docker
|
|
|
|
## Download
|
|
|
|
| Platform | Download |
|
|
|----------|----------|
|
|
| macOS (Apple Silicon) | [Download DMG](https://voicebox.sh/download/mac-arm) |
|
|
| macOS (Intel) | [Download DMG](https://voicebox.sh/download/mac-intel) |
|
|
| Windows | [Download MSI](https://voicebox.sh/download/windows) |
|
|
| Docker | `docker compose up` |
|
|
|
|
[View all releases](https://github.com/jamiepine/voicebox/releases/latest)
|
|
|
|
## Get Started
|
|
|
|
- [Installation](/overview/installation) -- download and install Voicebox
|
|
- [Quick Start](/overview/quick-start) -- get up and running in 5 minutes
|
|
- [API Reference](/api-reference) -- integrate voice synthesis into your apps
|