mirror of
https://github.com/jamiepine/voicebox.git
synced 2026-09-16 21:30:39 -07:00
- New ChatterboxTTSBackend wrapping ChatterboxMultilingualTTS (ResembleAI/chatterbox) - Supports 23 languages including Hebrew, forces CPU on macOS (MPS issue) - Monkey-patches torch.load for CPU loading, forces eager attention for compatibility - trim_tts_output utility cuts trailing silence/hallucination from Chatterbox output - Full engine integration: /generate, /generate/stream, model status/download/delete - Hebrew (he) added to supported languages in frontend and backend validation - Single flat model dropdown extended with Chatterbox option in both generation UIs - ModelManagement UI groups LuxTTS and Chatterbox under 'Other Voice Models' section
39 lines
879 B
Plaintext
39 lines
879 B
Plaintext
# FastAPI and server
|
|
fastapi>=0.109.0
|
|
uvicorn[standard]>=0.27.0
|
|
pydantic>=2.5.0
|
|
|
|
# Database
|
|
sqlalchemy>=2.0.0
|
|
alembic>=1.13.0
|
|
|
|
# ML models
|
|
torch>=2.1.0
|
|
transformers>=4.36.0,<=4.57.6
|
|
accelerate>=0.26.0
|
|
huggingface_hub>=0.20.0
|
|
qwen-tts>=0.0.5
|
|
|
|
# LuxTTS (voice cloning engine)
|
|
# piper-phonemize needs custom index (no PyPI wheels)
|
|
--find-links https://k2-fsa.github.io/icefall/piper_phonemize.html
|
|
# linacodec is a git-only dep of Zipvoice (uv-only source, pip can't resolve it)
|
|
linacodec @ git+https://github.com/ysharma3501/LinaCodec.git
|
|
Zipvoice @ git+https://github.com/ysharma3501/LuxTTS.git
|
|
|
|
# Chatterbox TTS (multilingual voice cloning, includes Hebrew)
|
|
chatterbox-tts>=0.1.0
|
|
|
|
# Audio processing
|
|
librosa>=0.10.0
|
|
soundfile>=0.12.0
|
|
numpy>=1.24.0
|
|
numba>=0.60.0,<0.61.0
|
|
|
|
# HTTP client (for CUDA backend download)
|
|
httpx>=0.27.0
|
|
|
|
# Utilities
|
|
python-multipart>=0.0.6
|
|
Pillow>=10.0.0
|