mirror of
https://github.com/jamiepine/voicebox.git
synced 2026-10-04 01:25:18 -07:00
Review follow-ups: mlx-audio's Model.from_pretrained fetches the S3 speech tokenizer from mlx-community/S3TokenizerV2 (~470 MB) separately from the chatterbox checkout, so _is_model_cached now requires both repos (same shape as the Hume backend's codec check) and the config's size_mb reflects the real footprint. backend.backends.chatterbox_mlx_backend is a function-level import that PyInstaller's graph will not see, so it is added to the Apple Silicon hidden-import list in build_binary.py and voicebox-server.spec next to mlx_backend.